Pith. sign in

Paper Citation Record · LEDGER

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

As of 10 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2608.04077.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04077 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:39:34.850226Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0e0489e4-4def-4311-89a1-b690d341cb8c · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Constitutional AI: Harmlessness from AI Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.692929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.692929Z digest=sha256:f7827019f31f2a151e25363541047906dbee250ff3b913fe2bd9ef76dde1ce70

Observation ff28b7f4-c0a7-4882-ad26-d0f5c327341b · outbound

This paper cites A.; and Terry, M.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables A.; and Terry, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.785146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.697760Z digest=sha256:6e2dd20f3fce642db941ceb06ee6da26e3a441d13a1b864d0fe58f611f21340a

Observation e5fbf2a1-d75e-486b-8828-2b3768a59c10 · outbound

This paper cites DISC-FinLLM: A Chinese Financial Large Language Model based on Multiple Experts Fine-tuning.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables DISC-FinLLM: A Chinese Financial Large Language Model based on Multiple Experts Fine-tuning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.701168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.701168Z digest=sha256:3c19fe662068968342461f9841b192caf57c07854f686331f583f76b6a1fe76a

Observation 292d2183-e124-4e8b-9926-67cdb67e7ef9 · outbound

This paper cites N.; Li, T.; Li, D.; Zhu, B.; Zhang, H.; Jordan, M.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables N.; Li, T.; Li, D.; Zhu, B.; Zhang, H.; Jordan, M

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.775289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.703990Z digest=sha256:d5e542931a1d39a13f96bb58ef3dd785c72c238fa781033364b933eb25d8bb87

Observation 715481de-f1aa-4784-9716-0375167468ba · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.767171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.706593Z digest=sha256:b47d8d6cb3e6512ab98759dab2cdad2222fd9920f0cab2d4656d1032e538037d

Observation 205a2af6-4d1f-450b-a694-701186f1ceb7 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 6

Resolution
verified exact
raw_fallback, observed 2026-08-08T00:39:35.591321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.710181Z digest=sha256:fb23ae892fa961aa19ac680006b138dbc7e5517e442e813ee1f94af8ba49ba03

Observation 9ef16854-2f69-480b-b599-fe7f05b64e5b · outbound

This paper cites E.; R \'e , C.; Chilton, A.; Narayana, A.; Chohlas-Wood, A.; Peters, A.; Waldon, B.; Rockmore, D.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables E.; R \'e , C.; Chilton, A.; Narayana, A.; Chohlas-Wood, A.; Peters, A.; Waldon, B.; Rockmore, D

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.759686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.712947Z digest=sha256:1311361af8c3f533426cac3dc06f7780d3314d5bfbbbca5217c192a4c5d9e4d4

Observation 3dcb4e1f-1932-4953-8c32-5dc31fcccc48 · outbound

This paper cites FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinEval: A Chinese Financial Domain Knowledge Evaluation Benchmark for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.715710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.715710Z digest=sha256:97175dddc6ec5e9341dcd63a4e153b7bb6cd2f9f4b388556194462115839589e

Observation 9ad630f9-9cd5-44a8-a45f-06550b4afa92 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.719154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.719154Z digest=sha256:a8d4a7512eca74d7c1d7f9a615b0a0dbd1ad3b470e4fc3609255b1a4840aaa9d

Observation 82722c69-3480-4d5e-89a3-fa737e5c1de5 · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinanceBench: A New Benchmark for Financial Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.722411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.722411Z digest=sha256:7c5e2bcef841fe31977734d9bcd0a104a507c2d082c2aaa14c1548d5a1d2480c

Observation 43397b8a-7e53-494a-83ad-ebd2ccb9a496 · outbound

This paper cites E.; Yang, J.; Wettig, A.; Yao, S.; Pei, K.; Press, O.; and Narasimhan, K.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables E.; Yang, J.; Wettig, A.; Yao, S.; Pei, K.; Press, O.; and Narasimhan, K

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.748658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.725837Z digest=sha256:4e8ad409e463abf6e7a97d3d80f5797a61b056e0c0cebd59df48360e1fa9e777

Observation c3727f3e-79cd-4e4f-af44-ff5d5e9f5e04 · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.729453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.729453Z digest=sha256:e38410068c34fa73a36b64e2ba2b3be809063ae525c1bfeccb5a11c2b860107d

Observation 37f84774-50fb-457a-b2b2-7dfaa74d9b33 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.740711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.733674Z digest=sha256:a7f352784baa7a7ac3cd6d6bc279992568b6d651eefc16ebccc9537d52e61b25

Observation fec3250d-6c14-44e8-aec2-606e17cdbf85 · outbound

This paper cites RewardBench: Evaluating Reward Models for Language Modeling.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables RewardBench: Evaluating Reward Models for Language Modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.737024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.737024Z digest=sha256:01d4d88b4afe047346fddca821c5e5ea75731a72a6e539333b85b086b05fffe1

Observation 2dfb6c3b-3496-4058-a2d0-7fc9d6b18843 · outbound

This paper cites CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.741055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.741055Z digest=sha256:f4c38f88930984c9688d8eb7a6f7dd6a0e268ebe11c62bb7a1e2ff066c55bc4d

Observation b91c7869-32d7-4f27-9cd2-585456c72a57 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.745767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.745767Z digest=sha256:0746467125f592280650a162a31cbbddb019d0563a3d4d26a359f9d4eaa6c343

Observation 680c13f9-7a68-4ed1-8936-1974e04d9c9c · outbound

This paper cites ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-08T00:39:35.432825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.748768Z digest=sha256:c9d80975097a514aa12791da7ed53efc793bfee0053d59c115654293f27a99f0

Observation 41fee5ac-b157-482e-aa9a-8150980958ec · outbound

This paper cites JobBench: Aligning Agent Work With Human Will.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables JobBench: Aligning Agent Work With Human Will

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.752504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.752504Z digest=sha256:fc518003eea9a7a6ec49c1099ac23b60773684e4cfbd09d2e4cce5a879b816a3

Observation 8eed73fb-c9a4-42e9-b4ce-fbf3a853ffdd · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.733076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.756993Z digest=sha256:de9d1a416020921594fbb6458a660f97a3017dcefcd77d9b7f826a12a318a04e

Observation 31e9824c-c8ef-4403-bb60-113d4d9226ec · outbound

This paper cites Y.; Deng, Y.; Chandu, K.; Brahman, F.; Bhagavatula, C.; and Choi, Y.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Y.; Deng, Y.; Chandu, K.; Brahman, F.; Bhagavatula, C.; and Choi, Y

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.724302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.760575Z digest=sha256:01a3c821d723b29c3661b0b7cbef3bf1bbf160b8c9ea33859e486ea6c8421faf

Observation cf680bf0-6096-43d9-9e00-a943cae13ef3 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.765278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.765278Z digest=sha256:2f9bddaeecd62583a4fbbfecaf6b2aa9b36be16add0d4bde9f501243ccfd3a74

Observation b2ea64d1-7ae5-4b21-a539-d14d4e553fec · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.716019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.768687Z digest=sha256:5f2ef04a48984533999a8ae355dbc5e1cc87188454b47f02d05408a67a75548b

Observation 81cf1719-acc9-4d98-acbd-df68591dabf3 · outbound

This paper cites AgentBench: Evaluating LLMs as Agents.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables AgentBench: Evaluating LLMs as Agents

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.772089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.772089Z digest=sha256:a839a0f414e99ef22e9914dd8de19e46664e17a056f278f82b96182dd83cc8a0

Observation bdd8ea85-1aa0-410a-9251-d3d142ddec80 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.707601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.775217Z digest=sha256:1c7097be5b7d7cc0cfad966824ec1616b5dea6658b8b61321647a8c08102fd88

Observation 9b3690fa-dd86-4197-a962-8c27a9431167 · outbound

This paper cites FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables FinResearchBench II: A Deep Research Benchmark with Consensus-Derived Gold Rubrics for Distinguishing Financial Report Quality

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.778096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.778096Z digest=sha256:ad81bdfa34c7cfed49095a5aec8456231aaedd52f3d9cadb385cccd0767eeb6d

Observation c1246f93-9f84-41db-b992-c951877e8222 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.699142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.782085Z digest=sha256:ff871bb6d47fac9f5adc66c844523f8728ba1556ecd33c5a83af92e55d96ce9f

Observation 3038f6ce-47d4-42b2-8b90-f7921749749c · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.691815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.784771Z digest=sha256:06d2658e17095b9fe62426989c88de04c09abc763a5308e10f6cb8b99b43ba83

Observation 3d34ca5f-c1e0-454d-9ae9-f11d1ebe5fba · outbound

This paper cites GPT-4 Technical Report.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables GPT-4 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.787813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.787813Z digest=sha256:6b3601977a671f7469a0eb2dde1a0ee13c516cc46ad7282e70c359a8d73a46a2

Observation 555ecf1d-a730-4e56-91ba-934108541986 · outbound

This paper cites L.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables L.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.683610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.790805Z digest=sha256:97c798ecd1d56bc24ca3dd26ff0d8f7e0adac8f29334ef01bcb789bc4bcb74b4

Observation 77965a63-7f2f-4934-883d-dbd1bbc416a3 · outbound

This paper cites GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables GDPval: Evaluating AI Model Performance on Real-World Economically Valuable Tasks

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.793186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.793186Z digest=sha256:9b44b1dcc871b3da2373c07c6681a439548b34dd14a4c786e917b3c7f26baf7c

Observation 2ba8720e-de77-441d-bc25-0bb5612bbbaa · outbound

This paper cites D.; Ermon, S.; and Finn, C.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables D.; Ermon, S.; and Finn, C

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.796664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.796664Z digest=sha256:6fbf95ae3c8d2b238d23e8e357b60f836b5b8752f48ebf94bb46aff53c941f1f

Observation e1975965-a9ce-4382-af5d-568ceb3cc374 · outbound

This paper cites L.; Stickland, A.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables L.; Stickland, A

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.671807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.799748Z digest=sha256:d6acb27a68e0d82d249b78272cc750dd5943e8965a5db98d65511a176adbc926

Observation ef025533-8f6e-40f5-8f6d-d8f3d3a31e42 · outbound

This paper cites S.; Chawla, K.; Eidnani, D.; Shah, A.; Du, W.; Chava, S.; Raman, N.; Smiley, C.; Chen, J.; and Yang, D.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables S.; Chawla, K.; Eidnani, D.; Shah, A.; Du, W.; Chava, S.; Raman, N.; Smiley, C.; Chen, J.; and Yang, D

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.662462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.802371Z digest=sha256:9d3a2063d011fb735c79b92d8eecd2e1bb844697116876317f9a4bb7bd0a8815

Observation 18828a47-4c2e-4f53-b9fe-e21e2de1e0bf · outbound

This paper cites F.; Qiu, X.; Whitehouse, C.; Alazraki, L.; Goel, S.; Barbieri, F.; Willi, T.; Mathur, A.; and Leontiadis, I.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables F.; Qiu, X.; Whitehouse, C.; Alazraki, L.; Goel, S.; Barbieri, F.; Willi, T.; Mathur, A.; and Leontiadis, I

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.805634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.805634Z digest=sha256:d514438812ad1ce14bb3a3e8a5e6058ebc6c9892595927111bc14271abc80482

Observation 2d9b5fde-e5b8-4b90-a9de-e53d3588eb76 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.809274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.809274Z digest=sha256:07a1354805a12d199e5eae705dff7b344cf8033f91218957a746b669bb7cd94c

Observation eaf13a77-dd98-40bb-a2d9-b349356bf4b4 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.812234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.812234Z digest=sha256:7e4587f6af1e88743a828533d08dc4307d5def1fa734ff876800efd358e5ea74

Observation c240abeb-97cf-457e-9f62-0faf239c4e3a · outbound

This paper cites R.; Zhang, S.; Sun, Y.; and Wang, W.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables R.; Zhang, S.; Sun, Y.; and Wang, W

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.652830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.814612Z digest=sha256:a885d4d13a38a10c3c7d2d6c9f62d38d4d8823194786b0323833d5c05d7199fc

Observation 994903cb-0052-4517-bc69-44ce23c6314d · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.644264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.818031Z digest=sha256:17d6f946e01d07cbc7f9a4a01fd06d8fee30d4139d4fc7c0f0de585469ba75e2

Observation 0fb0c3e0-5527-4f5f-9a73-0914d091ade0 · outbound

This paper cites BloombergGPT: A Large Language Model for Finance.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables BloombergGPT: A Large Language Model for Finance

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.821009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.821009Z digest=sha256:df37dc0cd1154f4e0260bcccde0f226cf97bd7f1c77bb6c6b5de79ad123d9426

Observation 0694b070-9be6-4301-bf50-1b91f6d33db8 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.823780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.823780Z digest=sha256:f8109e03e5c3024d8463356e2011c4af004b6b43050eaf366f98afa82323c7c6

Observation 29850e5e-9dfc-47d3-81c9-da242a46b030 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.636713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.826266Z digest=sha256:2ed1a838bb19525b51c966a0c5d3e2c1ae6e0166ff52daa15f690287e861d740

Observation 520d8522-0943-466a-8dfe-a5182b9142b1 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:39:35.629569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.828757Z digest=sha256:0f48ec6f3d90c7effaa3fc0708a6f97644ec509c37bffbe0124ed9f351e133db

Observation 96c4e350-66c7-4c5e-87f5-fc76ddad77ec · outbound

This paper cites J.; Cheng, Z.; Shin, D.; Lei, F.; et al.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables J.; Cheng, Z.; Shin, D.; Lei, F.; et al

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.621353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.831495Z digest=sha256:2c66e33fccff5b355fe8945eaad1c9daaa49c528e8af9404eb919dd4075cf96d

Observation 2bbe066f-1684-466d-9b35-57fe113b1f3b · outbound

This paper cites TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables TheAgentCompany: Benchmarking LLM Agents on Consequential Real World Tasks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.834018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.834018Z digest=sha256:e7c437e60176401022073dd72318248e3b7506d1d8368f0316e9f2b5564e90d1

Observation 5614d325-e0e8-413b-bb38-f520a481b47b · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.836759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.836759Z digest=sha256:3210c242cba01252218bb0d1e412e317a07137017a8c0e656981255a16c39d8c

Observation 55fabdb5-d135-4096-96fd-b7aa4720107c · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.839268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.839268Z digest=sha256:5d7ce6df6e5c6e63b000f3d298f296f486a2b8be4e0dea4e92ee8cf4af7c2ef3

Observation 2729b7a3-2b0c-43b9-9a03-93ffc16555c0 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.841766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.841766Z digest=sha256:6b1af47e3e9a16c1f082519eecd45c5bfee50c4c677b5e4f9d83100682a84a33

Observation 5d682386-6df6-4e0c-9e19-d3180f0d3798 · outbound

This paper cites P.; Zhang, H.; Gonzalez, J.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables P.; Zhang, H.; Gonzalez, J

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:39:35.613211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.844682Z digest=sha256:85364ffce75b46ea28c0c3a4afdf0b0375db30a1b864de4dbded4a8e80524d29

Observation 2351c981-e091-4280-9b88-18fdaca87113 · outbound

This paper cites an unresolved cited work.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables Unresolved cited work

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-08T00:39:34.942147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T00:39:34.847515Z digest=sha256:766abc6e3b4d5ab9f8e0e036a78321c5f6f08e4f53e96205aac2ab3bfca16399

Observation 323f78cf-7f11-45ea-8cd1-d91a038ad57a · outbound

This paper cites F.; Zhu, H.; Zhou, X.; Lo, R.; Sridhar, A.; Cheng, X.; Ou, T.; Bisk, Y.; Fried, D.; Alon, U.; and Neubig, G.

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables F.; Zhu, H.; Zhou, X.; Lo, R.; Sridhar, A.; Cheng, X.; Ou, T.; Bisk, Y.; Fried, D.; Alon, U.; and Neubig, G

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T00:39:34.850226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:39:34.850226Z digest=sha256:b3ea2842dffd15cc4877cee540c1eec8d0439c430aa2724b42524a8a1963726f

Pith citing papers

No inbound Pith citation observations are available.