Pith. sign in

Paper Citation Record · LEDGER

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2607.24273.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24273 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-31T19:30:47.313467Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d3899e85-923a-4eb6-a95e-4cec861ac791 · outbound

This paper cites an unresolved cited work.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.680312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.680312Z digest=sha256:819b6a6c3fcb2c1dc2641dc3438fbdf1d67747f6225105d8c7c0e09d1b233a24

Observation 2aa5536a-f6c7-4b3d-8507-761779d175cb · outbound

This paper cites Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.724972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.724972Z digest=sha256:dde284196a11f58644575359bcfac37ac26d35bab7e45a818534cfe732fd40c6

Observation 81702f14-babb-40d6-84cb-0712e6ed79e9 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Advances in Neural Information Processing Systems , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.788036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.788036Z digest=sha256:da75ec02862a1d82305d6aa380154cf0d121331fb4df39af9d41ee96610a0ac2

Observation bb90880a-6656-45ce-a8c1-54d8d0b40c37 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Advances in Neural Information Processing Systems , volume=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.852069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.852069Z digest=sha256:0882e8e9d2ab3eab03dabe45675a2003be04e643b54ac721ad7979e19f2f5925

Observation d4e45446-f2c4-42bf-bc89-e2a2f829ced5 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.910590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.910590Z digest=sha256:a1129be4f1a8cbfac8c632137cfe3ad06e1c65301fe347629a00652ee7518829

Observation 7dc76271-86b2-4b7b-86de-3ed745b33092 · outbound

This paper cites InsQABench: Benchmarking Chinese Insurance Domain Question Answering with Large Language Models.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models InsQABench: Benchmarking Chinese Insurance Domain Question Answering with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:45.989599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:45.989599Z digest=sha256:63ba6f4c5fb0b6a9a8edc1f4b37e1fcae8e64ac41e1143618f2c8a9e3c8a8447

Observation 47f6cb9c-2cb7-4ddc-9934-307efd54dbb6 · outbound

This paper cites INSEva: A Comprehensive Chinese Benchmark for Large Language Models in Insurance.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models INSEva: A Comprehensive Chinese Benchmark for Large Language Models in Insurance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.079681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.079681Z digest=sha256:859a178ff797c67c85a0e66f3246ed91cc4130e71ea7c4d7cc63fa967ae2567c

Observation a2789164-643b-4ed2-a5da-de53b96dd7f2 · outbound

This paper cites arXiv preprint arXiv:2511.07794 , year=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models arXiv preprint arXiv:2511.07794 , year=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.120045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.120045Z digest=sha256:dddddff4523ad44ee8f5957e50228d336076d7330d54654b384ffe8d665cb49f

Observation 3ce75931-1e8b-496f-8043-01de660f8373 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.164146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.164146Z digest=sha256:adf6084b857edfe41bca44218f8656dc01b8e204c958786957e93f7a01b086d3

Observation eabd5823-d10f-428c-bbdc-0398d81e138f · outbound

This paper cites British Actuarial Journal , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models British Actuarial Journal , volume=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.203448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.203448Z digest=sha256:ac2f06e5673cc18733c7b4ebb56c7298ce1088965f2ab3ce6c010da8e56ba3c0

Observation adbe69b1-e637-4da3-8073-e4ab71dfa331 · outbound

This paper cites an unresolved cited work.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.240586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.240586Z digest=sha256:50a6dd36c815d8ee2c269e06c96e2e06dfab848880d05f4f0590ccb499696376

Observation 2b51979f-0aa7-4ff6-82b3-0a3f9feb81e9 · outbound

This paper cites Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.283987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.283987Z digest=sha256:15a2f670bcbca4bb22b69991a67a94b3ad0a6fb08a9f01a27511c996fddda62a

Observation 27f08a79-426e-4b0e-9562-4fc517e0b255 · outbound

This paper cites Proceedings of the 33rd ACM International Conference on Multimedia , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 33rd ACM International Conference on Multimedia , pages=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.335481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.335481Z digest=sha256:4ed360409ef5a31081ba68ffb0f5dc9007d78a3a403272ce0dac8bb6c3fafc4a

Observation efeb2dcc-bacf-4c22-b4e8-a03a6b495ff5 · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.392709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.392709Z digest=sha256:6d22dfb0729ec4376eff02cc7fe68fb421b9d7c68d3f3de40eae3cca945e2eaf

Observation fbc213e4-baa1-44a6-a969-06b763ac9f6a · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.447495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.447495Z digest=sha256:e91cfdc2152379cdc1260d5a927a9d3242b431a293c3c4a6be51ca94aff05661

Observation 9b090139-8c17-4d57-86bf-2efe59280e09 · outbound

This paper cites 2017 , publisher=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models 2017 , publisher=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.500855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.500855Z digest=sha256:21ffee435cabee3e13393cf81e525b0ba6bdeeea88a0497bdbadc94eec71c47b

Observation b4ccb06d-117d-43b9-a11a-eb8cf344a0ba · outbound

This paper cites FinanceBench: A New Benchmark for Financial Question Answering.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models FinanceBench: A New Benchmark for Financial Question Answering

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.558715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.558715Z digest=sha256:ce19b655557c98821fa603dde091dc9269928d57eab4dee7847490fd8d352d85

Observation fb017d84-86a3-4d28-935e-c19f98f52cb6 · outbound

This paper cites Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.562005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.562005Z digest=sha256:4cd1ce42a8bc7a8bfb4ee4a431ad822d34ad02d01344b3ed6e6fa5578a679894

Observation 88a6c4db-4140-4b3c-a4af-65846e80a3f4 · outbound

This paper cites arXiv preprint arXiv:2603.07316 , year=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models arXiv preprint arXiv:2603.07316 , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.565434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.565434Z digest=sha256:c658035241baf92db8b7d3d3d7e799c2c367850401447b0d922f062d9b00117a

Observation d7b1547c-fef7-464c-9f3e-28b5abbe39dc · outbound

This paper cites Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Finch: Benchmarking Finance & Accounting across Spreadsheet-Centric Enterprise Workflows

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.613565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.613565Z digest=sha256:ef3bb27a7edb46525e43501e8043337bd118adf92d1084aeebfaea87589189aa

Observation f1de9ae5-2365-4ca8-8b34-be350bbc13fc · outbound

This paper cites Scientific Data , year=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Scientific Data , year=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.696129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.696129Z digest=sha256:1fc59f9b32c2fa8a1a9ac544a5189aff1e4e72793c2a097300c94b14bc5a3f95

Observation 3b4d9300-ab0e-4e01-9ed1-5c17d5a97b7e · outbound

This paper cites Proceedings of the 29th symposium on operating systems principles , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Proceedings of the 29th symposium on operating systems principles , pages=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.803582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.803582Z digest=sha256:3bb2110c43e0a53466c2ccb61191b41e3434ce841488d46fc09478c4ab0b34bb

Observation e2266c6f-e905-40b0-8074-c6ed2ce461cf · outbound

This paper cites From Abstract to Contextual: What LLMs Still Cannot Do in Mathematics.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models From Abstract to Contextual: What LLMs Still Cannot Do in Mathematics

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.892090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.892090Z digest=sha256:17cbe48e092e2dbb417a9e6c7f3b486b9496360fda878712d60e2bc1cf2e7b59

Observation 17b8c6e2-9cd0-4e8d-91e6-d48a5bb09d3f · outbound

This paper cites NYU Stern School of Business , year=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models NYU Stern School of Business , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:46.952783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:46.952783Z digest=sha256:ed5e5f1dd9981447f2de8800f2d964238f82ef6a5217dd1da133f6b208a75f71

Observation 5121772f-673e-4af6-91fa-0a5ed56458ca · outbound

This paper cites Journal of Artificial Societies and Social Simulation , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Journal of Artificial Societies and Social Simulation , volume=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:47.036841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:47.036841Z digest=sha256:e41c1cdee42359f850d2de453b4ddb21c343bcee31385a96329ea3c2867ccaaa

Observation 77800c3e-9e14-4d5c-a984-c24e11b9c71a · outbound

This paper cites British Actuarial Journal , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models British Actuarial Journal , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:47.129097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:47.129097Z digest=sha256:6a44368d08e83f2cb768e9ec42bb0d5fa5cd0e0fc6e2700abb8c774ee1a9561d

Observation b635c8ff-b8af-4100-8b50-4e21cbd3d73d · outbound

This paper cites Actuarial Practice Forum , pages=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Actuarial Practice Forum , pages=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:47.223107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:47.223107Z digest=sha256:d4ab1ef81c219a5b5eea3d6525a50acc3048cde79c6cae26d17fb8691109ace7

Observation 4056c7f6-5d7b-48af-8b5b-17e56aa9b0d5 · outbound

This paper cites Journal of Statistical software , volume=.

INS-ActBench: A Comprehensive Benchmark for Assessing Professional Actuarial Capability of Large Language Models Journal of Statistical software , volume=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-31T19:30:47.313467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T19:30:47.313467Z digest=sha256:ae76d5a0a73b2579c89ab031e06ad70f301a98dee2705524fa095bb7a8d07e74

Pith citing papers

No inbound Pith citation observations are available.