Pith. sign in

Paper Citation Record · LEDGER

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models

As of 23 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2509.10744.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.10744 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:41:26.749626Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved25
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8486a350-9521-4f10-84f8-264e20c14bf8 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.324388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.324388Z digest=sha256:7f943ae0001c1384e36f9043b17316cf00be65125ba36ba966cd436ad2ba52a7

Observation 42f280f6-3db1-4a6f-b8b0-d8da1ce95671 · outbound

This paper cites 2023.RADIATION AND CAN- CER BIOLOGY STUDY GUIDE.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models 2023.RADIATION AND CAN- CER BIOLOGY STUDY GUIDE

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.392622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.392622Z digest=sha256:73916d3f1973a4903275aad9e419574f15f80c11d0054d3d2b43c710484ca3e2

Observation 2723ce5c-c6d4-474e-854d-4209523dabdd · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.466641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.466641Z digest=sha256:ca4cad3082decd73abf7c971b43c619f196eb89e09ddcbc86baafda25744669e

Observation 9ed580f0-67bd-44f9-ab7b-ad35badd4874 · outbound

This paper cites Qwen Technical Report.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.542294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.542294Z digest=sha256:3418dbf90b672d62b54f915a52294619625013cbcfd52c2ead5e179e95b50028

Observation 38455af8-88ef-45e3-b4cc-7a583862dd47 · outbound

This paper cites Beattie, S.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Beattie, S

Reference 5

Resolution
verified exact
doi, observed 2026-08-04T17:44:06.657194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-04T17:41:24.599332Z digest=sha256:d71094711c63dac7f44c4d07423f2af4ceb7bd5b705ab6410a2a1c6bf963ad35

Observation 4361f049-0e8b-407b-a518-08ab7e77b44d · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.693090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.693090Z digest=sha256:b89ef2089452f7e710500df4f1a8c43d06fb648ca9bfcfb0f53665a1a3535933

Observation 3180879f-8562-4e9e-aceb-d2f3dc9be986 · outbound

This paper cites 2024.Science and Engineering Indicators 2024: The State of U.S.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models 2024.Science and Engineering Indicators 2024: The State of U.S

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.730975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.730975Z digest=sha256:052372123031a48b57d75065f2b456075f4111154c0207fbac5511fc1b3b6f84

Observation 989a7007-af5a-41f5-b050-a7c96927f59e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.794393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.794393Z digest=sha256:19d4d6029e0d953cc0987e5bef79499d489df63aabb6e9149234a73b1f638ba9

Observation a6c64ead-25fc-4b2d-9c18-e5671b6153c4 · outbound

This paper cites The Faiss library.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models The Faiss library

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.828179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.828179Z digest=sha256:14fc3445a935477e1acbe6344ed09730ac59a010b2171afecd75219b1cf3d949

Observation 4048d7d2-cec4-4399-aee4-bdcb8c762105 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.900523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.900523Z digest=sha256:d5565524ebc13b037d53d87e5c6dc74fb8b2e53b75ba39b51379e73003e644c6

Observation c675e996-8c89-482c-aac3-1b4f477fddc5 · outbound

This paper cites The Llama 3 Herd of Models.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:24.955443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:24.955443Z digest=sha256:fb3d2963cdfb3cf7c088bc55af963f36170a205a84f8f58f8d250a054b29bb5b

Observation 62aebe7e-b7c5-4516-ad99-d2d10bb53810 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.057261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.057261Z digest=sha256:892ef820925d909ef07c1255bc4e8e18d14cc2c73a5cfe839b9fea82e54a110b

Observation e6a80fa6-8f8a-4100-b825-882ace8c95dc · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Measuring Massive Multitask Language Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.150748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.150748Z digest=sha256:80b68d3c0425566291d6d74eda7c2a95f9c23858422aaa6daa6b93379ab0a0d6

Observation 12cefcf7-075d-48fd-b8d9-25def0368f62 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.253320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.253320Z digest=sha256:7f9782572b19def3283efdbca9fa219d17e00f7857065539b7ed403c5ec97e5e

Observation f1e518fe-2069-408d-915c-6c20be29c67d · outbound

This paper cites Mistral 7B.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Mistral 7B

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.356677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.356677Z digest=sha256:0808f7c779d5c098fcd6f3c4970b9b548c553180156bbe363cac24151925d0e5

Observation d27d0ddc-3885-4cf0-9321-4380fbebe63a · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 16

Resolution
malformed identifier
no resolver link, observed 2026-08-04T17:41:25.470163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.470163Z digest=sha256:330873cef5ab20a96c441d174c678bf9c194447cb6013363d41dfd340b3de24d

Observation 3d437974-6ad9-4d14-bf0f-ffd8a2cb1c7f · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.555929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.555929Z digest=sha256:5e7f432325e5ea794ab0fadef81665156f3e1ea45e837914a45e8a6cde783770

Observation 84c28f20-78fa-4c01-b41e-8b6ed0414fdd · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.608152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.608152Z digest=sha256:3cd59881ac310ae5fcb16cde3197f7dca3062e2f56bd92f45fdd2cd3a8df1dd7

Observation 8a08fa86-f3d1-48a8-b5bc-162b2240f87e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.685924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.685924Z digest=sha256:c1bd55f78a493432b775011a35cf0cd45f06e411852613af054eafedcd840fb2

Observation 9c914f5a-12a2-4d9f-9177-e6e9d074ff7c · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models OLMo: Accelerating the Science of Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.771827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.771827Z digest=sha256:f5fb56e72777c3e74eb4412145addb9dd2b097f74c7cbba39350708d86cde84b

Observation fe18cd74-ba37-409d-8e09-481a06146177 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:25.884961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:25.884961Z digest=sha256:17ec063870023ce6bce717d0b1aa272d925db56136c6814bd91b533105b65d49

Observation 7439b183-a787-42a9-8c11-efce408a164a · outbound

This paper cites AdaParse: An Adaptive Parallel PDF Parsing and Resource Scaling Engine.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models AdaParse: An Adaptive Parallel PDF Parsing and Resource Scaling Engine

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.013320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.013320Z digest=sha256:552452c2c7df02c6dc2beb08bbc72ad83874c68314ed3339309a13cabf986834

Observation 090749f7-9de1-42a0-8109-3e0e569522a0 · outbound

This paper cites Gemma 3 Technical Report.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Gemma 3 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.134486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.134486Z digest=sha256:da57a82928f0260e5553fb3aea058ab72f175af24d8190d4c11b82bf0345cd16

Observation dcd6ef8f-2025-4bdb-bd5d-88e958fd75f3 · outbound

This paper cites AstroMLab 1: Who Wins Astronomy Jeopardy!?.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models AstroMLab 1: Who Wins Astronomy Jeopardy!?

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-04T17:44:06.150232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-04T17:41:26.265139Z digest=sha256:d575d36a7482646d73460622c4a4a609ded35be2ee43ec02f9f6b67a269a333d

Observation cc4e81db-6201-47b3-8dab-57b9685da8b5 · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.408422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.408422Z digest=sha256:7da89567a6132d8ac1752fa408ae823bffacfb4cdde07fe43efc320a6d24accf

Observation d1ec3a66-d79d-44e4-8d7e-513ba62d7d8e · outbound

This paper cites an unresolved cited work.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.633666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.633666Z digest=sha256:ad01d41ae8174bf050cdce35d4e41d344d42d12d7bde68086f895a8d8053e7aa

Observation c24f4045-7edf-465f-b4bf-a04bb640e1fb · outbound

This paper cites TinyLlama: An Open-Source Small Language Model.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models TinyLlama: An Open-Source Small Language Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.749626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.749626Z digest=sha256:37273b63df6a4c0ccefdb4d4486995a614e3cd360a6b2093dbd7a810a454c859

Observation 7c317ed6-eaf3-4c24-9754-50e0d01248c6 · outbound

This paper cites InExtended Semantic Web Conference.

Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models InExtended Semantic Web Conference

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:26.516157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:26.516157Z digest=sha256:ff23afa5aff9822cce904d812b86154a5499fbcd2e4dddf334a8e6d0b09ea8e3

Pith citing papers

No inbound Pith citation observations are available.