Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:54.786375Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.21958.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:54.786375Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1d396de4-e817-47f7-b7e9-b7b3b91f1cd5 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Gpt-4 technical report,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65edc37e-dacf-487a-8eb9-0a7a5e2cfabe · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning The Llama 3 Herd of Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd7ad862-6986-472f-b77b-11bafee73dbb · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Qwen2.5 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f70421c-5b6b-4c09-9be7-b824e314303b · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning DeepSeek-V3 Technical Report
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12e8ce5-37f8-4164-a3b9-d300194a03cb · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A Survey of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef79f36e-cc32-4260-a0a6-7262ecace4c8 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Biomistral: A collection of open-source pretrained large language models for medical domains,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1c007bb-fd28-40b1-9964-705e026ebde3 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Finetuned language models are zero-shot learners,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18a4122d-9ada-4964-9f35-f92d2e582d6a · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Lima: Less is more for alignment,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e79c7fe-34f9-426a-82d1-216ead68a4bc · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning From quantity to quality: Boosting llm performance with self-guided data selection for instruction tuning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e7239ec-01b5-4fd5-a0fb-cc8895e06971 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning What makes good data for alignment? a comprehensive study of automatic data selection in instruction tuning,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d0896ea-74f5-4eac-acbb-09cbdd56b144 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Alpagasus: Training a better alpaca with fewer data,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc27c8c7-6cee-43ac-a8d4-cfe1331a3603 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Selfcheckgpt: Zero-resource black- box hallucination detection for generative large language models,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5dc06f3a-5663-44c2-9668-c87613474b15 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Knowledge conflicts for llms: A survey,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 15437f69-87e4-47c4-8a42-f433b53ffe59 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Does fine-tuning llms on new knowledge encourage hallucinations?
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71ed3bde-e1c3-4d1f-9037-9693f270a490 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning ConflictBank: A Benchmark for Evaluating the Influence of Knowledge Conflicts in LLM
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20f26d8c-0ee4-4337-b606-128bec430268 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Learning or self-aligning? rethinking instruction fine-tuning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85ba1d44-64d2-42ff-8657-1dcb06f79e16 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A survey of knowledge enhanced pre-trained language models,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5eda4091-bfb2-4790-a8af-5787af75c4ce · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Large language models on graphs: A comprehensive survey,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80d4977e-ed01-4933-a43c-c3e0624df5df · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Towards making the most of chatgpt for machine translation,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5cffbdb6-9366-4ffb-b8c1-5681d168da58 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Prompting large language model for machine translation: A case study,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9128a127-a49c-4d1e-a2a0-2bb23d176d33 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3420d96c-d2ef-4ba2-8100-5139a36d6b13 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Achieving >97% on GSM8K: Deeply Understanding the Problems Makes LLMs Better Solvers for Math Word Problems
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30beb639-c3e1-4cb2-a8b0-12c3c2e6535b · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A survey on aspect-based sentiment analysis: Tasks, methods, and challenges,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56586fd6-958b-4be5-b5c7-139e8b3b30c9 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Knowledge graph augmented network towards multiview representation learning for aspect-based sentiment analysis,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1aece9e6-f3ae-4bb0-b567-eed7ec9f7405 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Recommender systems in the era of large language models (llms),
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d421b72c-9311-4154-a1ac-09bdb0692777 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Collm: Integrating collaborative embeddings into large language models for rec- ommendation,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f5a873d-59a6-4d94-8478-78e58550e5b9 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning MEDITRON-70B: Scaling Medical Pretraining for Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93df9939-9697-4f2c-82a1-fa441beea830 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Zhongjing: Enhancing the chinese medical capabilities of large language model through expert feedback and real-world multi-turn dialogue,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9983f11c-f683-4774-ab5d-901caadd41ee · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning AlpaCare:Instruction-tuned Large Language Models for Medical Application
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30a58653-047e-48ec-a2a3-4fe50789ae20 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Chatdoctor: A medical chat model fine-tuned on a large language model meta-ai (llama) using medical domain knowledge,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9aae50ce-95df-4d2e-9950-8c8d242a8459 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Survey of hallucination in natural language generation,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be9b78a-ea5b-4fa4-b5da-09b26f18a6bb · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning AlpaGasus: Training A Better Alpaca with Fewer Data
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b58b89a8-d939-4cb0-8183-236b9f97704d · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Resolving knowledge conflicts in large language models,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a294b64-64df-46cc-b604-9f6b5b5d953e · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Knowing what llms do not know: A simple yet effective self- detection method,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a946d2de-0545-4d9e-9619-1d28cbe69367 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Tug-of-war between knowledge: Exploring and resolving knowledge conflicts in retrieval-augmented language models,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1824b8df-6ae1-4778-8440-f88a1f06165e · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Characterizing mechanisms for factual recall in language models,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 745e69b8-53e5-41f4-a0a1-a7a6d97279bd · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning 3ds: Decomposed difficulty data selection’s case study on llm medical domain adaptation,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873964fe-71f8-4e52-bf76-bf5b35d4140e · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Language models are few-shot learners,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 661449de-fe25-4386-b4b4-1cd1a8470264 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Rethinking the role of demonstrations: What makes in-context learning work?
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f051c0f-6d08-46fb-9b25-7dcb080c95a2 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning 60 Data Points are Sufficient to Fine-Tune LLMs for Question-Answering
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2e38901-2778-4804-99a5-ff5a5228a427 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Self-consistency improves chain of thought reasoning in language models,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0677557-808e-43f1-85c0-bea677a2471e · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Detecting hallucinations in large language models using semantic entropy,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5a59bdc-b0f9-4154-818e-174ffa8b5f7a · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Semantic uncertainty: Linguistic invariances for uncertainty estimation in natural language generation,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be37b4b8-76f4-41e5-954b-7cac5b4becb3 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A Survey on LLM-as-a-Judge
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5ac1c1b-7ca3-494e-a139-3afcdc72ae8b · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning M3-embedding: Multi-linguality, multi-functionality, multi-granularity text embeddings through self-knowledge distillation,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7dc1ae78-6f9b-4a9b-bcc2-7d591f5c5673 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Stanford alpaca: An instruction-following llama model,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e49bca8-9d8f-4c94-9205-feea6bf9d26b · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning MedAlpaca -- An Open-Source Collection of Medical Conversational AI Models and Training Data
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78567fff-a044-4a83-8c21-c352140193fa · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Medmcqa: A large- scale multi-subject multi-choice dataset for medical domain question answering,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 812a55b4-f5cc-4043-97bc-025fdd11c0d3 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning What disease does this patient have? a large-scale open domain question answering dataset from medical exams,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51a2eed3-fe37-43ba-8ce3-f345899f4217 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Pubmedqa: A dataset for biomedical research question answering,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c387b4e-dad7-4de2-bc1d-e14b4cbafcc0 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Measuring massive multitask language understanding,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6341f18b-71a1-487c-8741-1af0bf8f369a · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A benchmark for long-form medical question an- swering,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 095e1fbb-fb8f-4bba-aa2c-66e7cad9145c · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Rouge: A package for automatic evaluation of summaries,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21b2d988-8ac7-4cfc-acb1-962b3f9bc556 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Large language models encode clinical knowledge,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2683c4ee-deb2-478d-9793-31ce0d71355b · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Toward expert-level medical question answering with large language models,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7379532-a53f-4516-9fdf-895aefa8aac5 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Lora: Low-rank adaptation of large language models,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba76a535-49af-448f-979b-7ba51080afaa · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Towards building multilingual language model for medicine,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86989898-bdbe-41b4-aaa9-0de1b9785457 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Med-halt: Medical domain hallucination test for large language models,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cd20c5e-0dc7-40bc-a92e-5a9946dc073d · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Xuanyuan 2.0: A large chinese financial chat model with hundreds of billions parameters,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17ec2116-e2b9-44d5-9343-d0afda1769e0 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning Debertav3: Improving deberta using electra- style pre-training with gradient-disentangled embedding sharing,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4d5b5e99-c269-4fbd-85b1-2f6ef3172e31 · outbound
Resolving Knowledge Conflicts in Domain-specific Data Selection: A Case Study on Medical Instruction-tuning A broad-coverage challenge corpus for sentence understanding through inference,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.