Pith. sign in

Paper Citation Record · LEDGER

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2607.11997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.11997 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:50:47.352443Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b796874c-4845-4b22-9b05-eac8c00850f0 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.016497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.016497Z digest=sha256:25fc2ace13995d7e053fd8d78319faa21060bb02b1cb01808a3ee97fa37df372

Observation 9a6b27ab-4c0b-4105-930e-40b89b306940 · outbound

This paper cites an unresolved cited work.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.370017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.370017Z digest=sha256:e614524562efbc5ffc4801e6fc1ef1706e680a0e294fcc2fbb43cf77ac2174fb

Observation ff15ddd1-6727-4afc-9872-ad964480fa97 · outbound

This paper cites Magicoder: Empowering Code Generation with OSS-Instruct.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Magicoder: Empowering Code Generation with OSS-Instruct

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.625517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.625517Z digest=sha256:a85d085e4b9ecce54696467b9ad147b88aff83b9540630dcdc74df3a4800ffc7

Observation 1ad042f7-cf7b-46d1-a0e7-99a4b7bfdea5 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.683366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.683366Z digest=sha256:49d3cec1373cb85b790baa8ee2d53c10cbf3634ea16f69571cff52c54c87bb90

Observation 1b253a7f-37f8-451d-89b1-9d2b02ed5a58 · outbound

This paper cites Post-Hoc Reversal: Are We Selecting Models Prematurely?.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Post-Hoc Reversal: Are We Selecting Models Prematurely?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.761806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.761806Z digest=sha256:c8b85ed059ab57ee62d1e4e27b43879a1a6bef2432bccad3431965bf9faa6ccf

Observation f1d29d39-3815-471e-8ad2-1ef6ea2147a0 · outbound

This paper cites (How) Learning Rates Regulate Catastrophic Overtraining.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs (How) Learning Rates Regulate Catastrophic Overtraining

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.840290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.840290Z digest=sha256:f79c0953f01603c9b3722ddcd201bfd0e078af52be5407285dd67e061582478e

Observation 06b57ed0-044b-42d7-806c-9cbbd301789c · outbound

This paper cites Leveraging model soups to classify ICH images from the Mekong Delta.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Leveraging model soups to classify ICH images from the Mekong Delta

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:47.120762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:47.120762Z digest=sha256:d47401d5a90f9fa33f1bf0b152b33a9e5c12c59dcb76e22694c4ebe7737b569a

Observation 3de66465-e2a7-4063-a89f-5846ecb6dd93 · outbound

This paper cites DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:47.198790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:47.198790Z digest=sha256:d5428ed68cd3c6f98fba57f3d10edb35047340809d34e084dfda0f6984de8e4c

Observation dde66761-ade0-403c-a483-fde4b3d1c713 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Instruction-Following Evaluation for Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:47.289943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:47.289943Z digest=sha256:1aafa07679abcdd28a4fdcd00d2e7c26af93a7a4e60590d0d07fcf3837759eb7

Observation e7acefc7-7e31-4ac9-95ac-98600eebf8fc · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Evaluating Large Language Models Trained on Code

Reference 2001

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.150392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.150392Z digest=sha256:56f20ba9db573276099aa0c55d6a544efdce3afc73ba97551c9dee104835b66e

Observation 03f676bb-151d-4718-a2ea-a6e969067fd2 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Training Verifiers to Solve Math Word Problems

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.312944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.312944Z digest=sha256:5cbb5662c9a159bf5b406002fa52eb21597d8f7a3650b1b1bde40e0c545392d3

Observation 44d98251-a298-41ae-8a2f-6cb8f1d6fd65 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.228734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.228734Z digest=sha256:b670f9b1a02a9ff89d7027beb51523cc6c4e90235a4227f2aa58bce099acd213

Observation fa4101ca-5d6a-44de-a051-72d67ac4bfcb · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.535520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.535520Z digest=sha256:7fb94ca0ac5d421bcf4a55e3377cc2ea7993d15c31f49a1ed25b2a145e6eedc1

Observation 4133b9af-05ba-4708-8dc4-b794ed0739a3 · outbound

This paper cites ATM: Improving Model Merging by Alternating Tuning and Merging.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs ATM: Improving Model Merging by Alternating Tuning and Merging

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:47.352443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:47.352443Z digest=sha256:0fd853f479bca011eeb31d0399b5bbe565dd4f8e80e0a72811eebe37aa6c67aa

Observation ec3a490e-7255-444f-8403-9996ce34ccee · outbound

This paper cites Small Language Models are the Future of Agentic AI.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Small Language Models are the Future of Agentic AI

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.076212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.076212Z digest=sha256:ef7ffcd6f9b8faf05cd09264b4670e104117cf4d250a84cf7c52c9325731dc1e

Observation 295c6de6-2825-41f3-a601-d3454334ba90 · outbound

This paper cites From Memorization to Parameter Interference: How Overtraining Experts Harms Model Merging.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs From Memorization to Parameter Interference: How Overtraining Experts Harms Model Merging

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.438907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.438907Z digest=sha256:f66392c9f0b9828b1d7f1175edf73d6e10ed7a84ff57f3745ffc90b1c8a826a1

Observation c0052134-c04b-455f-9403-6e5c415f3c8c · outbound

This paper cites Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning.

Are we Merging the Right Models? Impact of Expert Training Duration on Model Merging for LLMs Aya Dataset: An Open-Access Collection for Multilingual Instruction Tuning

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-02T06:50:46.973903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:50:46.973903Z digest=sha256:c28031786ff50218db672abe826baefcbfd7a64133f2e9f23d4122748c58e720

Pith citing papers

No inbound Pith citation observations are available.