Pith. sign in

Paper Citation Record · LEDGER

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing

As of 18 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2505.08651.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.08651 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:52:37.602286Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:45:06.385103Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:45:07.351527Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation deceff0f-ed70-4cc8-8400-652d4fb1c73e · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.953784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.489618Z digest=sha256:91b3f9b93d2c576fc7b7144090023f39ef12696c413142bcbdd718a6290fa31c

Observation 2f20e7fb-2805-4d4e-a856-b5dad70f62cf · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.495059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.495059Z digest=sha256:0ed2492d188ed2c418b8b5d0817e3c82914b177552d7b0d6b75c4cfbefe7886f

Observation 974295d3-9c6a-4cad-a8a9-51533eda8b57 · outbound

This paper cites LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.500036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.500036Z digest=sha256:248e1dca82ca4f497237ed7519445a047cf355056ac39ec9e19dccb7379053ac

Observation 0cf48ab1-734b-4aa3-9d78-2fb294406ec6 · outbound

This paper cites Data Engineering for Scaling Language Models to 128K Context.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Data Engineering for Scaling Language Models to 128K Context

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.505381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.505381Z digest=sha256:ed505159085695751f0219a5fce52954276a7bb794e316b3ceae4ebf34ae24a8

Observation 5e5573b9-059d-4090-a526-067d1e223b38 · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.510928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.510928Z digest=sha256:74f065bbf20164a8767d58a5004c89db2b16da9fe2b898bdd29f7526d83c1920

Observation 655fc87d-6bec-4afb-8746-af449661fe36 · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.517238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.517238Z digest=sha256:c97b37235133dd7c137095246ab737aafa4dfb8459469ba09a072ff10d487326

Observation 7af3106e-68ac-4096-9bbd-6063a8da3b42 · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.522779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.522779Z digest=sha256:7dd576090ff9f4240663fde6cef32e9073220cbaaea10b27f601f8358949308f

Observation 6449a67d-54a7-4f3e-89ce-5193ab392bb8 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.926026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.527604Z digest=sha256:53fa6017535a9f04b8e77df7fedfd5126de92e9df9e71d5e72135cdda062a48b

Observation c49de547-a95e-4f53-b5cb-dc10a1c12874 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.909140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.532178Z digest=sha256:213edccebb40c43146e32b2ae73125b2d095495a386fde3b3cf79c2c92dc3c60

Observation e81f437d-a51a-4c0f-8f0b-fe8622031f86 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.892817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.537152Z digest=sha256:d62420d45cfe7100f84102f9960d170eb455b978b96b26f65aa7716cdff519d6

Observation 6eecb814-6264-47f8-ab56-a2556b693de7 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.542632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.542632Z digest=sha256:f97b867f1ab674b3bc57c3636f6050c53d757175b92288f95f753520cba3710a

Observation 8c00a651-6ae2-46a0-9446-361a3f813e24 · outbound

This paper cites Scaling Laws of RoPE-based Extrapolation.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Scaling Laws of RoPE-based Extrapolation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.548785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.548785Z digest=sha256:0535cb90916551dcad13bfa8e4e15d64b016654eb8be66ae7cde5ae5b5d265c9

Observation be6df86b-9675-4f1e-972c-c9b6e25478a8 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.876934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.554736Z digest=sha256:e468d5aba6352bd369d39849633e79f9beabce6b6275070f3d73f6b8f5c4af3c

Observation b3ed50ef-2f5e-4ca9-95c7-b4ea88d69f74 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.861765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.559820Z digest=sha256:3f44b24fcba03a2ee464c671a891eac209dd0925dae37a6f72da5b3817ffdf48

Observation 9eece4d8-5123-454a-a3f7-1e1c9d29c6ed · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.564517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.564517Z digest=sha256:b77cba6ed3bf45a27ed58c0508cfcbca14989315052c01c9edbe7132450d6be9

Observation 5ae7b4dd-c227-4bf4-8862-c41b06cc3100 · outbound

This paper cites When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing When Precision Meets Position: BFloat16 Breaks Down RoPE in Long-Context Training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.569688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.569688Z digest=sha256:25a947f3aa926afef5c9763b62049aa766ef4e1d8b53d1eadbd0048ef4425fd0

Observation adeebae4-3e4c-4f77-b4cb-5ba67f5b9a89 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.836419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.574958Z digest=sha256:dd6aa5ed03669d767ecea7966498469c79d68479bfd60ede6660f42ceeb6b8e3

Observation 464f27af-79e5-4c4a-97e3-ae8d6435568c · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.579689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.579689Z digest=sha256:5d989a29346ef95e733963a11a26aeee9dd1eb3280cdb5b3777bc23b59e45627

Observation 434607cb-0a15-4d2e-a708-628b6c21cb20 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Yi: Open Foundation Models by 01.AI

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.584431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.584431Z digest=sha256:b7ec08f322d41d959f95f954aa51d892be549611a6340ef5b45f279f6bd70500

Observation 658366cb-8d79-429f-b163-b2b2df6ec036 · outbound

This paper cites an unresolved cited work.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:52:37.820329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T21:52:37.590028Z digest=sha256:dd0affb90b00252d186c7924947fe190a3a22c384218c4c373159aedc5939374

Observation 7e1feec4-9ead-4fb1-bc1c-a8969ae82adb · outbound

This paper cites online" 'onlinestring :=.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing online" 'onlinestring :=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.596717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.596717Z digest=sha256:a78e9764613948bdf5da3d255e8f29c83dda71070c9a70346e777678a3c811c4

Observation de896502-c1a1-4f41-a67a-d5187f62430d · outbound

This paper cites write newline.

Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing write newline

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:52:37.602286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:52:37.602286Z digest=sha256:39eae10c5ccfea6e8e19edf1445a1709228f9c7fa0c61fcc50ad193b7eb040e7

Pith citing papers

Observation 18b3d858-4843-4de1-9cf0-4ddb4cf95d73 · inbound

SeqPE: Transformer with Sequential Position Encoding cites this paper.

SeqPE: Transformer with Sequential Position Encoding Scaling Context, Not Parameters: Training a Compact 7B Language Model for Efficient Long-Context Processing

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:45:07.462889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:45:06.385103Z digest=sha256:b0c45c997fc5bf281cfaf11adcf933defd2413a5b287e18e5fa61c8e5705c843