Pith. sign in

Paper Citation Record · LEDGER

MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.14731.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.14731 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:42:21.377519Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:13.092058Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e30a4de3-986a-4d57-bf11-6c0d9001326c · inbound

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache cites this paper.

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:21.377519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:21.377519Z digest=sha256:f74d8f2333f2f2a538ebcabb72904a9dd516ccf5f0edcfa5c8171632cbccb1c7

Observation 2434c60e-8e4b-408c-8625-af96bc073b3c · inbound

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference cites this paper.

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T17:55:09.824153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:55:09.824153Z digest=sha256:7d5787f9935062574fd4851af6e1ec783cf02478b90b935f046ea03c74145dfe

Observation e472c769-b3ee-4501-9645-99c78f93e942 · inbound

OjaKV: Context-Aware Online Low-Rank KV Cache Compression cites this paper.

OjaKV: Context-Aware Online Low-Rank KV Cache Compression MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:26:24.708939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T13:26:02.980973Z digest=sha256:ea62211031896b3e433dc7777af181f24bad52aaef66104fb7e6f81bf5f0edb4

Observation c96d1737-b759-4a46-9dd5-16974214fa8f · inbound

Quantization Dominates Rank Reduction for KV-Cache Compression cites this paper.

Quantization Dominates Rank Reduction for KV-Cache Compression MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:04.594047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T15:43:48.986845Z digest=sha256:6cf580fbac08fa5494ddd22e668dc06c869c4f29e7e3d0fcbed868c402a7ac86

Observation 2857c77d-554d-4a25-8550-0a7ba8acaf74 · inbound

Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales cites this paper.

Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:04.467078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T01:44:42.989053Z digest=sha256:daac42c182efab9685bfa312da25cf526399008011e8eafaa6d0b6d178037372

Observation 2e6f4999-a1af-4935-8382-db3ce7f018b5 · inbound

OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization cites this paper.

OSCAR: Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:18:16.471577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T12:16:07.797702Z digest=sha256:941cd22d6b165fd39a975d437d15b273419550a930ae494f4409c5059bbd9d47

Observation 962da862-1028-4ed9-8f89-857d9b0e5781 · inbound

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference cites this paper.

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:43:23.813351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T07:43:18.828740Z digest=sha256:e7646c286bf8169ff33c8c3128eb2809bdd1499acb8194f1f10b0f9df3631605

Observation 52de2b73-dddf-4db2-bd30-0f7cf05a5a1b · inbound

STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control cites this paper.

STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:57:26.283164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T18:34:20.645677Z digest=sha256:05406ba3cd5d0e861d191714d54b7ad4bf57938af195225cfff2ecefac6fe730

Observation 5a6ea966-1a82-47a5-a75a-00b743534a51 · inbound

Dual Dimensionality for Local and Global Attention cites this paper.

Dual Dimensionality for Local and Global Attention MatryoshkaKV: Adaptive KV Compression via Trainable Orthogonal Projection

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.093402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T21:22:05.095209Z digest=sha256:613a1afbf063b5902ff088a49a5c37b10922b22a7516408bee8db6f74d5378eb