Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T23:14:56.217903Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.15516.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T23:14:56.217903Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a0809f22-7d37-49db-bcae-391ce1b28604 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching ACM ICAIF ’25 Workshop on LLMs and Generative AI for Finance
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19eeaa77-46ec-4283-8dda-3a9df4e977ba · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching A Unified Approach to Routing and Cascading for LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e02edf72-1abe-4718-9950-6e60fbc097ef · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Auditing Prompt Caching in Language Model APIs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd748da0-4d6e-40ff-9f9d-b36300ca1cc9 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46dea7f4-22b0-4d06-a52d-df2dee069eef · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a4b7d72-aae2-46c4-93b9-407f4c3e6bdf · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching 500xCompressor: Generalized Prompt Compression for Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dbfb937-72ed-4373-8c73-3dc72786679e · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c176990-f89c-46d2-a762-9301e750ae2d · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Google; 15 authors
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a274ccf-32db-45c9-8427-627e218378d6 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Evaluation across OpenAI, Anthropic, Google on DeepResearch Bench
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d43ba17-ac2f-4da1-bcc0-d9bfa014d235 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Fundamental Limits of Prompt Compression: A Rate-Distortion Framework for Black-Box Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9066041-dd6e-4d9b-b383-804dfec0f00f · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching RouteLLM: Learning to Route LLMs with Preference Data
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0601da62-3a69-4f86-925e-e54011795ef7 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Salesforce AI Research + UIUC
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a733540b-2c0a-402e-9374-f26b32b22f98 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Accepted ICLR
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7a8842c-8dd5-4631-8fcf-d1e7c6c76164 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Converts code, documentation, PDFs, and images into a NetworkX knowledge graph with Leiden community detection and LLM-extracted concept edges
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf0e75c4-ceb1-4ce9-aa4e-8197b8d7d40c · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Shunyu Yao, Noah Shinn, Pedram Razavi, and Karthik Narasimhan.τ-bench: A benchmark for tool-agent-user interaction in real-world domains
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 619f889d-f4dc-42ee-af05-cc6a2eb03f25 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb75c710-b8c2-4b17-a3fb-8b8755367b24 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2801332b-4c07-439b-bc95-e961a678a4d8 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching "" Feed one observed document version
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 540d357a-8690-43e4-ac86-9461ff7113a7 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7f1a10d-9790-47e9-a803-8c988cda7288 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffe15007-a4c7-4182-82c4-3865fb982245 · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching NeurIPS 2025 Poster; 83% hit rate on production prompts
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f79c36e8-5ca0-49ed-ae22-5db0b9c5880d · outbound
Cache-Aware Prompt Compression:A Two-Tier Cost Model for LLM API Caching Prompt Compression in the Wild: Measuring Latency, Rate Adherence, and Quality for Faster LLM Inference
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.