Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2307.16039.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:05:12.709951Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T04:55:24.835950Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9adc86c2-8131-4a2c-ace1-61ac0989f332 · inbound
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dd7540e7-cb29-4bee-8b02-c108a7c09538 · inbound
Training Language Models to Self-Correct via Reinforcement Learning Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 060b73f1-9848-4154-a328-ac2b062ff445 · inbound
AdaMCoT: Rethinking Cross-Lingual Factual Reasoning through Adaptive Multilingual Chain-of-Thought Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 79639639-7383-4b54-8398-8dc8b3d1ca3e · inbound
Task Vector Bases: A Unified and Scalable Framework for Compressed Task Arithmetic Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4406abb-3375-45dd-9899-f7c59ffc709f · inbound
GenKnowSub: Improving Modularity and Reusability of LLMs through General Knowledge Subtraction Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fee93cf-f89e-4176-bb85-7f8223f0f1e8 · inbound
Semantic Aware Linear Transfer by Recycling Pre-trained Language Models for Cross-lingual Transfer Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18e7052e-52a3-43f1-a667-d4da8503d156 · inbound
The Multilingual Divide and Its Impact on Global AI Safety Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe79e325-7e55-4539-b120-e1ad3e1a14b9 · inbound
CausalAbstain: Enhancing Multilingual LLMs with Causal Reasoning for Trustworthy Abstention Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e19d6b6d-5385-400d-8660-d5b87b42977a · inbound
The Emergence of Abstract Thought in Large Language Models Beyond Any Language Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ee3b8ed-aa0d-41e0-91be-f4186e87631a · inbound
FineWeb2: One Pipeline to Scale Them All -- Adapting Pre-Training Data Processing to Every Language Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 044a60ca-8786-4521-92db-f6a9ac95860f · inbound
Text2Cypher Across Languages: Evaluating and Finetuning LLMs Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f4be4c7-7d94-47bf-b8bf-a289c9cc45c2 · inbound
The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18b34b9e-5ad8-43c2-b84c-b983a4f96d36 · inbound
Uncovering Cross-Linguistic Disparities in LLMs using Sparse Autoencoders Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd3debf8-77b6-4a9f-a89c-07734bb1bd2a · inbound
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62ee54a1-3688-45e8-97f0-c9fe9286fb83 · inbound
TASE: Token Awareness and Structured Evaluation for Multilingual Language Models Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.