Pith. sign in

Paper Citation Record · LEDGER

Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2404.12387.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.12387 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:43:41.818671Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:45.386126Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6b7d89b6-f32c-4a5d-a4cc-bf10d1d785b8 · inbound

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites cites this paper.

How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T20:58:59.171594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T20:58:58.849040Z digest=sha256:3918760ff1eb52da99c0438ea1b67bc52d524407fa19718363bfbf53b94a30d9

Observation e93eb41e-81e8-443e-a612-465ee0902820 · inbound

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs cites this paper.

VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T02:44:53.616921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T02:44:53.284345Z digest=sha256:aaf9091696f0ac9dbc3751657e15c1b46bdf17a5c4e4368e734fbab91dbb58cf

Observation f429b08d-65b1-41f1-9c39-98f38a7d72e6 · inbound

Long Context Transfer from Language to Vision cites this paper.

Long Context Transfer from Language to Vision Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:08:36.171513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T07:08:35.946669Z digest=sha256:423a38652efa4e46309c02bcca144711b3caf1ac49e88520c3ca451a58b8e584

Observation 5b9e54a3-21fd-4971-8b71-00a3a6b2bfea · inbound

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models cites this paper.

mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:20:36.412811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T06:20:36.235304Z digest=sha256:e2f6af255eb6ca604e0a5103623f4b6583348e3e008c79e3bd5501d6eb729051

Observation d5aa68bd-f30b-4e65-858d-7977993ceef8 · inbound

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities cites this paper.

From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:43:41.818671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:43:41.818671Z digest=sha256:3d89d243c760356d201cd295aae8f655071e81b99b69b32dd020b28df5d5bb2c

Observation 9b117915-a4df-4d8e-9b6c-4336376c7e33 · inbound

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers cites this paper.

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T01:03:12.205474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T01:03:12.205474Z digest=sha256:5a9ab62f1277e70e342ad062c32ec9e32aee7d8c1a36699175109784399b9165

Observation f7a866fc-febe-41f6-8dce-43297f6a59de · inbound

LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch cites this paper.

LLM360 K2: Building a 65B 360-Open-Source Large Language Model from Scratch Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T20:54:02.313497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:54:02.313497Z digest=sha256:221fa22ae7be0189128e16dfb562b1ea5cfb0ffc4fd6e953b70f3d2d33c1198b

Observation aa8832c5-e30d-471c-a17d-7504612c1fe3 · inbound

AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment cites this paper.

AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T00:02:26.260452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T00:02:26.260452Z digest=sha256:e6984b133bfa5e2a0f5ce229e5b60ccf2848811ce66c17c1586be56971beeed6

Observation c5ae3c7c-3872-4190-835f-b79be9d1e161 · inbound

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs cites this paper.

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:53:26.140705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T05:53:26.066674Z digest=sha256:2b576dcbba266fcf89453ea17e03f788e99484024f2e91de3c219620abaf2f9f

Observation 72595cfc-cc34-4337-9f94-be24665ba15d · inbound

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models cites this paper.

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T20:55:05.932609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:55:05.932609Z digest=sha256:e3c52f89b03bee3ed260b29d5f46eafb03a621e262e64a1671f01a4ccad787b1

Observation 3bf026fa-d614-4338-bdd4-9ba4f8da7982 · inbound

MMCircuitEval: A Comprehensive Multimodal Circuit-Focused Benchmark for Evaluating LLMs cites this paper.

MMCircuitEval: A Comprehensive Multimodal Circuit-Focused Benchmark for Evaluating LLMs Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:17.148602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:17.148602Z digest=sha256:7570a8a2a33f040441b0e268756e3c24ff5818de3e2c4983bbda15e957c71ba5

Observation e28eb4f0-7cfa-45e4-880c-b2695d6e4ce8 · inbound

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking cites this paper.

GraphThinker: Reinforcing Temporally Grounded Video Reasoning with Event Graph Thinking Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:50:17.299732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T20:48:44.933542Z digest=sha256:0c32a6df7360ed5f65051e235d643ef2f4f5cfda34c2c2fc7cef7df2a3e60e4f

Observation 9459c624-962d-4ac2-9a33-916af85596d5 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:45.387734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:2479bad2885f737c8958a0afec898121504b647f1f437e5869355f575e2ea919