Pith. sign in

Paper Citation Record · LEDGER

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

As of 16 August 2026, this Paper Citation Record lists 100 of 147 outbound references and 2 inbound Pith citation observations for arXiv:2412.06845.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06845 v6

Coverage vector

measured 100 of 147 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:27:18.375394Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T07:28:20.248452Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T07:33:07.543757Z

Reference resolution

100 of 147 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved99
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cbc60f7-5b3e-44b4-8633-1a7323d305a5 · outbound

This paper cites GPT-4 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.908700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.908700Z digest=sha256:018eff1cdb59045fac4028f58ac94f703516ab6481fde88cc271168824bcbe23

Observation f9025085-c100-493b-b23c-4e41eb1fffc7 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.913900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.913900Z digest=sha256:39d2bf9f087da10f4bd44d8d74bc03f2955cb1195917ff9cf0fea9bd67080038

Observation 7d826f12-3d0b-4108-bc29-4d7e2e867748 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A Family of Highly Capable Multimodal Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.917864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.917864Z digest=sha256:2a92f7e074aceb11bf31f0b315bf34fa247143bc0d8fa74ca36d3586890550b3

Observation 10ed27de-30c5-4417-88a0-505435db5fb8 · outbound

This paper cites The Llama 3 Herd of Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.922529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.922529Z digest=sha256:db943d4bb29cb3afc6e6f402f655dfd60f071db09d8057e79dbe2b12dc5299ee

Observation 59d72586-8095-46da-9571-cc457a7d749f · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.926903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.926903Z digest=sha256:c70fedc92bf711cf6d78f5a2d245d61bf174273fb04e842b7aee890fde9bedc1

Observation d39dd68f-bab5-4011-b266-6092d5b56fb3 · outbound

This paper cites Mistral 7B.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mistral 7B

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.931276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.931276Z digest=sha256:98c0cb5917bc074a9a09521fe7feb855a145009397fc29c1cef642a79c0afe77

Observation 0e7dc88a-ffcf-46c6-a8f3-42447b5e2d04 · outbound

This paper cites The Foundation Model Transparency Index.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Foundation Model Transparency Index

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.936174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.936174Z digest=sha256:fd07230680575f46244a9fe2d929cd34fe0e10aabe4ff5f5e48536b656843155

Observation ec3858cf-c5b1-48e2-b6a7-c71243e7a492 · outbound

This paper cites On the Societal Impact of Open Foundation Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the Societal Impact of Open Foundation Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.941292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.941292Z digest=sha256:31e1f41349df7f72e33e39de5162ac6c6736d056813625c39399fd37677d74db

Observation ba54148c-97a0-4644-892b-25d38fb8fca1 · outbound

This paper cites The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.946278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.946278Z digest=sha256:4a132206cba6b551d45bc53df95f5e1074c9bf7d00d141a12246860c1c6c2656

Observation 9c83d35f-f71a-4a7c-82e8-248a2b1f7bcf · outbound

This paper cites Qwen Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.950661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.950661Z digest=sha256:995553b7b9e45d6d2ec882e04a9a918edb815095fb547b88393232ac626d3e3e

Observation b66b9e42-7f90-4201-a427-1221c33f8bfd · outbound

This paper cites Qwen2 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Qwen2 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.955582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.955582Z digest=sha256:11ac1d8c038bc8db86dbe9a63d84d6fab1ca6b710978f14427b0a691e0703ce0

Observation 4b706cb5-2ab7-4e62-b10e-82f0109fbff9 · outbound

This paper cites Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.960659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.960659Z digest=sha256:e174438e5b6576c8cc345f2766c3449680d0f721b5458318be01ae55f675ca5e

Observation 959ccb6b-a24b-49e1-b2ad-e26ec2abd7c5 · outbound

This paper cites Advancing model pruning via bi-level optimization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Advancing model pruning via bi-level optimization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.971474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.971474Z digest=sha256:7be9516b0b4ac82b5ee86fac14713762fb948c2df15b2777276feda1ea8b3874

Observation 4326f0a1-abea-45fc-ae30-070bf34095b7 · outbound

This paper cites A generic layer pruning method for signal modulation recognition deep learning models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A generic layer pruning method for signal modulation recognition deep learning models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.977240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.977240Z digest=sha256:ae7d6d3ad066b9f97936de68e3dc4f36f654936d5bf1e7c1da36210ed73d1e99

Observation c9dee28d-ff30-4c0e-a0fe-cc5122462b92 · outbound

This paper cites Cross-layer graph knowledge distillation for image recognition.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-layer graph knowledge distillation for image recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.983689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.983689Z digest=sha256:0f6230c6f2e44ab4ede21893b9415beaa3395589650db84e2c982c12b65b58ea

Observation 6a9bef99-2617-4551-90e3-a2535d9c6b07 · outbound

This paper cites Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Peeling the onion: Hierarchical reduction of data redundancy for efficient vision transformer training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.988351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.988351Z digest=sha256:970cfaefdaf4321db4db28de74b197b6bc3bcbc0ae436bff2f05aa73d4544da4

Observation bbca38c9-3d9b-4605-a7d5-db71898fbbef · outbound

This paper cites Pruning foundation models for high accuracy without retraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning foundation models for high accuracy without retraining

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.992180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.992180Z digest=sha256:f213d16d6b9f28353d86b0e213494197cec0770830f7260f45afd2bba58aaada

Observation 8c30e468-be81-41a1-8e39-76f2099b05ec · outbound

This paper cites Sparse learning for state space models on mobile.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Sparse learning for state space models on mobile

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:17.996883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:17.996883Z digest=sha256:f11d590a482ce5f64a52d009d1f8fa6629b0b8912ae49ee6f9fad2dcf4300913

Observation ec0d0054-bb58-4d51-82b0-73d668d2164b · outbound

This paper cites Numerical pruning for efficient autoregressive models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Numerical pruning for efficient autoregressive models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.001431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.001431Z digest=sha256:2f55cc3c6557b4219642329b13b09ad6a2e9003da3b19f0e94b7ca9491433f3a

Observation 7ed4b091-0e04-4683-a625-60f728b84202 · outbound

This paper cites Lazydit: Lazy learning for the acceleration of diffusion transformers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Lazydit: Lazy learning for the acceleration of diffusion transformers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.006447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.006447Z digest=sha256:f1027c7275ce3b9e4b8c79b9e1f6aaf3b10daa80fb13038af6a2f137b8c582b8

Observation d071bf78-c14c-4c64-93ae-0314752c9376 · outbound

This paper cites Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.011095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.011095Z digest=sha256:97b1fc1050d73f5db4681964131dea68ad737e38894afe232dfeb8bd68b7c72a

Observation db35236b-feb4-4bfb-876d-72d7da0919dd · outbound

This paper cites Search for efficient large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Search for efficient large language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.015275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.015275Z digest=sha256:19139aaca21f2dd2fb754b03ad8ee1f225845727d5513aa87b80239456a5a5d5

Observation ac3ec23f-a669-4a1c-aa0a-de67d80a8e05 · outbound

This paper cites Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning parameterization with bi-level optimization for efficient semantic segmentation on the edge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.020380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.020380Z digest=sha256:8818922976db107023836f88f5f6e890be85b1ca08620ceb2a162786df4ba3f1

Observation a11a1879-0bbe-4ada-9530-a6752987ee55 · outbound

This paper cites COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement COMAE: COMprehensive Attribute Exploration for Zero-shot Hashing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.025591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.025591Z digest=sha256:f2a44f74340151577825d6f0f91e91305cb579f015e7bce069cb83d7f65d630b

Observation 836a6ad3-4c45-4f51-ad9f-eb9fb609cf32 · outbound

This paper cites Quartdepth: Post-training quantization for real-time depth estimation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Quartdepth: Post-training quantization for real-time depth estimation on the edge

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.030590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.030590Z digest=sha256:c6082564325d2154aa1099990f9c8e8bfce5e6c634de06efa8c53a53a81e555d

Observation 1ca5c2dc-61b6-40b8-8eec-023d89028509 · outbound

This paper cites Fast and memory-efficient video diffusion using streamlined inference.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Fast and memory-efficient video diffusion using streamlined inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.036705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.036705Z digest=sha256:ef45f058676bcb62698292f1caafc702f4aff2fa2e4007af417d9a620e35eb57

Observation da6348e0-ea44-47de-86f4-2c0607a1c85b · outbound

This paper cites Compiler-aware neural architecture search for on- mobile real-time super-resolution.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Compiler-aware neural architecture search for on- mobile real-time super-resolution

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.041450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.041450Z digest=sha256:715a563385f5d9e671c44b47c7735b3d32718d2bb8b4e621e76b80cf655c9da0

Observation 929d0e0c-8283-478e-a49e-002c0d2ba333 · outbound

This paper cites Achieving on-mobile real-time super- resolution with neural architecture and pruning search.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Achieving on-mobile real-time super- resolution with neural architecture and pruning search

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.045575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.045575Z digest=sha256:b4d90eaae5169ca354627813b4ed0e85cbab8ff2a472572cdf493021b37fc719

Observation 7264bf81-3b07-44dd-bf05-e01b7502d36d · outbound

This paper cites Towards real-time segmentation on the edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Towards real-time segmentation on the edge

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.050052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.050052Z digest=sha256:19aa9759fe424768635b1596df138fa890d08c5591b7cc29b31e46dc62e59399

Observation b344b1a0-2f93-4d9c-a25c-9ba8614f0daf · outbound

This paper cites Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pruning-as-Search: Efficient Neural Architecture Search via Channel Pruning and Structural Reparameterization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.054852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.054852Z digest=sha256:bcbf0c692a6e69f497828a34b85c40ba46edd775393aed0b8c5d6827cbab7a74

Observation 08e43dbf-ec4b-4690-940b-9eb3386b4d71 · outbound

This paper cites Exploring token pruning in vision state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring token pruning in vision state space models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.060294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.060294Z digest=sha256:59e3cb3a23f10ae7ecc3aca5a60cfe1e7e8e644ae9a265abb25231aa76d8aa93

Observation dacccc1f-49ac-4d78-afed-0c74e58f5741 · outbound

This paper cites Rethinking token reduction for state space models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Rethinking token reduction for state space models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.064480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.064480Z digest=sha256:c20e6a03134673dd324067db06ff62571f483cecc1a9cd9d8aff3fc6c3767496

Observation b7bc951a-ed9f-4c89-b642-45ebfe351496 · outbound

This paper cites Spvit: Enabling faster vision transformers via latency-aware soft token pruning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Spvit: Enabling faster vision transformers via latency-aware soft token pruning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.068696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.068696Z digest=sha256:fbd2b5a87d36842346cd225ff8626df645010fd172bdcc46453b4c930d54d5d1

Observation ab345c5d-4ba3-4cfe-8ce0-d16d64e0386c · outbound

This paper cites Efficient Reasoning with Hidden Thinking.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient Reasoning with Hidden Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.073656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.073656Z digest=sha256:d2f7b3a3f2852f9fa26531b957d4c563a603a0c2bd4db1396a3a3993aa6062e6

Observation 3abd59c6-13e2-419f-9512-6fc54eaa7876 · outbound

This paper cites Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Enabling Flexible Multi-LLM Integration for Scalable Knowledge Aggregation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-11T20:27:19.658885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T20:27:18.078272Z digest=sha256:305393a3838db42b74040ad003534a9ee1c09f22345be1343a75b3417679f7be

Observation d008724c-d0cd-40ac-82ac-18c98ba5a5d4 · outbound

This paper cites FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.082776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.082776Z digest=sha256:cdc98899bcb90135feee417688501c66e92da0bb55e15470381c50bbd2ad961b

Observation 1ff82d57-1316-4469-8488-c44a09c98adc · outbound

This paper cites Taming Diffusion for Dataset Distillation with High Representativeness.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Taming Diffusion for Dataset Distillation with High Representativeness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.088114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.088114Z digest=sha256:84e2a1b817fd6c1dee583f916ac394636cacf8b381e3373ef3ef2c589c3cd73a

Observation c9f477e5-0d08-4c4b-aa69-b690855049c3 · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba: A Hybrid Transformer-Mamba Language Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.092971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.092971Z digest=sha256:2b61f886be7ea8c463ed22ae5dd16e117175c20372c7acb022595fafa6580542

Observation b7abe382-b2de-44fa-8b08-62d13f66f576 · outbound

This paper cites Jamba-1.5: Hybrid Transformer-Mamba Models at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Jamba-1.5: Hybrid Transformer-Mamba Models at Scale

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.097840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.097840Z digest=sha256:63420461eb28be52ae2b7edd02d5b209cab69cc750019a649182efe09f18c72a

Observation fb18644c-6456-4998-95bb-4e0dba7e4456 · outbound

This paper cites Neural Machine Translation of Rare Words with Subword Units.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Neural Machine Translation of Rare Words with Subword Units

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.103285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.103285Z digest=sha256:be7d5864d908f60cd65eb87b0097343ea0c5f71be023b920eee41a46e71a2c7c

Observation ed8036f4-53ab-4c7b-bc3e-bfabe73f3b56 · outbound

This paper cites tiktoken, 2022.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement tiktoken, 2022

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.107814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.107814Z digest=sha256:1fb392fd89506ba2665914fe7f3fe7740e032b04a6667b5ec97d4c8b94f211e1

Observation 7d2336e0-a065-4959-a9cf-b398bcb0f0f1 · outbound

This paper cites SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SentencePiece: A simple and language independent subword tokenizer and detokenizer for Neural Text Processing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.112010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.112010Z digest=sha256:a839cb4176198d18dc00e9e72d298518d7a5ce3919cb261f32fe8f1ad55d7497

Observation 13db362d-04cd-4c7e-a853-44176c278e53 · outbound

This paper cites XLNet: Generalized Autoregressive Pretraining for Language Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement XLNet: Generalized Autoregressive Pretraining for Language Understanding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.116490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.116490Z digest=sha256:ce39858cc3df3b48d11804735e2d04b867ed79b41f53df4afcaa9a1051a2a15a

Observation b0fe366e-4d9f-48f9-aac8-894e386fa23a · outbound

This paper cites Summary of the tokenizers.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Summary of the tokenizers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.120711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.120711Z digest=sha256:e7af64a384b0d90fc6e29bb443207c7882528e39f587088e01c9b081bac63772

Observation 62274b73-20d3-4c2c-9611-9e441fe72d59 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.126245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.126245Z digest=sha256:9fec99dc13932fa7de33e1a526fd147ee9fd25932c3d40fc45b66e1a66d2c491

Observation 9de82e34-6ec3-4ffe-b736-97a01dcc7c2c · outbound

This paper cites Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.130591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.130591Z digest=sha256:945a5eaaa5de6780efffbed9612330f0da52f0c9dc2c3ff287026c8f6074554d

Observation 2946d674-5f97-4698-b98d-4e23395c60de · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.135247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.135247Z digest=sha256:11cc829b1a137365053feb9706945cf8df11edb741200c975efebe9f122e0e8e

Observation 966e4d78-04db-4a6a-b7e0-c018435e7efd · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.140032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.140032Z digest=sha256:e61dedd19aa6d9ddca1979a7cbf6c17b89a49bac1e4f3dc49e8b1fcb5ac85721

Observation 5797e93d-968f-47a2-b402-e609358f855c · outbound

This paper cites Mixtral of Experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mixtral of Experts

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.144550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.144550Z digest=sha256:993b21dbf9c945a9cdac2b3f5193b9a0a66e561c01319c34ce296c58473182d9

Observation 8dd8301c-3c20-493a-bd46-97f7816b8fd2 · outbound

This paper cites Language Models are Few-Shot Learners.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Language Models are Few-Shot Learners

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.148931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.148931Z digest=sha256:3a6484832478a0178f8cbf29403d589bbc1a19a2f3518c7846f1fe415b571ded

Observation cdb5b6d2-9fcb-493a-b2d7-ed5dc618b428 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.153116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.153116Z digest=sha256:af86ed30da7f0f512773d1dac121e19d1aca3c86d4aa21bdec8d9963830b1627

Observation 19f4fd42-2c78-4445-80a8-eb53c669554f · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.157473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.157473Z digest=sha256:51770f77c70e499d36b88b804b96036cf55b02a119c1d98db4e72d73273eabb2

Observation d74a08a5-22c9-4b86-8325-38d8a0183e85 · outbound

This paper cites CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement CCNet: Extracting High Quality Monolingual Datasets from Web Crawl Data

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.161664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.161664Z digest=sha256:8da4ae88836e4851012dc73bcc9338138455b962c63f5b94c271bdf2511e6635

Observation 35e4e164-1993-4a38-9b69-bff2f948200c · outbound

This paper cites mT5: A massively multilingual pre-trained text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement mT5: A massively multilingual pre-trained text-to-text transformer

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.166288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.166288Z digest=sha256:90457e850010cf58c60b6c9c2fc1bbc605bbff4006ec49ea88338785a50fbc99

Observation 5ed5ce60-564d-46d9-9747-a4faf0895b00 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.170671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.170671Z digest=sha256:bcd7f4f0a3b9e122ff308072b1167b43f56f8d57767a4d5012b77d4498f0b5db

Observation 56e9ca47-cc5f-482b-9cea-4abce246ec51 · outbound

This paper cites Cross-lingual language model pretraining.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Cross-lingual language model pretraining

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.175245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.175245Z digest=sha256:18a2019754ab09a35e38aec59d8fb2ba413f55eb73edc3af772b38b839b68b92

Observation ce2a3189-ff29-4dad-9d65-2d513948b7a8 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Evaluating Large Language Models Trained on Code

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.179395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.179395Z digest=sha256:a742691ba36a4ee8e00c77547d5e3eda21116c4c204d588d7ca2c99b99c6afc1

Observation f14ca8d5-1167-43f2-98e1-b7b94c07df9a · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.183453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.183453Z digest=sha256:a5abe2e408f7be40d124430af4c5bb12290732e664c25205e0a23f4641c5e093

Observation ec3b47f7-da73-41fe-a6ce-030286bc9153 · outbound

This paper cites How to Train Data-Efficient LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement How to Train Data-Efficient LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.187919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.187919Z digest=sha256:6a5bed395204ea8ea0c5be695a5db812837278595dcf52c3d37567cff691dfd9

Observation 8e2b32a2-7a26-473f-8553-26cf3fb73d34 · outbound

This paper cites A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.192841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.192841Z digest=sha256:1f8450e9ec74f86b532da44399fe2ab9191b2cd383351329f36cf24324c012f4

Observation 752f4963-cd4a-4c94-9156-ca9497cac1a6 · outbound

This paper cites Glam: Efficient scaling of language models with mixture-of-experts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Glam: Efficient scaling of language models with mixture-of-experts

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.197752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.197752Z digest=sha256:2fb700895be119292246ffb7ad797d93e8026a00c48610331cfaa54235e80246

Observation 02b4b7f7-1d41-430d-a3ad-41511d2bedf1 · outbound

This paper cites Deduplicating Training Data Makes Language Models Better.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Deduplicating Training Data Makes Language Models Better

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.202214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.202214Z digest=sha256:5f50574c1f1ef14fe3b8eba18b27ada4f26cbc48a187aa43c5800358493feb08

Observation c0174190-2c29-4481-aaf5-c239ae7031a1 · outbound

This paper cites Url normal- ization for de-duplication of web pages.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Url normal- ization for de-duplication of web pages

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.207533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.207533Z digest=sha256:e49fdf2b10dba51da27f9d106068e82d61895f73e40b0d5c773a088485f65b25

Observation 93160288-adc3-4018-b45a-9c52bc69cee3 · outbound

This paper cites DataComp-LM: In search of the next generation of training sets for language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DataComp-LM: In search of the next generation of training sets for language models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.212836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.212836Z digest=sha256:fc090f69cbf5ad30b70082deefff3020f00f38a1a51f137864c203f8c8c77d75

Observation f2dd4150-d875-4f87-a0d3-53cf7d538c23 · outbound

This paper cites Efficient online data mixing for language model pre-training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Efficient online data mixing for language model pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.217617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.217617Z digest=sha256:b6d49f444d78429fab351319c25dc3cd5a13c5bbbdae2497854dedaba531ea91

Observation 2fdf681f-5d41-4609-807f-a96e95c0b2c1 · outbound

This paper cites SlimPajama-DC: Understanding Data Combinations for LLM Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama-DC: Understanding Data Combinations for LLM Training

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.221452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.221452Z digest=sha256:2d049b90613c34b23312de426db6935ec23697ade2f6c2a181b4d63f0de47f31

Observation bf06d7aa-c143-4e16-8ff2-bd72429e5d1c · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemma: Open Models Based on Gemini Research and Technology

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.225753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.225753Z digest=sha256:19677b4d5ea7437c9820a099f95440bd8551f82ac1ea018edb2b04cd20ac5ec0

Observation bb1cfb66-2300-4402-a56b-9d649ed14911 · outbound

This paper cites Palm: Scaling language modeling with pathways.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Palm: Scaling language modeling with pathways

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.230402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.230402Z digest=sha256:b7e68e704a800e0091e5df0b8779fafdb6e0895bb23c2dd8006d23c79903bbe4

Observation 37066c23-d6cc-420f-a195-ed135180da2a · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.234748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.234748Z digest=sha256:e848b93c4cc10ecb9868a02b71b3fb4b52b005b00ad5758bb8119598c4afb7ef

Observation 06cd4c28-0fa0-4212-bf85-5c653126d384 · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.243855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.243855Z digest=sha256:1ad1579a908b4a6acd8d6a123413868ec74cb626aa2c1f22a60305461e27f131

Observation 0620b97a-ff06-4107-9b0f-07d1517fe1eb · outbound

This paper cites Llm-datasets: An open framework for pretraining datasets of large language models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llm-datasets: An open framework for pretraining datasets of large language models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.249188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.249188Z digest=sha256:dbc0b140f7512b8692458cb536f1258963492d17b032e1ff8309cbd4b5ea8171

Observation de883fd8-b005-422b-8a3f-4e9b7727f7cc · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement StarCoder 2 and The Stack v2: The Next Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.254010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.254010Z digest=sha256:5cd24ab0942478772aa385f9dd1bf7f91b60777dc235d77eee86cc4cd715f037

Observation dff75c77-9380-4d3a-837e-7652698887c5 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.258465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.258465Z digest=sha256:f74abbbe77461eea46602b1ea87118c1401b9cdfcae1d3697e279b87cd1b77ec

Observation 2f26dc05-8c6c-4aa6-a2fb-f4eb91b8330f · outbound

This paper cites Infinity Instruct.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Infinity Instruct

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.262501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.262501Z digest=sha256:39b5fa909820e0a1b2eac053258b184034e9cc447d450410e425dc373e34e40f

Observation fa758c31-56f6-4f35-ac65-29d3e288cb1d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.266679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.266679Z digest=sha256:6c034454e1b9e6d0a67019f96990cdfef416e56ab66b0ef8c339befc2f1a1bc7

Observation 023814d8-389d-448b-84ee-33afc67d6a77 · outbound

This paper cites Open Thoughts.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Open Thoughts

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.271271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.271271Z digest=sha256:b5f392c34132846b62e4f748954a2517984a38c72a550a9fe7db9b2d1fbdbe0e

Observation 28bba9e2-7f7e-4f33-9e1f-ffaae0274b79 · outbound

This paper cites OpenR1-Math-220k.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OpenR1-Math-220k

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.275117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.275117Z digest=sha256:ce1518f209989312521eacbc455eabaa2915c10a03e61c2a4b698103f7123fae

Observation 3ebb2631-8eb9-4d77-b7c2-40d75dfa6765 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Tianjun Zhang, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.279135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.279135Z digest=sha256:5c2c5cd9bb030f63d306052679c154ab31985b1f68432bfa36667ff620cbb2f4

Observation bc22fbdc-ae80-4eda-905d-b451cd60a763 · outbound

This paper cites Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.283265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.283265Z digest=sha256:c05763db1933e199da32e05a2ba5210f72e459ffd1a876f717f1ffdbbeec4cda

Observation 23c34f15-0efb-47fb-978c-a4d69d38317a · outbound

This paper cites ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.287868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.287868Z digest=sha256:3f318d3720429b6c7f175dde670bfdfe2679e3cbdc9cc3320dc626b9a589435c

Observation 3d4264eb-14bd-4476-a0b6-568706d1e07b · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2024.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Gemini: A family of highly capable multimodal models, 2024

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.292130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.292130Z digest=sha256:8d30daf7bcd922bdbc6f7d3704dc66416ec74271c7f56864e2e4b213cb122b4b

Observation 1000de85-5982-4acc-b44d-f827120fe87d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.295971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.295971Z digest=sha256:fc9f0c5bb18543632262f617e6864e616cbb6d64b67c4a0277dbea689596aabb

Observation d1079219-1573-4a92-b259-054ff5840111 · outbound

This paper cites DeepSeek-V3 Technical Report.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement DeepSeek-V3 Technical Report

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.301141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.301141Z digest=sha256:549fe7db55aac3153cbd9eccec474a442354b70c4a774d654a826ae5b4aee936

Observation 393e8302-5398-4bde-9bba-66ca674bfbcd · outbound

This paper cites Baichuan 2: Open Large-scale Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Baichuan 2: Open Large-scale Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.305532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.305532Z digest=sha256:d9eb3aaff74ae5b958b1e4741b3ce5e59f220acf005fd25fc23f3d6180063c80

Observation 77a95bc1-26f3-4888-8562-c411cbbd0479 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.310711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.310711Z digest=sha256:c9e37634fc8ce8f0268352b31e2e6862cefe1731797c3bca0c582e58da486366

Observation b7ef42a6-fbdf-46e0-876e-e209d2d9e310 · outbound

This paper cites Pythia: A suite for analyzing large language models across training and scaling.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Pythia: A suite for analyzing large language models across training and scaling

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.316066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.316066Z digest=sha256:60e572b909bf24afa9aa16408e3860f7130f12bc6a0482e4d302e2e7b0fd13d9

Observation 008db99f-adc3-4d05-a74f-15144c4ff43f · outbound

This paper cites GPT-NeoX-20B: An Open-Source Autoregressive Language Model.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GPT-NeoX-20B: An Open-Source Autoregressive Language Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.321064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.321064Z digest=sha256:a7c71527917e33b4cd0347c3d3acbcbe11d02792a1111f3cc31ff4c5f955d43c

Observation f1c78cca-ad87-4ed1-8a20-8e592feeaa80 · outbound

This paper cites OLMo: Accelerating the Science of Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement OLMo: Accelerating the Science of Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.325609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.325609Z digest=sha256:946ab19e89062769a345e6e2c93058631c34a93cc9b24b468695471e9d29b5b1

Observation 716216dd-7521-4f6d-b796-b0750c31e976 · outbound

This paper cites LLM360: Towards Fully Transparent Open-Source LLMs.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement LLM360: Towards Fully Transparent Open-Source LLMs

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.330173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.330173Z digest=sha256:99e99f75bb4a0b62a27315717fee24f3dc15123c27c8f634d8023509d9c21bba

Observation 935359d0-c036-455e-8821-cb4cd580a93e · outbound

This paper cites Code Llama: Open Foundation Models for Code.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Code Llama: Open Foundation Models for Code

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.334503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.334503Z digest=sha256:c45ee3f9a83e6fc70f27b2dcf4186e6e81a54d57d1e6fb33ae9e213cc645ef9d

Observation 68d82c91-e340-4f6b-897e-961caa53a6da · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.338879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.338879Z digest=sha256:82178655929920e0812b8d1c0d474e1771503d53335675ba92f89e47c02a7efd

Observation 4718b4d8-001f-4af4-969e-3465b1320517 · outbound

This paper cites Longformer: The Long-Document Transformer.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Longformer: The Long-Document Transformer

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.342788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.342788Z digest=sha256:0a37e622dbb7283d2bd44707ad29ab0f44864d9febba11d579c3d1fa712a627b

Observation 7cebbe5d-20f3-4220-970e-e17274678b59 · outbound

This paper cites SlimPajama: A 627B token cleaned and deduplicated version of RedPajama.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SlimPajama: A 627B token cleaned and deduplicated version of RedPajama

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.347059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.347059Z digest=sha256:3fadbdda3c011140094b5a844b68dab686d35ecf65df5bd593a9a471732cca48

Observation c2d8eb90-4199-4e43-8dfc-40bd293a72de · outbound

This paper cites RedPajama: an Open Dataset for Training Large Language Models.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement RedPajama: an Open Dataset for Training Large Language Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.350792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.350792Z digest=sha256:fd0c13dc1129ecc2e0db09ce8a45c75ca6e865de318ad779574fc12ffd74dfbb

Observation 8bdf83a3-2eb1-41cb-9444-fbdf07e88c10 · outbound

This paper cites SemDeDup: Data-efficient learning at web-scale through semantic deduplication.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement SemDeDup: Data-efficient learning at web-scale through semantic deduplication

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.355433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.355433Z digest=sha256:9dbd9ebbbacbf9023d9c6ebf7a3bedc625fdfd779170777f23d7e8bf890bd34c

Observation 311a437a-1c19-4b49-b689-4fbcbc90f178 · outbound

This paper cites an unresolved cited work.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.359443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.359443Z digest=sha256:6d2af4be7b4ca95468d413eb9ad5a49ce90db20fa3d993b41b6ffee4d3a4a083

Observation f2363aff-4900-4b7d-b92d-26ea6bcf49b0 · outbound

This paper cites The Curious Case of Neural Text Degeneration.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement The Curious Case of Neural Text Degeneration

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.363229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.363229Z digest=sha256:123f63779be79c5644968243ebe5ed02f1057b4bd602cc38db3a59965e17ebd5

Observation 35079eac-4071-48cf-93c1-6f310ea1a06c · outbound

This paper cites Mining of massive datasets, cambridge university press, cambridge, 2014.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Mining of massive datasets, cambridge university press, cambridge, 2014

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.367268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.367268Z digest=sha256:2d750ffb3ed6e75eb2d67001df25688cb1a46c46888814b24594c15c6ddaf12e

Observation 9a76da83-df64-4bd0-8ebb-1f1acb6abffb · outbound

This paper cites Introduction to common crawl datasets.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement Introduction to common crawl datasets

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.371325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.371325Z digest=sha256:03886d361b16702115fa4cb7cb57ebdc5c4ef42be75831efe0eb7361d8a4b6f3

Observation e0125241-7206-4192-a342-6b63db95f694 · outbound

This paper cites On the resemblance and containment of documents.

7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement On the resemblance and containment of documents

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T20:27:18.375394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:27:18.375394Z digest=sha256:c57334d64f596aa16be1e5529521b29f44642302e044d3560c18b6e08bb21c94

Pith citing papers

Observation 6e2ddd1e-3c4d-49ab-bc90-28ed41baa203 · inbound

Human Cognition in Machines: A Unified Perspective of World Models cites this paper.

Human Cognition in Machines: A Unified Perspective of World Models 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:12:26.076291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T08:12:15.663761Z digest=sha256:60257817e3acc2873976dd3944274336f859b3befc06bcfa30d4c747b11aec57

Observation bde30324-447d-45fa-b538-054dd951d0a9 · inbound

PhyWorld: Physics-Faithful World Model for Video Generation cites this paper.

PhyWorld: Physics-Faithful World Model for Video Generation 7B Fully Open Source Moxin-LLM/VLM -- From Pretraining to GRPO-based Reinforcement Learning Enhancement

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:33:07.545335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-20T07:28:20.248452Z digest=sha256:12378791fdc72bd924550a3b5531772f2f4c11230ba436cea87d7ae71069bca6