Pith. sign in

Paper Citation Record · LEDGER

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

As of 20 August 2026, this Paper Citation Record lists 100 of 123 outbound references and 18 inbound Pith citation observations for arXiv:2507.22448.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22448 v1

Coverage vector

measured 100 of 123 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:44:05.162957Z

measured 118 of 118 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:46:37.290952Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T01:46:26.624174Z

Reference resolution

100 of 123 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved98
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7f81a9a1-1676-4e54-901d-d3c7eee9c42a · outbound

This paper cites write newline.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.541501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.541501Z digest=sha256:485e0f4e82d637dbe3680d3180471be4894287c58c2298264ef22710804c2534

Observation 787993db-06c3-4a91-9758-6f50e2438e57 · outbound

This paper cites https://www.statmt.org/europarl/.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance https://www.statmt.org/europarl/

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.606975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.606975Z digest=sha256:d20de109604340a69902962478ea8c0fbc3502679d2d625ffebc55773911cd40

Observation 8a3a9768-82db-4bd1-9349-8040f2221bfc · outbound

This paper cites https://www.gutenberg.org/.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance https://www.gutenberg.org/

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.677129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.677129Z digest=sha256:17b64b351871b430a7170ea5651d93962b947bb9ecc4d9d54a6437891abdf9fd

Observation 07803083-faf7-4896-9dcc-a03b8ae48083 · outbound

This paper cites https://github.com/zeerakahmed/makhzan/.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance https://github.com/zeerakahmed/makhzan/

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.793814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.793814Z digest=sha256:950daa1915337f609e49525b6181ac865b2d14e607ab962abcfc9241f71e6427

Observation ea756849-0f6d-4db2-946f-049bbe8e8fdb · outbound

This paper cites AIME problems and solutions, 2025.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance AIME problems and solutions, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.822822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.822822Z digest=sha256:693cda869fa1ed71bfaf451bffe2cbd403c034868adce2dfc704a2a70fd16902

Observation bd3c80af-0051-43b2-bec3-13531e1e6a4b · outbound

This paper cites GQA : Training generalized multi-query transformer models from multi-head checkpoints.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance GQA : Training generalized multi-query transformer models from multi-head checkpoints

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.831020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.831020Z digest=sha256:e2a01b8b841bc8512a646506d77fb7c07112acf9bc07ee713144a65074bce186

Observation 06b68635-1547-49b0-8fe4-e013a0df1ce5 · outbound

This paper cites M ath QA : Towards interpretable math word problem solving with operation-based formalisms.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance M ath QA : Towards interpretable math word problem solving with operation-based formalisms

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.835598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.835598Z digest=sha256:b518d3ee0cada80fb647a61a7446d95558becc61ad98e49a3b6babdc9d68bb77

Observation 45981a6b-3ee7-4475-b4df-02ec2a67a340 · outbound

This paper cites Program Synthesis with Large Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Program Synthesis with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.839405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.839405Z digest=sha256:566e933fafb017d35b16b9d2c2b1548cd62d8179ea944c3f1aa505a632a39ad5

Observation 8e3c9258-662a-477c-938a-ab173b749ebe · outbound

This paper cites Llemma: An Open Language Model For Mathematics.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Llemma: An Open Language Model For Mathematics

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.843213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.843213Z digest=sha256:c03745d31ee39fc2bc40f3037929ec2f06729415c9977012e3961a43f5b0dd05

Observation d4ecc412-daae-4138-8cbb-188cac617ada · outbound

This paper cites MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.847321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.847321Z digest=sha256:5b0bd261088210a25538f8a16a125026d1e63ec35c0e8c1019e9d7d219166cb6

Observation 923fb2ac-2821-4365-bf8c-488e325ad290 · outbound

This paper cites Titans: Learning to Memorize at Test Time.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Titans: Learning to Memorize at Test Time

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.851523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.851523Z digest=sha256:ebe0ff3e1fa095723e181de1a1f63f89dc817ce34d5bf9373ca2a98327f7efb4

Observation a309cf25-b72f-439d-b055-ff719114e253 · outbound

This paper cites Smollm-corpus, 7 2024.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Smollm-corpus, 7 2024

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.855073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.855073Z digest=sha256:6d2d4d93f3a33877330921be22b8c0bdf7f8a73cc98f6f0e12fba8aaacaa725d

Observation f873f7a5-efc8-482e-9fb3-6c593a689fe7 · outbound

This paper cites Scaling optimal LR across token horizons.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Scaling optimal LR across token horizons

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.858691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.858691Z digest=sha256:e93b3ef060ec426e48251c9dea3a0faca7ca7e0e0aed94ca397cc92dedd13e9b

Observation 04e9b964-08d9-4122-8b3f-c58fab1b748f · outbound

This paper cites Ntk-aware scaled rope allows llama models to have extended (8k+) context size without any fine-tuning and minimal perplexity degradation, 2023 a.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Ntk-aware scaled rope allows llama models to have extended (8k+) context size without any fine-tuning and minimal perplexity degradation, 2023 a

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.862301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.862301Z digest=sha256:5dbf6f5fef6a219740154e5dd392815f1bd32724f6847ec4e2ec53903aa95079

Observation 04828cd5-b657-4183-b313-66031bb474b5 · outbound

This paper cites by parts.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance by parts

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.865621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.865621Z digest=sha256:8db4bc34a6a2fec3e420d16cc8bc6734626e6d36ee70c383cf918f1b2983e3a9

Observation 5b43db33-faff-4cf6-95fc-94e02262e82f · outbound

This paper cites Depthwise Hyperparameter Transfer in Residual Networks: Dynamics and Scaling Limit.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Depthwise Hyperparameter Transfer in Residual Networks: Dynamics and Scaling Limit

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.869255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.869255Z digest=sha256:b96ec29c40b4c29d6a545a7dcc06e620d1730a4e46bcf5728b35832846d0a84e

Observation 2a62b25e-3a83-44fe-9c12-9590fa7da06e · outbound

This paper cites On the resemblance and containment of documents.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance On the resemblance and containment of documents

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.872685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.872685Z digest=sha256:5a99fc20143d0d796bd9ab206adb7eb4634a050ca8ef11d3777d05d807f8aa8f

Observation 2270f3ec-43c5-44bb-b78c-17532c6ab9f2 · outbound

This paper cites An investigation of incorporating mamba for speech enhancement.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance An investigation of incorporating mamba for speech enhancement

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.875843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.875843Z digest=sha256:7065f1bb23bdf672c0ba87c44c9d2d9286a0e7c8436efa962bc0da150d654cfd

Observation fea037c4-e0f2-44f6-83db-7d76371b6564 · outbound

This paper cites Theoretical limitations of multi-layer Transformer.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Theoretical limitations of multi-layer Transformer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.879218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.879218Z digest=sha256:47e9c7a20aee570431179a22ece7c5b8d23a1103036cbd9448f103a76a8bb935

Observation 7b7cd417-8c87-453a-a002-534e7b9890dd · outbound

This paper cites an unresolved cited work.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.883072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.883072Z digest=sha256:590bf32757454fd534c916d3de489fffa9c370ca86e1fe2952d9b7db13e838fa

Observation d9aa81a8-71d8-45f0-9c3e-4771ccd6b90f · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Extending Context Window of Large Language Models via Positional Interpolation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.886379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.886379Z digest=sha256:e793c8e5c639fded1c4f5fd717244abc32c6feef74869f42a69697e3a0608f21

Observation 82a5db4f-5956-4d28-96fa-8824cf701442 · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance PaLM: Scaling Language Modeling with Pathways

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.890348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.890348Z digest=sha256:d37b74b33c437d3c56b591400961608f216c14f728c13a0c2f08a2bf66378e7b

Observation c5229a37-a067-42cd-97d4-824e9c69e947 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.894015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.894015Z digest=sha256:717d8c4f0f9cc87bd1d76c5e7db6abf49ce664e5d0402b9ec33c10415ee1fef4

Observation bcc88519-03eb-4b46-9d09-b2208b9b5970 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Training Verifiers to Solve Math Word Problems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.901334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.901334Z digest=sha256:0185b83a8737a65e7e34a9c54f7e39f1a20aa1cd879c866a702125e70e289c49

Observation 1e6d29a9-05b3-485b-b806-b2725688dee6 · outbound

This paper cites Okapi: Instruction-tuned large language models in multiple languages with reinforcement learning from human feedback.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Okapi: Instruction-tuned large language models in multiple languages with reinforcement learning from human feedback

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.905202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.905202Z digest=sha256:37aa9b311cd3783e56d70bfa1683003f11271a108c16ed8c985958fe6b1c889b

Observation 1bd4b413-912d-420b-8d04-8d07217651d4 · outbound

This paper cites Unsloth, 2023.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Unsloth, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.908379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.908379Z digest=sha256:48430f0826e80935c0ab027b8fff91917d8c4bb9ebec00cad5f2fce6f2c933e9

Observation f00624e6-5037-4f1b-86ad-17bdf18035b1 · outbound

This paper cites we also choose similar dimensions as modern Transformers, e.g. P =64 or P =128.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance we also choose similar dimensions as modern Transformers, e.g. P =64 or P =128

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.911411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.911411Z digest=sha256:3c43a8a965e9d5f1a39ea058dd4e9141d369df24410e6da864dba5565a6cc001

Observation 06e304bb-56d3-4f34-97ed-cbafdebd8b88 · outbound

This paper cites Flash A ttention-2: Faster attention with better parallelism and work partitioning.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Flash A ttention-2: Faster attention with better parallelism and work partitioning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.915204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.915204Z digest=sha256:1ae434384fc375d38fb657d5ea994c37fbc86e6964c95eca4241820a1c258509

Observation d229889f-f8dd-4734-92b5-44010efa1f21 · outbound

This paper cites Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Transformers are SSMs: Generalized Models and Efficient Algorithms Through Structured State Space Duality

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.918746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.918746Z digest=sha256:009a1626b59f34f6494758895600c76021ce277d444dbd1ae34f9586cffa42e4

Observation 157ac2de-34fe-40db-a501-72fb3e3c2c39 · outbound

This paper cites causal-conv1d: Causal depthwise conv1d in cuda with a pytorch interface.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance causal-conv1d: Causal depthwise conv1d in cuda with a pytorch interface

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.922246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.922246Z digest=sha256:80e971e8e07504de1ad60871066cb909b80cce9ce19cf21af19fa58a84539652

Observation 1391d620-ee6f-4a0c-9b19-b0c38e309ee1 · outbound

This paper cites Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.925289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.925289Z digest=sha256:d0c996f4d9fb090be88f7e1b5d524509fa634342d9a140f4607d5578c3c47abc

Observation 4c52a78a-a6e0-4ad8-a59b-a3e36ad36ad0 · outbound

This paper cites Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Cerebras-GPT: Open Compute-Optimal Language Models Trained on the Cerebras Wafer-Scale Cluster

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.928577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.928577Z digest=sha256:53da8d62d53431c375f026874d13adb507d593797e0a585f2c5655effc871c18

Observation b4efaa87-1a42-4004-ae5e-3b668982562a · outbound

This paper cites Don't be lazy: Completep enables compute-efficient deep transformers, 2025.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Don't be lazy: Completep enables compute-efficient deep transformers, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.932191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.932191Z digest=sha256:1e5e8f3c4e57a42d9fa7e3d7a358bc7f632e560872c8a2b7741e1d981f1a5189

Observation 4e4c4713-4027-417f-aa9d-b53bbadf77ac · outbound

This paper cites Hymba: A Hybrid-head Architecture for Small Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Hymba: A Hybrid-head Architecture for Small Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.935445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.935445Z digest=sha256:559ecc7242e2c3c67d99a282be77d88002bdc2f253959c62ad72d4035a0e26f8

Observation 38922801-4411-43c5-927e-eb7106fd46da · outbound

This paper cites The language model evaluation harness, 07 2024.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance The language model evaluation harness, 07 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.938969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.938969Z digest=sha256:5679c6fc7e2a60d25a1872e483924ecebde76bd319d04c7466a08466d6143617

Observation 9c271b76-004d-4fa5-af46-b1be1eb89733 · outbound

This paper cites A curated research corpus for agricultural advisory ai applications, 2024.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance A curated research corpus for agricultural advisory ai applications, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.942449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.942449Z digest=sha256:aec8739dc8bb1abe4d4103b3578017203ca6a97e09271f31f391f1a773c603be

Observation 4bf17fc8-973d-4244-92be-3a9a39845fa9 · outbound

This paper cites Zamba: A Compact 7B SSM Hybrid Model.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Zamba: A Compact 7B SSM Hybrid Model

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.945779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.945779Z digest=sha256:18fadb58077cc775c83e810bf31b10a191f87046da757e05e27ce41543c970cf

Observation dd845671-dcd8-46fd-b16e-c3a30e1ddd0d · outbound

This paper cites The Llama 3 Herd of Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance The Llama 3 Herd of Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.949357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.949357Z digest=sha256:33911c0c6a7983bd15428faba1de97db3fe9cab58e148c98745372d6faf5431e

Observation e2ef3f7b-c574-41f8-8831-2d2905325626 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.952490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.952490Z digest=sha256:bc13c697caffd91987118486651d19726b37ea075cf7e8103bad36db21352e60

Observation ee6fefc1-2714-422e-8224-7b75061dca0a · outbound

This paper cites CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.955924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.955924Z digest=sha256:20ba7c6a06648fd8df1166f4f2b2fca5a039433dd583d02ff44bbab0a4ce5f70

Observation 127ac7d5-5f02-4097-9a38-f7504bfa98b5 · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.959140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.959140Z digest=sha256:5ec5f16009e56102820d9824fedacf6219c471761dd4de6b3c9808e71a0aefae

Observation d58e7aae-90d8-4b57-8062-6f5f412e00bc · outbound

This paper cites InfiMM-WebMath-40B: Advancing Multimodal Pre-Training for Enhanced Mathematical Reasoning.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance InfiMM-WebMath-40B: Advancing Multimodal Pre-Training for Enhanced Mathematical Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.962316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.962316Z digest=sha256:c76a36674a97eef9b5110ae387438f5190943f1742cf03d0f95c880684d08a20

Observation ed0ee3a1-2d80-46c4-bc31-51183a07b4ef · outbound

This paper cites Simplifying transformer blocks.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Simplifying transformer blocks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.965689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.965689Z digest=sha256:20ae0bc4341d24ff0d2ccb5f486bb3a933f92aff2438773c7e8384b414abbe1a

Observation 6b820cea-be7f-49d1-a52d-49bc826f2736 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Measuring Massive Multitask Language Understanding

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.968529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.968529Z digest=sha256:12d725f16353b29500d65c726d03e2eff4f41b4a069dc237aacca1b0d1cbd6ac

Observation 1878db44-e39a-4c0f-8f6a-b08417fa6a1b · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Measuring Mathematical Problem Solving With the MATH Dataset

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.975054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.975054Z digest=sha256:6490611050f8cfdf8a4e09a4420aa2f84dcd443187773d0a6d446d5418c8f49a

Observation 4a6864a5-2df0-47a7-ae38-1afc831a6766 · outbound

This paper cites Clarification on how to interpret kernel size for conv1d (\#523).

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Clarification on how to interpret kernel size for conv1d (\#523)

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.978039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.978039Z digest=sha256:ebcafc29ef9cc693f1759e2f5b32690c2bb00e8e49cbe31648fa119a09a5bcd5

Observation f805150e-b425-4e6a-8802-56ddae5ce0c8 · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.981140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.981140Z digest=sha256:a758cc680be19f692b19171e200784db3792c663aff831eb209174f79b98f18b

Observation 0f838f95-5254-4696-9768-1be08b293133 · outbound

This paper cites YuLan-Mini: An Open Data-efficient Language Model.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance YuLan-Mini: An Open Data-efficient Language Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.984553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.984553Z digest=sha256:78202dd851eca8906f9436cf6efc997b7c623831132491e4ae4451ca1cb8d175

Observation 53049af4-bf01-4546-936b-d8c846683af7 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.991100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.991100Z digest=sha256:8714d6306364ccc347028a4f44f884dd167034766c662fe77fc777255f6c31e8

Observation eb15ca1f-8ad2-468d-8ac3-0005d335c76d · outbound

This paper cites Qwen2.5-Coder Technical Report.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Qwen2.5-Coder Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.994063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.994063Z digest=sha256:46da1de8b863219dad7ab199a755753bbe3845cfe9ea259f5bcc495da347bedd

Observation bc91d04e-28eb-4365-be2d-33df4ebe6f42 · outbound

This paper cites Teknium".

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Teknium"

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:04.998003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:04.998003Z digest=sha256:9ab6849725bed80f646343b708c834c9a1b2d8e1b4231997fc7422da1c59ef55

Observation d8989f89-5508-4755-ae7f-836d9061607a · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.001113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.001113Z digest=sha256:cb2e81e22fffa914fedf86e55d1e4658c9c407b83a278759d4bb9aff8171b05e

Observation 6780996f-4641-4c02-9c4d-7db16feb3695 · outbound

This paper cites Mistral 7B.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Mistral 7B

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.004353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.004353Z digest=sha256:f65b6be236e822e5c81eba57f23e79836fb855c7709d858fb83f40d2fe9efddc

Observation a94679ca-e4e1-4ee7-9121-f3b037feeb46 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Gonzalez, Hao Zhang, and Ion Stoica

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.007650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.007650Z digest=sha256:ff76dd496664719beefae394c8a79bb80c2202e654cb91cee2cdaba24d6bb486

Observation d60ab944-798a-43bc-a94f-7113a955211a · outbound

This paper cites Math-Verify: Math Verification Library.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Math-Verify: Math Verification Library

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.010727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.010727Z digest=sha256:068083c0350523269fe2bcdfd4f92ac148d8e529357eafa7a7a9ada4a8f250af

Observation f912cf6d-c2b4-4759-8e22-6ebfd0f9a369 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.013872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.013872Z digest=sha256:071f15576755baf0fdf66b83582866de4ea4c93bef5b4fe41e1a4f18be3cb11e

Observation ea05661e-d601-4855-8b72-48d053029e9c · outbound

This paper cites Hashimoto.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Hashimoto

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.017009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.017009Z digest=sha256:bb36c1786a85a799f82004c927d9cb9d9d866a8f2dacd562a104ca5a01149284

Observation 49962145-2704-4fcb-b1ef-4c39b15acc6b · outbound

This paper cites Jamba: A Hybrid Transformer-Mamba Language Model.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Jamba: A Hybrid Transformer-Mamba Language Model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.020186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.020186Z digest=sha256:d04c32e6168184a98bbd7ee4d0d37eed385e8ee172b1930529f5c4a4970d366b

Observation 01664cc4-cee7-4a14-b022-1c5cbf4ad88d · outbound

This paper cites Let's Verify Step by Step.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Let's Verify Step by Step

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.023784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.023784Z digest=sha256:814ceac181003a6f12ceefc21e7ba368ad0aed33c8a76f121fb62a31f164093e

Observation 0934111e-4d85-49e3-b35b-19a2fbaa5ee2 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.026994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.026994Z digest=sha256:0d0413098b2fb54a62b8b03ba5f400531dc043ff11fc9f57033b2a70d9a20d5b

Observation a9d9cc21-269b-4277-b869-987e90b27225 · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.030425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.030425Z digest=sha256:cf8a8ddfb1d651a2e864b675b1e6509f5bbeed29767875870107366401b16484

Observation 79f08128-3994-4a5e-9648-a943ce076329 · outbound

This paper cites DeepSeek-V3 Technical Report.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance DeepSeek-V3 Technical Report

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.033930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.033930Z digest=sha256:09c4ecca2097bc3332b91c55b1295123d9cff18b9ff43cd7856c3d4137a9a52a

Observation d56803e7-77f7-436f-9dee-bbdfa96710ff · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.037477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.037477Z digest=sha256:d02e51b076c6cbf0ac3f5210566432d52f2092058d5ce8f9ea8b710a458a44a3

Observation cff6b3fb-6a55-4bcd-94d8-8c1b63903a09 · outbound

This paper cites Is your code generated by chat GPT really correct? rigorous evaluation of large language models for code generation.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Is your code generated by chat GPT really correct? rigorous evaluation of large language models for code generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.040804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.040804Z digest=sha256:7abe0b6dd3d0898d03b8cefe145bcf93813558ae397e8996109ea36aa681a898

Observation 0b2cedc9-0998-4925-83ba-95eb4ff1793a · outbound

This paper cites FineMath: A Fine-Grained Mathematical Evaluation Benchmark for Chinese Large Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance FineMath: A Fine-Grained Mathematical Evaluation Benchmark for Chinese Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.044254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.044254Z digest=sha256:6864234b398456abfc0f43bfe6c6168bc5cab07b1ed904f7d7a227e5d79b6ded

Observation 45e7d6d2-774f-4515-9039-37326b740c80 · outbound

This paper cites VMamba: Visual State Space Model.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance VMamba: Visual State Space Model

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.048052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.048052Z digest=sha256:9d44b39e20171854d8b64200a8d544319d0bb96cc38ead84330f7c9d992993cc

Observation db1ac445-930a-4684-bbd8-88b298153cf6 · outbound

This paper cites LLM360: Towards Fully Transparent Open-Source LLMs.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance LLM360: Towards Fully Transparent Open-Source LLMs

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.051705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.051705Z digest=sha256:1d9ccb5b28dccdaf136df508911fae0926ea68337389bceaa8acb72b5d387626

Observation 2831f5a5-a1e4-45f5-803c-555773a4587d · outbound

This paper cites AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.055041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.055041Z digest=sha256:3a274bfb118528b3200ec605557fd0d1cffa2dd5c0f1a6b8205961936a1107fd

Observation b88f9cc7-7416-4d8d-b8e8-91e37a2868d3 · outbound

This paper cites Neural Thermodynamic Laws for Large Language Model Training.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Neural Thermodynamic Laws for Large Language Model Training

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.058346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.058346Z digest=sha256:ffba5bd55e99358a2926ece5cf1cfa6a3d40cdf20151ef7215a031fa806535c1

Observation 40fde3dc-73f3-4bf5-9b93-b0b98a2a431c · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance StarCoder 2 and The Stack v2: The Next Generation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.061741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.061741Z digest=sha256:533b1ad2a8af06e00951e339524bc6f978079813012ee3f58d020f961ecfdc68

Observation 9cebfcdc-1343-469b-9dbd-f933aebb1960 · outbound

This paper cites Finefineweb: A comprehensive study on fine-grained domain web corpus, December 2024.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Finefineweb: A comprehensive study on fine-grained domain web corpus, December 2024

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.064835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.064835Z digest=sha256:3353debcb3c37532ade2e2bb534dea5b45f57fb1c22d64d6bda41f67d8dc543c

Observation dc8be85c-b023-4bc2-a14e-d76ee16e124c · outbound

This paper cites Falcon2-11B Technical Report.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Falcon2-11B Technical Report

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.068137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.068137Z digest=sha256:32c78a11b8d95599f52a468906de06f55157b047c6309e27b70eb331c2f770f7

Observation 69aa7760-f631-4154-b6de-07e7b3a8a88f · outbound

This paper cites On the SDE s and scaling rules for adaptive gradient algorithms.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance On the SDE s and scaling rules for adaptive gradient algorithms

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.071691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.071691Z digest=sha256:0a641d02e8aa1cfd524d6cf218f7c0b8d78b56aad8f90c9f9bfdd1fd32aa3d66

Observation 8c31e340-b16d-4dd4-b068-15b4ee7a6e5d · outbound

This paper cites Peft: State-of-the-art parameter-efficient fine-tuning methods.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Peft: State-of-the-art parameter-efficient fine-tuning methods

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.074866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.074866Z digest=sha256:5ee2a497d1690710b1beed2926a2335baf1f1463f6d1d4673849d5ae542073fe

Observation ed2b363a-ac4b-4071-95f6-97eae5fb1bbc · outbound

This paper cites Characterizing state space model (ssm) and ssm-transformer hybrid language model performance with long context length.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Characterizing state space model (ssm) and ssm-transformer hybrid language model performance with long context length

Reference 78

Resolution
verified exact
doi, observed 2026-08-06T11:44:05.348325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T11:44:05.078266Z digest=sha256:d318f6402a777ed1103cf77bec36b7c21d27001326400b2f03ec0bcc55c5ff68

Observation a30eb55a-7539-45c5-be7e-bccbcc2c6261 · outbound

This paper cites Gptqmodel.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Gptqmodel

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.081530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.081530Z digest=sha256:7a54b21674b0278dddd99472d85a5cf629b6cd0546fde1d2c9c3d2e4f055576a

Observation fb7f6932-7122-4815-8adb-464c7b092de9 · outbound

This paper cites Oumi: an Open, End-to-end Platform for Building Large Foundation Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Oumi: an Open, End-to-end Platform for Building Large Foundation Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.084768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.084768Z digest=sha256:9b8815805cfc7f5599a9bd1f0236a29b874d3ac6c4422fc12d61629c5fb7803b

Observation d36d9ee2-e91b-4eb7-85c4-9bd8479c654c · outbound

This paper cites Training language models to follow instructions with human feedback.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Training language models to follow instructions with human feedback

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.088041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.088041Z digest=sha256:6df6a70f843114a11c3ca32a35d6f3481c65cb9f2081c04580a30029989997e5

Observation e41e80a7-9063-4176-bbee-9d575dc0f2fa · outbound

This paper cites OpenWebMath: An Open Dataset of High-Quality Mathematical Web Text.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance OpenWebMath: An Open Dataset of High-Quality Mathematical Web Text

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.091402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.091402Z digest=sha256:3fd3f8e247247abdbd85d4f0283f680b3d73243ed0f2a6afab30b9aa3c74c175

Observation c51db2a0-4b10-4fc2-8c76-e60b3f0a74db · outbound

This paper cites Let SSMs be ConvNets: State-space Modeling with Optimal Tensor Contractions.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Let SSMs be ConvNets: State-space Modeling with Optimal Tensor Contractions

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.094528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.094528Z digest=sha256:57d2c2a3bb3ae5de4ca66e83afd1f4af30062ba66b8e888b736bb7cfb0e8b8ca

Observation c8b09cda-0096-4a24-ba43-7a0e6b81e328 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.098462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.098462Z digest=sha256:1412a8cc03735ffc1e99962771fd0371582558455f16ce3dccc4094741be0e3a

Observation 3ee2a1ea-c56f-4a81-8837-99ad2a41e8ec · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.101847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.101847Z digest=sha256:fefd2827fd347ff2f55b5c55c614c93b5c9e909515a407fa063298a7426f0d82

Observation fa35d1c4-b80f-4b2a-8389-9444c428bcb2 · outbound

This paper cites Fineweb2: A sparkling update with 1000s of languages, 12 2024 b.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Fineweb2: A sparkling update with 1000s of languages, 12 2024 b

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.105443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.105443Z digest=sha256:80aa41517200884db0c5cb47dd3b239adb4feca8bb489d0c449911a08f6dd1bb

Observation bda3d50e-9dc8-42be-85d5-2a7da09f17c0 · outbound

This paper cites RWKV: Reinventing RNNs for the Transformer Era.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance RWKV: Reinventing RNNs for the Transformer Era

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.108362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.108362Z digest=sha256:1ac0bfc6bd8ed2e1876526fca67d069ec07526bb0babdc4f88faeacbbccd835c

Observation e01bc0f2-1fd3-4b6b-b4e3-f924a2f083cb · outbound

This paper cites Evalchemy , 6 2025.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Evalchemy , 6 2025

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.111733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.111733Z digest=sha256:e138d52cffcdd13b91e094f2ce1cbd634899182d8460e9174adb7d5abc036a7d

Observation 777610d2-4d3d-4562-a9c8-911ec6890b5e · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.114869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.114869Z digest=sha256:36d790679bd6987ad2900f7a9b092a0ad034ef28d82a71d080e8621657639605

Observation 393867cb-602c-4308-bba1-113d0c510b64 · outbound

This paper cites Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.118331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.118331Z digest=sha256:344d9c83c3c66554f2bf8e97de79ed5de7517bbed0f8bbc73226a3cb46c2262a

Observation 422be3ea-4046-47a7-a9ff-428f7ff8a7db · outbound

This paper cites Winogrande: An adversarial winograd schema challenge at scale.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Winogrande: An adversarial winograd schema challenge at scale

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.121691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.121691Z digest=sha256:cf69b54a66aab931303bd93964f49b716a268beaa2b8ad2fbbf73047018a9987

Observation 956151ee-d073-4565-a981-b1e217ec2aca · outbound

This paper cites Analysing mathematical reasoning abilities of neural models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Analysing mathematical reasoning abilities of neural models

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.124645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.124645Z digest=sha256:787e5865b11ce995fed366fe4a266af9202abab91e3856b0d51023654703af50

Observation 3628a4b7-7c89-44b7-9eca-a15340f7c846 · outbound

This paper cites Neural Machine Translation of Rare Words with Subword Units.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Neural Machine Translation of Rare Words with Subword Units

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.127589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.127589Z digest=sha256:cc504bf22c823d1debc28b059099bbede3d21823c2cf8756d4fbdf9debbc9b3c

Observation ec2cd26a-f6b6-43d5-984b-ed53622262a9 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.130974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.130974Z digest=sha256:4b9a3677b8f84c4a9255806ceed1145be53c241a344d1720021ee7bb53b6a2d9

Observation 0716ce06-0ba2-4ab1-9c3d-1ae8f576fd1f · outbound

This paper cites Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Power Scheduler: A Batch Size and Token Number Agnostic Learning Rate Scheduler

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.134546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.134546Z digest=sha256:3ee86d210d8c76857455c4ba873d1694e3114f30501ff854cfdb627eeefaa7b0

Observation ba6ed58d-d837-4a17-8923-0b9b59c9016d · outbound

This paper cites Language Models are Multilingual Chain-of-Thought Reasoners.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Language Models are Multilingual Chain-of-Thought Reasoners

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.138083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.138083Z digest=sha256:89a9ca6c191b3c9a156a86d63fbda9e214d04afa197602dd89d1c847d046e3d3

Observation a3304daf-b9de-49a4-9691-3c7c562a0ade · outbound

This paper cites Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.142116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.142116Z digest=sha256:e8e7a748e3ab5da7802cf8396471189f81aac204115acdfcc9cb6fc4cfc443d7

Observation 1141e35b-26e3-42e9-a0df-5200cd6e93fb · outbound

This paper cites Learning long sequences in spiking neural networks.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Learning long sequences in spiking neural networks

Reference 98

Resolution
verified exact
doi, observed 2026-08-06T11:44:05.272456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T11:44:05.145500Z digest=sha256:f4e40beb82c4854c373dd50a0ffd928524369063a79c08d1e34294d7d643218e

Observation e5ef4d0e-37f8-4d08-be1a-0cac25f6b8e1 · outbound

This paper cites Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Nemotron-CC: Transforming Common Crawl into a Refined Long-Horizon Pretraining Dataset

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.149322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.149322Z digest=sha256:33ed677eea7ceec2893f8f3fa74a942cf54bcb01a03fc81d4ec7b0bdec04dfa0

Observation d4d23035-cd8a-407e-9ca3-fd4b89f1c9d1 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.152682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.152682Z digest=sha256:31d68d96339aeb150f4aa8f13256607bc2206e3081fd3f6894878386f8dfe9dd

Observation 03cfb998-cc4a-40a3-b94f-f28f502efd42 · outbound

This paper cites Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.155947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.155947Z digest=sha256:a453395cc44cfb45ffebdef9d8c7d7a14148ce70202b7cc69f73d3f212b59164

Observation d187b681-39aa-4295-ba20-317f44cfdbb6 · outbound

This paper cites The falcon 3 family of open models, December 2024.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance The falcon 3 family of open models, December 2024

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.159813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.159813Z digest=sha256:2fb025a5be4cb5023cc4015423b3f874b5855e4854d30cb05fe36f5cc3a482b0

Observation 8978e1ff-d78c-423b-87e2-f5b82d67a399 · outbound

This paper cites Gemma 3 Technical Report.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Gemma 3 Technical Report

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.162957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.162957Z digest=sha256:b06f481658b378de0c94cb7caac4ac50e5f2a262886c8724a507d238d3fd136e

Pith citing papers

Observation 2704d2cb-c3e6-4f6a-be39-29bc7fb70629 · inbound

SpikingBrain: Spiking Brain-inspired Large Models cites this paper.

SpikingBrain: Spiking Brain-inspired Large Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:51:45.834693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T18:51:06.243305Z digest=sha256:f988c0a868bb8dda0f9199fa0e1e5f60d84bd4ad4656c9e85a46e1c5a93b081c

Observation aef477d0-1c57-42a5-a152-02b5b177e2b9 · inbound

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA cites this paper.

RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:54.139038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T04:55:03.309430Z digest=sha256:779a7e5b21b3f79d4095f1893f06566704b7392793d5060d2918172085afb008

Observation d605d63f-c9ad-4dfc-ba12-a6fefa308afa · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 129

Resolution
malformed identifier
arxiv_id, observed 2026-05-13T23:49:10.984633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:f4a5b9cc6c46890f40d0f9664689723f249c0ac205501f3a8a65fc4dfde2724d

Observation 73d6e007-626c-440c-acd8-ba40846cb89f · inbound

DiffuMamba: High-Throughput Diffusion LMs with Mamba Backbone cites this paper.

DiffuMamba: High-Throughput Diffusion LMs with Mamba Backbone Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T21:20:37.542843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:20:37.542843Z digest=sha256:c9e340a665330b257c669a41c714b1a2ee7fee02331e9e4e8f7f85cbbe62b75e

Observation b36332b5-c734-4800-9644-ca9698e276f0 · inbound

When Perplexity Lies: Generation-Focused Distillation of Hybrid Sequence Models cites this paper.

When Perplexity Lies: Generation-Focused Distillation of Hybrid Sequence Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T17:21:21.494194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:21:21.494194Z digest=sha256:e7bc3e294a6bed43dd1af621007b91fda4cd15986fe69bf53195858a3d50683d

Observation 17a632e8-d8fe-45da-aa80-d13aa97c9ee9 · inbound

S0 Tuning: Zero-Overhead Adaptation of Hybrid Recurrent-Attention Models cites this paper.

S0 Tuning: Zero-Overhead Adaptation of Hybrid Recurrent-Attention Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:23:21.391466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T22:20:38.938305Z digest=sha256:e7a1de936c6d47969de50df0c3f1f5600ed3cfbb6724a48df71dea38275a4381

Observation 4dcf20f3-d19c-4764-ad29-598907d510cb · inbound

Super Apriel: One Checkpoint, Many Speeds cites this paper.

Super Apriel: One Checkpoint, Many Speeds Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:01:24.853106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T02:27:11.553553Z digest=sha256:0cc4e68c7eac293016ca5cdae00047a082176a039a6f14737f7d39fa854be484

Observation dbf90b1d-db8e-40bc-ad1d-7030b0754a37 · inbound

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference cites this paper.

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:06.834096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T12:18:23.898779Z digest=sha256:a36bda38d587a7ca1830ed43315e8cd1cd6768d7eeb8521d35a3e46ea440e179

Observation 509c05d9-b8bf-4656-803c-e219c63b6005 · inbound

Component-Aware Self-Speculative Decoding in Hybrid Language Models cites this paper.

Component-Aware Self-Speculative Decoding in Hybrid Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:06:40.648369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-09T18:40:17.017805Z digest=sha256:793cec6415e1631bd59c729997c33c9cef12c87e705a1ea014657298c0e30795

Observation 1f426a32-4a49-41d5-af73-67d496314406 · inbound

XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity cites this paper.

XL-SafetyBench: A Country-Grounded Cross-Cultural Benchmark for LLM Safety and Cultural Sensitivity Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:41:11.065650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T11:08:22.322879Z digest=sha256:35f88d48888e282499b63cff356237557e3a2ca55e3bdbd17d10151f8b0fec27

Observation 22d69211-ca82-44fe-a18e-d700c1bd060d · inbound

LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models cites this paper.

LLMForge: Multi-Backend Hardware-Aware Neural Architecture Search with Infinite-Head Attention for Edge Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:03:20.524249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T13:58:55.899958Z digest=sha256:e52bad34f993e3f18517ce5a36a5a1afa1633936d78a2993bee50a8e1abe31b5

Observation 7c32fa78-ac57-4f48-8389-41d785d5e3a1 · inbound

Forget Attention: Importance-Aware Attention Is All You Need cites this paper.

Forget Attention: Importance-Aware Attention Is All You Need Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:06:21.045608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T14:38:40.948032Z digest=sha256:13cab8a6c529f2f9ce4e65b82726afb1f6c9b1f84e2a31fc852cefe6788a16c8

Observation 1008ae7b-3e77-4e5b-9f84-1fcfff013c6e · inbound

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency cites this paper.

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T01:46:26.630023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T11:34:24.019560Z digest=sha256:a55b9bd46c52c8bf8ff2107f520446f9540ffebdc611eaf348abc538c8627463

Observation 4e091ae8-936e-4da1-9189-ce5823b5a42a · inbound

Morphing into Hybrid Attention Models cites this paper.

Morphing into Hybrid Attention Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:44:28.064214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T05:56:51.447893Z digest=sha256:da353ea45868c991068b8a491875d3a9fe39e4a616dbd172f56187847dcbcc97

Observation dff1143d-d849-420a-9631-c9316323ac43 · inbound

Train Smarter, Not Longer: Memorization-Guided Data Reuse for Efficient LLM Training cites this paper.

Train Smarter, Not Longer: Memorization-Guided Data Reuse for Efficient LLM Training Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T10:45:46.618668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T10:45:46.618668Z digest=sha256:25434674ad3f6e71bff16c5e7b4619a96b6214e82821a1fddb389a37fd2f889d

Observation 371f08b4-78ee-45ea-9589-85c464010d0c · inbound

Memory for Large Language Models cites this paper.

Memory for Large Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T02:37:54.548638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:37:54.548638Z digest=sha256:cdb757cbd8ae0a32e2c4fd04cccb038c7325b1bf484c79e6472e6409d13f0d0f

Observation 42ad4e1c-28de-438b-88fe-d1b6f1ef4a65 · inbound

CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models cites this paper.

CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T00:51:15.651274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:51:15.651274Z digest=sha256:8b9932e219fcb35ad1c4d7232d0ff34bcdaf65bca606002b9f1f3bd27e2c6ec8

Observation cbf796e2-675f-4d36-b664-e6d9a6f43f49 · inbound

CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models cites this paper.

CurveFP: Co-Designing Numerical Representation and Product Arithmetic for Language Models Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:46:37.290952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:46:37.290952Z digest=sha256:fd0399ad5e6a8052ee6cd5b654cb60c8e6b0c975e9a693ec4dd54dd64b152b96