Pith. sign in

Paper Citation Record · LEDGER

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning

As of 23 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 3 inbound Pith citation observations for arXiv:2506.02911.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02911 v1

Coverage vector

measured 76 of 76 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:17:01.431576Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T18:48:59.629102Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:56:14.438810Z

Reference resolution

76 of 76 outbound references displayed

  • verified exact3
  • verified fuzzy47
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 96474c8c-d8e5-4fc0-8303-02681b2abdd3 · outbound

This paper cites Challenges in unsupervised clustering of single-cell rna-seq data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Challenges in unsupervised clustering of single-cell rna-seq data

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.707112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:53.958421Z digest=sha256:469f9302aff43bafabf5ae5326073871af2037e0f6bbb209f274b5370f3a16ec

Observation 8e01376e-887b-4bf9-8b7d-49477baf1075 · outbound

This paper cites Current best practices in single-cell rna-seq analysis: a tutorial.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Current best practices in single-cell rna-seq analysis: a tutorial

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.627531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.023869Z digest=sha256:53f02fff05bd679407288b07223f5308a90348187dce424da1bdb05aa20aa54a

Observation 5717a930-e9b8-43d0-bb34-c2a913b6669d · outbound

This paper cites Integrating single-cell transcriptomic data across different conditions, technologies, and species.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Integrating single-cell transcriptomic data across different conditions, technologies, and species

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.553872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.128475Z digest=sha256:a32a5a751625fd4270c199c8ecf8e3fe1635a3ac72e693f3edfb054643d4b965

Observation 054f247f-850a-4c3b-8dc8-6cf0c8983e40 · outbound

This paper cites From louvain to leiden: guaranteeing well-connected communities.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning From louvain to leiden: guaranteeing well-connected communities

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:54.209205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:54.209205Z digest=sha256:374afea20bc1f29dcff21a27b793b1589caaab42403fe7ca1136027607e277f8

Observation bbbb5311-4b20-4e96-b1c3-ba4d198de141 · outbound

This paper cites Com- prehensive integration of single-cell data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Com- prehensive integration of single-cell data

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.447145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.292901Z digest=sha256:13a0f5250d67512e188face6e0cfbfa7bc025cd624b695405516aff6b5e211cd

Observation ece4455e-2a93-4239-83f4-c4ff31dcf799 · outbound

This paper cites Reference-based analysis of lung single-cell sequencing reveals a transitional profibrotic macrophage.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Reference-based analysis of lung single-cell sequencing reveals a transitional profibrotic macrophage

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.362763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.380106Z digest=sha256:5c952c6405976d497990efe992b7bf0a069773d709bd73da9d42198751e825b7

Observation 112bed47-9f3e-4031-b4cf-080df221e0ae · outbound

This paper cites Eleven grand challenges in single-cell data science.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Eleven grand challenges in single-cell data science

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.277431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.474896Z digest=sha256:3f8275ab825e477a7704609f05b0c206e6fee9fad68b2c7c864e1aeb1145062a

Observation b665e430-4b9e-4488-88c2-f03e24550502 · outbound

This paper cites A comparison of automatic cell identification methods for single-cell rna sequencing data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning A comparison of automatic cell identification methods for single-cell rna sequencing data

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.188155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.555903Z digest=sha256:b49517c156b538caedce18c6f45346bd92abf61c686c7c6c9028eaad1bddc6a7

Observation 7f50af10-5d1e-4121-ba93-f111d41cf5f8 · outbound

This paper cites scbert as a large-scale pretrained deep language model for cell type annotation of single-cell rna-seq data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning scbert as a large-scale pretrained deep language model for cell type annotation of single-cell rna-seq data

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:54.644380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:54.644380Z digest=sha256:76aff013261178394f69e852cb45ca2f8803eb8a8b17efc4925c903aeb80a700

Observation 297f3e38-a97e-40bc-81cd-30610bfbe306 · outbound

This paper cites Transfer learning enables predictions in network biology.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Transfer learning enables predictions in network biology

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.105838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.732048Z digest=sha256:729e467bca44cf3ada042f0abfc4e9e7500665c149c8ef7e632c0f5d971ef3f4

Observation f586aefe-ae76-4bc7-927b-0c5b1e9fdd19 · outbound

This paper cites Large-scale foundation model on single-cell transcriptomics.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Large-scale foundation model on single-cell transcriptomics

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:54.818615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:54.818615Z digest=sha256:2742ef7b9a7fce308e6fdb554ca2194bca388230c761ef0dd7f6181a0d37583a

Observation 8d3f4009-20f8-484a-9b5f-a734ac630ed8 · outbound

This paper cites Assessing gpt-4 for cell type annotation in single-cell rna-seq analysis.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Assessing gpt-4 for cell type annotation in single-cell rna-seq analysis

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:06.025255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:54.941842Z digest=sha256:d257f22c689ba9a7236f22057b160b4046bb40427d56bd02c8c23f6d32938adf

Observation e6e6f8dc-ca3d-4aa2-80bf-3ac343704c2e · outbound

This paper cites Cell2sentence: Teaching large language models the language of biology.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Cell2sentence: Teaching large language models the language of biology

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.940756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.007031Z digest=sha256:623c77bf0ef160afd4f427ec867fafd5d5e7f42187741d97140bcafff918225a

Observation b34b11f4-4909-4044-bdf7-87b52819f27c · outbound

This paper cites scelmo: Embeddings from language models are good learners for single-cell data analysis.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning scelmo: Embeddings from language models are good learners for single-cell data analysis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.871178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.088057Z digest=sha256:89a206b02eec15ff93265cad970ab679572b0df51c3eebe9ae227ce35555674b

Observation 60a29201-8f2c-4ce0-96c0-5bfe3aec2136 · outbound

This paper cites Simple and effective embedding model for single-cell biology built from chatgpt.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Simple and effective embedding model for single-cell biology built from chatgpt

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.792010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.155805Z digest=sha256:31287f95183b10afe92117035606fc661b9c0ca7170376565a0a54b73e76c55d

Observation 34c0df96-8d0d-4f7f-91e5-efe7fcf123f4 · outbound

This paper cites Langcell: Language- cell pre-training for cell identity understanding.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Langcell: Language- cell pre-training for cell identity understanding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.708576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.277527Z digest=sha256:62b7657a3b2bca246c5eddce9c3d118a31c815ced2676e216b1b90a2021accc0

Observation efcb3d14-48ee-4304-b950-af8d68c60081 · outbound

This paper cites Multimodal learning of transcriptomes and text enables interactive single-cell rna-seq data exploration with natural-language chats.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Multimodal learning of transcriptomes and text enables interactive single-cell rna-seq data exploration with natural-language chats

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.646763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.423672Z digest=sha256:4da236be607c841ad34f359c31d891675b3b34bfaf64293d16bba5ae4cc2725a

Observation 0c283730-2fe5-4c9e-bd8e-4d62930734c9 · outbound

This paper cites A Multi-Modal AI Copilot for Single-Cell Analysis with Instruction Following.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning A Multi-Modal AI Copilot for Single-Cell Analysis with Instruction Following

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:17:02.652491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.560425Z digest=sha256:af65e98fb24e1e27f549fe547b68d1062c4bd99ac8ca2875c218fff46b79ee53

Observation 6ae63305-8ee6-4157-bc6a-e8d6cf22e131 · outbound

This paper cites Language-Enhanced Representation Learning for Single-Cell Transcriptomics.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Language-Enhanced Representation Learning for Single-Cell Transcriptomics

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:17:02.581789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.668673Z digest=sha256:0d5ed33e84b65e2dc998acc34ffb5877f6dcc172b32815959511f89326b13e35

Observation c4ef9fbd-bcf7-40ce-9627-72261937b22e · outbound

This paper cites Au- tomated methods for cell type annotation on scrna-seq data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Au- tomated methods for cell type annotation on scrna-seq data

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.575377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.775102Z digest=sha256:bb19dea70bcd3e506db675bc2fc587a2e9582afea8a63e9291edc2a2b2e0e992

Observation 8f439dd7-d1a4-49b1-9e48-1a70186cfd37 · outbound

This paper cites Opening the black box: interpretable machine learning for geneticists.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Opening the black box: interpretable machine learning for geneticists

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.514660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:55.901198Z digest=sha256:630a2506899084dc1fe3dbbad67ecc17b4e73c4a28175eecdab8bd6819c8fcce

Observation d49b6564-8e10-4e70-ba55-519949e89df7 · outbound

This paper cites OpenAI o1 System Card.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning OpenAI o1 System Card

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:56.021056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:56.021056Z digest=sha256:67d65cde17126682118bb9e20e9478ac0d29adc619d4ca94b4da6ab6b678830e

Observation 90619cce-0b95-45d5-9b04-dc6f7e367744 · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:56.122164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:56.122164Z digest=sha256:6e7a146b290ec90acf117899ba014a699e7259b99dd1a738260f000ffbc6bb1c

Observation 5745bcd5-c6f8-4eba-934f-d6517a3067ae · outbound

This paper cites Evaluating large language models through role-guide and self-reflection: A comparative study.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Evaluating large language models through role-guide and self-reflection: A comparative study

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.446961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.227893Z digest=sha256:17a738e82f9e0a8db73e2cdef01e5cbd04f03faaf3ebb9b0a2e0713a3cf8df4b

Observation 427131eb-f085-4143-b24d-323cbb8d8b24 · outbound

This paper cites A mathematical model for curriculum learning for parities.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning A mathematical model for curriculum learning for parities

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.374235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.345818Z digest=sha256:6397215b748004722f0ab243702a02ccff44c379516f9e51953c5032d329b9cb

Observation 26372481-4fbc-4e29-a570-33c28b71b145 · outbound

This paper cites On curriculum learning for commonsense reasoning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning On curriculum learning for commonsense reasoning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.294552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.471034Z digest=sha256:e415d8be59d925f87ffece7f756f89d1ef1ecb1abf15c5275742c95f0a8a822d

Observation 43593b7a-07ca-4c48-b31a-ecafb66ba442 · outbound

This paper cites Defining cell types and states with single-cell genomics.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Defining cell types and states with single-cell genomics

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.202061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.570603Z digest=sha256:826186e2ac8b1a3030b28d5b36deb0b8dd71a4a7ae8b7ae61d42a0b127c92e29

Observation 57acc0c8-80ed-455e-a986-50b0d57ea465 · outbound

This paper cites A scalable scenic workflow for single-cell gene regulatory network analysis.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning A scalable scenic workflow for single-cell gene regulatory network analysis

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.139571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.685166Z digest=sha256:880f717e14bec1cbfad63be6164745a2ce02970392a414ff13917b648468027c

Observation 27ebcc3a-11c2-4a64-a219-ecf88475805e · outbound

This paper cites sctenifoldnet: a machine learning workflow for constructing and comparing transcriptome-wide gene regulatory networks from single-cell data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning sctenifoldnet: a machine learning workflow for constructing and comparing transcriptome-wide gene regulatory networks from single-cell data

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:05.063963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.760635Z digest=sha256:cdd581eb401136e325f577ed4d319cf7a79be83e218fc630b6f6b24758adf6e9

Observation dc8306e7-3d04-4120-9652-e7abf495d114 · outbound

This paper cites scgen predicts single-cell perturbation responses.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning scgen predicts single-cell perturbation responses

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:56.875825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:56.875825Z digest=sha256:28d66e84389f1c5567dc291b20de20d72481d86c7ecd8965975a02364dc137d9

Observation 6b066e76-ffd8-493c-91bb-7fb62ef7a9ae · outbound

This paper cites Machine learning for perturbational single-cell omics.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Machine learning for perturbational single-cell omics

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.926656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:56.952864Z digest=sha256:5b043fc67d03f35646eeb29c63d4f492fefe9148a8df085024e3c3228d2abcb7

Observation 1b2f109f-6467-48b6-a73b-6645c7dd7143 · outbound

This paper cites scgpt: toward building a foundation model for single-cell multi-omics using generative ai.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning scgpt: toward building a foundation model for single-cell multi-omics using generative ai

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.063468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.063468Z digest=sha256:60add17bd4762c81714549690424861e8a20a88af9c8c7322ee4771eff959633

Observation eda8a69c-b5f6-476f-b959-bb2c4be43738 · outbound

This paper cites GPT-4 Technical Report.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning GPT-4 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.194200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.194200Z digest=sha256:b0ed24c149e8c8f643b6baa27880e6955982aab3d44779327053cdfd8cae6eea

Observation cb408bc8-ac01-4dac-8bf9-fb54abd69618 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning LLaMA: Open and Efficient Foundation Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.288599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.288599Z digest=sha256:5e8a7119ae51f06a5e80e41d94d8442a6d88228c7896d2c8899ec71d73a218a3

Observation 97d807b1-ce16-4f1e-bdcd-a78fcf4e5027 · outbound

This paper cites Large language model instruction following: A survey of progresses and challenges.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Large language model instruction following: A survey of progresses and challenges

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.790859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:57.410806Z digest=sha256:69d80d66893475b558e1bb47ea39bc957781b4cfb08dffa50a89ab5192ab7898

Observation ac48ee0e-d853-41e8-8ec5-a551e9c10bbc · outbound

This paper cites Language models are few-shot learners.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Language models are few-shot learners

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.491425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.491425Z digest=sha256:61589fab6a3883ae8fffe0c3f36441e9aad5d89c489cd4cdd520604492ec301f

Observation 538e0425-740d-4428-9715-772a013f3622 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Chain-of-thought prompting elicits reasoning in large language models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.593851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.593851Z digest=sha256:60940adb44a21629f7e15d579d99f6d617d929b981268fd85519776a4f5ade14

Observation 0889c62f-50a5-4ff8-b7b5-f043d70a0f62 · outbound

This paper cites Large language models are zero-shot reasoners.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Large language models are zero-shot reasoners

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.686326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.686326Z digest=sha256:bd7e5a87e586910b835655538adf9d5580b3f7e7581f500a69180fba3c97ee9d

Observation 46a8d3a9-012c-426c-8328-dbb73c2cf2e2 · outbound

This paper cites Cz cellxgene discover: a single-cell data platform for scalable exploration, analysis and modeling of aggregated data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Cz cellxgene discover: a single-cell data platform for scalable exploration, analysis and modeling of aggregated data

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.604384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:57.785727Z digest=sha256:46b738f2c30ed3d43f78deedfc4a859ad57371c9c74f61616fbf7960822c32b5

Observation b0e8e1be-c5f3-4b0c-97a4-92a7e1afc757 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.867996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.867996Z digest=sha256:f9adf0eaf8026480cb116aa24cf19bae5882329261cdb73333d6142a71bad586

Observation 52923677-86c9-471d-bb43-fdd2fffafabf · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:57.980859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:57.980859Z digest=sha256:938a2b79df3ff7c62d70b616a7a7ce300faeb7310f89a0a99f2c92a9455b8e51

Observation 56a0ed5f-a6f2-4bc9-ac70-490bafecb94b · outbound

This paper cites Qwen Technical Report.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Qwen Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.069243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.069243Z digest=sha256:03c8c1a7cc525029ed4d5573cd9cde9b4e4a4255ef943a2718a3a3a5be198486

Observation 410eaf93-cd9d-4b28-ac9f-28cd17bbe857 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.226235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.226235Z digest=sha256:61c3bd67838cb972e1b96cf55944943b560c91f0f89762e294a00c18a16253df

Observation 5f5dd9ac-b4da-402f-9528-d31211a0bfa9 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.400074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.400074Z digest=sha256:3f291a315913278ee8ebf18bd60fef04daf71a0f51f596d389053be0e1dfa145

Observation 456875cc-aab0-41bf-88a4-46e23986080c · outbound

This paper cites DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.570880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.570880Z digest=sha256:c4ddb7a9c29485390fbba85312a8704f302a7f5f39e8d24275e03cd90da5e5ef

Observation d3eb3b00-292c-4421-856b-7d9ee4b3945e · outbound

This paper cites Supervising the search process produces reliable and generalizable information-seeking agents.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Supervising the search process produces reliable and generalizable information-seeking agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.682242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.682242Z digest=sha256:dddb2dda664b1af5155f0864fe004d95bda9fe375beab43e2a08ed7f1b39fdb9

Observation e33be5f4-da44-49b8-895f-66da14033761 · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:58.839930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:58.839930Z digest=sha256:cbb4cd5acead4684b7c22df6bb780aad00e8950f2ab496031c866cbeaf7a7912

Observation ad069c54-e81d-47ab-be4c-631b872fab63 · outbound

This paper cites The role of ontologies in biological and biomedical research: a functional perspective.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning The role of ontologies in biological and biomedical research: a functional perspective

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.519626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:58.962617Z digest=sha256:97ae3eda14569cf01dc76c35d69d989960ab795e1d5afce80e3c26f923c9b145

Observation b95b1192-41da-4a1d-8e5a-12099638567e · outbound

This paper cites Gene ontology: tool for the unification of biology.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Gene ontology: tool for the unification of biology

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.459346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.043025Z digest=sha256:c4244d3b7aa6952cb1d426576bd611794f3857a03a47baa97c3f5c99c90700db

Observation 3dcc2852-3a6e-49bf-82cc-ffd3dcd0bcc4 · outbound

This paper cites The Impact of Reasoning Step Length on Large Language Models.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning The Impact of Reasoning Step Length on Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:16:59.137930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:16:59.137930Z digest=sha256:39935126560107de034728a09d78889259b0044f3bfba7e91dc5c785cb5a7ed5

Observation 095a7ad4-357b-44b2-9a39-c234102d7d8e · outbound

This paper cites Opportunities and challenges for chatgpt and large language models in biomedicine and health.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Opportunities and challenges for chatgpt and large language models in biomedicine and health

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.390642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.212979Z digest=sha256:2fc15c2d7248833818662847e4e18eb42c976d119f5bc157f7d401415e35d1a7

Observation 1fe3ab29-a6ba-419e-adbf-c1dd12806dab · outbound

This paper cites Single cell rna sequencing of human microglia uncovers a subset associated with alzheimer’s disease.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single cell rna sequencing of human microglia uncovers a subset associated with alzheimer’s disease

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.286171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.300689Z digest=sha256:aea6c0ddb27ef1b71bfe58278d5b8841d20aa991102ab5343cad0c92a3f72306

Observation a6127949-2ba2-4861-acf0-359482a84946 · outbound

This paper cites Single-cell rna-seq analysis reveals cell subsets and gene signatures associated with rheumatoid arthritis disease activity.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell rna-seq analysis reveals cell subsets and gene signatures associated with rheumatoid arthritis disease activity

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.193112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.384244Z digest=sha256:afd35cffa39022041105f5245952dfd9ed6616a9190795b4248d76779fbcfee2

Observation 5e1b902a-d9e3-43dc-8c05-d61fa4d7633c · outbound

This paper cites High- resolution single-cell atlas reveals diversity and plasticity of tissue-resident neutrophils in non-small cell lung cancer.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning High- resolution single-cell atlas reveals diversity and plasticity of tissue-resident neutrophils in non-small cell lung cancer

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.134146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.474042Z digest=sha256:f1b153fddeb3e92cdaf9eed1e2f6194f07a5cb85047423d1a970fbef86abec42

Observation 1f394338-bf2d-4b6b-bb56-c2bdea754bcd · outbound

This paper cites Cells of the human intestinal tract mapped across space and time.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Cells of the human intestinal tract mapped across space and time

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:04.080111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.571636Z digest=sha256:c0cafd2d93f1486d46332e223303d76054a8b0d2ee67771c2db1b2209be37fa3

Observation c2d467af-8292-455f-93ab-090062646b3b · outbound

This paper cites Persistent t cell unresponsiveness associated with chronic visceral leishmaniasis in hiv-coinfected patients.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Persistent t cell unresponsiveness associated with chronic visceral leishmaniasis in hiv-coinfected patients

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.954065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.631363Z digest=sha256:be9ad9b0c5bac38871df82b41c3ce93b9523736f23f5cb4ad6a63340e8bdacaf

Observation fe3d8965-122b-4830-8a66-84e20374ce99 · outbound

This paper cites Single-cell multi-omics analysis of human pancreatic islets reveals novel cellular states in type 1 diabetes.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell multi-omics analysis of human pancreatic islets reveals novel cellular states in type 1 diabetes

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.866585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.702650Z digest=sha256:a6440dba68c83043fa0c500f5a042feb0420437e0b4b6685b6cdd80bc5b3e049

Observation d2eb3e37-3678-4444-a463-7af50a0c2974 · outbound

This paper cites An integrated cell atlas of the lung in health and disease.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning An integrated cell atlas of the lung in health and disease

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.771923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.793262Z digest=sha256:3429f2949f14b45da71c336b082c03b0c059f9c125d433c5171bad78b5e3621a

Observation 6df45982-5f83-4e13-bea6-fc152f5b731c · outbound

This paper cites Single-cell transcriptomics of the human retinal pigment epithelium and choroid in health and macular degeneration.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell transcriptomics of the human retinal pigment epithelium and choroid in health and macular degeneration

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.669440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.887942Z digest=sha256:0df369d9d49f32c1242eb16da74a2bea25825b30d94ae512b99f54153467f799

Observation cffba966-6346-4b2e-81fb-cc797b91df07 · outbound

This paper cites An atlas of healthy and injured cell states and niches in the human kidney.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning An atlas of healthy and injured cell states and niches in the human kidney

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.584176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:16:59.958371Z digest=sha256:20631d7d9651b981649bb897a428b0e31ecff78971836b971d5c7899cac38276

Observation c2558940-4ce9-43f5-8612-24154dc35015 · outbound

This paper cites Ovarian cancer mutational processes drive site-specific immune evasion.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Ovarian cancer mutational processes drive site-specific immune evasion

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.507627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.028529Z digest=sha256:fa8f657c99787b4ee68d678d06eca6cda8723e278c7e9f059cb9884dfdf4d395

Observation c09b4e6b-7884-42c8-9393-e6082662b9bd · outbound

This paper cites Single-cell multiomics reveals increased plasticity, resistant populations, and stem-cell–like blasts in kmt2a-rearranged leukemia.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell multiomics reveals increased plasticity, resistant populations, and stem-cell–like blasts in kmt2a-rearranged leukemia

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.438412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.128438Z digest=sha256:761a2e87bbc32df324b16618b7cde93700f74fb3403a8dcd4f74f6d5f497509a

Observation a435344f-2d4c-4dfa-af24-a62a453ef8a4 · outbound

This paper cites Single-cell atlas of common variable immunodefi- ciency shows germinal center-associated epigenetic dysregulation in b-cell responses.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell atlas of common variable immunodefi- ciency shows germinal center-associated epigenetic dysregulation in b-cell responses

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.360178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.205996Z digest=sha256:1ffe3eae702f01299df64bf8c69f4eda73617847500e74cfe2027be70389cdd4

Observation 683ee068-e380-48ad-861d-1a6b66801ef5 · outbound

This paper cites Single-cell resolution characterization of myeloid-derived cell states with implication in cancer outcome.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Single-cell resolution characterization of myeloid-derived cell states with implication in cancer outcome

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.290176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.264891Z digest=sha256:5c10ddbbe1287fe3c2b1ca9aa13ff330f2f951012d9e5cc7415d6acfd81062fe

Observation 3a89a888-d1ea-4b28-afad-320b2c929129 · outbound

This paper cites Distribution-independent cell type identification for single-cell rna-seq data.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Distribution-independent cell type identification for single-cell rna-seq data

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.225274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.379519Z digest=sha256:1b28ca019ca02f584ec47a1e0ce515809708c7ce87905eccbdf01edae18238e5

Observation 90ad1f8f-8708-44ad-92a1-a114baebff62 · outbound

This paper cites Celler:A Genomic Language Model for Long-Tailed Single-Cell Annotation.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Celler:A Genomic Language Model for Long-Tailed Single-Cell Annotation

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:17:01.798590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.466224Z digest=sha256:4c4cd2b8bbc64ba13374ddb979584e64a724c8a5b6dff9e62993ff8384cf7358

Observation 1d78876e-2c5a-4db4-8bc7-69a775814f40 · outbound

This paper cites Trl: Transformer reinforce- ment learning.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Trl: Transformer reinforce- ment learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:00.590584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:00.590584Z digest=sha256:aa423329936a28262a837391f3eb52997cd16be1b3c6fe5127b874f77cf34a65

Observation a8fb9d8c-c22b-4bc9-a22a-5384a4e27167 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Lora: Low-rank adaptation of large language models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:00.618313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:00.618313Z digest=sha256:703aaf35df2af076c96865271cfbdbe6a475ce4f0f2302e3cacddd098eb0b986

Observation 84ca8230-ddc9-4526-b4e1-92b88305c321 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Qwen2.5: A party of foundation models, September 2024

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.125187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:00.692671Z digest=sha256:a24b286643ed269b836828316581b0d1621a81b4e84b726f0603c326b8da8c6a

Observation 780c8b66-c312-430a-a861-bd189270c019 · outbound

This paper cites Qwen2 Technical Report.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Qwen2 Technical Report

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:00.827840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:00.827840Z digest=sha256:eed697b95250e3b74c1ce3ed3cd021560f9b8271c530c85c0f3c7cfd8becf89d

Observation 3a9f7399-e5b9-461a-b03e-1e1fcbf73ffe · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:00.927387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:00.927387Z digest=sha256:6eab07a23be2f64ee81dd8d35835cf5f1d4d1607012e7e906cfe8d67a4d08a2c

Observation 21b79bf0-b282-4cad-90d6-5d7256f32b32 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning Gonzalez, Hao Zhang, and Ion Stoica

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:01.042121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:01.042121Z digest=sha256:0b3c132606b335d409d9030beb3a24d341d4fc36f2153f38e26f832a742ed0f3

Observation 3a5959f8-91ed-4d46-8caa-0833888a3c12 · outbound

This paper cites chain-of-thought.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning chain-of-thought

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:03.021654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:01.116167Z digest=sha256:86371d4cc28bad52bd10863e6e4f79330a9350e85e02c42688448a9c33df47dd

Observation a5b343b4-af07-43e3-86f4-c301089be1bd · outbound

This paper cites These genes alone are not sufficient for cell type identification.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning These genes alone are not sufficient for cell type identification

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:02.932795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:01.204194Z digest=sha256:1448f9d7441a11d7137235b1c2b0fff45da11bc9a29fd301a1af5982843a18a7

Observation b95fb062-9d1d-4ae4-a2be-d2003e91aec2 · outbound

This paper cites • Some cell types are more specific (e.g., IgA plasma cell, activated CD4-positive T cell), while others are broader (e.g., plasma cell, vein endothelial cell).

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning • Some cell types are more specific (e.g., IgA plasma cell, activated CD4-positive T cell), while others are broader (e.g., plasma cell, vein endothelial cell)

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:02.852002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:01.325480Z digest=sha256:253d42f44d7e0463404efa3c5026bc8375a0ecf27d2e4c2bf9fac07af2fcf41a

Observation 924e6c46-b86a-4c62-bbb5-89ed2d0e5b92 · outbound

This paper cites T-helper 1 cell.

Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning T-helper 1 cell

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:17:02.764815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T11:17:01.431576Z digest=sha256:e1e5931cc532aae9b7af30efa659a94cfa4c327423c9a5629e753590b0e37c60

Pith citing papers

Observation 931a8e8e-ad3b-42e9-a1f7-a90113ca7e37 · inbound

Gene-R1: Reasoning with Data-Augmented Lightweight LLMs for Gene Set Analysis cites this paper.

Gene-R1: Reasoning with Data-Augmented Lightweight LLMs for Gene Set Analysis Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T18:48:59.629102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:48:59.629102Z digest=sha256:0b8fa1d9b97eb27a1d6b9dcc2a990c88871785d01a5eebb14fca2a42a6aae23b

Observation 5f3b5f4c-df96-4550-80b1-9a364ae5c934 · inbound

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches cites this paper.

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:14.440258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T17:35:01.285534Z digest=sha256:05a6cd406b975abbd9d471011457fbd993c12ad7b8d81847e80a3f0215890114

Observation 2f77efaf-05ed-45d9-b62b-22ddd243659e · inbound

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches cites this paper.

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches Cell-o1: Training LLMs to Solve Single-Cell Reasoning Puzzles with Reinforcement Learning

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:35:34.111016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-01T07:22:25.349398Z digest=sha256:78fef02accb30a0e34b270175429bedda9eb3c5b57ef7b23531187e786b0db60