Pith. sign in

Paper Citation Record · LEDGER

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models

As of 14 August 2026, this Paper Citation Record lists 100 of 142 outbound references and 0 inbound Pith citation observations for arXiv:2507.09955.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09955 v1

Coverage vector

measured 100 of 142 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:47:40.258067Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 142 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 92074b83-7e64-41e9-867f-3f7a37855124 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.866977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.866977Z digest=sha256:38e77a68c927bb6b9d624f428f5a1aa1381b98dcf9cd27a35c6e706f30f000c2

Observation f1dc8827-3bf3-4305-b6db-a38dfb845eae · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.918113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.918113Z digest=sha256:38e2a674b737bcc6df6421bf46adb26389dbf6f19c1f461e996347b81a984312

Observation ddb35eec-f457-4a72-8535-cdc1633a5159 · outbound

This paper cites China’s cheap, open AI model DeepSeek thrills scientists,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models China’s cheap, open AI model DeepSeek thrills scientists,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.968566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.968566Z digest=sha256:caa4db577cd9ec97c9cfb99406bc24309b292c57102ff80dcbc234441cd36152

Observation 1ca8a6d7-63cb-47bc-8eae-b139caf3a9f8 · outbound

This paper cites What to know about DeepSeek and how it is up- ending A.I. - The New York Times,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models What to know about DeepSeek and how it is up- ending A.I. - The New York Times,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.971862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.971862Z digest=sha256:1c93c6a962a2cf96b9c9514a9b889f108318c50a595e1e8cbe0deb60c099631e

Observation ec9b6116-5b32-452f-8788-10c257fde08f · outbound

This paper cites What is DeepSeek, and why is it causing Nvidia and other stocks to slump? - CBS News,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models What is DeepSeek, and why is it causing Nvidia and other stocks to slump? - CBS News,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.975132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.975132Z digest=sha256:785fad976b81e5cb0f944cf21464036852a244c00b00a0631911494793c5dbe8

Observation 6bf67c6e-1566-4bc2-9c5e-3db0bfb0f2a6 · outbound

This paper cites A brief overview of ChatGPT: The history, status quo and potential future development,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A brief overview of ChatGPT: The history, status quo and potential future development,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.978543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.978543Z digest=sha256:736c96e26758978bac00772144b1b50017ab1e63a548df0b13ffeb0f67e965bf

Observation eba4530e-881a-41f6-a531-b1552c9ea905 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.981924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.981924Z digest=sha256:6a69249de90213431d5cb294a17de8152782a5f1473f19d57b32c00e507cf215

Observation 2b61b4bd-a087-43a4-945e-ec269b2b1b24 · outbound

This paper cites Training language models to follow instructions with human feedback,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Training language models to follow instructions with human feedback,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.985035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.985035Z digest=sha256:388ba9bb7f9e28889095c6cf1b3086484a75b94b805a5c8a60a711657eea652d

Observation 508a59c2-21f9-4c20-a3fc-0d2b40398938 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Learning transferable visual models from natural language supervision,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.987912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.987912Z digest=sha256:50240118a2944cbdb27657500c3ddffc222508c9fa028779c6a6782699af4e2c

Observation 02ce712d-d8bf-4ceb-92d2-05a162d081c5 · outbound

This paper cites Pre-trained language models for text generation: A survey,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Pre-trained language models for text generation: A survey,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.990600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.990600Z digest=sha256:a5f6e0eb0878ba0481e1aca295366d10e68eb840ef4639d864b08805072af3e4

Observation 83795e88-4990-4172-96bd-a827f183ac0c · outbound

This paper cites Recent advances in natural language processing via large pre-trained language models: A survey,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Recent advances in natural language processing via large pre-trained language models: A survey,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.993254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.993254Z digest=sha256:f6cedf4735b428801271a24a8b593780d3db06c4d329509f259481297b6ed123

Observation 3a6f7e12-ce04-448d-8bcb-ff9005580b30 · outbound

This paper cites Large language models versus natural language under- standing and generation,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Large language models versus natural language under- standing and generation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.996069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.996069Z digest=sha256:25641e094f4c891290efcc49bff6c0920e746c91ea012f818a91fe5cd9a3f9fb

Observation 4c7cfa2c-1b45-4230-8687-4602466361a5 · outbound

This paper cites Application of deep belief networks for natural language understanding,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Application of deep belief networks for natural language understanding,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:39.999322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:39.999322Z digest=sha256:4b1bc283b3fb329c1347154d7bb7ba7eb2591f79538f4fd6e3526eca17778c50

Observation 5fdae1a1-a2f1-450e-91af-3713ab1fccc5 · outbound

This paper cites PaLM: Scal- ing language modeling with pathways,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models PaLM: Scal- ing language modeling with pathways,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.002306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.002306Z digest=sha256:b175cf6c76b30e5ec9e0932392ef09b31649f5f281e7d7a0738fef53319866c2

Observation 519e6f24-fb65-40e1-a2c4-2f61018178a5 · outbound

This paper cites Vision-enabled large language and deep learning models for image-based emotion recognition,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Vision-enabled large language and deep learning models for image-based emotion recognition,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.005510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.005510Z digest=sha256:91d83dba92ce90b3c0833a2b999d45e6e44e1c317f911bb95613d599c126274b

Observation d017a179-eb73-4246-886b-fdd7e7dd74e9 · outbound

This paper cites A survey on deep multimodal learning for computer vision: advances, trends, applications, and datasets,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A survey on deep multimodal learning for computer vision: advances, trends, applications, and datasets,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.008318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.008318Z digest=sha256:1d4c15c36e97185d559738e99b0f9b6b5a24b560be49d85b268d7203e4217267

Observation deda322d-2d2c-4541-8983-a349d4bf0ecc · outbound

This paper cites Deep learning models for digital image processing: A review,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Deep learning models for digital image processing: A review,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.011232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.011232Z digest=sha256:7c96fc4bb2092387786a3c8397be0c6fe310f3836ac88dcb3f5697965f26d688

Observation 2fc1e6c6-0333-4274-b3c4-e08a81bec49c · outbound

This paper cites AI Action Summit (10 and 11 february 2025),.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models AI Action Summit (10 and 11 february 2025),

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.014079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.014079Z digest=sha256:009a6e862dbb22ed9062b623eef8efb1d9950319d72b2dadd2f42eb058d39526

Observation 24109014-e54a-44fe-bf46-cc24a7a7ec99 · outbound

This paper cites Can ChatGPT replace traditional KBQA models? An in-depth analysis of the question answering performance of the GPT LLM family,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Can ChatGPT replace traditional KBQA models? An in-depth analysis of the question answering performance of the GPT LLM family,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.017311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.017311Z digest=sha256:d72edf6e59d39e4cdf973f2953197b620c0477a39a6eebb21cc835fc1048a361

Observation 1f0f5a32-f24b-4004-95da-f8f49d7e4160 · outbound

This paper cites Reasoning with large language models for medical question answering,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Reasoning with large language models for medical question answering,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.020675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.020675Z digest=sha256:b7529a0a350fa0bfa610bb88df51728fcb37fa5ac6f428690cfb09b34a408b5c

Observation 3dd5671d-8234-4873-9d1c-963414a8d4f3 · outbound

This paper cites Proactive conversational agents in the post-ChatGPT world,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Proactive conversational agents in the post-ChatGPT world,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.024306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.024306Z digest=sha256:229871197e6fc8d073a291457e654ef0b10c53bb3a2d5e418e83053f1d7c925e

Observation 7a795c53-071b-4131-864c-d5ba0da1af46 · outbound

This paper cites Unlock life with a chat GPT: Integrating conversational AI with large language models into everyday lives of autistic individuals,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Unlock life with a chat GPT: Integrating conversational AI with large language models into everyday lives of autistic individuals,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.027453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.027453Z digest=sha256:0385c2bbac5037c6957207791e37740eadea7b7e4ca150e6befc20d486c3aa9e

Observation a81e4371-5062-4a89-9e24-9881f4d81f18 · outbound

This paper cites A contemporary review on chatbots, AI-powered virtual conversational agents, ChatGPT: Applications, open challenges and future research directions,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A contemporary review on chatbots, AI-powered virtual conversational agents, ChatGPT: Applications, open challenges and future research directions,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.030681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.030681Z digest=sha256:1ea81d899ee03e2d6210a6a11337639888f9300d36e418d382ceaafa91ce3e3a

Observation 9f2fcb8d-761d-405e-b986-e18e73fc5598 · outbound

This paper cites Self-collaboration code gener- ation via ChatGPT,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Self-collaboration code gener- ation via ChatGPT,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.033602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.033602Z digest=sha256:947fb5a0adb01e35509f8093be05aea0e18ed9cec2af32f4709308e3801110f7

Observation a49b61a1-00f1-4f2d-838c-2dd4b6ac04c5 · outbound

This paper cites Is your code generated by ChatGPT really correct? rigorous evaluation of large language models for code generation,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Is your code generated by ChatGPT really correct? rigorous evaluation of large language models for code generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.036683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.036683Z digest=sha256:bb90c29540d21fcd22f80471029c9a4418e905ddc27c16aea8dd777691841a83

Observation 9aff0331-9273-4b15-96b8-cf5ad8a720cd · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Gemini: A Family of Highly Capable Multimodal Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.039690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.039690Z digest=sha256:cbe5f2ac7a043cee62a6dc010569873c1ee4fc5ae57e862481f5646479056ee4

Observation bd157fda-7fea-47c5-acbb-68166b010aaa · outbound

This paper cites Introducing Claude 2.1,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Introducing Claude 2.1,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.042967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.042967Z digest=sha256:57bd232844cc9c527dbc76292a3e15ad9727d8e9ba6c70e27465d3c436d21c98

Observation ecb943fa-ca36-4da2-a099-bf97c0efb468 · outbound

This paper cites Introducing llama 3.1: Our most capable models to date,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Introducing llama 3.1: Our most capable models to date,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.046496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.046496Z digest=sha256:99461fec016b529dbbdf0b549e5cbeea59a53114e1da4a7e6ffacb8ea63bd7e9

Observation 3133a507-f0d3-4e44-b195-d0c8dcd64725 · outbound

This paper cites Mistral 7B.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Mistral 7B

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.049924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.049924Z digest=sha256:1bf43c8872da9997ef3ea057ca1ef80d8d121d3a8ed1f9df5e2ec136702430c9

Observation a54af932-c4a9-4673-bdd3-d6824d564078 · outbound

This paper cites A comprehensive survey on transfer learning,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A comprehensive survey on transfer learning,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.053192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.053192Z digest=sha256:24d1e5799634a44bcfc2c432f583a0165f1fd1b04203f04a59b9e0cee901558c

Observation 00d1e4ec-8184-42e7-b560-b368e6e37033 · outbound

This paper cites Multimodal learning with transform- ers: A survey,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Multimodal learning with transform- ers: A survey,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.056373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.056373Z digest=sha256:904c40746bb3bbbdb2ae31227c11b7b56a6bf205798959db21eb7999da91ae20

Observation f2dcf991-e4c8-47ed-adaa-c9243f38109f · outbound

This paper cites GPT-4 Technical Report.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models GPT-4 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.059461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.059461Z digest=sha256:e9bb21fdab4a92d22c3b0bf49a1c4d71edb65ad8ae7dd11dd1936a7262cf51a3

Observation 7e67bb5d-6465-4742-a8b4-71d3a9b07a48 · outbound

This paper cites DALL·E 2,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DALL·E 2,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.063226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.063226Z digest=sha256:d14812950c7656ad0e9b05cf3e0dd1ace862b4883cda11f5c72720e66a1edaec

Observation aca591a8-f1e8-48e4-9bba-79439d258f36 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.066287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.066287Z digest=sha256:f6a297c7d8bc1251ddf2f731746d6899142c6d959d3facfe37c3ae11bda0c755

Observation 9ce2e707-82b4-46d2-aa04-6ab53b787977 · outbound

This paper cites Introducing OpenAI o1,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Introducing OpenAI o1,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.069629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.069629Z digest=sha256:e1144aacdae78c3444f90cb35cfede1100e74b7b55c2807548372b140f20fffe

Observation 34f17dc2-7197-4a22-a9b8-fcc9569466b8 · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Math-shepherd: Verify and reinforce llms step-by-step without human annotations,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.073209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.073209Z digest=sha256:62557e2bec506c1bb325389990f2f3ea3089239743292cdd1bdc20ff04e54f1e

Observation 8f557111-f919-4bf4-9116-2a49950fd7e4 · outbound

This paper cites Grok 3 beta — the age of reasoning agents,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Grok 3 beta — the age of reasoning agents,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.076107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.076107Z digest=sha256:b7295a9c691dcf77fa4067ed3eb7d3df4e1d892fa6277fab7bf2ce67515cb29e

Observation b86e704d-e0c2-469c-adcc-37dedcda7723 · outbound

This paper cites DeepSeek-V3 Technical Report.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeek-V3 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.078990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.078990Z digest=sha256:fab0d569811b6c139c8eb0c7c1fd508891ec1240fc05ec51c146bec7c8da045c

Observation 9490a70d-5e07-44a4-ba95-d1bcb6814d45 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.082238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.082238Z digest=sha256:41bddf0f0384c13aaf0f5b1bc4cc5b78c744c26d2bf15402d767ef6d99f4fc1d

Observation 8a9a38f1-681d-44a5-aa31-c58902be3ce9 · outbound

This paper cites DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.085404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.085404Z digest=sha256:f982d59bacb5383bc7c52e4fe6490bea3a4bcaae0e8816144bef6c35eb0d8e1d

Observation 3437fb8f-f996-46c8-be8c-ebfba3597bad · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.088517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.088517Z digest=sha256:f3ac5044c1b0e3f96dfe54fb33cb1b164f8c6511b6c2bff33a912fec6b1a06f5

Observation b8baf641-84e4-4f4e-914a-b8451b77ace5 · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.091431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.091431Z digest=sha256:0c62ce93c6a34eb23b3ae0a9743469a701786531804f3d8f2be9ba406b456f94

Observation 7dd3a60f-51a4-4f0c-85d8-83b0ef82e23c · outbound

This paper cites Hidden markov models,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Hidden markov models,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.094795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.094795Z digest=sha256:18a15cb8943a87c9db432b5b3d21689892547c3b628f367eecba3bfadde05161

Observation ce9febf7-f751-4191-bf72-4d7855790936 · outbound

This paper cites Large language models in machine translation,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Large language models in machine translation,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.097597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.097597Z digest=sha256:869e7ca479cc6d1f5fd39a028d05dc8d49a2a0de6bd55e7399c87beeae910d1a

Observation adba33ed-a799-4310-91d1-7bb725631250 · outbound

This paper cites Dis- tributed representations of words and phrases and their composition- ality,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Dis- tributed representations of words and phrases and their composition- ality,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.100318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.100318Z digest=sha256:553b1f7fea61074e26f3ea6ce2be829e88dce997cedf69ef8150687fc507b91d

Observation 7058d49c-513f-49dc-81db-d5a65068f3a9 · outbound

This paper cites Glove: Global vectors for word representation,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Glove: Global vectors for word representation,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.103188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.103188Z digest=sha256:9f6615c922c421619506319f5d03d1184273292be62419cff95caa19d21ce083

Observation 8fee7647-9e6a-4695-aa7d-28f3e2eeb5cb · outbound

This paper cites Fundamentals of recurrent neural network (RNN) and long short-term memory (LSTM) network,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Fundamentals of recurrent neural network (RNN) and long short-term memory (LSTM) network,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.105871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.105871Z digest=sha256:c49238a77de944ad63b94525615ab69ac5615736d854ac574703f1cbae6c7cf2

Observation 1e5f51cd-e2fd-4d0c-a7fe-f888b12f7b99 · outbound

This paper cites Long short-term memory network for learning sentences similarity using deep contextual embeddings,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Long short-term memory network for learning sentences similarity using deep contextual embeddings,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.108903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.108903Z digest=sha256:ad0f819fac44c83dbad3e30761f733950c7ad96c0442a3b333683ae60a6bf94b

Observation f84e484a-0dce-4131-bab1-4678a40553af · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.111522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.111522Z digest=sha256:4d9ba59c10f80de31cf601931a539e7420bf7a64ef23047770fe2ae803c67b98

Observation 66b9917d-d3e5-4c0d-8f3c-5288bb9da120 · outbound

This paper cites DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.114788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.114788Z digest=sha256:b3b337c10ef0945cce5d53f3135649de0372d95be7ed32fe75fd816d3599cfa3

Observation db441385-dba8-4447-9b2c-a49a203eb58e · outbound

This paper cites BioBERT: a pre-trained biomedical language representation model for biomedical text mining,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models BioBERT: a pre-trained biomedical language representation model for biomedical text mining,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.117740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.117740Z digest=sha256:fbe9a0d195f26327048eefe99c7bd7316f05a584ff3224ff0de0d12e9c5f5484

Observation 7fee2620-5985-46c4-be11-03b89ae19137 · outbound

This paper cites ALBERT: A Lite BERT for Self-supervised Learning of Language Representations.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models ALBERT: A Lite BERT for Self-supervised Learning of Language Representations

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.120391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.120391Z digest=sha256:f7700bbd9f18abf14304b7858945cf732bb00381ecbd78ec3e96b8bdcf46760c

Observation b8082bd0-aa27-4449-a25a-26f29579803c · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A General Language Assistant as a Laboratory for Alignment

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.123293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.123293Z digest=sha256:c33eccb0b9d10c1fa6e6bf0fe63baf2ae527ca526336e028930eb1b1c6d6fffc

Observation 63014fb7-1378-4262-90f6-05ccfd027a62 · outbound

This paper cites ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models ERNIE 3.0: Large-scale Knowledge Enhanced Pre-training for Language Understanding and Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.126152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.126152Z digest=sha256:78f9dbc05b6486d3285da330c3dd694aae7a3541d3c1ab4f73bd8e218fec38bc

Observation e52f074a-28cd-46c2-8629-c421bc329d4f · outbound

This paper cites Multitask Prompted Training Enables Zero-Shot Task Generalization.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Multitask Prompted Training Enables Zero-Shot Task Generalization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.129359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.129359Z digest=sha256:150a60f2174a512baa81f36fa7bdbdc73ac726184f5f0c10902c4f06b35edbd8

Observation 6b3552a3-4045-4668-8fa4-655c3fc1d823 · outbound

This paper cites Language models are unsupervised multitask learners,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Language models are unsupervised multitask learners,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.132643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.132643Z digest=sha256:41636ee374e079ad742981001a970b9702d1a62b028680f43ed6aa9b14568469

Observation 1f730cbd-470c-4678-a7a7-227779c2a3f1 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Evaluating Large Language Models Trained on Code

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.135490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.135490Z digest=sha256:d8fe6f14fadb7771a1d2810eb8e415b0e5045200b0113e6adda45927c490d6ed

Observation f4f5d346-a886-4313-a195-0061b974b6d9 · outbound

This paper cites CodeGeeX: A pre-trained model for code generation with multilingual benchmarking on HumanEval-X,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models CodeGeeX: A pre-trained model for code generation with multilingual benchmarking on HumanEval-X,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.138212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.138212Z digest=sha256:85e388f08e11392c449652d57d27f8584787ec41dd48ddd93de0a7067a68865d

Observation 6258d46d-85ec-4677-a4bb-953cb8ea637e · outbound

This paper cites Llama: Open and efficient foundation language models,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Llama: Open and efficient foundation language models,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.140738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.140738Z digest=sha256:685f867598776ef639bd4d0a08678d29ec0992fe2cba099adbe671cee1f1fe8b

Observation b1ab39a1-a3df-475d-bb31-ef1909d9025a · outbound

This paper cites Pythia: A suite for analyzing large language models across training and scaling,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Pythia: A suite for analyzing large language models across training and scaling,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.143391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.143391Z digest=sha256:0eeb9083aa264121da6a42bc02957023257db8496c5064621ca74c890c5373e7

Observation 5eb97d9f-b6e5-4389-b8c8-40da58628730 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.146017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.146017Z digest=sha256:56772e418786ed909f5ce396332c036ee527486d3149a68c6dccd644ba8f763f

Observation d38b566d-63f4-456f-b389-6a858cc0e2fa · outbound

This paper cites Language models are few-shot learners,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Language models are few-shot learners,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.148792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.148792Z digest=sha256:85cca98d9371c59eb9ba61ed1947d7ecf850cb68bcbfc5c0c2d1cc25887b8817

Observation 3fd0c0a3-47aa-403d-955c-eb5d81ec8ae7 · outbound

This paper cites Scaling Language Models: Methods, Analysis & Insights from Training Gopher.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Scaling Language Models: Methods, Analysis & Insights from Training Gopher

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.152032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.152032Z digest=sha256:c423ed8afb7ee4e97fc3eff3d4b26da44ffaa9b115b1cd1704627b12afa7d24c

Observation 0e49895d-b10f-489c-9cd3-e4eb4ba82d1c · outbound

This paper cites BLOOM: A 176B-Parameter Open-Access Multilingual Language Model.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models BLOOM: A 176B-Parameter Open-Access Multilingual Language Model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.154938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.154938Z digest=sha256:c7886475d0e592182c818f6712d558ec81522f662dd144fab53bc8493b636d69

Observation f2786884-7275-4457-8eec-dd3161cefb43 · outbound

This paper cites Training Compute-Optimal Large Language Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Training Compute-Optimal Large Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.157903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.157903Z digest=sha256:5fb78fc309a1d52abb1d19630520a8818ef5e4354ec7b44e98164f08cf1a6b9b

Observation 671319ef-14ac-4f6f-b4be-2e9148ced7ff · outbound

This paper cites GLaM: Efficient scaling of lan- guage models with mixture-of-experts,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models GLaM: Efficient scaling of lan- guage models with mixture-of-experts,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.160712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.160712Z digest=sha256:c7cac7c4bd3fe1e2c67e2340e0e41bcfb0b5c24aacb96483cdda68e869d4534f

Observation 5ab8cee8-0eec-4f59-87b7-36e36fa06d30 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models LaMDA: Language Models for Dialog Applications

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.163443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.163443Z digest=sha256:936fb39ced2ff013cae9474f8f058a246e42358f1eb6ca37ec29ac08425da2b1

Observation 3dfbd059-797e-47a7-b746-58c67c178a86 · outbound

This paper cites Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Using DeepSpeed and Megatron to Train Megatron-Turing NLG 530B, A Large-Scale Generative Language Model

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.166161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.166161Z digest=sha256:3cf7da9d852059da4f6aa3b1f0b0916946a21958a7420cd67956ab99f24e826c

Observation 4a830d00-7af0-4933-a88f-f9346d932b0b · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models OPT: Open Pre-trained Transformer Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.169086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.169086Z digest=sha256:d6e93e740190be4fc44f2431e58263581d29cc2e1964677f206b626faae52369

Observation f8f6dc85-ec5c-42d2-8e92-ed43b30c1bb3 · outbound

This paper cites BloombergGPT: A Large Language Model for Finance.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models BloombergGPT: A Large Language Model for Finance

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.172271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.172271Z digest=sha256:ea62ad20ce1b80b7aa5ea7371ab2b03fa05567c1fc97a056fd95577b226c3f5d

Observation 4b73b2ee-4a4f-44e4-b24c-edbc1ae19fc3 · outbound

This paper cites The Llama 3 Herd of Models.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models The Llama 3 Herd of Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.175437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.175437Z digest=sha256:950439629eb5101963375ee6ee00902b4eba6e1ee9ec3675d816d4d19dac73d7

Observation ee7a3292-c3d1-4a62-b7f8-9e5a2048d886 · outbound

This paper cites Qwen Technical Report.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Qwen Technical Report

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.178384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.178384Z digest=sha256:522ad997af3bcb6cf860ca0835b25ffd0628914c66786a738873f3f2492bdc1b

Observation ba90cc22-65af-40c1-9b83-523597597de8 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.180958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.180958Z digest=sha256:7fe807d28f49644d818953f870b2ca37b5d7327ce99cab737e3a8d08b82074df

Observation 44c00772-270d-4177-8086-3de63693540f · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.183987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.183987Z digest=sha256:e09753d3f5a34c4bd3ecaace46de4f8158060e39510984d0adf14d946f4b3141

Observation e7603c62-fa59-4b2e-90e8-18c29d0c25b3 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Gemma: Open Models Based on Gemini Research and Technology

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.186821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.186821Z digest=sha256:eade1c6d230dc3f7282b8c1f47d733c3b389f968c2d9cc463bfd87097bf85111

Observation 44726ed4-049b-4ad1-9a96-def5fc077e9a · outbound

This paper cites Grok-2 beta release,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Grok-2 beta release,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.189816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.189816Z digest=sha256:13289414068ed96f09a198299b83d8e370b2b20938428004a2654c50d7b5b108

Observation 9379c27f-8ebb-419c-9e8f-8dc731e1414a · outbound

This paper cites Llama 3.1: An in-depth analysis of the next-generation large language model,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Llama 3.1: An in-depth analysis of the next-generation large language model,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.192786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.192786Z digest=sha256:521efe4a9f2292cdcd6fc185fbdbdb4e40581d7663b768ac1b2c9959a20f3d07

Observation 862aef8b-caa3-41cf-a7fc-12cb91dbe380 · outbound

This paper cites PredRNN: A recurrent neural network for spatiotemporal predictive learning,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models PredRNN: A recurrent neural network for spatiotemporal predictive learning,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.195263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.195263Z digest=sha256:4b3927426a8b29214a1afe4d8da7aff125ed13be507caee9fec5af57ac217c23

Observation 933fd5ff-37ab-463c-911c-e2a4a16dc593 · outbound

This paper cites Efficient and effective training of sparse recurrent neural networks,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Efficient and effective training of sparse recurrent neural networks,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.198436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.198436Z digest=sha256:b60c4f2bc4d2e7190c41b6e486fb8104378269a1d9d1e0ff7be79c0aceed4393

Observation f254d36e-4046-4813-a601-19801a87718e · outbound

This paper cites Attention is all you need,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Attention is all you need,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.201057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.201057Z digest=sha256:901dda5c5079514fc18415ca5f504cc6fe029e1e99d96526d46fd67dd6d9ffe3

Observation 2db19066-3c93-40d1-9c73-2bf8d7a14b65 · outbound

This paper cites A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o1.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o1

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.204182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.204182Z digest=sha256:7554660744e338360377178c1e8e5c4d18c7d29119368c9262c179bdc48da269

Observation 5dbc411f-0794-47ed-89ec-fbcfafba1714 · outbound

This paper cites Star: Bootstrapping reasoning with reasoning,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Star: Bootstrapping reasoning with reasoning,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.206881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.206881Z digest=sha256:9fb4f0efb211376a715796d1f082a729b62c354733fa8dc9e8825057ef1d6c0a

Observation 60d6e2b2-8357-4380-9b77-2c4b9d9fc6b8 · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.209380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.209380Z digest=sha256:25862a211e9363e0fd4fad2e4aa0948bc30574064b39cfc3242784f97ae3daaf

Observation c45363bb-1725-427c-95da-0c9a3ea2b716 · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.212052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.212052Z digest=sha256:19e81c906ada2357707b487f19e33bbe2041df4e3e00d9e54ff75a3a5e834e34

Observation 1352d2a4-96dc-40f8-b34c-4790bfb18015 · outbound

This paper cites Making Large Language Models Better Reasoners with Step-Aware Verifier.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Making Large Language Models Better Reasoners with Step-Aware Verifier

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.214785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.214785Z digest=sha256:4429af6978d26318657f60f55668d23fed44a7f75075c3e4f078170f4f397a8b

Observation 02fd27ff-8c44-4c2b-ba6e-5bb016739e1d · outbound

This paper cites An empirical analysis of compute-optimal inference for problem-solving with language models,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models An empirical analysis of compute-optimal inference for problem-solving with language models,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.217670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.217670Z digest=sha256:747380c210e1f4042cded0a22d114ab716a2734950413095a22b88b09a3959df

Observation e4b5ef6b-ffaa-40e9-8de9-4f232cb26193 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.220308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.220308Z digest=sha256:bc18993cb381033e4268c29cbc47c44aa8a7a4cd2044c370b5518a53870bb829

Observation 2cae8d5f-4135-4923-bc17-5d42cee0e4d3 · outbound

This paper cites Dynamic programming and stochastic control processes,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Dynamic programming and stochastic control processes,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.223250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.223250Z digest=sha256:4052c2fae37287c9a236d0bff2cd9f9cdad53dc4f60525b266ebed9984b08e2a

Observation f89b80f9-c662-4759-ae6b-f9a7c713db12 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Chain-of-thought prompting elicits reasoning in large language models,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.225928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.225928Z digest=sha256:50530b6b59963d83482ca7d5e804f194046f127ca9eef62accc82f1a83c1b11e

Observation 39672797-a393-4fc9-a70a-0f83732d6f1e · outbound

This paper cites Pangu-Agent: A Fine-Tunable Generalist Agent with Structured Reasoning.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Pangu-Agent: A Fine-Tunable Generalist Agent with Structured Reasoning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.228486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.228486Z digest=sha256:8adb207b9e117c21e1629befda3e41d52590bafe4cd86893192081c8e42de582

Observation a2ca418b-5482-445a-a33d-b10ff5ea8de7 · outbound

This paper cites Qwen2.5: A party of foundation models,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Qwen2.5: A party of foundation models,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.231791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.231791Z digest=sha256:7b9a98e9de0f965fbcf7c168277ab7545a17c4e92a47e8bc21c16802381f84e7

Observation ac06e6c2-2968-4c6c-9a9b-bb0660368604 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.234718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.234718Z digest=sha256:8939bab17fd2757bd7506a8c2605e16142f79c9f2052d20b7fcbbf07cb82d309

Observation 531becb3-1767-4013-aa74-131cf861ce33 · outbound

This paper cites Visual instruction tuning,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Visual instruction tuning,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.237901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.237901Z digest=sha256:5408b6a16f94df35eb6333f702548b1045b4884e789e7ded288b1fbf6890afa5

Observation 6aa9439b-c1e3-48fc-a7f6-7d03c5e71df0 · outbound

This paper cites Sigmoid loss for language image pre-training,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Sigmoid loss for language image pre-training,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.240685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.240685Z digest=sha256:821bf1e97bf333df8a2947f2dae849e340608d9a56045580f29f271c144ba084

Observation 2f52e35c-5f34-4937-bc79-1a0115ca072e · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.243919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.243919Z digest=sha256:7ec05dc8b677f80e173406b24ac954f8d8f6ee8d6b88ccc53fe87fa81753a5c9

Observation eb68c668-028c-4847-b115-9b70ec0b2552 · outbound

This paper cites Roformer: En- hanced transformer with rotary position embedding,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Roformer: En- hanced transformer with rotary position embedding,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.246864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.246864Z digest=sha256:d4b60f8dd08f0345ab6801eb3bff14860dc3ac697c82b6eac9dd06fec6b552a8

Observation 18f656ea-0aaf-4e33-936c-c30e23dae5ff · outbound

This paper cites MHA-Net: Multipath hybrid attention network for building footprint extraction from high-resolution remote sensing im- agery,.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models MHA-Net: Multipath hybrid attention network for building footprint extraction from high-resolution remote sensing im- agery,

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.249549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.249549Z digest=sha256:17693469973dd3c2b8abef4fa883532ec7b991e7d1312253cec63054ee29c592

Observation a1c15df1-44df-4760-8fd7-c7bc65460b6e · outbound

This paper cites Improving Transformers with Dynamically Composable Multi-Head Attention.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Improving Transformers with Dynamically Composable Multi-Head Attention

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.252282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.252282Z digest=sha256:8cb8b4f44182d280fb0165f743c461f6af9de050f57438e841eacf69c22367c2

Observation f37ec339-0b23-4a03-bf5b-9a5ccd011380 · outbound

This paper cites Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.255004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.255004Z digest=sha256:4d43e9670c2b11a780794a453cb516d4db5d1d39f820d42caafc4746120d35c0

Observation a585dbb0-4ad6-4bb8-8d12-ddad2ac716d3 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Fast Transformer Decoding: One Write-Head is All You Need

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T17:47:40.258067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:47:40.258067Z digest=sha256:f25d0dd9d7ebf67a4c0bf52f2dae40043eb4f86956fab55f6eaea54f5462280a

Pith citing papers

No inbound Pith citation observations are available.