Pith. sign in

Paper Citation Record · LEDGER

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning

As of 15 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 0 inbound Pith citation observations for arXiv:2507.02200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02200 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:39:42.744296Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

72 of 72 outbound references displayed

  • verified exact1
  • verified fuzzy41
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7fa9976f-1977-4bce-a214-6a78e4fdb26a · outbound

This paper cites GPT-4 Technical Report.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:35.967746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:35.967746Z digest=sha256:aa6bf3984ad485ba4fc5ceaba3e22b91634c029e98a4ba3afc45d1f9092249af

Observation 5107c995-1937-4009-be2b-4bca8a1b6c69 · outbound

This paper cites GPT-4o System Card.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning GPT-4o System Card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.044235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.044235Z digest=sha256:b66fd7769efb50606fb6b110d8ed47bb6a39fc335fe3eab9a99f9063bd6cf044

Observation d7c765c3-961a-4228-a99d-1084a7fd394a · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.050654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.050654Z digest=sha256:0e5c7aba59c8ac17a8f083686c480deb6155c50bc992b0a7f8d1af22037be078

Observation 69a4a4f2-1e40-4a5b-bfda-ec51446ed499 · outbound

This paper cites Qwen Technical Report.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.056493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.056493Z digest=sha256:da524e20e44170788e1b8cc335c0c9acaa70ad39de342828d90976f4cb5e5545

Observation ded81192-5e4a-4fa7-a375-9ab38931b5af · outbound

This paper cites Vqa: Visual question answering,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Vqa: Visual question answering,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.060758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.060758Z digest=sha256:647f3b2ed7c353fb8a4a5f41c7f7127b628a4e0a81c6a0f34e1d410ff141b605

Observation d7a75d1b-8a04-4854-ac5e-06e4197a948f · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Flamingo: a visual language model for few-shot learning,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.159865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.159865Z digest=sha256:8c6476f7f5cddb96fbdfd78cccd298391023ba3db41a93739c9ead3289486c18

Observation 46ae1e91-5a3a-4873-acfe-4f3702a5bae5 · outbound

This paper cites Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.276623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.276623Z digest=sha256:e4a97f18eb049b8bcbee5ec844118046e87c448bf6206a5d620215b707b4b700

Observation a63c5b57-2e2b-4c12-a056-4a76c4cf19c8 · outbound

This paper cites Bliva: A simple multimodal llm for better handling of text-rich visual questions,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Bliva: A simple multimodal llm for better handling of text-rich visual questions,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:50.938126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:36.391472Z digest=sha256:6cb7295670d6417fa7468553d810a055904e71cb605f219759ac3e2a71a596dc

Observation 99baa04f-ce55-447c-b321-4863963dd2ad · outbound

This paper cites On the Automatic Generation of Medical Imaging Reports.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning On the Automatic Generation of Medical Imaging Reports

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.460302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.460302Z digest=sha256:ac2f4af4fcc6f614619a8e5338be8641705cb762dfab2c90828e4eb971c5f245

Observation 3cd4a321-5884-4a88-87b1-26e048bd4ba5 · outbound

This paper cites Mmtn: multi- modal memory transformer network for image-report consistent medical report generation,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Mmtn: multi- modal memory transformer network for image-report consistent medical report generation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:50.699702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:36.526931Z digest=sha256:669c4b94c402c2b03c227d51acab549c21e6ca39f97db7008b29379c1eb4ba69

Observation 9d71ccbb-954c-482e-8b81-666501ec5ecf · outbound

This paper cites Medblip: Bootstrapping language-image pre- training from 3d medical images and texts,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Medblip: Bootstrapping language-image pre- training from 3d medical images and texts,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:50.468712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:36.611391Z digest=sha256:84dd46bea7e845d92bcc210f56daa3e9a05a2685924efeaeb0dd92d27cebe317

Observation 106f0d4d-3c0b-4633-8576-17462267029c · outbound

This paper cites Cxpmrg-bench: Pre-training and benchmarking for x-ray medical report generation on chexpert plus dataset,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Cxpmrg-bench: Pre-training and benchmarking for x-ray medical report generation on chexpert plus dataset,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:50.312214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:36.695632Z digest=sha256:96e3d81d607310b9dad0afa13e927d8102116a231cb94d8b9114f3f46dca8e27

Observation cdb269cb-b2e4-4c84-8496-dac5bfbdaa65 · outbound

This paper cites TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.782313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.782313Z digest=sha256:ec520abbcc9658f5f31616bd9042846f46f98d2487bee88dd3840254c692d29f

Observation a0c23f38-0caf-46c4-91e9-39e8418b7a07 · outbound

This paper cites DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:36.874218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:36.874218Z digest=sha256:08485cbd96492cefd7e0db15c383bf957277aea1985691d4b242bf8de4c8fff3

Observation 464a6115-9253-43a5-93fe-6970c3902ff7 · outbound

This paper cites Vary: Scaling up the vision vocabulary for large vision-language model,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Vary: Scaling up the vision vocabulary for large vision-language model,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:50.145467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:36.985825Z digest=sha256:95d2b4bdcdd02676ea33641d02e2714f1c8f40c06ae30ca903cd03d22a080472

Observation 28ac8605-6fd0-46f2-bb60-72bcd742971a · outbound

This paper cites mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:37.113140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:37.113140Z digest=sha256:8116dc8efb1c4db93522203394dfc941a7ad81564ff62cd04f9c9a85f465e1d3

Observation a060338a-bfa2-405d-91f5-108686d35f23 · outbound

This paper cites General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:37.205231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:37.205231Z digest=sha256:9af055d0e577371c2b84eb71ce0eb126da6d0d927f11fcdf14fad76c96e294db

Observation 70a3a502-64d8-4460-b958-b48458241f0c · outbound

This paper cites EventSTR: A Benchmark Dataset and Baselines for Event Stream based Scene Text Recognition.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning EventSTR: A Benchmark Dataset and Baselines for Event Stream based Scene Text Recognition

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:39:43.263965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.290321Z digest=sha256:96e8723e0f65c14f0e6809d008e63effe96a10e6d7065bf247c80995c7905136

Observation 45d74202-06e2-475b-8129-4a7ae80b28a2 · outbound

This paper cites Recurrent vision transformers for object detection with event cameras,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Recurrent vision transformers for object detection with event cameras,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:49.915023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.370022Z digest=sha256:60b7a4b7b76529144f5525a38efc90e0bbdd8158b04f6e44665677a398a08498

Observation 8f7053d6-5ce5-4466-8c35-efcdbe6ea73a · outbound

This paper cites Spatiotemporal aggregation trans- former for object detection with neuromorphic vision sensors,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Spatiotemporal aggregation trans- former for object detection with neuromorphic vision sensors,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:49.606913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.439740Z digest=sha256:4ef3cf226e9551d428f9c9793832af52def036aa13268a5ce9ff87fdecd5062e

Observation 97467366-9bc3-45b7-ba2b-b61016a283d8 · outbound

This paper cites Scene adaptive sparse transformer for event-based object detection,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Scene adaptive sparse transformer for event-based object detection,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:37.503471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:37.503471Z digest=sha256:5d3abaf83718c4902314b8019dc8315256ca4fafdfba58c4d9ef9b551aa00154

Observation 5f9d9a52-a7ff-4246-9554-8dfe8789bdcf · outbound

This paper cites Object detection using event camera: A moe heat conduction based detector and a new benchmark dataset,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Object detection using event camera: A moe heat conduction based detector and a new benchmark dataset,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:49.388795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.585998Z digest=sha256:056d0485fdf4bebb5ab09a761a2441b40d2583a494fc255e800a79ee7d05370f

Observation e8cdda75-d600-44eb-9a46-65c251444208 · outbound

This paper cites Dynamic Graph Induced Contour-aware Heat Conduction Network for Event-based Object Detection.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Dynamic Graph Induced Contour-aware Heat Conduction Network for Event-based Object Detection

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:37.647234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:37.647234Z digest=sha256:2e2d753bc64dd39c66e2f8b29ab8a7857225454dac1cd302c6e45ce1d51c373e

Observation 550adb1e-0927-4ddc-9289-e90ff112750d · outbound

This paper cites Frame- event alignment and fusion network for high frame rate tracking,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Frame- event alignment and fusion network for high frame rate tracking,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:49.091115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.712469Z digest=sha256:c61e7e812cf78ef020a84cb4db49c40288df5f771c16ed77e9120076a569cd83

Observation 7804fd71-ddf7-4e37-800e-f6894de30ddb · outbound

This paper cites Asynchronous tracking-by-detection on adaptive time surfaces for event-based object tracking,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Asynchronous tracking-by-detection on adaptive time surfaces for event-based object tracking,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.919115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.802857Z digest=sha256:7871a3fc3fd79b55bddf65aaac061c95e4b9973b234e34423a0df75b722c7c76

Observation 2cba4211-3b84-4bd3-92b1-9e3d498dd82f · outbound

This paper cites Object tracking by jointly exploiting frame and event domain,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Object tracking by jointly exploiting frame and event domain,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.772480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.886263Z digest=sha256:6a5ac373c5900ff9c499f9b2e7b871394696ab116296fc4e411c5b5d307e6fec

Observation 41fac88c-c70f-4c61-8947-483ecaaeff46 · outbound

This paper cites Event stream-based visual object tracking: A high-resolution bench- mark dataset and a novel baseline,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Event stream-based visual object tracking: A high-resolution bench- mark dataset and a novel baseline,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.636857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:37.951195Z digest=sha256:ab4341dc3470ab17f6080cbfa68c8c00422eeca20871549893750d1f0390bff5

Observation 0b804b5f-7b24-49f2-8317-c2d403e94ba9 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:38.040843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:38.040843Z digest=sha256:9d534be292b372dc8ce9d557ccdc883a2b5c9eea4160914e77db52708fdaf6a3

Observation 77921fcc-e1aa-404d-9a39-914b8b5409f3 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:38.129720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:38.129720Z digest=sha256:dc848d4ff2e3b95edb005403928b0fd23bccdd824eab273de8dff699d93d3993

Observation 2138f408-6212-4bae-b27a-f53b86a1b773 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:38.201262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:38.201262Z digest=sha256:a197f9512eda380860b7ad2b6cfe598186bff98dd4b5907a516ea358aed32da4

Observation e36d2bff-e199-4082-a772-299525206509 · outbound

This paper cites Toward understand- ing wordart: Corner-guided transformer for scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Toward understand- ing wordart: Corner-guided transformer for scene text recognition,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.483904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.307102Z digest=sha256:1e1e0018c5f126329865329d3306dc1416b12da558c4be3ae4b4a7a7927e3a18

Observation fb3583c6-146e-4fef-ac2c-068d3fa29b91 · outbound

This paper cites Icdar 2015 competition on robust reading,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Icdar 2015 competition on robust reading,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.341480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.397542Z digest=sha256:b72368335d521f9673a4b049a8610e4615f96e6bbddc162585ee95db8cbb5837

Observation 8edd767c-a5c9-4b36-9d0f-e9945ec5bb20 · outbound

This paper cites Large-scale multi-modal pre-trained models: A comprehensive survey,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Large-scale multi-modal pre-trained models: A comprehensive survey,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.218824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.491872Z digest=sha256:5256c41ab04ab7eddca2fa4d7a86cf37e236ac168a6bda4760353f8130897092

Observation 196379df-241b-4110-b8d1-ac85355d3ec8 · outbound

This paper cites Towards reasoning in large language models: A survey,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Towards reasoning in large language models: A survey,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:48.065677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.570046Z digest=sha256:99b1f9769e2370ea4eb80208cc3866905db560604cd65867f996ea07a1e87778

Observation bf01319e-44c9-4475-b5e4-4ea7f06a732d · outbound

This paper cites Scene text detection and recognition: The deep learning era,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Scene text detection and recognition: The deep learning era,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.959478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.664732Z digest=sha256:f32a32aaff260839c67ae9200a31f259e55acacda42afacb2491f4493494019e

Observation da666ce1-c4de-402c-8e0d-3616fd11b40d · outbound

This paper cites End-to-end scene text recog- nition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning End-to-end scene text recog- nition,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.852463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.796614Z digest=sha256:aeed363b913779356647b4581e9d90d5d3c2ccbbf3e9d0ad649069783789b753

Observation 6d19681a-8604-4785-9c04-0a7b4c1f7e6f · outbound

This paper cites An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning An end-to-end trainable neural network for image-based sequence recognition and its application to scene text recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.759833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:38.902784Z digest=sha256:d01776dc1d257bbd4bea7a2df3fea91651a06d8a081ede479ae5eaa12e54982e

Observation 62f4259b-a431-4cf8-83c8-e4218970b495 · outbound

This paper cites Scene text recognition from two-dimensional perspective,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Scene text recognition from two-dimensional perspective,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.601237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.005100Z digest=sha256:6a689dc5fd460664b490cbbacd3a11df5d3a97755cda97232d7339d9a98e8075

Observation 20c20456-8da3-4b44-ad9c-fbeef2eebaa0 · outbound

This paper cites Spotlight text detector: Spotlight on candidate regions like a camera,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Spotlight text detector: Spotlight on candidate regions like a camera,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.497734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.134674Z digest=sha256:9e6c70110035c3a66a5e2440b228e033324c607600bc398eb7faa3b229ab23c0

Observation ef4e2c86-136b-42ed-b4f0-4c14c740effc · outbound

This paper cites Multi-modal in-context learning makes an ego-evolving scene text recognizer,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Multi-modal in-context learning makes an ego-evolving scene text recognizer,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.310112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.272309Z digest=sha256:6ecceb5f8a4586425df6689723a0e1d8238cd6cf788e26f45f113986d2316964

Observation 84fcfc40-da24-4549-ba4a-26470426d56a · outbound

This paper cites Self- supervised character-to-character distillation for text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Self- supervised character-to-character distillation for text recognition,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:47.103138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.373243Z digest=sha256:3d053bd99f21c5d9e0045185a75f9c1d1204e410e544ce4d22dc1a87e3241c71

Observation af216b40-5a6e-4eae-9d72-7055bd1ca2c3 · outbound

This paper cites Self- supervised implicit glyph attention for text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Self- supervised implicit glyph attention for text recognition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.859814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.496053Z digest=sha256:75ce82f8baab9eec98ca2bc374e4d70fa39224266749e7f71f2ac082c13b60da

Observation 5693d0ac-a012-4fb9-b99f-63431d1ff9f7 · outbound

This paper cites Cdistnet: Perceiving multi-domain character distance for robust text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Cdistnet: Perceiving multi-domain character distance for robust text recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.680741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.556835Z digest=sha256:84addca7db3199e49d2cec7c51d31fd2eff0399e3a191d1d16b450c7ab31f36e

Observation 385fcf6a-d6ab-4a3f-bb01-1779c474a50c · outbound

This paper cites V olter: Visual collaboration and dual-stream fusion for scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning V olter: Visual collaboration and dual-stream fusion for scene text recognition,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.538441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.616706Z digest=sha256:af6c2e2565b153e43526976818ee3252a8437e82573f650643f98228bb07dc33

Observation f008b21b-d096-4c3e-b641-6c4541103e41 · outbound

This paper cites Image as a language: Revisiting scene text recognition via balanced, unified and synchronized vision-language reasoning network,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Image as a language: Revisiting scene text recognition via balanced, unified and synchronized vision-language reasoning network,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.335956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.712104Z digest=sha256:911a9a29869ecf1f145e8535b0e31d17aa6795a0c67e7b9b2c342535c5c95f4b

Observation 7682d2f6-2d3b-4e26-9ebf-25df9e34d9e8 · outbound

This paper cites Multi-modal text recognition networks: Interactive enhancements between visual and semantic features,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Multi-modal text recognition networks: Interactive enhancements between visual and semantic features,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.139522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.788971Z digest=sha256:71a05e17d63f8ee738815c0eaeaf86a640baad4a6b767a4114132746b5a60169

Observation 4b03f2fa-8add-454b-93cb-539377e921e6 · outbound

This paper cites Levenshtein ocr,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Levenshtein ocr,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:46.016506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:39.916145Z digest=sha256:5d260dbec8bcaa41a5a136d6bfec70b61e47de5439392c3a3b666f660688ca55

Observation b84bb222-b6fb-48b4-a8c8-c2a74ad3a4cb · outbound

This paper cites Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Read like humans: Autonomous, bidirectional and iterative language modeling for scene text recognition,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:45.793509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:40.044376Z digest=sha256:0dd6b089bd9317d5a73b16a90145fa21d5e6d3b74b9439c34926326a1163163b

Observation 2e0837db-127b-4244-8fea-b8590730ab73 · outbound

This paper cites Language models are unsupervised multitask learners,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Language models are unsupervised multitask learners,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.132801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.132801Z digest=sha256:962d6da981553cef588d71653ae40669db09cd0918c65baa5110f9aae3971d9e

Observation 0ca50e7b-9fd9-4b85-aba8-e44f03a93fd6 · outbound

This paper cites Language mod- els are few-shot learners,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Language mod- els are few-shot learners,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.228451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.228451Z digest=sha256:d973ed9303ca2e281549e58a51604f338c3acb03a407f38f5ca4311053bad1fc

Observation 72e3ebaf-c919-46fb-aea4-212ee8438bdf · outbound

This paper cites Palm: Scal- ing language modeling with pathways,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Palm: Scal- ing language modeling with pathways,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.310423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.310423Z digest=sha256:0189744620ce87ebd0a8b72200893c58690be47bb4875ffc1310300b8f48e235

Observation 5810a6b8-dd53-46bc-9000-3cc4857e785e · outbound

This paper cites Training language models to follow instructions with human feedback,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Training language models to follow instructions with human feedback,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.376174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.376174Z digest=sha256:4cd424045515a9b5fbef1df60fb1150a906771f5a831423761f483e508f3bb87

Observation a636d4c8-a326-47a0-aa97-2634dcb5ff83 · outbound

This paper cites Linin: Logic integrated neural inference network for explanatory visual question answering,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Linin: Logic integrated neural inference network for explanatory visual question answering,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.486852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.486852Z digest=sha256:6c3ab7bbcad33935d1d4bfb951eeab0ff7afc11d0c2a064e2a0ad34b621d6d52

Observation 4cc48566-8ec0-4aee-ad92-abd6467b89c9 · outbound

This paper cites Large lan- guage models are zero-shot reasoners,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Large lan- guage models are zero-shot reasoners,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.554946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.554946Z digest=sha256:121994e7ecec1894c3426faef29bee4361d7b6205ac62820b8913355b04e7b81

Observation 615ad96f-fc98-43bd-a9f8-bb90c3a106f1 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Tree of thoughts: Deliberate problem solving with large language models,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.639627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.639627Z digest=sha256:c8db1a5ab953e8099ef439cdb0f14f30540b3c2025ae7945aab5befd82f9eac4

Observation c72af958-f8e0-4fc7-89e4-3b49ae2d46ab · outbound

This paper cites Topologies of reasoning: Demystifying chains, trees, and graphs of thoughts,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Topologies of reasoning: Demystifying chains, trees, and graphs of thoughts,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:45.584020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:40.703978Z digest=sha256:a026fdfa48740c483d59df14371d884a3a36627c77a07ace8cc1cba45889a1a7

Observation cb12a887-39d4-4183-8cd1-c661dd95a98f · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.789469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.789469Z digest=sha256:550b6c8a803635bf3c50d28665d82ba5b322060baf9c9eb5cc9851ec81496515

Observation 2e373ab5-2292-4a3d-817d-38a9b1ace140 · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Improve Vision Language Model Chain-of-thought Reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.883150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.883150Z digest=sha256:9acff816c882958dd6ba2cba92c6133c0105ab4e7fbac8effea1e92be7e7c4d9

Observation a7d91313-ae43-4a8b-8141-7462d19a51a0 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:40.974488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:40.974488Z digest=sha256:12710443cd18164bc9ce0f21e346ca7ba417a9a2d8ba1ed14a28d962a4a8ca6a

Observation 2c9d24e2-8ec7-4beb-8004-6c6f42a63507 · outbound

This paper cites Event- based vision: A survey,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Event- based vision: A survey,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:45.423082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.076628Z digest=sha256:a73e43d741cc2eb6b800507dccaf2a128736ee0613e8dc1e4c34bbd9ee792d47

Observation 6ee778a0-9391-4e3a-b04c-7e15333e9604 · outbound

This paper cites Evcslr: Event-guided continuous sign language recognition and benchmark,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Evcslr: Event-guided continuous sign language recognition and benchmark,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:45.229099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.154754Z digest=sha256:fe132f55d705530af57b709f29d7607939f61b6b018cc58fb9ba38630299aa78

Observation fecc42b6-d763-48d2-a6e9-99b4128b4bfa · outbound

This paper cites Masked autoen- coders in 3d point cloud representation learning,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Masked autoen- coders in 3d point cloud representation learning,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:45.020282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.249084Z digest=sha256:569c7595007d4bc03693bdd1c9e2e88d8b5dc69e16d0fc1cad9bed9df503eadb

Observation 65232be1-9059-4abf-9223-9e0264bfd0ee · outbound

This paper cites Hardvs: Revisiting human activity recognition with dynamic vision sensors,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Hardvs: Revisiting human activity recognition with dynamic vision sensors,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:41.320774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:41.320774Z digest=sha256:24e02bb70c02ed4f98984ea10313d662b18781de180c06f88c44c85481e8b882

Observation 0bfa0891-d9e3-4c23-81d5-2ea8022f5ba3 · outbound

This paper cites Semantic-aware frame-event fusion based pattern recognition via large vision–language models,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Semantic-aware frame-event fusion based pattern recognition via large vision–language models,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:44.842282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.496344Z digest=sha256:1994300c2e52a835d9d48d825cd6bde784f93993ca5488e45e2d60fba0cb778f

Observation 5eda1917-6ace-4cfa-99d0-e6f6abf1cbfb · outbound

This paper cites Scene text recognition with permuted autoregressive sequence models,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Scene text recognition with permuted autoregressive sequence models,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:44.671540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.683677Z digest=sha256:f115c54f5d3a0b67da900315ccb6db4cacfb6591f1decffdf2d955802703cf95

Observation d06b9a01-fa6d-43e0-8100-3f71f69e4a72 · outbound

This paper cites Multi-granularity prediction for scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Multi-granularity prediction for scene text recognition,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:44.466263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.829838Z digest=sha256:781f5ef85f059c0e80ab1e6c6386fc016d9664e0b1b4f1e5730b4a4af562baa4

Observation 8cafcde9-ea20-43f0-ba2d-f0ff02ad0de0 · outbound

This paper cites Lister: Neighbor decoding for length-insensitive scene text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Lister: Neighbor decoding for length-insensitive scene text recognition,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:44.259711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:41.992610Z digest=sha256:f446902a720f0e5ff17ec8b7b82d992b1bd16be07afd3f1477966731ef9d1ee0

Observation 39235da8-8806-4ba1-9eb0-fd660525f0e8 · outbound

This paper cites Reading and writing: Discriminative and generative modeling for self-supervised text recognition,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Reading and writing: Discriminative and generative modeling for self-supervised text recognition,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:44.040365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:42.135929Z digest=sha256:e27738fecc9defbefac0f56c2990598415ca390398c9219ac0f13062f3bd7d58

Observation 4495c532-f27d-4056-9d36-c7ecd8735462 · outbound

This paper cites Esim: an open event camera simulator,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Esim: an open event camera simulator,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:43.786942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:42.310624Z digest=sha256:c008fba40ec4fb318b449485b441bc32a0a2eb93d8c16c88c50f65356f6e9a16

Observation c5dd5101-020e-453a-8981-e2aeabcbc84c · outbound

This paper cites Decoupled Weight Decay Regularization.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Decoupled Weight Decay Regularization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:42.464800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:42.464800Z digest=sha256:eab959ea4b9649f12f4e94825303c9f2e6231a499524f389d35a18b9d3485835

Observation 1d742370-781c-4273-8ece-3ba08ae1bf2c · outbound

This paper cites Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Synthetic Data and Artificial Neural Networks for Natural Scene Text Recognition

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:42.588318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:39:42.588318Z digest=sha256:16570093073df23787a5b1d1d805b1f9b7f2928397e55b61a057922fa3cfc9e5

Observation 467c50bd-27e1-4450-af13-9a226f1731d7 · outbound

This paper cites Synthetic data for text localisation in natural images,.

ESTR-CoT: Towards Explainable and Accurate Event Stream based Scene Text Recognition with Chain-of-Thought Reasoning Synthetic data for text localisation in natural images,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:39:43.633541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:39:42.744296Z digest=sha256:b94d1848e8551be03c1157fe5c01110c8871f7372c4af29b92f3b4c542ae0503

Pith citing papers

No inbound Pith citation observations are available.