Pith. sign in

Paper Citation Record · LEDGER

Recognize Anything: A Strong Image Tagging Model

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2306.03514.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.03514 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:36:27.109914Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:58:38.244809Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f77e605a-4257-48f1-bf6e-74856ab78195 · inbound

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks cites this paper.

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks Recognize Anything: A Strong Image Tagging Model

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:20:16.231345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T06:20:15.656356Z digest=sha256:8f7155cb892a2f8acc28a98b4a4392d33cd8b378163db736a39a0cc0d4399046

Observation 2bb2d1cd-a897-4e51-8ef6-a3cfab578efc · inbound

A Survey on Hallucination in Large Vision-Language Models cites this paper.

A Survey on Hallucination in Large Vision-Language Models Recognize Anything: A Strong Image Tagging Model

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:10:10.262114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:10:10.186950Z digest=sha256:5b0a0b27fe46640f71cc9f238fe967cc01f0cc25c691b0a483b2d43476f2b9e0

Observation 4351ae15-7939-46f5-93d9-5f9415e714d2 · inbound

Consistent Video Colorization via Palette Guidance cites this paper.

Consistent Video Colorization via Palette Guidance Recognize Anything: A Strong Image Tagging Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T20:36:27.109914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T20:36:27.109914Z digest=sha256:435576080fc5344fac15a2ff5dd6cc3c1010ceefdce870bad8b4e500eaa6f8e9

Observation 545cf9cc-3c2d-4110-9fd7-fa9c111dc442 · inbound

Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation cites this paper.

Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-09T11:52:15.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:52:15.003082Z digest=sha256:a273b304dbfb99da0b05139f108e8e47c04a5f46bbe2b1e55b0fc74a6f150018

Observation 30e699ae-cadf-4b19-8586-7092962bdd79 · inbound

VACE: All-in-One Video Creation and Editing cites this paper.

VACE: All-in-One Video Creation and Editing Recognize Anything: A Strong Image Tagging Model

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:53:54.113310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T00:53:53.855965Z digest=sha256:539c1db87164c74ee528c6c4606dfdb58686a2513beaf99aa2727825295e8fb2

Observation db08e31d-29fd-4975-a579-a9bb6949759d · inbound

A Woman with a Knife or A Knife with a Woman? Measuring Directional Bias Amplification in Image Captions cites this paper.

A Woman with a Knife or A Knife with a Woman? Measuring Directional Bias Amplification in Image Captions Recognize Anything: A Strong Image Tagging Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:55:15.097067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:54:15.135595Z digest=sha256:8532f37b8eddb2714e62e5594be98396291ea460ed80e02e28b84e7b4447ce0f

Observation 36865902-00ca-4044-89f6-bdc67c795e50 · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Recognize Anything: A Strong Image Tagging Model

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.924000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:7c668a33d27d1806725bf04df3d20b837caed82a20f0f701f62b09045cf25673

Observation 35cdf453-fea3-4095-8941-1fcc73ad8d74 · inbound

What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation cites this paper.

What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:35.758398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:35.758398Z digest=sha256:5788dca400021269b991c005d9fd1fbfd0f7ae7595a002cb1db1b4ab191e06f2

Observation baae690b-501f-4003-bd4b-b06b3901c93b · inbound

ADAM: Autonomous Discovery and Annotation Model using LLMs for Context-Aware Annotations cites this paper.

ADAM: Autonomous Discovery and Annotation Model using LLMs for Context-Aware Annotations Recognize Anything: A Strong Image Tagging Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:16.194374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:16.194374Z digest=sha256:3f3f82d3f1813f578249ad13e1e5b937ad00ac82f9158adefbdc22ecabd197b5

Observation 5f0f4907-fe50-4cc2-89aa-f235015a8176 · inbound

Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph cites this paper.

Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph Recognize Anything: A Strong Image Tagging Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:58:56.201780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:58:56.201780Z digest=sha256:1746b8f004c3606961e82e75b8d4e1e9a034f10388248f52636d3152a5e6a825

Observation fff8ffb9-0e83-4198-b399-5c03635fc775 · inbound

Reinforced Visual Perception with Tools cites this paper.

Reinforced Visual Perception with Tools Recognize Anything: A Strong Image Tagging Model

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:05.091735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:27:05.091735Z digest=sha256:ab3c8dd4ceb4533c705878a10e05fbac1418f27a99cbbaed512d0002f13779fd

Observation be448d09-7cc2-4571-bb94-51e5a993bebc · inbound

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation cites this paper.

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.396778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T05:10:44.608959Z digest=sha256:62d0afd565f99c32ff43b36e837e1ee537bdbe4883ffda34bc5f87937809beb0

Observation a95c5bef-7906-4e3e-9d71-a8682a003bf3 · inbound

Empowering NPC Dialogue with Environmental Context Using LLMs and Panoramic Images cites this paper.

Empowering NPC Dialogue with Environmental Context Using LLMs and Panoramic Images Recognize Anything: A Strong Image Tagging Model

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:26:03.724298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T01:45:42.756787Z digest=sha256:a7f83876db8fbad2fe3b174d3eb2f1487e2176f3e91af2dcf3d932baa140ec38

Observation 1bca43ee-80b4-4498-837c-8f35d9944663 · inbound

Vista4D: Video Reshooting with 4D Point Clouds cites this paper.

Vista4D: Video Reshooting with 4D Point Clouds Recognize Anything: A Strong Image Tagging Model

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.263787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:07:10.070757Z digest=sha256:37115fb57e17ead2e9883ff3bd7f6a0466eee0d72675e07caad4a6e2c50dcdc0

Observation 28b87019-9fae-4f06-9162-8828411f7cb8 · inbound

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding cites this paper.

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding Recognize Anything: A Strong Image Tagging Model

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:10:53.994295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:35:57.843351Z digest=sha256:80823c4a033ba1361074c7f5f11cd0f0e35c46925fbb474630c9533561919274

Observation 928a528b-2563-47d9-873f-f3b0b8833e91 · inbound

SR-Ground: Image Quality Grounding for Super-Resolved Content cites this paper.

SR-Ground: Image Quality Grounding for Super-Resolved Content Recognize Anything: A Strong Image Tagging Model

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:34:40.136791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:34:17.056685Z digest=sha256:4e3a0e9d2828303efd98b73b4fae32822358fa2ceeb587ef67dc38d26b2fed8e

Observation b0d85f2a-4280-4640-87f9-8a85c30d95fc · inbound

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs cites this paper.

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs Recognize Anything: A Strong Image Tagging Model

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:16:13.673065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T17:24:02.843529Z digest=sha256:4411818076ed40ab9a116ea7de39ab1b26cdff77e1908931877b4831b530a68a

Observation 37f6e3cc-df5a-4c54-afb8-df89e4b5768b · inbound

SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture cites this paper.

SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture Recognize Anything: A Strong Image Tagging Model

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:58:38.248195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:13:49.396855Z digest=sha256:1dd7651194ed8581bd4cadfbeb3775f10f945c13d7843118e19836a6ea148a32

Observation c6f0ebdf-5b56-4461-809a-c50d35f44a99 · inbound

Embodiment Meets Environment: Toward Context-Aware, Safe Physical Caregiving Robots cites this paper.

Embodiment Meets Environment: Toward Context-Aware, Safe Physical Caregiving Robots Recognize Anything: A Strong Image Tagging Model

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:55:49.243148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T00:56:51.666226Z digest=sha256:dea12ff4c2d672d64fce78d89cffb99eb1fad94a44b7fdabe89f4c2cb31fa0ae

Observation dc268afb-10ac-4d4c-8bcc-ab301db38d8a · inbound

StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation cites this paper.

StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation Recognize Anything: A Strong Image Tagging Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T16:26:39.443485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T16:26:39.443485Z digest=sha256:97b5f5d2b105ffc20c601012e00e4e1c03c60fe2ad7bb9e6c30c5cb85cd93b2c

Observation bc185dca-cc97-409c-be44-371d52aaa6eb · inbound

Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution cites this paper.

Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution Recognize Anything: A Strong Image Tagging Model

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-01T22:34:46.291494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T22:34:46.291494Z digest=sha256:8b6a88dd2927e1224a493ce11f9c9832e494665f27c1ff6122ff7110ce1a8a61