Pith. sign in

Paper Citation Record · LEDGER

Recognize Anything: A Strong Image Tagging Model

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2306.03514.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.03514 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:36:27.109914Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:58:38.244809Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f77e605a-4257-48f1-bf6e-74856ab78195 · inbound

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks cites this paper.

Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks Recognize Anything: A Strong Image Tagging Model

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:20:16.231345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T06:20:15.656356Z digest=sha256:984f5e1b04290c566eabdbf07abee68ec61450d7edeb02f3a6c88c5bc1293477

Observation 2bb2d1cd-a897-4e51-8ef6-a3cfab578efc · inbound

A Survey on Hallucination in Large Vision-Language Models cites this paper.

A Survey on Hallucination in Large Vision-Language Models Recognize Anything: A Strong Image Tagging Model

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:10:10.262114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:10:10.186950Z digest=sha256:4ba5102a6b5a268c68c546e917af4a5da1e82a5f321fde23baa493a2e6d57968

Observation 4351ae15-7939-46f5-93d9-5f9415e714d2 · inbound

Consistent Video Colorization via Palette Guidance cites this paper.

Consistent Video Colorization via Palette Guidance Recognize Anything: A Strong Image Tagging Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T20:36:27.109914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T20:36:27.109914Z digest=sha256:f6efab50a387426708ed5c1c89331335c9967c3839a3547a91fa4a8af93ba9e5

Observation 545cf9cc-3c2d-4110-9fd7-fa9c111dc442 · inbound

Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation cites this paper.

Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-09T11:52:15.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:52:15.003082Z digest=sha256:72b0869a7302f0fdee2b76e7decb56a73a8ad49e4d6a3f2d7be27eda8ae97116

Observation 30e699ae-cadf-4b19-8586-7092962bdd79 · inbound

VACE: All-in-One Video Creation and Editing cites this paper.

VACE: All-in-One Video Creation and Editing Recognize Anything: A Strong Image Tagging Model

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:53:54.113310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T00:53:53.855965Z digest=sha256:143838c1fe80a9c37e66b191040561b1ebae7817d4abcfd905a3e75904867a4c

Observation db08e31d-29fd-4975-a579-a9bb6949759d · inbound

A Woman with a Knife or A Knife with a Woman? Measuring Directional Bias Amplification in Image Captions cites this paper.

A Woman with a Knife or A Knife with a Woman? Measuring Directional Bias Amplification in Image Captions Recognize Anything: A Strong Image Tagging Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:55:15.097067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T23:54:15.135595Z digest=sha256:d5a28fc3193751c2c8c23cbe7e2bbd7a02b0140a0b4f4ccbc15ad7700b47b1de

Observation 36865902-00ca-4044-89f6-bdc67c795e50 · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Recognize Anything: A Strong Image Tagging Model

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.924000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:306ceffbca5e9048b8e804c012e55ee19a9e002c2b789d00c76f404fdf8071be

Observation 35cdf453-fea3-4095-8941-1fcc73ad8d74 · inbound

What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation cites this paper.

What You Perceive Is What You Conceive: A Cognition-Inspired Framework for Open Vocabulary Image Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:35.758398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:35.758398Z digest=sha256:6e4dbaee665ddc5d731faa84d4014827cf54bd3b5c3f4cb474980fb6cab7047c

Observation baae690b-501f-4003-bd4b-b06b3901c93b · inbound

ADAM: Autonomous Discovery and Annotation Model using LLMs for Context-Aware Annotations cites this paper.

ADAM: Autonomous Discovery and Annotation Model using LLMs for Context-Aware Annotations Recognize Anything: A Strong Image Tagging Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:16.194374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:16.194374Z digest=sha256:d30ea8109bcc752e78ec8162f5cf6c2d97066f37b5183648851c2a9ea6794d8f

Observation 5f0f4907-fe50-4cc2-89aa-f235015a8176 · inbound

Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph cites this paper.

Open-Vocabulary Indoor Object Grounding with 3D Hierarchical Scene Graph Recognize Anything: A Strong Image Tagging Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:58:56.201780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:58:56.201780Z digest=sha256:9e145cc343cfe4c67b4f186dbf262082ee846f626632c83c819b14c5ddc20008

Observation fff8ffb9-0e83-4198-b399-5c03635fc775 · inbound

Reinforced Visual Perception with Tools cites this paper.

Reinforced Visual Perception with Tools Recognize Anything: A Strong Image Tagging Model

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:05.091735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:27:05.091735Z digest=sha256:1cc40e64213c7d2d78c62be5392eb5b4ad15c4b4ee74b097f2140f7ef8798c25

Observation be448d09-7cc2-4571-bb94-51e5a993bebc · inbound

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation cites this paper.

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation Recognize Anything: A Strong Image Tagging Model

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.396778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T05:10:44.608959Z digest=sha256:75f3a694046bcc33282d70ee4da9a5651cfec517af5ce77cd18320bf638551e1

Observation a95c5bef-7906-4e3e-9d71-a8682a003bf3 · inbound

Empowering NPC Dialogue with Environmental Context Using LLMs and Panoramic Images cites this paper.

Empowering NPC Dialogue with Environmental Context Using LLMs and Panoramic Images Recognize Anything: A Strong Image Tagging Model

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:26:03.724298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T01:45:42.756787Z digest=sha256:19ed7dcbac02acbf8de2efce9cd1a8a65c8771143eb6bd6aaa0c25ab681f316b

Observation 1bca43ee-80b4-4498-837c-8f35d9944663 · inbound

Vista4D: Video Reshooting with 4D Point Clouds cites this paper.

Vista4D: Video Reshooting with 4D Point Clouds Recognize Anything: A Strong Image Tagging Model

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.263787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:07:10.070757Z digest=sha256:1bc06bd621f12edb91bb629271f04babab220160fccc97786ac152a33fc0a835

Observation 28b87019-9fae-4f06-9162-8828411f7cb8 · inbound

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding cites this paper.

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding Recognize Anything: A Strong Image Tagging Model

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:10:53.994295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:35:57.843351Z digest=sha256:5a2cf81f3566eed941b122ef258b4bc8ad198255378daa0900f18c73eef71bf0

Observation 928a528b-2563-47d9-873f-f3b0b8833e91 · inbound

SR-Ground: Image Quality Grounding for Super-Resolved Content cites this paper.

SR-Ground: Image Quality Grounding for Super-Resolved Content Recognize Anything: A Strong Image Tagging Model

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:34:40.136791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:34:17.056685Z digest=sha256:f8baa290e8fab1a901c50e42946023d82ae422526a5054b5c5b691e721867b04

Observation b0d85f2a-4280-4640-87f9-8a85c30d95fc · inbound

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs cites this paper.

Expanding Spatial and Temporal Context for Robotic Imitation Learning With Scene Graphs Recognize Anything: A Strong Image Tagging Model

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:16:13.673065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T17:24:02.843529Z digest=sha256:0c8eca971ac1f9f7de0e9e5df4db6965b449b44063d2522c6536129fb6be7a82

Observation 37f6e3cc-df5a-4c54-afb8-df89e4b5768b · inbound

SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture cites this paper.

SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture Recognize Anything: A Strong Image Tagging Model

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:58:38.248195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:13:49.396855Z digest=sha256:ae06d5db4e2b7b6c0affbc8dcd47d57e5f9ca63cc775917dd8eb257aadef0b77

Observation c6f0ebdf-5b56-4461-809a-c50d35f44a99 · inbound

Embodiment Meets Environment: Toward Context-Aware, Safe Physical Caregiving Robots cites this paper.

Embodiment Meets Environment: Toward Context-Aware, Safe Physical Caregiving Robots Recognize Anything: A Strong Image Tagging Model

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:55:49.243148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T00:56:51.666226Z digest=sha256:38efbd25d9389d9590f75b04f5618257d45b0ecaa553a6be43659eed5235d1f0

Observation dc268afb-10ac-4d4c-8bcc-ab301db38d8a · inbound

StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation cites this paper.

StructuredEdit: Constraint-Aware Graphic Design Editing via Differentiable Parameter Propagation Recognize Anything: A Strong Image Tagging Model

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T16:26:39.443485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T16:26:39.443485Z digest=sha256:8e845f858f79a667aa2e12451f1b42ed18c71ed52065bdd7f48c8047b98cfa5c

Observation bc185dca-cc97-409c-be44-371d52aaa6eb · inbound

Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution cites this paper.

Efficient Difficulty-Aware Dynamic Routing for Diffusion-Based Real-World Image Super-Resolution Recognize Anything: A Strong Image Tagging Model

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-01T22:34:46.291494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T22:34:46.291494Z digest=sha256:499f4d7b5900a1d929ff816cdd3eea86401bd59c2a8d7ef8a2f3d154bd16e73c