Pith. sign in

Paper Citation Record · LEDGER

Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2311.06783.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.06783 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T04:31:25.697475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T19:52:01.984738Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 439d8daa-2fbe-4583-963d-724f143185e0 · inbound

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels cites this paper.

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 288

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:35:48.052998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T16:35:47.826165Z digest=sha256:1504a04f4d82f4953b814ea9e01fef0d9c3d3ede1c816b10637c85307d91938f

Observation 62cddf9d-c106-438b-aaa0-d36e497ee0cc · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:05:03.790735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:60d83b7dbaf91fcf3912e468b3198b66cb29c10e10bba8dc4a60f1c2e44927ae

Observation 8a7f8e73-2e58-47a8-aaee-975629daa6a0 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.317986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:0707f4bfbe9bd3a68d372a4d6626aa547228a984ca5cd3cbef3f40c9bf18e2b5

Observation 802eabf2-ac88-4d93-b729-f46580100167 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.986923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:79b01d1b99af93b231dd1c68a57b3183a0f92c92e995cda4fe25798a0d2c43eb

Observation d46953ed-94ab-4b7f-9676-3bdd2604202f · inbound

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset cites this paper.

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:41:45.598200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:41:45.598200Z digest=sha256:2254379d3220bbe8bc274a03bedbdd3b620e6d0967f53f0715e22a7c9be5540d

Observation 2fd25818-8080-4507-bf87-18931a4ee56d · inbound

Grounding Degradations in Natural Language for All-In-One Video Restoration cites this paper.

Grounding Degradations in Natural Language for All-In-One Video Restoration Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:47.659414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:47.659414Z digest=sha256:18387feb81a79a281a5bc9e4a4ea7246deb27a3d23edd26ac3d794abf034c417

Observation cd0eb886-c7c4-4d11-9518-3a0d4a945e7e · inbound

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA cites this paper.

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:55.252815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:55.252815Z digest=sha256:b59f38a8debc7165a0281fd19d68e535d07d916037f78269fa5c4f5418b6fb8a

Observation 93f23fb1-bc83-47b0-bb33-4330b1a11487 · inbound

SR-Ground: Image Quality Grounding for Super-Resolved Content cites this paper.

SR-Ground: Image Quality Grounding for Super-Resolved Content Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:34:40.116069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T05:34:17.056685Z digest=sha256:f876ed6fb8bdb3bf5452a9cd2ba3eabf88134d35ba9c73e1f0ddc169660e50bd

Observation f225b0fa-d5d9-4574-8a25-a4ac93456240 · inbound

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis cites this paper.

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:13:25.910608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:13:25.910608Z digest=sha256:d2c4d26e3f56d0e6e668fdfc5d168570244c6745adf962396dd4c96cbf0c6050

Observation a19f7c65-42e9-4c50-97ba-0fa6fdccf506 · inbound

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis cites this paper.

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T04:31:25.697475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T04:31:25.697475Z digest=sha256:8fc2c18ab4595b022519c580266b57b193ee94aa38077f2b9595dbad3db66230