Pith. sign in

Paper Citation Record · LEDGER

Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2311.06783.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.06783 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:13:25.910608Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T19:52:01.984738Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 439d8daa-2fbe-4583-963d-724f143185e0 · inbound

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels cites this paper.

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 288

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:35:48.052998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T16:35:47.826165Z digest=sha256:2312fadb96aab54ffef76caa400a240a955529a6829b212d976dcba1e587520c

Observation 62cddf9d-c106-438b-aaa0-d36e497ee0cc · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:05:03.790735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:4ef0ebf9fa79f1d98ecae16dc717d8264796e103baae9932fe4788069b119788

Observation 8a7f8e73-2e58-47a8-aaee-975629daa6a0 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.317986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:c4548d3b7292207586c025a5d8542ed4c22d49000012fe7bfe90606688ba9de5

Observation 802eabf2-ac88-4d93-b729-f46580100167 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.986923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:b6612055b9fe27d6fef9d206e5af681a6f88fd8befb6e4ba5c70032271b23b67

Observation d46953ed-94ab-4b7f-9676-3bdd2604202f · inbound

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset cites this paper.

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:41:45.598200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:41:45.598200Z digest=sha256:2254379d3220bbe8bc274a03bedbdd3b620e6d0967f53f0715e22a7c9be5540d

Observation 2fd25818-8080-4507-bf87-18931a4ee56d · inbound

Grounding Degradations in Natural Language for All-In-One Video Restoration cites this paper.

Grounding Degradations in Natural Language for All-In-One Video Restoration Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:47.659414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:47.659414Z digest=sha256:18387feb81a79a281a5bc9e4a4ea7246deb27a3d23edd26ac3d794abf034c417

Observation cd0eb886-c7c4-4d11-9518-3a0d4a945e7e · inbound

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA cites this paper.

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:55.252815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:55.252815Z digest=sha256:b59f38a8debc7165a0281fd19d68e535d07d916037f78269fa5c4f5418b6fb8a

Observation 93f23fb1-bc83-47b0-bb33-4330b1a11487 · inbound

SR-Ground: Image Quality Grounding for Super-Resolved Content cites this paper.

SR-Ground: Image Quality Grounding for Super-Resolved Content Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:34:40.116069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:34:17.056685Z digest=sha256:56af7d1dd1870b15f1082fa665af0d2aa49462c128bdda510dfd4ee113b1dbf4

Observation f225b0fa-d5d9-4574-8a25-a4ac93456240 · inbound

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis cites this paper.

PRISM: Priority-aware Rubric Internalization via Structured Multimodal Data Synthesis Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:13:25.910608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:13:25.910608Z digest=sha256:7751693fa47df7e763740dc21cb73e4b83b224ba6396ef273c3ce889ba96d525