Pith. sign in

Paper Citation Record · LEDGER

Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.10676.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.10676 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:49.386461Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:16:12.329267Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 65c0f6e8-2b51-4951-8a28-cdba04adcb98 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.160370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:db71907e736c639f34e413feafb2c3a2b3cb49ce8effc8c719ff87b6ade4f8a6

Observation 52e9c42c-8a56-447f-96d7-bc50b11b75d8 · inbound

In-the-wild Audio Spatialization with Flexible Text-guided Localization cites this paper.

In-the-wild Audio Spatialization with Flexible Text-guided Localization Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:49.386461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:59:49.386461Z digest=sha256:8fb5baa631bb9223a83112644828caa33355b689025a293bb0b2da54a69e970f

Observation 21c7ed62-4dbc-43e3-8164-65268c863d4d · inbound

ASAudio: A Survey of Advanced Spatial Audio Research cites this paper.

ASAudio: A Survey of Advanced Spatial Audio Research Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 162

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:55.277149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:54:55.277149Z digest=sha256:6f6ef3faa552e9749b6ea66d73c4016dabd822736acf01dd52ce07b1b1bbd58e

Observation b466d441-2685-4965-b333-0df6934711ff · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:59.530861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:59.530861Z digest=sha256:a50572caa071c0d1dea8a8778e93263315efd5597e6a8a15d4d8a63c4b6e1a36

Observation 22f7ffa9-5bba-4abd-80c3-24147f580c6e · inbound

Bagpiper: Solving Open-Ended Audio Tasks via Rich Captions cites this paper.

Bagpiper: Solving Open-Ended Audio Tasks via Rich Captions Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 569

Resolution
unresolved
no resolver link, observed 2026-08-03T04:22:52.703592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:22:52.703592Z digest=sha256:6f6a66233908ba3e311e54403ccbd0edf1d3007c8c8c27a70f9428648e4bf615

Observation d24123ec-0ecc-4ac0-978b-27bf350c98a2 · inbound

Bagpiper: Solving Open-Ended Audio Tasks via Rich Captions cites this paper.

Bagpiper: Solving Open-Ended Audio Tasks via Rich Captions Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 569

Resolution
unresolved
no resolver link, observed 2026-08-04T06:13:30.182634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:13:30.182634Z digest=sha256:4894de1794bd632500c16628b752a201affb9b5089381765aa9ef36d148e8433

Observation 05ee09f0-d164-45c2-988a-941acd3d7775 · inbound

FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips cites this paper.

FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:40:51.721552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T18:57:21.434793Z digest=sha256:5924fc387c39b27fda5e2d9d81a5cba0099e7e0931055549b69ca477a2b27a59

Observation 42f0b113-d308-4a44-9556-dbcffd6d9fc2 · inbound

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer cites this paper.

Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.331074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T21:17:48.886421Z digest=sha256:a70df5d12bb71fba27373138c9e920be1a1a28e3b29d26e01cd1355220313d50

Observation aa741522-7280-4e57-a108-c61ee2375d0e · inbound

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models cites this paper.

Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models Both Ears Wide Open: Towards Language-Driven Spatial Audio Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T11:20:17.997004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:20:17.997004Z digest=sha256:62de6b02fefc190f24304622ef50c9b79cadbf3e852590374453b8b9b770abc6