Pith. sign in

Paper Citation Record · LEDGER

BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2309.00916.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.00916 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:36:19.730415Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:49:52.412623Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3c98430e-cf81-410b-9ab9-3f518cb53e09 · inbound

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey cites this paper.

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:19.730415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:19.730415Z digest=sha256:6ac484df42bfaa71bac13c5fec727875937cbec7922be4f1098f0cf8f26332b0

Observation 3c34eff9-41c3-4103-8f60-ee7f2a3239ba · inbound

SparQLe: Speech Queries to Text Translation Through LLMs cites this paper.

SparQLe: Speech Queries to Text Translation Through LLMs BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T22:05:23.866383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:05:23.866383Z digest=sha256:fd9da7657e2b11046a27146d58129ca1d89ee854bf353a71b2a2c3f2109cb2bf

Observation 9ee2869b-0654-4519-8eb0-b6f4ff4d55d2 · inbound

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples cites this paper.

Teaching Audio-Aware Large Language Models What Does Not Hear: Mitigating Hallucinations through Synthesized Negative Samples BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:05.217051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:05.217051Z digest=sha256:6053f66218c780e5297ccf64692678e43bf3aece953138e5b99b32912c06e4ab

Observation 94f98aa1-3d84-46a6-a31c-e2f6c2a8f51a · inbound

Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving cites this paper.

Enhancing Generalization of Speech Large Language Models with Multi-Task Behavior Imitation and Speech-Text Interleaving BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:30:32.148354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:30:32.148354Z digest=sha256:c156cc77473877f9050edb5b8b28047d08df42c0c23da0f141693ee1ac380ba0

Observation a1560d01-4f81-4119-825a-81f305426c83 · inbound

Towards Reliable Large Audio Language Model cites this paper.

Towards Reliable Large Audio Language Model BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:45.561089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:45.561089Z digest=sha256:b064689ee93ebb37af08d87376a3e99014060539295f6eeb12397eda57b177ce

Observation 123184ec-ddfa-45d5-854a-097c2d554f96 · inbound

TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment cites this paper.

TESU-LLM: Training Speech-LLMs Without Speech via Unified Encoder Alignment BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:58:41.157874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:58:41.157874Z digest=sha256:58ce9e5c0856e738f9ce9e8839343ceb4642f0216a545a09b52bd853464f6c72

Observation f30eebb6-7ab4-4f99-9d5f-b9f959587e64 · inbound

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks cites this paper.

MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:41:59.626563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-19T02:41:52.996457Z digest=sha256:3c64f76d4ca97bd20a596c8e6de662d6081ca47e73aa12071b2910ab72483b4e

Observation c278616e-00d3-432b-aeb3-9c587e7b02bb · inbound

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs cites this paper.

Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T16:46:49.234700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:46:49.234700Z digest=sha256:7b38b33e4eb65ee826b936c6d5aabec0f2272a55d3fc46c44097687bb533bd5a

Observation b1a4e76c-ff2a-4310-91be-faa1d2d905be · inbound

DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality cites this paper.

DESAMO: A Device for Elder-Friendly Smart Homes Powered by Embedded LLM with Audio Modality BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T16:07:49.885206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:07:49.885206Z digest=sha256:0d007072fc55359bcf208986e52d244ed8a72f9f04510ebc2f0658afa6fb33bf

Observation 74145097-ce6d-4288-a66a-0038203c7a13 · inbound

Enhancing Speech Large Language Models through Reinforced Behavior Alignment cites this paper.

Enhancing Speech Large Language Models through Reinforced Behavior Alignment BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:24:23.518043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T22:23:52.392075Z digest=sha256:6794dd6edd5e97f412410b90346c9249d3b746a8ae9de1af1f2f63fd17f3b94a

Observation 29adec66-3801-409b-ae2b-d5541b2a59f6 · inbound

Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision cites this paper.

Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:42:06.739946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:42:06.739946Z digest=sha256:631396be815bfba108e3178d9f63f4343204d427c10e76a7e79aaad5e708d6de

Observation ab3104e4-f65f-48cc-bd16-1de81931578b · inbound

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models cites this paper.

Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:46:12.325651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T14:29:18.348031Z digest=sha256:f97df81660a20c87ff46c9f684c7785357f57918d60072bbc1c73299f09f95d1

Observation 0a7ccca2-9b03-4994-8abc-ea1e838c5398 · inbound

A Survey of Audio Reasoning in Multimodal Foundation Models cites this paper.

A Survey of Audio Reasoning in Multimodal Foundation Models BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:09:24.328813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T02:08:06.976461Z digest=sha256:cf1eb52c5395bf310bcbec936106267d944eb8896a1a5133d9bcaeeaf4593740

Observation 5b986a84-e9ef-4973-a800-cacddc3888f7 · inbound

AuRA: Internalizing Audio Understanding into LLMs as LoRA cites this paper.

AuRA: Internalizing Audio Understanding into LLMs as LoRA BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:57:38.861357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T14:18:00.550824Z digest=sha256:3284057aaa3b9d1482f72e812c39f55c444e1055f896164717051ab0b13038df

Observation 372b9ece-7fe3-49a2-9164-4a13e2350174 · inbound

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation cites this paper.

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:18:12.809529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T08:18:23.182355Z digest=sha256:0b0364cb999b9901d4d0dc353a88699f97f9657723f35bd6514b22146e26cd8a

Observation 917018ec-780e-4b06-b7c8-e9ee054eb45f · inbound

CORTIS: Text-Only Adaptation of Spoken Language Models for Task-Oriented Voice Agents cites this paper.

CORTIS: Text-Only Adaptation of Spoken Language Models for Task-Oriented Voice Agents BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:29:39.091692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T13:23:59.250301Z digest=sha256:95fa0e26d0ff3526e42fc5ed8c0e7ab9d82990641bf23b564e9e8d074b100e2d

Observation 2302c956-6cb5-45fa-b8a3-3581ee28ef9c · inbound

Uncertainty-based Debiasing and Unlearning for Decontamination cites this paper.

Uncertainty-based Debiasing and Unlearning for Decontamination BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:49:52.413925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T05:57:44.576411Z digest=sha256:93bb37008ac405addc5c6b5d4adfa223366cf9bc73b2acfae34ecd9ff27a8aaf

Observation 0f35bbe4-350f-4beb-bfd6-a448f097bb72 · inbound

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models cites this paper.

Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T05:20:00.587754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:20:00.587754Z digest=sha256:3c39d980a42007058451f22996c6dcb4fbb3f2e3cbfb8d489ca621b9f920564f