Pith. sign in

Paper Citation Record · LEDGER

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness

As of 17 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 2 inbound Pith citation observations for arXiv:2507.18119.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18119 v2

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:41:57.488433Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:03:09.632009Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T20:59:47.893961Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eb64e32c-ff90-48d9-b9f2-575e4fb635e2 · outbound

This paper cites On the landscape of spoken language models: A comprehensive survey,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness On the landscape of spoken language models: A comprehensive survey,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.914459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.365862Z digest=sha256:e0cea275f44b0eca600e4107f7de18f9a0c66b7d59999125115db5ab1ac35174

Observation 68bbeba8-9c7c-467d-b81d-bd99ce2c822f · outbound

This paper cites Speechgpt: Empowering large language models with intrinsic cross-modal conversational abilities,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Speechgpt: Empowering large language models with intrinsic cross-modal conversational abilities,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.901450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.371356Z digest=sha256:c45282fd5820f9bd5f299b074c61a325f125b9cde37a3489088dbce7638cb985

Observation 24041366-a50b-40ad-a380-5a6e1234c62c · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Moshi: a speech-text foundation model for real-time dialogue,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.888581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.376159Z digest=sha256:9cb230563a294c14f18b5ccce4d0f3f694dd6353f827e2677e098d09b8074368

Observation 72c5bfa6-bcc3-430b-a2bc-3d0e4a717200 · outbound

This paper cites Speechgpt 2.0-preview,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Speechgpt 2.0-preview,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.875689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.381247Z digest=sha256:60044a95f9952c79400e94449682ad90608de10124a6385f8cbe97c47e7e9fe9

Observation 938b6b44-702d-478d-a7f9-d18e29f44c28 · outbound

This paper cites Llama-omni: Seamless speech interaction with large language models,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Llama-omni: Seamless speech interaction with large language models,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.862931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.385844Z digest=sha256:dae408e1e5c260064b599fa2a7b1ec6b6fa63766f724ae7863d3e939e8beb293

Observation e72106f9-554c-4ef0-8642-7a89c502a69f · outbound

This paper cites Freeze-omni: A smart and low latency speech-to-speech dialogue model with frozen LLM,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Freeze-omni: A smart and low latency speech-to-speech dialogue model with frozen LLM,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.849408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.390739Z digest=sha256:ea7849f51911f5856ba9a32dfd7aace9cc2ae18de9a3d617a762d0ec84a9ae28

Observation 36bed796-faa9-4a2e-96fc-0c967461f38a · outbound

This paper cites Slam-omni: Timbre-controllable voice interaction system with single-stage training,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Slam-omni: Timbre-controllable voice interaction system with single-stage training,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.836513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.396459Z digest=sha256:01dfea4b205f3b72dd143d45fa9ef8b51a656f51d6c82dcde3920872c4b0901f

Observation 02c73c13-ce25-43d2-8f96-5a2a7c3f9521 · outbound

This paper cites Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Glm-4-voice: Towards intelligent and human-like end-to-end spoken chatbot,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.823488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.400457Z digest=sha256:a8873078da7dc96ab49477ac2d5d0d6f826fc1e8e74eafc031c435e111f9261f

Observation cc859baf-0b64-426d-a228-f8fba81a3694 · outbound

This paper cites Minicpm-o 2.6: A gpt-4o level mllm for vision, speech, and multimodal live streaming on your phone,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Minicpm-o 2.6: A gpt-4o level mllm for vision, speech, and multimodal live streaming on your phone,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.810776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.404284Z digest=sha256:851aeb2affb0d346130ab5e544d9db217247cb19e75fdc2e8d733ef117d9b92c

Observation cdb7173f-a781-48aa-8211-75e898235a4a · outbound

This paper cites Baichuan-omni-1.5 technical report,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Baichuan-omni-1.5 technical report,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.796923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.408621Z digest=sha256:6b56ca0f8f64cbbaf22bb298fcb40e91a0065f265f44eeb660e2eddf1d3b68ad

Observation 000ff57d-2b87-4200-8928-eda716400dae · outbound

This paper cites Salmonn-omni: A codec-free LLM for full-duplex speech understanding and generation,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Salmonn-omni: A codec-free LLM for full-duplex speech understanding and generation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.782628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.412961Z digest=sha256:f4c531c2560903e289ed851643dc7243816d3d964144023b61e3de3325a4c9b7

Observation d6181bd6-0b70-41ca-a7aa-c952f014718c · outbound

This paper cites Minmo: A multimodal large language model for seamless voice interaction,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Minmo: A multimodal large language model for seamless voice interaction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.768963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.417459Z digest=sha256:be8e3e32d76bcc6bbdd56ff893730d6587ddb5c07f5dcbfd11a3321abac3a121

Observation 9e6aff3e-aff5-4316-b407-9cb0517125a5 · outbound

This paper cites Qwen2.5-omni technical report,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Qwen2.5-omni technical report,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.755597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.421720Z digest=sha256:9cce0f5189eca2a5bede9cf6ff44d43e718029993cbed9ac5cdef7614708815a

Observation b5cd58e1-1896-4480-91a6-0998d2991aa4 · outbound

This paper cites Step-audio: Unified understanding and generation in intelligent speech interaction,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Step-audio: Unified understanding and generation in intelligent speech interaction,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.742003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.426101Z digest=sha256:b52daf29621936b7a826eb14be016909a8ad26757994f64659a9f328208ae890

Observation b5606be8-d984-47f7-be92-d774c481c6c4 · outbound

This paper cites Step-Audio-AQAA: a fully end-to-end expressive large audio language model,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Step-Audio-AQAA: a fully end-to-end expressive large audio language model,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.727510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.430435Z digest=sha256:b5693ad4724ab2b53750031b26ac639c54b60b6fa8df53e7a1a11959fdbbee5b

Observation 7c8a943b-d78f-4531-92a9-a43779cb87a7 · outbound

This paper cites Llama-omni2: Llm-based real-time spoken chatbot with autoregressive streaming speech synthesis,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Llama-omni2: Llm-based real-time spoken chatbot with autoregressive streaming speech synthesis,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.713857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.435030Z digest=sha256:537740fe24fec585d6f0dea6cc032abdb1e231d08e07b368185eed0d7c31fcdc

Observation f7d4abe7-6ffa-4d74-8890-cef3efe68d10 · outbound

This paper cites Deeptalk: Towards seamless and smart speech interaction with adaptive modality-specific moe,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Deeptalk: Towards seamless and smart speech interaction with adaptive modality-specific moe,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.698773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.439127Z digest=sha256:4d7d5b73de03ecfc9083a99a9bae7305a68fef8dacda2b83a9db3f3ef7bfe2cb

Observation 16f671c4-e59e-470c-aac9-dc8dc68bbab7 · outbound

This paper cites BoSS: Beyond-semantic speech,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness BoSS: Beyond-semantic speech,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.685148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.443222Z digest=sha256:8cf3b198630345b871f310b74ff9face82590d8f500527a9cc761baca6551bad

Observation 29a2c081-b8f8-430c-a9cf-a6f7328a828f · outbound

This paper cites V oila: V oice-language foundation models for real-time autonomous interaction and voice roleplay,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness V oila: V oice-language foundation models for real-time autonomous interaction and voice roleplay,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.670291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.447315Z digest=sha256:e637d359a697b7a877f65ae318e1dac4204dce0e77cdb30f770578079f3537cb

Observation 9f8af00c-cc09-4c45-bfea-071d41d386b0 · outbound

This paper cites Kimi-audio technical report,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Kimi-audio technical report,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.657231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.451156Z digest=sha256:c94862d3c33cfe070f867116c7813cf1b8f42c8b2c3961a261b6ad276cb718c7

Observation 0243e4ed-6eda-4fc8-b4ab-4ada7e33f86d · outbound

This paper cites GOAT- TTS: llm-based text-to-speech generation optimized via A dual-branch architecture,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness GOAT- TTS: llm-based text-to-speech generation optimized via A dual-branch architecture,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.643553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.455369Z digest=sha256:5fc9ae179ff36f289e681cd2b816f2c59e57d16610f3a623116096a3bca4afb4

Observation 1bb63ac0-1f4d-412c-a07b-5700418282e0 · outbound

This paper cites TELEV AL: A dynamic benchmark designed for spoken language models in chinese interactive scenarios,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness TELEV AL: A dynamic benchmark designed for spoken language models in chinese interactive scenarios,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.630056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.459831Z digest=sha256:10dbffc1ad019194c1449c1a0bc46c6ca316d495194bae269cd75aad1a927dba

Observation 4b62431f-b158-4d8f-b63a-611371451226 · outbound

This paper cites Qwen2-audio technical report,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Qwen2-audio technical report,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.616374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.464545Z digest=sha256:2723ae832bb741c4461fee7421137c9c30defd7ea29bcf43b399db3a5e2d2716

Observation 4385ee2b-2e76-42bf-8e1e-6c072eb367d2 · outbound

This paper cites Baichuan-Audio: A unified framework for end-to-end speech interaction,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Baichuan-Audio: A unified framework for end-to-end speech interaction,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.601830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.468284Z digest=sha256:41d94373c8049f7d30a8b2b61ec44a4ba5f80151779254115b853730877ea7fe

Observation b846dc80-2c58-42c1-adda-536d03c04710 · outbound

This paper cites Telechat technical report,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Telechat technical report,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.586620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.472468Z digest=sha256:c0f01bd49d8c1208635345874db5b6ede7d94dc3f6674c2fcf80f5946d2a8015

Observation a872b40c-1ecb-4efb-8603-fad8c29bbf85 · outbound

This paper cites Audiochatllama: Towards general-purpose speech abilities for llms,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Audiochatllama: Towards general-purpose speech abilities for llms,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.572007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.476480Z digest=sha256:1a61996c0708d3534d539997023c12339771818c00d750edf16bac5d2ddaf123

Observation f5806a61-abdf-41cb-b52c-e4bb51384651 · outbound

This paper cites BLSP: bootstrapping language-speech pre-training via behavior alignment of continuation writing,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness BLSP: bootstrapping language-speech pre-training via behavior alignment of continuation writing,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.557553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.480339Z digest=sha256:a7531430383034e174d33585641bff72fb43265f07e9d423294b111bdc40bd85

Observation 6db03b0e-9444-4583-a2ec-65646a9a624d · outbound

This paper cites Wav2prompt: End-to-end speech prompt learning and task-based fine-tuning for text-based llms,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness Wav2prompt: End-to-end speech prompt learning and task-based fine-tuning for text-based llms,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.542387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.484253Z digest=sha256:1502a0f28280f8b4ccc48ca1b7c299901010e3999db8302fc97ba29d15b56153

Observation 3bac4abc-912c-46b0-87bb-31a4d20a8f0f · outbound

This paper cites DeSTA2: Developing instruction-following speech language model without speech instruction-tuning data,.

GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness DeSTA2: Developing instruction-following speech language model without speech instruction-tuning data,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:41:57.527293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-06T14:41:57.488433Z digest=sha256:52a33ba6351704382dfc73aa11a82aac951ab4f519d5b04f68e318cb44b5f99f

Pith citing papers

Observation 114b1b5d-ca08-4e4f-99be-214cde165239 · inbound

BridgeTA: Bridging the Representation Gap in Knowledge Distillation via Teacher Assistant for Bird's Eye View Map Segmentation cites this paper.

BridgeTA: Bridging the Representation Gap in Knowledge Distillation via Teacher Assistant for Bird's Eye View Map Segmentation GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness

Reference 2008

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:59:47.953809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T20:59:43.785209Z digest=sha256:65d2c8f3a6a5a18b3dc2027f73209c5ac3d548b6ccb955fa2f098f6a6a53fa22

Observation 5affa23d-c094-4c16-896c-d006ce8796cc · inbound

OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue cites this paper.

OSUM-EChat: Enhancing End-to-End Empathetic Spoken Chatbot via Understanding-Driven Spoken Dialogue GOAT-SLM: A Spoken Language Model with Paralinguistic and Speaker Characteristic Awareness

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-05T21:03:09.632009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:03:09.632009Z digest=sha256:26ebe37531962fab9caef4eadfc08bc316c31094e9b6d99689249d54b52b1a09