Pith. sign in

Paper Citation Record · LEDGER

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?

As of 13 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 1 inbound Pith citation observation for arXiv:2506.06756.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06756 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:53:46.177712Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:53:41.686588Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:53:46.575085Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved7
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 619034f8-a4d3-4385-a877-97c322ddd4f0 · outbound

This paper cites Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:53:46.714939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:41.686588Z digest=sha256:da39ec38210d6e264e5f6457110e7bca8c8100ba09e0647ff527ee74e744f67a

Observation befc544c-3856-4c93-8b2d-e616bde42fcd · outbound

This paper cites , xT }, where each frame xi ∈ Rd, T is the total number of frames, and d is the dimensionality of each frame.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? , xT }, where each frame xi ∈ Rd, T is the total number of frames, and d is the dimensionality of each frame

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:53.775356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:41.820304Z digest=sha256:55e5594943c087a78c1fb2411c8ff9003bcf8f380ad44d0fa0ff72cc0ce5cc4e

Observation ab3d14b1-325e-4ec9-b4c0-631604418605 · outbound

This paper cites an unresolved cited work.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Unresolved cited work

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T05:53:53.408576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:42.053609Z digest=sha256:57f173c49a9386a13d98e6a7698a02cf7efd1f3308af6773250c6bf7334dd215

Observation 5ac88124-4318-47b6-8e23-b59ffe1412ee · outbound

This paper cites an unresolved cited work.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:53:53.024690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:42.206192Z digest=sha256:b3dda889d6c463c179087e3a4b9b0522ad061fa102ae055a4d15c8f2941107db

Observation 598dadb5-7ded-4ab6-9dd9-e34a8290dd8d · outbound

This paper cites The authors also gratefully acknowledge the support of IndiaAI and Meta through Srijan: Centre of Excellence for Generative AI.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? The authors also gratefully acknowledge the support of IndiaAI and Meta through Srijan: Centre of Excellence for Generative AI

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:52.775425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:42.392482Z digest=sha256:160b82367495e6db6d09616561cb17918d95be8d56fd84c82695f08761d33481

Observation 65ed964c-16db-4e40-97dc-e2ca37f0cd04 · outbound

This paper cites A survey on speech large language models,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? A survey on speech large language models,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:42.552426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:42.552426Z digest=sha256:525a205d5cbf858b0414db19210fbb432696bce8379182cc0e46683ccb2911d4

Observation 77671e89-220a-4470-ba43-3449a82916bd · outbound

This paper cites Pengi: An audio language model for audio tasks,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Pengi: An audio language model for audio tasks,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:52.399755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:42.756657Z digest=sha256:1f8145ec683318b95ccb1c95d49b90bacf0cce3252404cf7d2d5748d70143148

Observation ca7920ca-83cc-4357-9eeb-7638162454d6 · outbound

This paper cites Listen, think, and understand,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Listen, think, and understand,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:42.917701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:42.917701Z digest=sha256:cd6403b38e3e5a2325a4a4a507972fdfb0d115885c6fdd6259bce8e78c470304

Observation 1ccf13a9-e7fd-4aab-a400-2f012904c747 · outbound

This paper cites Neural codec language models are zero-shot text to speech synthesizers,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Neural codec language models are zero-shot text to speech synthesizers,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:52.046837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:43.059579Z digest=sha256:7b9609baf0d885a3733dd4c0be1c3b839751c87f99164507e2de74e3361c075c

Observation a70db387-9f55-42b0-9d19-982eb978689a · outbound

This paper cites Mobilespeech: A fast and high-fidelity framework for mobile zero-shot text-to- speech,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Mobilespeech: A fast and high-fidelity framework for mobile zero-shot text-to- speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:51.657404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:43.197717Z digest=sha256:51268f61486b063d5d8aa8819d94405a062a02f5f1d340b8d55e717e22f6fdfa

Observation 32c170c8-4a14-4e84-bbe3-78709ccb905f · outbound

This paper cites Can rag- driven enhancements amplify audio llms for low-resource lan- guages?.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Can rag- driven enhancements amplify audio llms for low-resource lan- guages?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:51.306554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:43.338242Z digest=sha256:165edcc81ef1d10a6421840e9f8e99be8de22af0db9273370b5f176b5527b63a

Observation 6dc123cf-8fa7-46cc-94c1-c9d7f1a98db5 · outbound

This paper cites A White Paper on Neural Network Quantization.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? A White Paper on Neural Network Quantization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:43.538083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:43.538083Z digest=sha256:fbfec1564363ba5239233faa529df7db6723815be9307d62df8dcf840778f1a1

Observation d24e1cdc-a9a6-476b-9ce6-ddc7c8faee7f · outbound

This paper cites Sv-deit: Speaker verifica- tion with deitcap spoofing detection,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Sv-deit: Speaker verifica- tion with deitcap spoofing detection,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:50.922281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:43.692101Z digest=sha256:3574cfdff24b47f665f343ae10c795a1544155c957688a629b3029a978a43570

Observation c17f2288-a9ee-4e5d-ba25-0e9aee2b4d45 · outbound

This paper cites Faking fluent: Un- veiling the achilles’ heel of multilingual deepfake detection,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Faking fluent: Un- veiling the achilles’ heel of multilingual deepfake detection,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:50.535046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:43.839750Z digest=sha256:97da54c72388985c902a5578a3b72c697e54abbbd91bfa13e4bff2277c66127b

Observation 0e6ed026-db15-44d2-8ac6-c656b1418ab9 · outbound

This paper cites Context encoded multi-modal attention network for detecting audio spoofing,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Context encoded multi-modal attention network for detecting audio spoofing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:50.192227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:44.011055Z digest=sha256:92409d9c9d5443f7c171837540ca1d6df041afc9b802572baf41f7cc3f4975ee

Observation a26f2131-a477-4e15-84f1-3345db629dac · outbound

This paper cites GAMA: A large audio-language model with advanced audio understanding and complex reasoning abilities,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? GAMA: A large audio-language model with advanced audio understanding and complex reasoning abilities,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:49.886715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:44.215530Z digest=sha256:34327f188f02043280b5c27cc5af8f5c5450434bc0a2bcb9de71a1371ac5df4c

Observation 973e30d4-adff-4e7b-8470-f14e4fe45fb9 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? LLaMA: Open and Efficient Foundation Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:44.376678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:44.376678Z digest=sha256:58ceef2951714d5736b2cc86029e4897859bc73693ec3f48d9bc2472d6c16392

Observation 96c5930e-a4f0-4934-a6ad-19a6570dcaad · outbound

This paper cites Joint audio and speech understanding,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Joint audio and speech understanding,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:49.537664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:44.580432Z digest=sha256:fd748a5557604afd58159432bab05a3935fe3c7f4ef6bb7843185d18b0f46094

Observation 62632fbf-3ca5-48e1-9c5c-321b350e58a4 · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Robust speech recognition via large-scale weak su- pervision,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:49.189022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:44.729191Z digest=sha256:a3b449f20bd935e4a6eeb28e286076e2eaeaf01bb4373a27a59e6a657331cc2d

Observation ea9b714b-767a-4334-ba29-1bd3b7e6cb14 · outbound

This paper cites MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? MERaLiON-AudioLLM: Bridging Audio and Language with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:44.932527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:44.932527Z digest=sha256:90f453f8cadd1281a215c74666f0a8ed22749207e6f81adc2961ed3f6443bc68

Observation ba26a92e-fbcc-4d97-8043-21e6e3f0c0eb · outbound

This paper cites Sea-lion (southeast asian languages in one net- work): A family of large language models for southeast asia,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Sea-lion (southeast asian languages in one net- work): A family of large language models for southeast asia,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:48.885947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:45.065096Z digest=sha256:fdf8d414efe4838a51c441130ccf3c724f75df0949fe2400ca259f70db23c5b9

Observation a047bf3c-cd4e-4e5d-a4f9-7df2d0718f40 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:53:45.230382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:53:45.230382Z digest=sha256:a3c372a853aa2c8ea1805aaafce0ce1a9633f19f00c23a618db9fcfce4d8d270

Observation c2e97566-4099-4ab9-8ebe-24ca57a30d7d · outbound

This paper cites SALMONN: towards generic hearing abilities for large language models,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? SALMONN: towards generic hearing abilities for large language models,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:48.506318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:45.391416Z digest=sha256:e31fdd001beb62ff3b97a8db78e2a46806be55734cef123bb4fda9abfb80bb6a

Observation 67d529dc-35b7-4f9d-a12c-fae02b9b6bd9 · outbound

This paper cites Vicuna: An open-source chatbot impress- ing gpt-4 with 90%* chatgpt quality,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Vicuna: An open-source chatbot impress- ing gpt-4 with 90%* chatgpt quality,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:48.105285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:45.595885Z digest=sha256:7348394df5bc3cc61ac37438cded212daa720c80f88daab97c1b088282e2ab10

Observation 9cee30bb-7b37-41d0-ab5e-78f52075e62d · outbound

This paper cites Asvspoof 2019: Spoofing countermeasures for the detection of synthesized, converted and replayed speech,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Asvspoof 2019: Spoofing countermeasures for the detection of synthesized, converted and replayed speech,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:47.756321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:45.744556Z digest=sha256:8a6d98d4e56ff5a0adc80d3cedf6b03c86054490665a059ac45029d6be275f7b

Observation 5d189bd7-4b2e-4152-9824-c8b07c066f67 · outbound

This paper cites Does audio deepfake detection generalize?.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Does audio deepfake detection generalize?

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:47.455935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:45.943188Z digest=sha256:fbb75affd2e47dfa1163e6572d7a154fb0bb60a2ba7facc3f6f391cb9f8368d4

Observation 845c7959-fd43-4df2-bc3f-1b4adb6d35b7 · outbound

This paper cites Wavefake: A data set to facilitate audio deepfake detection,.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Wavefake: A data set to facilitate audio deepfake detection,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:53:47.164985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:46.177712Z digest=sha256:a81c901e8cddea7f53c2da4cdfc354bcc2a064f937c9fe50a1fcaa4855387fee

Pith citing papers

Observation 619034f8-a4d3-4385-a877-97c322ddd4f0 · inbound

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? cites this paper.

Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection? Can Quantized Audio Language Models Perform Zero-Shot Spoofing Detection?

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:53:46.714939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T05:53:41.686588Z digest=sha256:da39ec38210d6e264e5f6457110e7bca8c8100ba09e0647ff527ee74e744f67a