Pith. sign in

Paper Citation Record · LEDGER

VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2404.06674.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.06674 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T05:54:18.393875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:39:41.912271Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6092f1ee-f92a-4a94-96ec-7c9a045bf1c2 · inbound

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models cites this paper.

Seed-TTS: A Family of High-Quality Versatile Speech Generation Models VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:26:37.489334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T12:26:37.300599Z digest=sha256:fb60ddf0202312f9516ea0d806ecf23d2d0d87882a87f040fa0fdfd431b130b6

Observation c2974013-dc9e-4bfe-ba11-fa1f12cffff2 · inbound

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training cites this paper.

Metis: A Foundation Speech Generation Model with Masked Generative Pre-training VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T05:54:18.393875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T05:54:18.393875Z digest=sha256:f591058f73eb6c73281591cb2e23874fa4e33dff8d4cad00c6f6f598cccea68c

Observation b4e0abf5-deb8-4310-a6aa-de5bb23ebfa5 · inbound

Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement cites this paper.

Vevo: Controllable Zero-Shot Voice Imitation with Self-Supervised Disentanglement VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T13:26:19.150092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:26:19.150092Z digest=sha256:fb81d015a88b168c67fa21b828e10a58e8b3a396264b78a922dd5973260a6c4c

Observation 7cc95112-ee4b-4565-a734-262055912543 · inbound

SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment cites this paper.

SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:14:07.384296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:14:07.384296Z digest=sha256:e3c5c21f77881a0001c787f0a42e2db8d0d9d5d9cbd70475638e58eed8e34ae4

Observation 6501fa4b-73ab-4f26-90f4-c28890034c58 · inbound

RIVET: Robust Idempotent Voice Attribute Editing cites this paper.

RIVET: Robust Idempotent Voice Attribute Editing VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:49:25.460914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T18:50:54.353334Z digest=sha256:7236b3195d78f5d971fe228ecb08fae9884d68d476c6e4061aaeed1b7b29daf6

Observation 166d11cd-d61f-4147-9f75-5aa46ec3399a · inbound

Encoder-Decoder Manifold Alignment for Idempotent Generation cites this paper.

Encoder-Decoder Manifold Alignment for Idempotent Generation VoiceShop: A Unified Speech-to-Speech Framework for Identity-Preserving Zero-Shot Voice Editing

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:39:41.913850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T11:19:21.734052Z digest=sha256:5909b83608b7e958816dd8783a0f8990a778b0ae198b94157783055722cf1b3b