Pith. sign in

Paper Citation Record · LEDGER

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles

As of 13 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2501.03181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03181 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:41:36.351183Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0cd0b04b-9cb8-4660-a2ff-e5416f9b7ad3 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:37.087147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.132625Z digest=sha256:218687e4194a4d327c5f174abc51c6baf876e8fc8c8f5a28757a0e98212ac6b2

Observation 4c1c9f97-fd41-4ca4-afa3-388462644bff · outbound

This paper cites D.; Junior, A.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles D.; Junior, A

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:37.072363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.138383Z digest=sha256:ff1382d5edd2fe0187ddfd42ce9c0462226cf738842bd1225fc7c234b87f0be6

Observation e8f132ba-2ff2-4845-8880-75f30ed24922 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:37.057806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.143911Z digest=sha256:0a3ffd7d9765ada191fe48a15d07f326fcd5872a65dac347dabb468099cdd23f

Observation bc795638-fa69-4bdf-84c3-7e82450275dc · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:37.042970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.148931Z digest=sha256:44b6c7c66998f6739a73a6e81b0fc8174df2f6a248652525bbe3bcea691eccdd

Observation 370281cf-ac18-4132-aa8d-03d4df1aba5a · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:37.028124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.153441Z digest=sha256:459a8325f9603f3797576e81994d1f5b278af6c9e21b3e390722b3f5649ac103

Observation d45c52d2-4315-4401-9aeb-9f65a73f683e · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:37.012214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.157815Z digest=sha256:7eecd88348720ef4a766451932bce2a2f99cf798974b1f643c39c76c0e5d4384

Observation e569554e-38c3-4e38-827f-78b80cae1557 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.996987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.163087Z digest=sha256:bb8b5a2369029ad3708541ac3cbbf80f172370955e3cc0ec74c8e5766d9657f2

Observation 7a9b4400-df5a-4ef6-93c1-7c2aa3b52999 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.982446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.167444Z digest=sha256:89412388306583b71ca6383353185d1bee32b7a06bb8819ed34cf50f766e8c51

Observation 213bf79e-889b-48ba-8ad5-1488e3bb7316 · outbound

This paper cites J.; Laver, J.; and Gibbon, F.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles J.; Laver, J.; and Gibbon, F

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.965738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.171829Z digest=sha256:9ed7098dd5cd8d3db4b8477c694bfc4a7da8d4170219f659de3a18fcf0913e57

Observation a7d28ccc-4a4a-4d16-99fa-691fa88b482d · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.950518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.176993Z digest=sha256:bbad8b30e405e66bc5e53e06de4d269d01bd8e68e1ae655881fe78a9c3de8759

Observation 6a420c99-2087-4b50-9afe-01b93b56e2f1 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.935990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.181437Z digest=sha256:7c9d95bed9dbf265ba01be84b2f1103709319472ae0b730554b45294162d6118

Observation 2e561cb2-75f7-4f63-8ced-106c91d58529 · outbound

This paper cites Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.186653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.186653Z digest=sha256:d9d6c428ee4ba1a6f74ed58eb8b479980be07ac58786df16b217bddc24e063ce

Observation 8ad4b229-10ea-46f5-92ff-4cefb56dd5e2 · outbound

This paper cites VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.192653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.192653Z digest=sha256:8525ef43b53a42c049622898a555bcbf4cfe73fcff8dddb12a790fbe14d68b4c

Observation f9ecc135-25ee-4f94-8f73-49566124acec · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.920130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.197715Z digest=sha256:6feded834ea099416ecf8abfefbbebddd0b9f1d2d40dee1280bcba925d9beb09

Observation d0140481-2f20-4b52-9968-c7e144247b2b · outbound

This paper cites S.; et al.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles S.; et al

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.905855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.201693Z digest=sha256:0d1e2a8b125a54847760dd8f9e00f169d35aed2e930d9414753f66d5fbd31a82

Observation 1b1eccc1-a6be-45da-9578-c46b99003b11 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.891333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.205911Z digest=sha256:91d4ee440b607c9c451f187aed25c05af320d1d4a9a6186e02186d7cb6d7ef00

Observation 6443626d-8eb3-4468-a1fe-b7a79294c5a5 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.877211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.210724Z digest=sha256:15d8ccba7f8d3a7b27da9170d7d08c9acd74e461e4a19e975e1cb4a59db5b63e

Observation 6490c36d-b69b-4771-981e-fa10fb891513 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.862626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.216138Z digest=sha256:61d362dd9acf95364368927485f72a19c391cabf270fcf4e4cb6ffde6784b4a8

Observation e649a9df-0a16-46cb-87fc-44e52961b477 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.848562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.220509Z digest=sha256:e716fcb4c66ee1f5fc7e741cf7bc4c6f39e5ba42ed40a550cbecd40c394bd987

Observation bca8b896-fcb0-4077-a98a-d51a3a8e00c2 · outbound

This paper cites PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles PhotoMaker: Customizing Realistic Human Photos via Stacked ID Embedding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.226840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.226840Z digest=sha256:2015383ee1e7d8d91ac19c5921f4f588722c50c7f2663ae08bcbf654907033df

Observation 9b4e0f90-4d39-4ac6-a5a7-633f4cf1fbe4 · outbound

This paper cites R.; and Russo, F.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles R.; and Russo, F

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.835037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.231284Z digest=sha256:b8c3df0507a17d783ea2b8d31698a71f5c948fcbd9e8282e18d5626caa12c075

Observation a7e7f835-91ae-4a6b-991e-62694e2db416 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.820369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.235359Z digest=sha256:e2f3e48aef5609b81beeb7826e28471f0f4130b37075041f26538bc063b14ef3

Observation 9e01b7ae-2aa8-4ac2-8300-9277a1d020f2 · outbound

This paper cites B.; Yang, E.; and Hwang, S.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles B.; Yang, E.; and Hwang, S

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.805787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.239468Z digest=sha256:cb9e467c888a367aa15ebb4ed6e1a774faf5e7ab5913f0a1af6436ac8cc330d0

Observation e7b37353-40e3-47e5-ae16-97ff7acc4002 · outbound

This paper cites S.; Morerio, P.; Mahmood, A.; et al.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles S.; Morerio, P.; Mahmood, A.; et al

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.790428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.245095Z digest=sha256:6383c7fe3a28b61d05601272e10c126c25f3bf340e32445cfc8c16014553c182

Observation 859b9f78-d0d5-48ab-a872-84b4d20ea4e9 · outbound

This paper cites EXPRESSO: A Benchmark and Analysis of Discrete Expressive Speech Resynthesis.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles EXPRESSO: A Benchmark and Analysis of Discrete Expressive Speech Resynthesis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.249771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.249771Z digest=sha256:25859b78fc1096b5ee3cf62e2630497438642fe595c75629194ebc515a4812a7

Observation 83d2d1da-cc32-4446-8d57-a3a8c9fcd4e3 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.776154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.254893Z digest=sha256:bf5de8f5e477a40de0c4436744d5c90f497162ab641aceed67e5b36c4167de6c

Observation ff7b5364-6e67-41e2-8e14-cf6fd77dfee1 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.761405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.259539Z digest=sha256:23afb87902c63303bd45e3ba6179f64b5d7a723a6e6ad22155626a4dac482cdc

Observation 68c217f8-b477-4d00-a6f6-1b230cea27e9 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.264122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.264122Z digest=sha256:db3aa5301f814ff71fef0466a981630661365fd0e654d90e985533f38e9f6006

Observation b92f1731-c817-423a-abb2-7130a389c4d2 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.746789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.269217Z digest=sha256:296714addd8ec3bc4add3cf16238ac8601f72ff6d2eaef431a3db02ce7223054

Observation f07d4c8f-263c-4414-b209-fc9300debf63 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.731703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.273603Z digest=sha256:78d5575eeacc73daddd1fff58aeb4e9a370e986b74aa1e09c320088232037dec

Observation 34464685-00c5-40bd-b593-65d44ca4a78c · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.718006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.278152Z digest=sha256:484cbbe9a99ef846343580d289bcd35bf766d04fe75e82db1a03d22f0c255851

Observation 958d0713-330a-48ec-bb17-eb3ef55fe7c1 · outbound

This paper cites G.; Alrashoud, M.; et al.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles G.; Alrashoud, M.; et al

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.701996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.282505Z digest=sha256:9dfa5fcff8950e86d8c29b7e6c370fb4f446de98023d77c305675f99b3a40fbd

Observation 2a9363cf-2b0d-4ab3-805f-e9b0d6d1733f · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.687199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.287114Z digest=sha256:cbc58c4a9a7e50684a7f8b005a225030ab630e57c59c5d1d7786b2b988a9bd5b

Observation 2c025e73-ffbe-4ee1-ba4b-8fabdf13b662 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.672398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.291708Z digest=sha256:6a5f7dc88cfc921c7bbc46946b2e53893bad5bf9967a10c3abdb374faa562cce

Observation ab33fd14-6c68-45f3-8cb5-4f3847c96b45 · outbound

This paper cites FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles FaceChain-ImagineID: Freely Crafting High-Fidelity Diverse Talking Faces from Disentangled Audio

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:41:36.431025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.296090Z digest=sha256:4aa2fbc6cacb70dcf64976751e9ee2fad87f85e914b5d013eb04f9a316f3c67f

Observation 2e150eb5-7ff7-4d11-8889-fdb2aca014de · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.656378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.301064Z digest=sha256:5b3976ff664f67e2029c92af751afc2ea6258fee592234cd2562048dd98a74e6

Observation cf387bbf-4908-448c-aa4d-445cbbf9fb8f · outbound

This paper cites InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles InstructTTS: Modelling Expressive TTS in Discrete Latent Space with Natural Language Style Prompt

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.305336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.305336Z digest=sha256:df900048eacd3700e5be5ac84bc601eaf9fc9aa228e4410008761000ceb68f18

Observation b9fa1489-4802-4ae7-84d6-fdd8b9f99561 · outbound

This paper cites B.; Liang, P.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles B.; Liang, P

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:41:36.640536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.309424Z digest=sha256:3edf0b18c3dec4c8da447450852a21dee72a8218f11317e57064e61c09a84546

Observation 571f9b92-c45c-4a6c-9c2e-278b7a7da0ce · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.313840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.313840Z digest=sha256:a3fe20ac48eddcb113601f0841836e8afd28bba8299cac80bf757ba190c7c977

Observation d1210931-c7cd-4225-b7a7-4a6eeb7ad5af · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.624566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.319416Z digest=sha256:4402e89865ef333e07cb5d9a4c8c7779c83556d07c1eb2a736854b6d4ba4f75e

Observation cad85ad3-c416-4df9-8b5d-b976b5068b48 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.607585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.325033Z digest=sha256:7be68953b61e731a180c2925f6f7393ba9919bbd75a91efea5d6e709bfb9c415

Observation c8c0ec08-566c-40ae-bb11-0d4a9aa0350a · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.592283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.329492Z digest=sha256:47011ce57fa380f94cc5582335e8e5501da822bc583b845b0e87edb217d32fdc

Observation d835e132-281f-430b-b7f9-8972c6784be9 · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.578244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.333820Z digest=sha256:3c9ddb97f0f828f06c274286909570635e9e76c35971e60bb860b3c0fb8065f7

Observation cb0cc680-f348-47f6-a658-ba9264dd1e5b · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.561815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.338027Z digest=sha256:fce2770ab52fd40c978e1ef2bbcbfb53bd16e97f3424555d5eedc75138ca8b92

Observation b42a7bf9-f3d8-48aa-91fa-894be044462d · outbound

This paper cites an unresolved cited work.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-10T22:41:36.547121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-10T22:41:36.342375Z digest=sha256:d9fa8b4ea6470b7371389ac53d46444b47632ea41014f29cc5cfd4100e612f3b

Observation 5f4b2ccb-86c0-4855-af29-4d57f3e7ba18 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles , " * write output.state after.block = add.period write newline

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.346197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.346197Z digest=sha256:fe7040172218c46f56f281390473af36c2246d543506178082143e280ae92245

Observation 0a573d03-4024-43aa-a807-665f0c373ba4 · outbound

This paper cites write newline.

FaceSpeak: Expressive and High-Quality Speech Synthesis from Human Portraits of Different Styles write newline

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T22:41:36.351183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:41:36.351183Z digest=sha256:0a4a8edd26228e418976ab9c963dc4412092a8cc1864dc029df04d66e1805801

Pith citing papers

No inbound Pith citation observations are available.