Pith. sign in

Paper Citation Record · LEDGER

InfinityHuman: Towards Long-Term Audio-Driven Human

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 4 inbound Pith citation observations for arXiv:2508.20210.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20210 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:19:37.295188Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:13:11.504526Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:45:40.337791Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc0a63cf-7900-428f-a39b-d6c8e77fb643 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

InfinityHuman: Towards Long-Term Audio-Driven Human , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.050321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.050321Z digest=sha256:0fea2dd60d0e541d0b927407b105dfecf4fd0a19400d83bb28f9673e0a07b5e5

Observation 6505ceef-5c88-4522-8425-d8ed5a63c410 · outbound

This paper cites write newline.

InfinityHuman: Towards Long-Term Audio-Driven Human write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.056293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.056293Z digest=sha256:64abed3f72eb00786f12436ae27b892a96d3baae15cba7bc216d246b3affb995

Observation 271442c1-62b2-4a68-aea6-e7ae265cbd28 · outbound

This paper cites MAGI-1: Autoregressive Video Generation at Scale.

InfinityHuman: Towards Long-Term Audio-Driven Human MAGI-1: Autoregressive Video Generation at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.062536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.062536Z digest=sha256:d864428a93f2e27fe6dfa9d856c69a3fa65aa78bebbc599f0a217be2a0215836

Observation 36bc65c8-c9c2-4d2d-af15-bb7aad95c0f7 · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

InfinityHuman: Towards Long-Term Audio-Driven Human Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.068815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.068815Z digest=sha256:6e26cdd12ee9fe637e6b8a06f3c6ea8c9267d938e089a75d3aeac7fcb626a89a

Observation 0ed0deee-0350-4de7-94b2-808957e9e457 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.248653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.073887Z digest=sha256:56ab060527572d797cca0e0075a9670fa11c7bf37b62f40579a1ed25b19e26d6

Observation 23e80ff6-605b-4c51-b118-967cc6ce4aab · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.231867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.078799Z digest=sha256:1483c986d393c54d2302126e3de4b6c852332a6fb9a4a39085dd7f23683f6ef2

Observation 6ba1d3e8-553d-4672-8be4-43b9c8007171 · outbound

This paper cites E.; Fang, Y.; Lee, H.-Y.; Ren, J.; Yang, M.-H.; et al.

InfinityHuman: Towards Long-Term Audio-Driven Human E.; Fang, Y.; Lee, H.-Y.; Ren, J.; Yang, M.-H.; et al

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.215824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.083837Z digest=sha256:2c58403e023eeb69fb6cd1db79e84e696dedfae28d3fa947103a4c04d16feddf

Observation 69e9fe5b-c627-4661-b3b4-3e72b6611d20 · outbound

This paper cites HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters.

InfinityHuman: Towards Long-Term Audio-Driven Human HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.088684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.088684Z digest=sha256:666d08ef499ad1efa1782bd80f1d50a633eddbe6f456c1dbe8a695794056955c

Observation af6ce6b3-4224-4969-821f-bcbbb626aca4 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.199900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.094997Z digest=sha256:f2484968d0bb84b13de6d83fb7ba19f1040b6e00a205f035fae245a304efaec2

Observation 038f89b7-c932-4c4b-8e8e-2c3d2d0df5ea · outbound

This paper cites S.; and Zisserman, A.

InfinityHuman: Towards Long-Term Audio-Driven Human S.; and Zisserman, A

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.183644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.100643Z digest=sha256:1d6a4675ec0c78d5e16c3a2e51bef85b0f4798f274f99428f70f2ef9aecdb3e5

Observation c85cec5c-d3de-468e-b6cd-8fe99dcf3647 · outbound

This paper cites Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer.

InfinityHuman: Towards Long-Term Audio-Driven Human Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.110895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.110895Z digest=sha256:c94325521883176387b63dd95687617337a48fa280d3457415eff9c13f012915

Observation 2a2ed21f-c783-4b36-8408-d28c9f114b26 · outbound

This paper cites HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.115944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.115944Z digest=sha256:669fc765de76b8d67f5ef9098a06ee887b26eb4a241e25db5b87ee5786177f3d

Observation fde95110-063c-4cbc-b40e-69b6b689c989 · outbound

This paper cites OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.120737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.120737Z digest=sha256:f33bd315680e976998ff239dda47ecc3b0ea03bccf0fe959c8ff28d82ed62031

Observation be44d807-5be1-4d44-ab87-3bdd660a9c7b · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

InfinityHuman: Towards Long-Term Audio-Driven Human StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.125808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.125808Z digest=sha256:1f31a83932b55328af98e79c448690dfba3b9304bba2d72ce27c71288450e3f6

Observation 5d6bbbce-33c5-4d1f-a207-93c7d5aa1711 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.130452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.130452Z digest=sha256:7e04d12cbb83c871cab30fd258438ec3f35dd806730e17001e883a6bdc674491

Observation 74d7fde3-0a2b-426a-b704-d1d1e071aba8 · outbound

This paper cites Classifier-Free Diffusion Guidance.

InfinityHuman: Towards Long-Term Audio-Driven Human Classifier-Free Diffusion Guidance

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.134859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.134859Z digest=sha256:92b8f128c53c48f3237c5456e79d363a0ad069015c0e3e9419622b58ff2f5c31

Observation 2f00f6c6-edfe-4e6b-a774-de02a241b6e2 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.156136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.139672Z digest=sha256:326ef7348b34fb47c0a958d27478a170ee0a5dd049211daab2ebdb3ae9b4d045

Observation 16202775-438d-4101-9eef-a2e57bc02574 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.144059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.144059Z digest=sha256:000b0879c9104ab788e1ed2db0c638f2d3b53012f273a7c308f1de88b1d566c0

Observation 8ef05fc4-daad-43b3-84d3-de7da141afac · outbound

This paper cites ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving.

InfinityHuman: Towards Long-Term Audio-Driven Human ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.149071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.149071Z digest=sha256:a35e8724f08a1b406db1024af1c967f3831ecf10545672e7180e27ac862b9949

Observation 146a4c50-26b6-4f83-a965-e321afbfb3c1 · outbound

This paper cites Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency.

InfinityHuman: Towards Long-Term Audio-Driven Human Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.153799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.153799Z digest=sha256:b814f8b7165c1ad484bbd9eecb2412fad1dfb0d1c3b768acd27594b335d48dc8

Observation 469de640-625a-4dff-b7be-60a476dd0985 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.140059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.158574Z digest=sha256:c32db1bd09533fac7c167900a7f0d9ce0cb08fd6254a832de07b359186db997a

Observation aefc38d3-8410-471e-ae5e-979f2c89fec7 · outbound

This paper cites Sapiens: Foundation for Human Vision Models.

InfinityHuman: Towards Long-Term Audio-Driven Human Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.163455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.163455Z digest=sha256:970444ac3b04a77d8a82f768f17d22c287af3bd684897436f735be5d2b29683d

Observation 2b649be8-018a-4de3-b28b-05f9b18454b2 · outbound

This paper cites Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.168237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.168237Z digest=sha256:53fae6decc4c5aa64b1576da13ca87098b15b848eebfa7ae21518236975844c0

Observation 3022cb00-cc3f-403b-b724-35d80d10d2ff · outbound

This paper cites CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention.

InfinityHuman: Towards Long-Term Audio-Driven Human CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.173088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.173088Z digest=sha256:21f631b874b3d246158f7c2874a282aefd4a3d2ba1b5b859ecd12cb291a6e2f6

Observation efb1e719-f16c-46fd-9d6d-61eceaba2a69 · outbound

This paper cites OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models.

InfinityHuman: Towards Long-Term Audio-Driven Human OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.178523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.178523Z digest=sha256:ebca06286315b1daae49aa204e7633d4b74112525ddf746f9bbb63e0731b653a

Observation 20f435bd-b508-4dfa-addd-d92d80bb6b1a · outbound

This paper cites Flow Matching for Generative Modeling.

InfinityHuman: Towards Long-Term Audio-Driven Human Flow Matching for Generative Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.183577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.183577Z digest=sha256:eeaea7457dafd304112b3f3c765339d0bd19a637479b89d7924758e02f9a7ac5

Observation 202e6d05-139c-4a7a-b6a7-0f719cdf91d8 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.188569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.188569Z digest=sha256:b5bae8d0a86f210110bf320888dd874e1e9abb39a16e1aa547ffe2a44e13b798

Observation 2718cb59-cc64-459f-b56e-103e47549e3a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.125171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.193017Z digest=sha256:f6545f746263002ee5e4a54ab0adb47d1da30b760a5da3216b761555f734f16d

Observation d7d657ff-ebc3-4af5-8c10-adf6cf5f7986 · outbound

This paper cites FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling.

InfinityHuman: Towards Long-Term Audio-Driven Human FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.197844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.197844Z digest=sha256:ee4608350f2f399a320ce81e3cf21f46556102d324ffb9188259f7b1f8c5c5e0

Observation 97676dc0-2473-447e-b5e5-e0efa5e7aeac · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.109624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.203126Z digest=sha256:f7dbfd15d069a531fc1ad5d2b8c4d0dd061286126f7f83d9c9913cfa1d5cef7c

Observation c1f3cd11-6ade-422c-b551-322c9f20e18a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.093372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.209356Z digest=sha256:313b8843fc72b5847c39a76b7a61ecc49edc34055d9f1ad503d18958233764a3

Observation d34993e8-472d-4343-8392-b985e2748253 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

InfinityHuman: Towards Long-Term Audio-Driven Human Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.213971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.213971Z digest=sha256:c718e9f498c1909a1fc9d641093874f885e28a11278ffbf46ad77a3b49733d6b

Observation 272fa9aa-e990-4612-add2-f17b4de8a0d7 · outbound

This paper cites V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.223978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.223978Z digest=sha256:f05f4c249d6eb12b1bc4689d5ab1d1252d8e3ecdb08260099800e04a079b0adb

Observation d25195cc-2719-4df0-8b08-247ecd523f08 · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

InfinityHuman: Towards Long-Term Audio-Driven Human Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.229734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.229734Z digest=sha256:a3c003f901df1e75dc5a2105c0d86e64c42e7e941af31d11483750b7afba2785

Observation 62331e57-43f5-4b1b-a284-cc7be6a613b5 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.077176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.234878Z digest=sha256:810cd9c6dac8a73e03cf9adda291ae06380beb26abc87cd0278236b511d0e37f

Observation 5cf5dd96-bd4c-4fee-82a1-fe661de31133 · outbound

This paper cites FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis.

InfinityHuman: Towards Long-Term Audio-Driven Human FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.239985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.239985Z digest=sha256:bf27472d80d863869d34160b83b8a3eec05b78793ab731a8106a871e5470b2e8

Observation 2728a245-1ee2-4b8b-9c40-ab9ca16ade0a · outbound

This paper cites AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.244894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.244894Z digest=sha256:aacececff1407ff8c759b743f7508f471c795d8bb2badac81ee97c0205d73991

Observation f037a761-4b70-49a6-82b6-336aaf4a3a45 · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

InfinityHuman: Towards Long-Term Audio-Driven Human Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.250325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.250325Z digest=sha256:39f46a18bb66af9540c8a04f8d44843253d8c1dc84bc49af0d7fb73d515dbf6e

Observation 38bb490f-c1fa-4a8d-83aa-6c91a5e74898 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.255149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.255149Z digest=sha256:e21b2334312b398e5f2398b9c8666308b72669080513becbc9d3dcdf1e770817

Observation 62b552ec-c597-4ea0-8810-5b3f829d2845 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.259854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.259854Z digest=sha256:f0e8ef849e914d84f691cb6cb7259878a30c5b74f8d4f749a00408c85667253c

Observation bc0897b5-7993-4954-b38f-c6b837c17fcf · outbound

This paper cites T.; Durand, F.; Shechtman, E.; and Huang, X.

InfinityHuman: Towards Long-Term Audio-Driven Human T.; Durand, F.; Shechtman, E.; and Huang, X

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.061494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.265364Z digest=sha256:d626f77fdd347762c8e9ba12ced8f7d226f4dfb665a39f4fc468734d4fe0eb76

Observation aa22d217-42a1-4175-86d8-d5c0af9172a9 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.045258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.270656Z digest=sha256:48bc4adc8e294a924bae8f2e6198f0bbfcf098a1bfbc9d4f8bf8608302899f07

Observation 712e5916-70f3-47b7-ab08-bff605e41d50 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.028130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.274971Z digest=sha256:e614551e4f7bf37c96a25ba44770f6fc5feb61e6f957226af35ad98d9eba50fa

Observation 5457650e-b5c7-4784-a037-f3279b68171a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.013438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.280329Z digest=sha256:46e4b75cb3e883b91a0008312be8e3256fbde63ad212d332a659ef3f32e1365c

Observation 72e7c357-8f3a-4075-abd4-21cdf2c8c831 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.995922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.285521Z digest=sha256:457c9705d940172f05dd65790195d5bb177319671cea2adfed75fa9e844a7d5f

Observation 5a86f957-c397-499e-a732-7b44102fb00f · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.979057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.289764Z digest=sha256:7d8bff7640c84770e937845e0088e9ff0feefcf44eb44f8120fa45a4a7146bfc

Observation 7ff90854-b684-44a0-9008-1925e1374859 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.961578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.295188Z digest=sha256:565156e7c3ad9ceb55568c3f99231080d516f9137e3c1c9e6b4a9df5c7e68445

Pith citing papers

Observation 8cc9086a-b1ed-47f3-9a91-3f6e7527df13 · inbound

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body cites this paper.

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:08:36.256770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T22:04:07.403410Z digest=sha256:58229be574be7e8e49430d8651ae5a07da1d2b6578ffef908aedffefbb30e07b

Observation b2c3cf4f-83f6-44c6-af86-5a5b00ca820b · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.355150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:d67b82cc44688b7a15c8ead59754440c56a6055feb9f583724d471b579800751

Observation ea9b4d82-3dad-4742-a484-16aae52bc771 · inbound

Generate Your Talking Avatar from Video Reference cites this paper.

Generate Your Talking Avatar from Video Reference InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:29.794020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-07T05:32:04.519820Z digest=sha256:28eff2431a7572e23a60298f062c1638129cd93bbefdf81ed35148ba9e4b02df

Observation 2200ada8-547f-435e-8314-93b52070d8c4 · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:40.339046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:57d3e14dce43f2dce7c8c7987e2d842eb937214d97de861d339c47862f50c221