Pith. sign in

Paper Citation Record · LEDGER

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

As of 14 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2412.02419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02419 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:34:10.895977Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:16:44.323653Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:16:44.447784Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy60
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ca2d14f-68e6-416e-8beb-072713cd51e4 · outbound

This paper cites Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.790910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.605320Z digest=sha256:79fceda3ece8b0629d6a92ef50244b5844d3bac0a74de96272e6577e69726eba

Observation 5915c4ce-a3a8-4a78-9186-f1c5c2c6e941 · outbound

This paper cites Style-controllable speech-driven gesture synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style-controllable speech-driven gesture synthesis using normalising flows

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.780191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.609753Z digest=sha256:336847e43b269632f0d12d0db9c9f6272ca84c3e6299fdbbd47fe17c1b3f4bc7

Observation 23fa6aa1-13a1-4474-81cc-37bca8c711a9 · outbound

This paper cites Listen, denoise, action! audio-driven motion synthesis with diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Listen, denoise, action! audio-driven motion synthesis with diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.769829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.613798Z digest=sha256:f782faac188e8dc37ec3c462d3b721256e09e6af439e8d78120f2cb610e2cdfd

Observation 56da8f1c-79f6-436a-be1c-d6fb598e5c18 · outbound

This paper cites Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.758838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.618075Z digest=sha256:d1cea7c502c64ab3a3e5e4acbae40f1cc201e2149d35bc2c50567d81c47f132e

Observation 20fae4cf-0e19-4bf7-a8e5-5b3370613036 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.747711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.622171Z digest=sha256:bd2baa7d38bc2bafdcbad61c46cf90c92aa947762754ccc0edfe03ace7616f2a

Observation 8cb414e2-3ad3-4685-80cf-eba842005ca4 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.736630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.626457Z digest=sha256:7d9f4c6430c0286faebb3ab0ff66392c3fa486dd86050203f2deaf471b53532a

Observation b397dc9e-a53d-4c1d-9d7f-a9140d0dd09e · outbound

This paper cites Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.725245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.630449Z digest=sha256:b3c9e65d58e0bd5ffbcda6ed98f2dc4497cac2a5ebf4ad748e385959991c8267

Observation 6431bad8-6118-4472-876a-11ca2642cacd · outbound

This paper cites Digital life project: Autonomous 3d characters with social intelligence.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Digital life project: Autonomous 3d characters with social intelligence

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.711947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.634808Z digest=sha256:5d9a74057a2c70d55b1945d32ec10dced9a6ca9cbce24d5bfe77703cf44ab739

Observation 7e103cd7-d7ae-428b-84aa-57f58c0f276f · outbound

This paper cites Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.700250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.638849Z digest=sha256:8de194d2295e1ae059b8656840029fd30c47035d86c5d5e7372e26622868cfa2

Observation cfb72b07-3375-4e42-8bd1-870df924531b · outbound

This paper cites Beat: the behavior expression animation toolkit.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Beat: the behavior expression animation toolkit

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.687482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.642619Z digest=sha256:42dbfd3f30123b82a49256ed1f67aeb4def74128b25cddc053f5be5a38931665

Observation 85d61874-e5c6-4922-96d3-b600001ea7ed · outbound

This paper cites Taming diffusion probabilistic mod- els for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Taming diffusion probabilistic mod- els for character control

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.674539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.646002Z digest=sha256:7b9e671321bc971083ec0e72dc06f54f00cc44009fbf18c9500d6e8f6a88383d

Observation 02990db2-beda-4498-86b3-5acf292fe63c · outbound

This paper cites Executing your commands via motion diffusion in latent space.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Executing your commands via motion diffusion in latent space

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.662865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.649538Z digest=sha256:b146f6e37436fb4cc5dcd05190ffa57ffb27349ab77a47e74a51cf8fb758969b

Observation 3999d157-0ad4-48c8-8792-b022fbee905e · outbound

This paper cites Black, and Timo Bolkart.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.651114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.652577Z digest=sha256:005d6bfdc850881c01b79d77bf1ebfd6b5f13c6f0e31ed00dfa46cc76eb071bc

Observation d942fe81-7f19-4eeb-8422-60fc06bfc535 · outbound

This paper cites The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.639164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.655939Z digest=sha256:2ff0a659d0d5fb3e81daf21cbbab3e299ac79611aebe87bd629110cd24f4e5e3

Observation 8d99c66d-a64f-4890-9654-6d40300cbef8 · outbound

This paper cites Diffusion-based co-speech gesture genera- tion using joint text and audio representation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Diffusion-based co-speech gesture genera- tion using joint text and audio representation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.658853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.658853Z digest=sha256:08032fef5a661ff388ef57f2420bafac2a0af380e191c55d024eb0a5e9740a4d

Observation 3abea151-4294-4c25-b01d-f7eaf2b334c2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.662380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.662380Z digest=sha256:ab14f1fa2cc0909e9e56e6331534f0f30d5fd12b8d5b4e7f7c2985820da932d7

Observation d96f74c4-816d-45e0-8b73-21673e297df5 · outbound

This paper cites Freemotion: A unified framework for number- free text-to-motion synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Freemotion: A unified framework for number- free text-to-motion synthesis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.618609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.666246Z digest=sha256:7a199bd51b57a67c56e55b803741c6f49da616d8c49398ef75f3de1402cb4a43

Observation 1a8b866c-d5e1-4e38-bcd5-f656c1de7739 · outbound

This paper cites Troje, and Marc-Andr ´e Carbonneau.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Troje, and Marc-Andr ´e Carbonneau

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.606790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.669724Z digest=sha256:6c8b03c1ca6aae1f05780dcc5dc41514193b14241428b1a7ae40c89fefe5107a

Observation 41df565f-9d18-483f-93ee-3c04dfccf4db · outbound

This paper cites Remos: 3d motion- conditioned reaction synthesis for two-person interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Remos: 3d motion- conditioned reaction synthesis for two-person interactions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.595162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.673292Z digest=sha256:f3517587df5ce43fe4e603d6d7746c6c760322ddc3dfdc5c30d3e9aef331125a

Observation be8d6683-299f-4191-ad28-cc5dbb6023c7 · outbound

This paper cites Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.582801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.677265Z digest=sha256:0a86be8229c82baa3fee77d3140ab8055219675c45d3bcf48662ae5ae9f1ccd9

Observation f65e4349-e081-4312-bacb-6d623a6315ab · outbound

This paper cites Learning speech-driven 3d conversational gestures from video.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning speech-driven 3d conversational gestures from video

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.571181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.680789Z digest=sha256:bb8b705a1766c18df09d65f7939d4a2f03bfc7778b0efbdd35bced90e8fdd21d

Observation 439a4759-23c4-4b32-8f36-fc848410ef8a · outbound

This paper cites Evaluation of speech-to-gesture generation using bi-directional lstm network.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluation of speech-to-gesture generation using bi-directional lstm network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.558985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.684444Z digest=sha256:156bcc578e022d5c45bf9ff3db9b8fb7c1714358502d649d609013e1e6650375

Observation 6316e41f-a6d5-4520-957f-96fb48617ede · outbound

This paper cites Moglow: Probabilistic and controllable motion synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Moglow: Probabilistic and controllable motion synthesis using normalising flows

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.547685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.687870Z digest=sha256:a1474e1b0617a7312b23828ff240928ad11198ce80f1cd58e3bdb4822aa8c0a2

Observation 6880d816-5fdb-4531-a03d-3c4ebe8d07d4 · outbound

This paper cites Denoising dif- fusion probabilistic models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Denoising dif- fusion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.536260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.691524Z digest=sha256:7370d09ff71f841069117f3d9550743aad848f96f821f542098e89d8d800b2cd

Observation 3d47a550-f56b-4174-abb6-ac35337d1631 · outbound

This paper cites Dead blending, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Dead blending, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.522980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.695463Z digest=sha256:1da6263f730de3f8f80e82b17a2e87ec103918343cc4ea944584872ba1535e09

Observation cc08d7be-1e85-419d-9115-5f77e7e8c519 · outbound

This paper cites Phase- functioned neural networks for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Phase- functioned neural networks for character control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.511947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.698968Z digest=sha256:42022443beaa656a51775b92098f0ff50a4bc447577294d28fe9221c158b334c

Observation 2b0e30cf-8f77-4624-a3e6-e99e4ea8d345 · outbound

This paper cites Example- based control of human motion.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Example- based control of human motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.500817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.702423Z digest=sha256:a269fae7cb77ee679df90b1955e9388106bbce8e70fb7ef2f07caa9d91d2a715

Observation d1972587-0b70-413a-a0b4-cd858bc3d2e3 · outbound

This paper cites Robot behavior toolkit: generating effective social behaviors for robots.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Robot behavior toolkit: generating effective social behaviors for robots

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.489364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.706354Z digest=sha256:32f6af4d7212e8e548dd3f5b7d635d0732eac6721c2c1c408c4fae43fceadf81

Observation e76a4e53-20ee-4230-b662-17e640e00c88 · outbound

This paper cites Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.478397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.710164Z digest=sha256:4450cd18a71d6db828d55348b8e11c5ed92b389a96b22b8ab2fb501db3eb9385

Observation 5975b0a1-b9a4-4922-b180-6af8b50f6144 · outbound

This paper cites Intermask: 3d human interaction generation via collaborative masked modelling, 2024.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intermask: 3d human interaction generation via collaborative masked modelling, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.467350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.713926Z digest=sha256:a9255fe7af0320c6989ef406faaa802a37f0893a822c83f246c7479588704402

Observation 6a1026dd-d786-468d-ba8d-5105a70365b1 · outbound

This paper cites Synchronized multi-character motion editing.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Synchronized multi-character motion editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.456925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.717512Z digest=sha256:c87397f4ea2b5db844392c21ebf68783de8ab584becdc987fd727a250966629f

Observation b151996e-cce2-4bae-b3c5-f427d7b187d9 · outbound

This paper cites Tiling motion patches.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Tiling motion patches

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.446327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.721459Z digest=sha256:821d29f250b44603a932abab8626f50c73661d14766296206a3eb98e66123db6

Observation f4c80f5a-1a47-4127-ae36-802a4891f356 · outbound

This paper cites Towards a common framework for multimodal generation: The behavior markup language.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Towards a common framework for multimodal generation: The behavior markup language

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.435790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.725418Z digest=sha256:2714a518fa8b4b843535e8cbce65d6559d84d687ccf6e55baf6555b34fd03991

Observation 3ee368d6-8506-447a-8370-2ed9a34fc80c · outbound

This paper cites Gesticulator: A framework for semantically-aware speech-driven gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesticulator: A framework for semantically-aware speech-driven gesture generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.424954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.729119Z digest=sha256:b0568516d2781b1e5c12767022d4df6360dbe7e07b17797b689bd8d984392d5e

Observation fe930589-3c4b-4eac-8d7d-633d2cfbad82 · outbound

This paper cites The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.413976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.733295Z digest=sha256:1010f48af9f16678d0ed2c7cf026c57741210eeeb62e4b0892db9062987f1aab

Observation db9a7417-ebc6-4442-8197-e2d8f5ffe4ba · outbound

This paper cites Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.402783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.737923Z digest=sha256:93bc33dbddb04f9be7f2e14b577c232b8f9e248afd883fa0f25b3f031cd3d9bf

Observation 07d536e3-ea5a-4b05-81ed-b0a8cc5e7334 · outbound

This paper cites Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.391368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.742001Z digest=sha256:f870ec64905e82f1b4ced8ba71827a696d2314c008ce67b9e8054429f2123712

Observation 34fd5287-c8b0-4620-ad91-b95e190c0122 · outbound

This paper cites Two-character motion analysis and synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Two-character motion analysis and synthesis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.380112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.745937Z digest=sha256:d24dddcd3eb86fcac7a820352c608ddcb8cf10706f0b72746c3d3d11f6932114

Observation 88e21c0f-5319-46d7-a528-85150a04f76f · outbound

This paper cites Motion patches: building blocks for virtual environments annotated with motion data.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Motion patches: building blocks for virtual environments annotated with motion data

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.369173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.749738Z digest=sha256:1a17e9b315562094086b579de45f8e1069557a70274cfffd9077f4d05f461597

Observation eeb58929-60ac-4dcb-bb5f-cb6f417baeec · outbound

This paper cites Gesture controllers.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture controllers

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.358230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.753522Z digest=sha256:3a82e9d8f95d21af669b1abc0ec543df8eda189d7146e066441e1233ceb8e611

Observation ce76ea1f-9bcb-4a30-83b7-e9f25ee0032a · outbound

This paper cites Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.346673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.757026Z digest=sha256:5c01d615b9f7caacda4d3eb95bc5095d14d7172d24b4bd59e0ada75e7e0b02ad

Observation 33e9020d-0cf5-4ca8-b632-16a7e2524a53 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.335671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.760840Z digest=sha256:46624f00c2b11e5bcbe095a19e5afa74a4444795e489c1d19a3044a0166d2210

Observation 66f12494-a44e-4e61-aecc-0c913f74421f · outbound

This paper cites Intergen: Diffusion-based multi-human motion genera- tion under complex interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intergen: Diffusion-based multi-human motion genera- tion under complex interactions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.323941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.764298Z digest=sha256:b2904bc24bef9a52942b0f01311c231c576fbea8668ab3a9f552727fa9734eba

Observation b7ec29bf-bd6c-4ff3-87d4-1650789ade19 · outbound

This paper cites Com- position of complex optimal multi-character motions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Com- position of complex optimal multi-character motions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.311617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.768474Z digest=sha256:523d1fa3b4146a7debb31528b9949c1956cf2a9a72115d7f9ee51a51594686b0

Observation ec01dcc6-bd0c-4b0d-a6d3-caf1a887e39e · outbound

This paper cites BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.771875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.771875Z digest=sha256:1263cdc8286860e50df3ea1a591580e9e6f72d048634cde7e43dd1556566239b

Observation 2ed6c7a3-56e8-4729-b2b3-af2a888a2c7c · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.300739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.775706Z digest=sha256:680dbce5fab5b638c862b862ec39e45d8f37893f1f9d956941ec360f726d5c0e

Observation 896ea1c6-8479-480e-8f25-f37af429c758 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.289500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.778805Z digest=sha256:ad4550e374462178654d53b1349425e6f9c7205f93afb55f6501dc1f97b1591f

Observation 492e9b10-72f3-4075-9c87-15620a75b76e · outbound

This paper cites Learning hierarchical cross-modal association for co- speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning hierarchical cross-modal association for co- speech gesture generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.278004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.782159Z digest=sha256:d3e6fa4c5fa5d294ac1a2cd6056f877324d930d9f459617531c5bfb86b1e4084

Observation 87d76ea7-3213-4a20-9f4d-760622c75bc4 · outbound

This paper cites librosa: Audio and music signal analysis in python.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model librosa: Audio and music signal analysis in python

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.265161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.785431Z digest=sha256:8119f2d81ce7a8264c94eb73449b3ba6c7fe1b30fdfd60d30b48bbbee8020b42

Observation 841a0d39-cf2a-4aa3-b837-0cdeb2fc7b2f · outbound

This paper cites Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.788440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.788440Z digest=sha256:654f5c81f1028d5d0317879d7e91b4d0a9f2b6f654a832d09d95bda238156dfe

Observation 3d7bf52f-16fe-4278-9957-e82fc9ed806d · outbound

This paper cites Gesture modeling and animation based on a proba- bilistic re-creation of speaker style.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture modeling and animation based on a proba- bilistic re-creation of speaker style

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.246159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.793221Z digest=sha256:f461036c71cc85aff1021ac54904face1e9cee26aa1c73baee6e220f8d180c26

Observation dbacaf9f-79b6-4c46-8db8-274cf3022c59 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.234149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.797880Z digest=sha256:0b232313246ae8c518043821fb40a5cf55548a6d88fa6afa1d08343816db545a

Observation 6b842028-841c-4509-abbd-437d226e06e3 · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.221478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.801777Z digest=sha256:6088d80de98afb3413905cc5701f37732eae50a1ba599eaedb87389eb308e0f8

Observation 2cb3c7b4-ee91-49dd-a6fa-dd0cbacbb2e0 · outbound

This paper cites A comprehensive re- view of data-driven co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model A comprehensive re- view of data-driven co-speech gesture generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.209841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.805434Z digest=sha256:030b0f067bf23f172ac9d0e113500dd681aff4f1794ec4a629864fa8309ff813

Observation a6cdc883-08ad-40df-8e89-2e719926fbe7 · outbound

This paper cites Librispeech: An asr corpus based on public do- main audio books.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Librispeech: An asr corpus based on public do- main audio books

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.197919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.809497Z digest=sha256:9948d30bd439a0533d537c8e7febeeabf71e781a59552341011e871d6674decc

Observation 96373b2b-9d16-4261-be07-a9b8a5389ead · outbound

This paper cites Bodyformer: Semantics-guided 3d body gesture synthesis with transformer.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Bodyformer: Semantics-guided 3d body gesture synthesis with transformer

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.186332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.813195Z digest=sha256:b48831dd3640e6171a4396c9d69fa50050b9955c7682d7cca6d6fbd4578abe50

Observation 06947c17-cb57-4ad1-a79e-9497e1f20f13 · outbound

This paper cites Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.175473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.817013Z digest=sha256:d14670654c44b80a4e05f73ce3f6f66b6b8629ddfac458ffcbf79d466911f363

Observation 9608dfc9-9723-4269-bd87-dc81c22c86ef · outbound

This paper cites Hierarchical text-conditional image gener- ation with clip latents, 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Hierarchical text-conditional image gener- ation with clip latents, 2022

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.820676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.820676Z digest=sha256:b3dad0b40c47752a1dbff2e8e86df5b9e7dc5e6bd3e7180e5c098508b8209d0c

Observation 320c13ba-d1ba-42b0-9e1f-905dc283375d · outbound

This paper cites Human motion diffusion as a generative prior.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffusion as a generative prior

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.824254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.824254Z digest=sha256:3745ab9b81db75217e1a0097bbc9a5507c0c24ba3a845badb110a795edb78ed9

Observation f186f572-c4df-4ef5-9c3c-59cd0b75dff6 · outbound

This paper cites Interaction patches for multi-character animation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction patches for multi-character animation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.150778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.828174Z digest=sha256:e131d76d8ca0e7b5b9c4cd850487777ad6db1f385fbb6461bc1fb876f94a1c55

Observation 0730844b-9f8c-4d86-bf71-2a7a2509860e · outbound

This paper cites Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.140373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.831698Z digest=sha256:03a3596cc5ee4c5a3f89aab98a4e022b90afb44a194d0049294b47f0f6ec9e5f

Observation 025536bc-6071-4b24-b77e-2d7cd30f4927 · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.130124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.835159Z digest=sha256:b39c1cad56035bea4d5740c83cad6574226f820a6d211b7c6e33426fc03e821e

Observation 032f199d-881a-410c-a726-83cede472aae · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.118915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.838922Z digest=sha256:36f11d7c47ec01513b62160dd8f5e81e31452d1b0b6c3d770387bd0d95dd662c

Observation 1f82c318-2fb5-4e16-8299-1414a0979751 · outbound

This paper cites Human motion diffu- sion model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffu- sion model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.842737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.842737Z digest=sha256:cc34950223652915a406a955b8ee85f5773be186c84d460fce51026e4ebef90b

Observation 8480fd60-50da-43cc-9e08-ab52fbfc4716 · outbound

This paper cites Attention is all you need.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Attention is all you need

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.846546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.846546Z digest=sha256:22498cbb2c6c90a8d8fbe2908b264743515e9cb262f9a9209a1d009066f87216

Observation 9904a46b-a906-44db-bb81-bda659acea85 · outbound

This paper cites Gesture and speech in interaction: An overview, 2014.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture and speech in interaction: An overview, 2014

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.093731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.850351Z digest=sha256:ec69c970940db315c9261f42b6e60f9d5c00eea105b1228a05d63ae68ff59d7f

Observation 0cfc6845-f777-4876-9fc4-187caeb14324 · outbound

This paper cites Generating and ranking diverse multi-character interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Generating and ranking diverse multi-character interactions

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.082699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.854232Z digest=sha256:cff33f80b465f1bc2d24ceed6835ca9244d7af36c697c7aa09c0b835d6a802c4

Observation 04ebfbf7-9c3e-4e2b-9f7b-4cad06788d4e · outbound

This paper cites Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.070716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.858168Z digest=sha256:db0633bac3fbaab0b97427d0cdd29d0033a9ab70bf02f3ace88ae159d39a2c0d

Observation 1857373f-ab58-4012-b90b-9c575ef60273 · outbound

This paper cites Speech ges- ture generation from the trimodal context of text, audio, and speaker identity.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech ges- ture generation from the trimodal context of text, audio, and speaker identity

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.058176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.861661Z digest=sha256:8ebe8767a093d10cec8ffb65036301834f524c85b976f6a4fad8c8572fff3569

Observation 6c8e1412-677c-4b5e-a378-08917eb488c0 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.865334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.865334Z digest=sha256:1c0b71d6ef35796c4079ebc8c73bb95180043c5f02620449420f5f50653ec139

Observation b08a51bb-227b-4d0a-a766-a8bb1b413897 · outbound

This paper cites Speechtokenizer: Unified speech tokenizer for speech language models, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speechtokenizer: Unified speech tokenizer for speech language models, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.046810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.869270Z digest=sha256:d08390fc5238f8242a29a84c139d096983ddf59390d09dae3061b6af01725a40

Observation e9b5eec0-f331-46e2-bc64-4fc7626ba157 · outbound

This paper cites EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.872691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.872691Z digest=sha256:cf693484e692041933c57910238d98b3ebaec065c8808c31a4fb31a1011af0ca

Observation bed212dd-405f-48ee-b3e1-43f298d8d68e · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.034625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.877279Z digest=sha256:2b904055ef87f636ce2c0f3da35fb4f6b457f935278b18ff631b17c2184a3779

Observation 537878f9-1104-4be0-ac14-303447a336c2 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.021563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.880990Z digest=sha256:eabc2fca4aca799e146dc1291756b11e03c50d1ad5ed0607c5fd2a1b0a014e5b

Observation adbe8dce-c156-489a-9358-81100f6ab823 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.005552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.884500Z digest=sha256:4b26beed2bd41489b7d9d6a3b3584106effee8dba88d44ad60c1fbebde21a6eb

Observation 180a1b52-9edf-4be0-af4c-74965d9361a1 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.993214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.888908Z digest=sha256:654a9949312ef01a0bb45707f9bb3b6364184152a7b1699cef0611526997530c

Observation 53797734-b0a1-43f3-9f72-245a532149a3 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.982324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.892402Z digest=sha256:ad568c088e84dcfc4eaabc38967f0183eba0e6f7155c5141d97d8e0c0b522a12

Observation a563ca93-52e9-4dc4-90a4-6e4250b80133 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.970852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T23:34:10.895977Z digest=sha256:02f53ef8a7132ce1273c94e4d0e2f34c2c99380e757e3c5ff0475ae746cffdb2

Pith citing papers

Observation 72db793a-a47a-4f04-b638-14853299fecc · inbound

MotionPersona: Characteristics-aware Locomotion Control cites this paper.

MotionPersona: Characteristics-aware Locomotion Control It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T12:16:44.518506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T12:16:44.323653Z digest=sha256:875098508f4bf9068cbfdde5804ab513a9736cfde9778201a89d4557bd4fdd1e