Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:55:28.944960Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 2 inbound Pith citation observations for arXiv:2506.11144.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:55:28.944960Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:06:47.301483Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T22:20:22.357042Z
51 of 51 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fd8d86da-6153-4bef-a97b-4378a606b53a · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation A general theoretical paradigm to understand learning from human preferences
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fa5b9a3-079e-49ef-af72-95a0684e3a38 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation wav2vec 2.0: A framework for self-supervised learning of speech representations
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b2f9db-ad79-4432-bd5b-8233d8d2d7e9 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation SkyReels-V2: Infinite-length Film Generative Model
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b8ab91f-1943-4a55-96c5-781a00fd4a7b · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Echomimic: Lifelike audio- driven portrait animations through editable landmark conditions
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 31c0ca70-889b-4879-8337-4a1f37d8e2dd · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Out of time: automated lip sync in the wild
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fa74461-f8fc-4e81-820f-7b24069f5283 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccbba998-7073-4628-b767-729d3c09c679 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 332b2d9a-3fe5-4db6-b829-f34eee398ab6 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a64a6aa-ae2a-4e2d-ac61-1cd4ad56309a · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Diffted: One-shot audio- driven ted talk video generation with diffusion-based co-speech gestures
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f3cb1880-620a-49d7-abef-6e4645e6b613 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Loopy: Taming audio-driven portrait avatar with long-term motion dependency
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4de104e3-24dc-47c3-9f22-8e2c4e85cdc7 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Elucidating the design space of diffusion-based generative models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14601b1c-aea7-419d-99d0-34d6dd353218 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Auto-encoding variational bayes, 2013
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ef71f1-9dc5-4609-b324-a1291948ba06 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation OpenHumanVid: A Large-Scale High-Quality Dataset for Enhancing Human-Centric Video Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd030be6-b309-445a-b29d-afe56d43af43 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation T2V-Turbo: Breaking the Quality Bottleneck of Video Consistency Model with Mixed Reward Feedback
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55c403f1-9e1d-4dc4-8708-d78ecdbdda0d · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Cyberhost: A one-stage diffusion framework for audio-driven talking body generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6e47c66c-d5ac-4cfb-81e4-cd383f6c223d · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec72267-a68f-48f7-bd25-d5955335aa6a · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Flow Matching for Generative Modeling
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a37f8ded-f857-4f57-9170-1f9ee3402428 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Improving Video Generation with Human Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3b91f74-c925-4196-a87f-0d3bf56f71de · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaee10a4-b814-4bed-be3d-ab01b6ab6454 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67473cb3-1792-452b-b3c9-f9add8a36d7b · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation OpenELM: An Efficient Language Model Family with Open Training and Inference Framework
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5df5be53-4302-4b5e-aa28-5d2486ae2e29 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Echomimicv2: Towards striking, simplified, and semi-body human animation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55458c22-2b77-4943-ab82-be5c5862081f · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Simpo: Simple preference optimization with a reference-free reward
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 560225d4-646b-4d12-8ba8-9bf455d38ca6 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Clip-dpo: Vision-language models as a source of preference for fixing hallucinations in lvlms
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d0e261d7-f504-4929-bf90-a71fd56def66 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Scalable diffusion models with transformers
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f412f38c-0a59-4fe2-a4e8-94f7499ef526 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation SkyReels-A1: Expressive Portrait Animation in Video Diffusion Transformers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f1fd874-d2a5-4f25-aaa8-6e7570cbc331 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Direct preference optimization: Your language model is secretly a reward model
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53c8f23a-562d-4125-af33-9fd5bf8e15ed · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Seaweed-7B: Cost-Effective Training of Video Generation Foundation Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84fdc69f-6180-43d1-82f3-e60177f829ef · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation First order motion model for image animation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb78e81a-9ec2-45fa-aa05-8f363e9ed10a · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Motion repre- sentations for articulated animation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c0e90711-c221-422b-9c4c-c10a9716e5ee · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Learning to summarize with human feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71ee290a-41c2-40a6-a35c-52be9b948044 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3317897e-905e-4580-b747-3b0fcb39c634 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b28877d-56fe-4a91-9b86-b40291324f9c · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Cambrian-1: A fully open, vision-centric exploration of multimodal llms
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 812fa1c1-c70e-4371-a564-bfb6b70afd5c · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Fvd: A new metric for video generation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9de4a076-298c-4f18-845f-79c9c37e7b38 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Diffusion model alignment using direct preference optimization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab4b4deb-7c2c-47d2-8177-452149814295 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c039a11-5378-406b-8f97-09dfc2d9857c · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d15514-4ce7-4003-b194-d045990134d5 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Video-to-Video Synthesis
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31529f22-2b38-45ac-bb66-f124081eb677 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation MoCha: Towards Movie-Grade Talking Character Synthesis
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3623e821-9b74-4eb0-8bca-2c23fbf8d40c · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b96501ee-2d9b-4a1d-b447-6e00bd365e02 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02dfe58c-2446-4726-9098-ea9aa4e97b5b · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 973dce22-a8ec-4689-adcc-bebf3488a4f9 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Imagereward: Learning and evaluating human preferences for text-to-image generation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8b87c68-d613-4395-a5e6-d84a82c53b51 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Qwen2.5-1M Technical Report
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38fb32d1-51cb-41d7-a3e7-27574f913fd6 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91a34f10-0f46-42b2-b700-0a2d7e954b6b · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43bb5afe-d99e-421f-b48d-a980d648cfd6 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf331162-f1b4-4f4e-a36a-2d3f7c10ba00 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Learning to compare for better training and evaluation of open domain natural language generation models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 81a56989-01fb-48eb-b3bd-a3f7b69ee758 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Taming diffusion models for audio-driven co-speech gesture generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b5ee12e-bb9b-47aa-84d2-20ad52b4d177 · outbound
AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation Vlogger: Make your dream a vlog
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 80b09f0d-f1bb-4a0f-88ba-5793b7779ac2 · inbound
FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2220945c-701f-4271-b0de-e9a73bb81630 · inbound
EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.