REVIEW 9 cited by
AnimateLCM: Computation-Efficient Personalized Style Video Generation without Personalized Video Data
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper introduces an effective method for computation-efficient personalized style video generation without requiring access to any personalized video data. It reduces the necessary generation time of similarly sized video diffusion models from 25 seconds to around 1 second while maintaining the same level of performance. The method's effectiveness lies in its dual-level decoupling learning approach: 1) separating the learning of video style from video generation acceleration, which allows for personalized style video generation without any personalized style video data, and 2) separating the acceleration of image generation from the acceleration of video motion generation, enhancing training efficiency and mitigating the negative effects of low-quality video data.
Forward citations
Cited by 9 Pith papers
-
Yume: An Interactive World Generation Model
A diffusion-based video model generates extendable, keyboard-controlled walkthroughs from a single input image, using quantized camera actions as text prompts.
-
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
OSA-LCM distills a portrait video diffusion model into a single-step generator that matches the quality of a 20-step teacher on FID/FVD, enabling near real-time talking-head generation.
-
SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device
SnapGen-V prunes, searches, and adversarially distills a video diffusion model down to 0.6B parameters that generates a five-second, 512x512 video on an iPhone 16 Pro Max in under five seconds.
-
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
A diffusion model trained progressively from Kingdom to Species generates more accurate fine-grained animal images, including rare species with as few as one training sample.
-
RAIN: Real-time Animation of Infinite Video Stream
RAIN makes real-time character animation and video style transfer practical on consumer GPUs by grouping streamed frames into shared noise levels and running cross-noise-level temporal attention.
-
Accelerating Video Diffusion Models via Distribution Matching
A few-step video generator trained with video GAN loss plus 2D score distribution matching matches or beats prior 4-step video distillation methods.
-
Single Trajectory Distillation for Accelerating Image and Video Style Transfer
A consistency-distillation method trains from a fixed partial-noise start and uses a replay bank plus DINO-v2 adversarial loss to improve few-step image and video stylization.
-
Reinforcement Learning: From Algorithms To Foundation Models
A dissertation uniting the author's published results: non-exploitable Nash-DQN policies and the FightLadder benchmark for games, plus diffusion/consistency-model world models for RL — a compilation rather than new results.
-
Generative Diffusion Modeling: A Practical Handbook
A handbook that unifies notation and practical guidance for diffusion, score-based, consistency, and rectified-flow models, including distillation and reward fine-tuning, but with some mathematical errors.
Discussion (0). Continue with ORCID to comment.