HF Daily Papers
· Papers
MemLearner: Learning to Query Context memory for Video World Models
Video World Models are interactive video generation models that predict future world states based on user actions and history video frames. A critical challenge in video world models is the lack of memory, causing inconsistent generated scenes over extended durations. Previous methods explored rule-