Latent Space·2d agoRunway’s WorldPrompt and the Engineering of Real-Time Worlds#runway#world-models#worldprompt 2 sources 15 min
Hacker News · security·2d agoOpus 5.5 is good at explainer videos#claude-opus#video-generation#ai-agents 2 min
Hugging Face daily papers·3d agoWanPE: Towards Cinematic Prompt Enhancement for Modern Text-to-Video Generation#text-to-video#prompt-enhancement#wanpe
Hugging Face daily papers·3d agoViRDM: Taming Representation Distribution Matching for Few-Step Causal Video Generation#virdm#video-generation#diffusion
Hugging Face daily papers·4d agoTraining Object Permanence in World Models#world-models#object-permanence#video-generation
Hugging Face daily papers·4d agoAll modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation#video-generation#diffusion#cross-attention1
Hugging Face daily papers·4d agoThe Past Frames the Future: Memory for Autoregressive Video Generation#video-generation#autoregressive#memory
Hugging Face daily papers·5d agoWorldCrafter: Consistent Video World Model with Implicit 3D-aware Memory#world-models#video-generation#3d-memory 2 sources
arXiv cs.AI / cs.LG / cs.CL·5d agoGenerative Tutorial: Towards Live Contextualized Visual Instructions for Physical Tasks#generative-ai#augmented-reality#visual-instructionAI research
OpenAI News·5d agoHiggsfield AI ships new video features in a day with GPT-6 Astra#openai#gpt-6#gpt-6-astra 2 min1
Hugging Face daily papers·6d agoGAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation#gae#3d-generation#latent-space
Hugging Face daily papers·6d agoVideoGen-Agent: Reinforcing Video Generation Agents#videogen-agent#video-generation#reinforcement-learning
Hugging Face daily papers·8d agoRewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling#reward-model#video-generation#reinforcement-learning
Hugging Face daily papers·9d agoOmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation#benchmark#dataset#omnivbench1
arXiv cs.AI / cs.LG / cs.CL·9d agoVideo DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation#diffusion-transformer#hybrid-attention#inference-efficiencyAI research
Hugging Face daily papers·10d agoVideo DeltaNet: A Video-Native Hybrid Attention for Livestream Video Generation#diffusion-transformer#distillation#inference-efficiency
arXiv cs.AI / cs.LG / cs.CL·10d agoDreaming the Sound of Contact: Leveraging Video and Audio Generation for Zero-Shot Force-Aware Manipulation and Data Generation#audio-generation#data-generation#force-controlAI research
Hugging Face daily papers·11d agoCan MiniMax-H3 Reason About the Physical World? An Evaluation of Omni-Modal Generative Model#evaluation-benchmark#minimax-h3#omni-modal1
arXiv cs.AI / cs.LG / cs.CL·11d agoPhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control#autoregressive-model#image-to-video#motion-controlAI research1
Hugging Face daily papers·12d agoZing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control#5b#autoregressive#open-weights
Hugging Face daily papers·12d agoPhysStream: Streaming Physics-Grounded Video Generation with Structured Scene Memory and Fine-Grained Motion Control#autoregressive-models#image-to-video#motion-control
arXiv cs.AI / cs.LG / cs.CL·12d agoA Chosen Future Can Still Be Rewritten: Causal Writability in Video Models#interpretability#model-editing#physics-simulationAI research
Hugging Face daily papers·13d agoLynnReal-Omni: Native multi-modal Video Generation for Agentic Visual Workflows#agentic-workflows#diffusion-transformer#multimodal1