Wan
Also known as: Wan 2, Wan2
12stories this week
12last 30 days
13all time
Timeline
- Oct 8, 2026 · Research paper · 1 sourceWorldAlign: Decoupled 4D Reward for World-Consistent Video GenerationFaithful visual world simulation requires generated videos to maintain 4D world consistency, encompassing both static and dynamic consistency.
- Oct 8, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.905-beta: Sandboxing is here!We're introducing Windows, Mac and Linux sandboxing in Unsloth!
- Oct 8, 2026 · Research paper · 1 sourceMemory Forcing: Attendable Mid-Horizon History for Streaming Video GenerationAutoregressive video diffusion enables causal video streaming without a bidirectional pass over the full clip, but existing few-step systems usually retain only the opening and most recent frames in a fixed-size KV cache.
- Oct 8, 2026 · Research paper · 1 sourceTowards Unified Evaluation of Prompt Enhancers for Video GenerationTo address this gap, we introduce PEBench, the first unified benchmark for direct PE evaluation across text-to-video, image-to-video, and reference-to-video prompt enhancement.
- Oct 8, 2026 · Research paper · 1 sourceParametric Trajectory Distillation for Few-Step Video GenerationWe introduce Parametric Trajectory Distillation (PTD), which lets the student parameterize teacher trajectory segments as polynomials and learn from teacher guidance along its own predicted path.
- Oct 8, 2026 · Research paper · 1 sourceiCATS: Fast Video Generation via Interaction-Aware Sparse Attention and Timestep-Adaptive SparsityTraining-free sparse attention offers a practical acceleration solution to Diffusion Transformers (DiTs) via reducing computations without fine-tuning.
- Oct 7, 2026 · Research paper · 1 sourceFluid-Gen-Zero: Grounding Pretrained Video Generators in Physics without TrainingWe present Fluid-Gen-Zero, a training-free framework for physics-aware fluid-object interaction video generation that decouples physical reasoning from appearance synthesis.
- Oct 7, 2026 · Research paper · 1 sourceGRACE: Generation-aware latent compression for efficient video generationTo address this, we propose Generation-Aware Latent Compression for Efficient Video Generation (GRACE), a two-stage framework that compresses a pretrained video autoencoder while keeping it compatible with the pretrained DiT.
- Oct 7, 2026 · Open-source release · 1 sourceunslothai/unsloth v0.1.904-beta: Train your own Decision modelTurn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
- Oct 6, 2026 · Research paper · 1 sourceBackend-Agnostic Sparse Attention for Fast High-Resolution Visual GenerationTo tackle these challenges, we propose BASA, a backend-agnostic sparse attention, which brings the best of both worlds: visual quality and practical acceleration.
- Oct 6, 2026 · Research paper · 1 sourceOpenWAM: An Open Framework for Composable World-Action ModelsWe introduce OPENWAM, an open world-action modeling framework built around a common causal robot-video foundation and configurable video-action interaction.
- Oct 6, 2026 · Open-source release · 1 sourcehuggingface/diffusers v0.41.0: Diffusers 0.41.0: QwenImage 2.1 pipeline and more> This release brings Qwen-Image 2.1 to Diffusers, with text-to-image generation, image editing, native transparency, and LoRA training.
- Jun 13, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.13DeepSeek V4 — context parallelism & sparse-attention kernels: Building on the v0.5.12 Day-0 path, v0.5.13 extends DeepSeek-V4 to context-parallel serving and adds its sparse-attention kernels: