State space models
Also known as: Mamba, SSM, state space model
10stories this week
12last 30 days
25all time
Timeline
- Oct 8, 2026 · Research paper · 1 sourceSACQ: Structured Decoding with Memory-Conditioned Refinement for Long-Horizon ForecastingWe present SACQ, a plug-in structured prediction head that replaces flatten readout while keeping the encoder unchanged.
- Oct 7, 2026 · Research paper · 1 sourceCMP-IRRT*: A Perception-Assisted Height-Adaptive Planner for Quadruped RobotsWe propose a perception-assisted height-adaptive planning framework based on CMP-IRRT, a Channel Mamba PointNet-guided Informed RRT planner.
- Oct 7, 2026 · Research paper · 1 sourceHAN-Mamba: Hierarchical Selective State Space Networks for Multi-Scale Financial Volatility ForecastingShort-horizon realized volatility forecasting requires the integration of market information that evolves at incompatible temporal resolutions, from second-level order book dynamics to weekly regime drift.
- Oct 7, 2026 · Research paper · 1 sourceFedSSMCoOp: SSM Encoders for light-weight Federated Prompt Learning for Few-shot ClassificationTo overcome this issue, we propose FedSSMCoOp, a federated few-shot image classification framework that enables multimodal learning while preserving data privacy.
- Oct 7, 2026 · Research paper · 1 sourceGeoPrior-Mamba: Structured Process Priors with Mamba for Fine-Resolution XCO2 ReconstructionReconstructing fine-resolution column-averaged dry-air CO2 (XCO2) fields from sparse satellite observations requires models to infer spatial structure that is only weakly constrained by direct measurements.
- Oct 7, 2026 · Research paper · 1 sourceEM-SNN: Efficiently Modulated Spiking Neural Network for Remote Sensing Image DehazingTo address this challenge, we propose the Efficiently Modulated Spiking Neural Network (EM-SNN), a dedicated spiking framework tailored to remote sensing image dehazing.
- Oct 6, 2026 · Research paper · 1 sourceMaRK: Markov-adapted Recurrent Kernels for Dynamic Operator Conditioning in State Space ModelsWe introduce MaRK (Markov-adapted Recurrent Kernels), a dynamic operator-conditioning framework that maps context vectors directly into bounded modulations of a frozen SSM's recurrence ($A$), read-in ($B$), read-out ($C$), skip ($D$), and discretization ($Δ$) parameters.
- Oct 6, 2026 · Research paper · 1 sourceRSJEV: Discriminative Remote Sensing Scene Classification with Multimodal Large Language ModelsTo address these limitations, we propose RSJEV, a one-pass multimodal decision framework for remote sensing scene classification.
- Oct 5, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.31.0Fast restart: the new vllm preload CLI launches the weight-cache daemon that keeps post-quantized weights resident in GPU memory across engine restarts (#56680), now with data parallelism (#57386), MTP draft models (#57312), a /health endpoint (#58552) and a readiness wait (#58370).
- Oct 5, 2026 · Research paper · 1 sourceTIDES: Implicit Time-Awareness in Selective State Space ModelsSelective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step TildeΔ a learned function of the input.
- Sep 30, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.18.0: Release 5.18.0Nemotron 3 Diarization is an open-weight streaming speaker diarization model designed to determine "who spoke when" in real-world audio.
- Sep 22, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.30.0This release features 762 commits from 315 contributors (104 new)!
- Sep 9, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.29.0MRV2 also gained CUDA graph memory profiling for KV cache auto-sizing (#53306), batch-sharded sampling that cuts per-step logits memory by 1/TP (#50465), prompt embeds (#42963), extracthiddenstates speculation (#49811), padded FULL cudagraph dispatch for uniform decode under spec decode (#53407), and DP-sync skipping before EAGLE/MTP draft prefill (#53694).
- Aug 26, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.28.0DeepSeek V4: sparse MLA now works end-to-end for plain decode, MTP, and DSpark speculative decoding (#51538), joined by AMD Quark NVFP4 support (#47972), reasoning-effort prompts and mappings (#50580), sparse top-k metadata kernel optimizations (#52084, #51967), narrowed eager CUDA graph regions (#51430, #52401), and ROCm enablement on gfx11 and gfx950 (#47017, #52212).
- Aug 22, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.18| Model | Type | PRs | Cookbook |
- Aug 10, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.27.0Kimi K3 support with a full stack landing in one release: core model files and kernels (#50089, #50000), Python (#50093) and Rust (#50104) frontends, AttnRes kernels (#50090), DeepGEMM support (#50458), compressed-tensors quantized checkpoints (#50500), DSpark AR fusion (#50242), and an option to shard the shared expert instead of replicating it (#50656).
- Aug 10, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.15.0: Release: v5.15.0Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for agentic use cases.
- Jul 25, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.16DSpark: confidence-driven speculative decoding: A new speculative algorithm.
- Jul 11, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.25.0Building on quantized-model support from the previous release, MRv2 is now the standard execution path, with new support for EVS (#46535), realtime embeddings (#46762), prefix caching for Mamba hybrid models (#42406), multimodal-prefix bidirectional attention (#46942), and dynamic speculative decoding compatible with full CUDA graphs (#45953).
- Jun 26, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.14Full release notes by category below.
- Jun 13, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.13DeepSeek V4 — context parallelism & sparse-attention kernels: Building on the v0.5.12 Day-0 path, v0.5.13 extends DeepSeek-V4 to context-parallel serving and adds its sparse-attention kernels:
- Jun 10, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.11.0: Release v5.11.0DiffusionGemma is engineered to reduce the sequential bottlenecks of standard causal language models by employing an encoder-decoder architecture specifically optimized for inference speed.
- May 29, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.22.0DeepSeek V4 maturity: DeepSeek V4 received a major hardening pass this cycle — the model was reorganized into a dedicated vllm/models/deepseekv4/ package (#43004, #43039, #43073, #43077, #43149), gained NVFP4 fused MoE support (#42209), full + piecewise CUDA graph (#42604), and MTP speculative decoding (#43385).
- May 15, 2026 · Open-source release · 1 sourcevllm-project/vllm v0.21.0Transformers v4 deprecated: This release formally deprecates transformers v4 support (#40389).
- May 5, 2026 · Open-source release · 1 sourcesgl-project/sglang v0.5.11Speculative Decoding V2 by default: Spec V2 (with overlap scheduling to hide CPU overhead) is now the default, materially reducing per-step CPU cost for EAGLE/MTP/DFLASH paths: #21062