AION
Technique

State space models

Also known as: Mamba, SSM, state space model

10stories this week
12last 30 days
25all time

Timeline

  1. Oct 8, 2026 · Research paper · 1 source
    SACQ: Structured Decoding with Memory-Conditioned Refinement for Long-Horizon Forecasting
    We present SACQ, a plug-in structured prediction head that replaces flatten readout while keeping the encoder unchanged.
  2. Oct 7, 2026 · Research paper · 1 source
    CMP-IRRT*: A Perception-Assisted Height-Adaptive Planner for Quadruped Robots
    We propose a perception-assisted height-adaptive planning framework based on CMP-IRRT, a Channel Mamba PointNet-guided Informed RRT planner.
  3. Oct 7, 2026 · Research paper · 1 source
    HAN-Mamba: Hierarchical Selective State Space Networks for Multi-Scale Financial Volatility Forecasting
    Short-horizon realized volatility forecasting requires the integration of market information that evolves at incompatible temporal resolutions, from second-level order book dynamics to weekly regime drift.
  4. Oct 7, 2026 · Research paper · 1 source
    FedSSMCoOp: SSM Encoders for light-weight Federated Prompt Learning for Few-shot Classification
    To overcome this issue, we propose FedSSMCoOp, a federated few-shot image classification framework that enables multimodal learning while preserving data privacy.
  5. Oct 7, 2026 · Research paper · 1 source
    GeoPrior-Mamba: Structured Process Priors with Mamba for Fine-Resolution XCO2 Reconstruction
    Reconstructing fine-resolution column-averaged dry-air CO2 (XCO2) fields from sparse satellite observations requires models to infer spatial structure that is only weakly constrained by direct measurements.
  6. Oct 7, 2026 · Research paper · 1 source
    EM-SNN: Efficiently Modulated Spiking Neural Network for Remote Sensing Image Dehazing
    To address this challenge, we propose the Efficiently Modulated Spiking Neural Network (EM-SNN), a dedicated spiking framework tailored to remote sensing image dehazing.
  7. Oct 6, 2026 · Research paper · 1 source
    MaRK: Markov-adapted Recurrent Kernels for Dynamic Operator Conditioning in State Space Models
    We introduce MaRK (Markov-adapted Recurrent Kernels), a dynamic operator-conditioning framework that maps context vectors directly into bounded modulations of a frozen SSM's recurrence ($A$), read-in ($B$), read-out ($C$), skip ($D$), and discretization ($Δ$) parameters.
  8. Oct 6, 2026 · Research paper · 1 source
    RSJEV: Discriminative Remote Sensing Scene Classification with Multimodal Large Language Models
    To address these limitations, we propose RSJEV, a one-pass multimodal decision framework for remote sensing scene classification.
  9. Oct 5, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.31.0
    Fast restart: the new vllm preload CLI launches the weight-cache daemon that keeps post-quantized weights resident in GPU memory across engine restarts (#56680), now with data parallelism (#57386), MTP draft models (#57312), a /health endpoint (#58552) and a readiness wait (#58370).
  10. Oct 5, 2026 · Research paper · 1 source
    TIDES: Implicit Time-Awareness in Selective State Space Models
    Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step TildeΔ a learned function of the input.
  11. Sep 30, 2026 · Open-source release · 1 source
    huggingface/transformers v5.18.0: Release 5.18.0
    Nemotron 3 Diarization is an open-weight streaming speaker diarization model designed to determine "who spoke when" in real-world audio.
  12. Sep 22, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.30.0
    This release features 762 commits from 315 contributors (104 new)!
  13. Sep 9, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.29.0
    MRV2 also gained CUDA graph memory profiling for KV cache auto-sizing (#53306), batch-sharded sampling that cuts per-step logits memory by 1/TP (#50465), prompt embeds (#42963), extracthiddenstates speculation (#49811), padded FULL cudagraph dispatch for uniform decode under spec decode (#53407), and DP-sync skipping before EAGLE/MTP draft prefill (#53694).
  14. Aug 26, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.28.0
    DeepSeek V4: sparse MLA now works end-to-end for plain decode, MTP, and DSpark speculative decoding (#51538), joined by AMD Quark NVFP4 support (#47972), reasoning-effort prompts and mappings (#50580), sparse top-k metadata kernel optimizations (#52084, #51967), narrowed eager CUDA graph regions (#51430, #52401), and ROCm enablement on gfx11 and gfx950 (#47017, #52212).
  15. Aug 22, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.18
    | Model | Type | PRs | Cookbook |
  16. Aug 10, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.27.0
    Kimi K3 support with a full stack landing in one release: core model files and kernels (#50089, #50000), Python (#50093) and Rust (#50104) frontends, AttnRes kernels (#50090), DeepGEMM support (#50458), compressed-tensors quantized checkpoints (#50500), DSpark AR fusion (#50242), and an option to shard the shared expert instead of replicating it (#50656).
  17. Aug 10, 2026 · Open-source release · 1 source
    huggingface/transformers v5.15.0: Release: v5.15.0
    Muse Glimmer, released today, is Meta’s new multimodal model, especially designed for agentic use cases.
  18. Jul 25, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.16
    DSpark: confidence-driven speculative decoding: A new speculative algorithm.
  19. Jul 11, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.25.0
    Building on quantized-model support from the previous release, MRv2 is now the standard execution path, with new support for EVS (#46535), realtime embeddings (#46762), prefix caching for Mamba hybrid models (#42406), multimodal-prefix bidirectional attention (#46942), and dynamic speculative decoding compatible with full CUDA graphs (#45953).
  20. Jun 26, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.14
    Full release notes by category below.
  21. Jun 13, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.13
    DeepSeek V4 — context parallelism & sparse-attention kernels: Building on the v0.5.12 Day-0 path, v0.5.13 extends DeepSeek-V4 to context-parallel serving and adds its sparse-attention kernels:
  22. Jun 10, 2026 · Open-source release · 1 source
    huggingface/transformers v5.11.0: Release v5.11.0
    DiffusionGemma is engineered to reduce the sequential bottlenecks of standard causal language models by employing an encoder-decoder architecture specifically optimized for inference speed.
  23. May 29, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.22.0
    DeepSeek V4 maturity: DeepSeek V4 received a major hardening pass this cycle — the model was reorganized into a dedicated vllm/models/deepseekv4/ package (#43004, #43039, #43073, #43077, #43149), gained NVFP4 fused MoE support (#42209), full + piecewise CUDA graph (#42604), and MTP speculative decoding (#43385).
  24. May 15, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.21.0
    Transformers v4 deprecated: This release formally deprecates transformers v4 support (#40389).
  25. May 5, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.11
    Speculative Decoding V2 by default: Spec V2 (with overlap scheduling to hide CPU overhead) is now the default, materially reducing per-step CPU cost for EAGLE/MTP/DFLASH paths: #21062

Often appears with