AION
Model

MiniMax models

Also known as: MiniMax-M1, MiniMax-M2

1stories this week
1last 30 days
5all time

Timeline

  1. Oct 8, 2026 · Research paper · 1 source
    Chronos Enables Code Agents to Reason over Software Evolution
    We introduce Chronos, a test-time framework that makes this connected history available to large language model (LLM)-based code agents.
  2. Aug 22, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.18
    | Model | Type | PRs | Cookbook |
  3. Jun 29, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.24.0
    MiniMax-M3: Added support for the new MiniMax-M3 model (#45381), with a fast follow-on of BF16/FP8 indexer via MSA (#45892), MXFP4 support (#45896), FP8 sparse GQA (#45744), and extensive AMD/ROCm tuning — mxfp8 MoE/linear on gfx950 (#45725), fp8perchannel for bf16 weights on MI300X (#45854), FP8 KV-cache fix (#45720), and packed-modules mapping (#45794).
  4. May 29, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.22.0
    DeepSeek V4 maturity: DeepSeek V4 received a major hardening pass this cycle — the model was reorganized into a dedicated vllm/models/deepseekv4/ package (#43004, #43039, #43073, #43077, #43149), gained NVFP4 fused MoE support (#42209), full + piecewise CUDA graph (#42604), and MTP speculative decoding (#43385).
  5. Feb 24, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.9
    TRT-LLM NSA Kernel Integration for DeepSeek V3.2: Integrate TRT-LLM DSA kernels for Native Sparse Attention, boosting DeepSeek V3.2 performance by 3x-5x on Blackwell platforms with trtllm for both --nsa-prefill-backend and --nsa-decode-backend

Often appears with