AION
Repository / library

PyTorch

7stories this week
11last 30 days
35all time

Timeline

  1. Oct 10, 2026 · Opinion / analysis · 1 source
    Is there a better option than llama.cpp for 4GB VRAM for Higher tokens/sec?
    I want something that utilizes my system resources more efficiently—for instance, by managing my RTX with 4GB VRAM more intelligently—and delivers higher tokens-per-second, all without the burden of heavy dependencies like PyTorch.
  2. Oct 8, 2026 · Open-source release · 1 source
    huggingface/trl v1.15.0
    SFT, DPO, KTO, GRPO, RLOO and Distillation now score tokens with a fused LM head: a Triton kernel projects the hidden states through the LM head in tiles and reduces to per-token log-probs and entropy directly, so the [batch, seq, vocab] logits tensor is never built.
  3. Oct 8, 2026 · Research paper · 1 source
    The Operator Mismatch Problem: Deploying BEV Perception with Portable GPU Compute
    We present BEVPIPE, a framework for deploying multimodal BEV perception pipelines using portable GPU compute APIs and integrating them with production inference runtimes.
  4. Oct 7, 2026 · Research paper · 1 source
    KGATE : a Knowledge Graph Embedding Training Environment
    Knowledge graph embedding (KGE) models encode the entities and relations of a knowledge graph into a low-dimensional latent space, enabling tasks such as classification or link prediction.
  5. Oct 5, 2026 · Open-source release · 1 source
    speridlabs/iris-3b
    speridlabs published the model iris-3b on Hugging Face.
  6. Oct 5, 2026 · Model release · 1 source
    perplexity-ai/pplx-decider-v1.1-27b
    perplexity-ai published the model pplx-decider-v1.1-27b on Hugging Face.
  7. Oct 5, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.31.0
    Fast restart: the new vllm preload CLI launches the weight-cache daemon that keeps post-quantized weights resident in GPU memory across engine restarts (#56680), now with data parallelism (#57386), MTP draft models (#57312), a /health endpoint (#58552) and a readiness wait (#58370).
  8. Oct 2, 2026 · Open-source release · 1 source
    ray-project/ray ray-2.59.0: Ray-2.59.0
    Ray Data LLM & Ray Serve LLM are GA/Stable: the LLM APIs graduate to general availability this release (\#65194), alongside an upgrade to vLLM 0.27.0 (\#65351).
  9. Oct 1, 2026 · Product / feature launch · 1 source
    unslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UX
    This release brings faster navigation, shareable run settings, and clearer errors to Unsloth Desktop.
  10. Sep 30, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.14.1: PyTorch 2.14.1 Release
    This release is meant to fix the following regressions and silent correctness issues:
  11. Sep 29, 2026 · Open-source release · 1 source
    NVIDIA/TensorRT-LLM v1.3.0rc29
    Expose Nemotron-H vision-language LoRA configuration for supported inference paths #19151
  12. Sep 2, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.14.0: PyTorch 2.14.0 Release
    <tr><td><strong>A preview of our rewritten NCCL backend for PyTorch</strong>, ported from torchcomms, implementing the full collective contract with nonblocking communicators and eager communicator splitting and advanced features such as fault tolerance and windows designed as a drop-in replacement of existing NCCL c10d backend</td></tr>
  13. Aug 10, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.27.0
    Kimi K3 support with a full stack landing in one release: core model files and kernels (#50089, #50000), Python (#50093) and Rust (#50104) frontends, AttnRes kernels (#50090), DeepGEMM support (#50458), compressed-tensors quantized checkpoints (#50500), DSpark AR fusion (#50242), and an option to shard the shared expert instead of replicating it (#50656).
  14. Jul 8, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.13.0: PyTorch 2.13.0 Release
    <tr><td><strong>torchcomms</strong>, a new communications backend for PyTorch Distributed, improves fault tolerance, scalability, and debuggability for large-cluster training.</td></tr>
  15. Jun 26, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.14
    Full release notes by category below.
  16. Jun 18, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.12.1: PyTorch 2.12.1 Release, bug fix release
    This release is meant to fix the following regressions and silent correctness issues:
  17. May 15, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.21.0
    Transformers v4 deprecated: This release formally deprecates transformers v4 support (#40389).
  18. May 13, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.12.0: PyTorch 2.12.0 Release
    <tr><td><strong>Batched linalg.eigh on CUDA</strong> is up to 100x faster due to updated cuSolver backend selection.</td></tr>
  19. May 5, 2026 · Open-source release · 1 source
    sgl-project/sglang v0.5.11
    Speculative Decoding V2 by default: Spec V2 (with overlap scheduling to hide CPU overhead) is now the default, materially reducing per-step CPU cost for EAGLE/MTP/DFLASH paths: #21062
  20. Apr 27, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.20.0
    CUDA 13.0 default: Default CUDA wheel on PyPI and vllm/vllm-openai:v0.20.0 image switched to CUDA 13.0; architecture lists and build-args cleaned up (#39878), and CUDA bumped to 13.0.2 to match PyTorch 2.11.0 (#40669).
  21. Mar 23, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.11.0: PyTorch 2.11.0 Release
    <strong>FlexAttention</strong> now has a <strong>FlashAttention-4</strong> backend on <strong>Hopper</strong> and <strong>Blackwell</strong> GPUs
  22. Jan 21, 2026 · Open-source release · 1 source
    pytorch/pytorch v2.10.0: PyTorch 2.10.0 Release
    <td> Reduced kernel launch overhead with <strong>combo-kernels</strong> horizontal fusion in torchinductor </td>
  23. Nov 12, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.9.1: PyTorch 2.9.1 Release, bug fix release
    This release is meant to fix the following issues (regressions / silent correctness):
  24. Oct 15, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.9.0: 2.9 Release Notes
    <td>Updates to the stable libtorch ABI for third-party C++/CUDA extensions</td>
  25. Aug 6, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.8.0: PyTorch 2.8.0 Release
    <td>High-performance quantized LLM inference on Intel CPUs with native PyTorch</td>
  26. Jun 4, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.7.1: PyTorch 2.7.1 Release, bug fix release
    This release is meant to fix the following issues (regressions / silent correctness):
  27. Apr 23, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.7.0: PyTorch 2.7.0 Release
    <td>Torch.Compile support for Torch Function Modes
  28. Jan 29, 2025 · Open-source release · 1 source
    pytorch/pytorch v2.6.0: PyTorch 2.6.0 Release
    This release features multiple improvements for PT2: torch.compile can now be used with Python 3.13; new performance-related knob torch.compiler.setstance; several AOTInductor enhancements.
  29. Oct 29, 2024 · Open-source release · 1 source
    pytorch/pytorch v2.5.1: PyTorch 2.5.1: bug fix release
    This release is meant to fix the following regressions:
  30. Oct 17, 2024 · Open-source release · 1 source
    pytorch/pytorch v2.5.0: PyTorch 2.5.0 Release, SDPA CuDNN backend, Flex Attention
    We are excited to announce the release of PyTorch® 2.5!

Often appears with