AION
Organization

NVIDIA

Also known as: Nvidia

32stories this week
52last 30 days
72all time

Timeline

  1. Oct 11, 2026 · Opinion / analysis · 1 source
    I trained a 102M recursive BitNet-v2 model from scratch: 64K context, trained on less than 5B tokens
    Hiya, I’m releasing Recursive BitNet N-Gram 102M, a small experiment combining ternary weights, shared transformer layers, and hashed n-gram embeddings, trained with a whooping budget of 100€
  2. Oct 11, 2026 · Opinion / analysis · 1 source
    Imagine M5 Ultra + this: NVIDIA GeForce RTX 5060 runs macOS 15 with Metal 3 acceleration through unofficial driver
  3. Oct 11, 2026 · Opinion / analysis · 1 source
    AMD Reportedly Raises GDDR6 Prices for Board Partners
    Is an AMD price incoming as well?
  4. Oct 10, 2026 · Funding / M&A · 1 source
    Nvidia in talks to acquire US 'open' model startup Reflection AI
  5. Oct 10, 2026 · Opinion / analysis · 1 source
    No more RTX 5090
    Nvidia reportedly halts GeForce RTX 5090 production in favor of AI data center and professional GPUs — impending supply drought expected to drive up prices, RTX 5080 24GB rumored as new gaming flagship
  6. Oct 10, 2026 · Opinion / analysis · 1 source
    NVIDIA reportedly discontinuing RTX 5090, GB202 GPUs to be reserved for RTX PRO series
  7. Oct 10, 2026 · Opinion / analysis · 1 source
    Strata with Qwen3.8 Flash Next UD-Q4_K_XL
    Most of the benchmarks I've seen are using IQ2 or IQ3 quants, so I wanted to see how Unsloth's UD-Q4KXL performs instead.
  8. Oct 10, 2026 · Model release · 1 source
    nvidia/agile_one_s_place_ssd_n17_24050
    nvidia published the model agileonesplacessdn1724050 on Hugging Face.
  9. Oct 8, 2026 · Opinion / analysis · 1 source
    5 Steps to Create SimReady Assets for Robotics with Frontier AI Models
    Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision...Preparing CAD assets for robotics simulation requires more than converting geometry to OpenUSD: developers must configure and validate materials, collision geometry, joints, and other physics properties before testing robot behavior.
  10. Oct 8, 2026 · Opinion / analysis · 1 source
    Building Reliable Data Analytics Agents: Lessons from the KDD Cup
    The NVIDIA KGMON team placed second in the KDD Cup 2026 Data Agents competition with a system built around a simple idea of making an agent's harness smaller,...The NVIDIA KGMON team placed second in the KDD Cup 2026 Data Agents competition with a system built around a simple idea of making an agent’s harness smaller, clearer, and easier to verify.
  11. Oct 8, 2026 · Research paper · 1 source
    GLIO2: A GPU-Parallelized Tightly-Coupled LiDAR-Inertial-GNSS System for Robust and Real-Time Global Localization and Mapping
    We propose GLIO2, a tightly-coupled LiDAR-Inertial-GNSS system whose GPU-parallel front-end jointly optimizes scan-to-multiscan LiDAR, IMU pre-integration, and raw GNSS measurements in a single sliding-window factor graph, sustaining real-time operation on edge hardware.
  12. Oct 8, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.905-beta: Sandboxing is here!
    We're introducing Windows, Mac and Linux sandboxing in Unsloth!
  13. Oct 8, 2026 · Research paper · 1 source
    Traceable World State: A Provenance-Aware State Representation and Deterministic Replay Framework for Robotic Systems
    We present Traceable World State (TWS), a middleware-neutral semantic representation and reference runtime for provenance-aware robot world state.
  14. Oct 8, 2026 · Opinion / analysis · 1 source
    AI breakthroughs in robotics won’t change your life any time soon
    This would be Tesla’s Optimus, an AI-powered humanoid robot that Elon Musk, the company’s CEO, believes will be “not just Tesla’s biggest product ever, but probably the biggest product ever,” headed to work on factory floors and, later, in our homes.
  15. Oct 8, 2026 · Research paper · 1 source
    DivMoE: Fine-Grained MoE Upcycling via Cross-Domain Expert Composition
    DivMoE introduces domain-specialized fine-grained expert initialization, deriving experts from dense models that have undergone domain-adaptive continual pre-training, and diversity-constrained routing, a hard structural constraint guaranteeing that each token activates experts from distinct domain groups.
  16. Oct 8, 2026 · Research paper · 1 source
    Read What Matters: Query-Adaptive Quantization for KV Caches
    KV-cache entries are stored before their future queries are known, but each decoding query needs precision in different places.
  17. Oct 8, 2026 · Research paper · 1 source
    Demonstrating Arena 5.0: A Photorealistic ROS2 Simulation Framework for Developing and Benchmarking Social Navigation
    Building upon the foundations laid by our previous work, this paper introduces Arena 5.0, the fifth iteration of our framework for robotics social navigation development and benchmarking.
  18. Oct 8, 2026 · Research paper · 1 source
    GATOR: Generative and Agentic 3D Object Reconstruction From Casual Images
    We present GATOR, a generative and agentic framework that recovers textured object assets and their scene-relative pose from one or more images.
  19. Oct 8, 2026 · Research paper · 1 source
    SatFix: Absolute Visual Localization of UAVs in Satellite Maps from a Single Oblique Image
    We study absolute metric UAV localization within a provided geo-referenced satellite region, recovering continuous map position and viewing heading from a single oblique image or a short multi-view clip.
  20. Oct 7, 2026 · Research paper · 1 source
    The Missing Fourth Term for the Emulation Tensor Memory Equilibrium (TME) Model: The Residue Deconstruction Cost
    The Tensor-Memory Equilibrium (TME) model of "FP8 is All You Need (Part 1)" calculates the execution time of Ozaki Scheme II emulation of fp64 as the maximum of a tensor-core term and a High-Bandwidth Memory (HBM) traffic term, plus a per-output reconstruction term.
  21. Oct 7, 2026 · Opinion / analysis · 1 source
    The Machines that Make the Machines
    However, the process of assembling GB300 trays requires skilled physical labor in factories across the world.
  22. Oct 7, 2026 · Opinion / analysis · 1 source
    Multimodal open d1 decision models for the edge
    Best decision model under 10B on the Decision Index 0.2.1: d1-3B scores 48.57, ahead of every 4B and 9B model and of Decider 35B-A3B (47.11).
  23. Oct 7, 2026 · Product / feature launch · 1 source
    Scaling Decision Optimization to 100 Million Variables and Beyond with mPDLP in NVIDIA cuOpt
    NVIDIA cuOpt GPU-accelerated decision optimization can already deliver speedups of more than 10x over CPU...
  24. Oct 7, 2026 · Opinion / analysis · 1 source
    Faster Scientific Image Analysis with NVIDIA cuPhoton
    Observatories and telescopes, lasers and X-ray light sources, and other high-throughput instruments generate image data faster than CPU-bound pipelines can...Observatories and telescopes, lasers and X-ray light sources, and other high-throughput instruments generate image data faster than CPU-bound pipelines can process it to support timely decisions.
  25. Oct 7, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.904-beta: Train your own Decision model
    Turn any text or vision LLM into a Jev-style decision model in Unsloth, with decision accuracy going from 30% to 80%.
  26. Oct 7, 2026 · Research paper · 1 source
    Reproducible LLM Inference Benchmarking: A Sequential Isolation Protocol for Regression Testing
    Reproducible benchmarking of Large Language Model (LLM) inference is challenging because repeated measurements can vary with execution and system state.
  27. Oct 6, 2026 · Product / feature launch · 1 source
    Introducing Mistral Large 4: Le chonk
    Introducing Mistral Large 4: Le chonk
  28. Oct 6, 2026 · Opinion / analysis · 1 source
    Scale Bitwise-Deterministic Pretraining with NVIDIA Megatron Core
    Bitwise determinism makes large-scale pretraining easier to debug, validate, and resume reproducibly.
  29. Oct 6, 2026 · Opinion / analysis · 1 source
    How DOCA GPUNetIO Unifies GPU-Initiated Networking Across the NVIDIA Software Stack
    GPU applications increasingly need networking and data movement to behave like first-class GPU-controlled operations rather than host-driven services.
  30. Oct 6, 2026 · Open-source release · 1 source
    unslothai/unsloth v0.1.903-beta: New Browser + Voice Cloning
    This release adds a browser inside Unsloth (browser use coming very soon), so files, web pages and pages the model writes open right beside your chat.

Often appears with