Technique

RLHF

Also known as: reinforcement learning from human feedback

1stories this week
1last 30 days
2all time

Timeline

  1. Oct 7, 2026 · Research paper · 1 source
    A Scoping Review and Experimental Study on Reinforcement Learning from Human Feedback for Human-Robot Collaboration
    Human-Robot Collaboration (HRC) can facilitate mass customisation in Industry 4.0, with Reinforcement Learning from Human Feedback (RLHF) representing a promising approach for developing safe AI-based robots.
  2. Jul 27, 2026 · Open-source release · 1 source
    vllm-project/vllm v0.26.0
    New Inkling model family with a full support stack: base modeling (#48799), piecewise CUDA graph support (#48822), Hopper FA4 relative attention (#48858), MTP=1 speculative decoding (#48869), LoRA (#48884), and standard ModelOpt NVFP4 quantization (#48990).

Often appears with