PEFT
7stories this week
10last 30 days
11all time
Timeline
- Oct 8, 2026 · Open-source release · 1 sourcehuggingface/trl v1.15.0SFT, DPO, KTO, GRPO, RLOO and Distillation now score tokens with a fused LM head: a Triton kernel projects the hidden states through the LM head in tiles and reduces to per-token log-probs and entropy directly, so the [batch, seq, vocab] logits tensor is never built.
- Oct 8, 2026 · Research paper · 1 sourceWhere to Adapt Matters: Layer-Selective Fine-Tuning for Capability RetentionParameter-efficient fine-tuning (PEFT) enables large language models (LLMs) to adapt to specialized tasks, but often at the cost of degrading general capabilities acquired during pretraining.
- Oct 8, 2026 · Research paper · 1 sourceStability-Plasticity Balance via Singular-Vector Selection in LLM Continual LearningDomain-specific continual adaptation of LLMs risks catastrophic forgetting, creating a fundamental tension between acquiring new capabilities and preserving those learned during pretraining.
- Oct 6, 2026 · Research paper · 1 sourceAre Parameter-Efficient Fine-tuning Methods Really Different?Parameter-efficient fine-tuning (PEFT) offers many parameterizations, yet their methodological and functional differences remain unclear.
- Oct 6, 2026 · Research paper · 1 sourceMemFLoRA: Memory-Floor LoRA for CNN Adaptation at the EdgeThis paper introduces Memory-Floor LoRA (MemFLoRA), a low-rank CNN adapter built around a memory-first design principle rather than a direct application of transformer-oriented LoRA.
- Oct 6, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.19.0: Release v5.19.0EmbeddingGemma 2 is a multimodal embedding model from Google built on the Gemma 4 architecture.
- Oct 6, 2026 · Research paper · 1 sourceDynamic Positional Attention Modulation for Parameter-Efficient Fine-Tuning of Large Language ModelsIn this work, we propose DyPAM (Dynamic Positional Attention Modulation), a PEFT method that adapts how positional information contributes to attention by operating directly on the query and key representations.
- Oct 1, 2026 · Product / feature launch · 1 sourceunslothai/unsloth v0.1.902-beta: Command Palette + Desktop UI/UXThis release brings faster navigation, shareable run settings, and clearer errors to Unsloth Desktop.
- Oct 1, 2026 · Open-source release · 1 sourcehuggingface/peft v0.21.2This is a PEFT release fixes an issue that prevented encoder-decoder models to work when using Transformers ≥ 5.18.0.
- Sep 29, 2026 · Open-source release · 1 sourcehuggingface/peft v0.21.1This is a small PEFT release to enable Tensor Parallel (TP) to work properly with PEFT.
- Jun 15, 2026 · Open-source release · 1 sourcehuggingface/transformers v5.12.1: Patch release v5.12.1Updated the lower bound for PEFT and a fix for auto tokenizer to properly resolve the mistral tokenizer (when mistral-common is installed).