In-context learning
Also known as: ICL
16stories this week
17last 30 days
17all time
Timeline
- Oct 8, 2026 · Research paper · 1 sourceCompile the Table: Query-Calibrated Operator Compression for Tabular In-Context LearningWe propose QCOC (Query-Calibrated Operator Compression), which exploits the exchangeability and repeated use of in-context examples by compiling their full KV cache once into compact memory shared across subsequent queries.
- Oct 8, 2026 · Research paper · 1 sourceDo LLMs Learn from Rewards in Context? : Rethinking the role of reward in In-Context Reinforcement LearningLLM agents increasingly improve at inference time by accumulating experience in context rather than by updating parameters.
- Oct 8, 2026 · Research paper · 1 sourceSFT-as-Context Mitigates Forgetting in Supervised Fine-TuningWe introduce SFT-as-context, a training-free method in which the parent model uses the SFT model's response as context to answer the query.
- Oct 7, 2026 · Research paper · 1 sourceWhat can linear attention learn from nonlinear teachers in-context?Linear attention is a tractable model for understanding the mechanisms governing in-context learning in transformers.
- Oct 7, 2026 · Research paper · 1 sourceBEANS-Next and ROOTS: Broadening Audio-Language Capabilities for BioacousticsIn this work, we introduce BEANS-Next, a benchmark grounded in a taxonomy of bioacoustics tasks spanning acoustic perception, biological category recognition, scene understanding, and in-context learning.
- Oct 7, 2026 · Research paper · 1 sourceWhen to Unpair: Regulating Pairing Dependence in Medical Visual In-Context LearningVisual in-context learning (ICL), well suited to label-scarce medical imaging, uses support image-label pairs to demonstrate input-output mappings, while the labels collectively indicate the requested task.
- Oct 7, 2026 · Research paper · 1 sourceEfficient Provably Private Classification with a Tabular Foundation ModelHere we introduce PrivTab, an easy to use tabular foundation model for differentially private classification that embeds a privacy mechanism within its architecture.
- Oct 7, 2026 · Research paper · 1 sourceLeaner Transformers Can Easily Learn to ClusterRecent work shows that transformers can exactly perform Lloyd's algorithm for $k$-means clustering with $n$ points in $d$ dimensions with an embedding size $d{\textsf{emb}} = d+k$ (thus, requiring attention projection matrices of size $(d+k)^2$).
- Oct 6, 2026 · Research paper · 1 sourceSpatial Induction Heads: In-Context Learning of Multidimensional Cellular AutomataWe introduce spatial induction heads, two-layer gather-and-match circuits in which the first layer reconstructs the relevant spatial neighborhood and the second matches the resulting configuration against earlier occurrences.
- Oct 6, 2026 · Research paper · 1 sourceGeneICL: A Tabular Foundation Model for Bulk TranscriptomicsTowards this end, we introduce GeneICL, a 4.2M-parameter tabular foundation model combining a semi-synthetic pretraining prior built from measured bulk expression profiles with a parameter-efficient recurrent architecture.
- Oct 6, 2026 · Research paper · 1 sourceTowards In-Parameter Memory Augmentation for Large Language ModelsRecently Large Language Models (LLMs) and LLM-based agents increasingly need to incorporate knowledge acquired after pretraining, e.g., domain facts, user preferences, documents, and interaction experience.
- Oct 6, 2026 · Research paper · 1 sourceThe Standardization Trap: Certifying Joint Label Processing in Tabular Foundation ModelsLinear regression and kernel smoothing offer tractable explanations of in-context learning: in both, the features determine the weight assigned to each context label.
- Oct 6, 2026 · Research paper · 1 sourceTICDA: Tabular In-Context Data AttributionWe introduce TICDA, a method that measures the influence of every demonstration in the context directly from linear surrogates trained on TFM latent embeddings, in a single forward pass and at negligible cost.
- Oct 6, 2026 · Research paper · 1 sourceContinuous Memory MachinesTo that end, we introduce the Continuous Memory Machine (CMM), a recurrent architecture with matrix-valued short- and long-term memory states serving distinct functional roles.
- Oct 6, 2026 · Research paper · 1 sourceAdaptive Mean Estimation by In-Context Learning: A Gradient-Flow AnalysisPrior Fitted Networks (PFNs) such as TabPFN now rival established statistical procedures across prediction and estimation tasks.
- Oct 6, 2026 · Research paper · 1 sourceAdversarially Trained Linear Transformers Are Optimal Robust In-Context Learners for Gaussian MixturesSpecifically, we show that, for a family of Gaussian-mixture classification tasks, a sufficiently deep linear transformer adversarially trained across tasks can asymptotically attain the robust Bayes error on previously unseen tasks through in-context learning from clean demonstrations.
- Sep 29, 2026 · Research paper · 1 sourceIn-context Robot Learning Made Simple: A Democratized Recipe for Manipulation TasksWe study robotic in-context learning (ICL), an emerging paradigm that enables robots to infer and execute tasks from visual demonstrations.