AION
Research paperMLOps, Tooling & Infrastructure · Large Language Models1 source · Oct 6, 2026

Harness Engineering for Software Engineering via Modular Executable Dev-Primitives

Building on Dev-Primitives, we propose HERMES, a Harness Engineering framework for software engineeRing via Modular Executable Dev-PrimitiveS, which instantiates these primitives at repository scale through a dependency-aware dynamic activation mechanism and a bug diagnosis mechanism that maps execution evidence back to the components that must be revised.

Key points

  • Large language models (LLMs) equipped with terminal access have demonstrated strong capabilities in automating software engineering tasks.
  • However, existing agents remain brittle on long-horizon workflows, where they must repeatedly reconstruct program state scattered across source files, configurations, tests, dependencies, and runtime behavior, leading to increasingly long interaction histories, context explosion, and semantic drift.
  • To address these challenges, we introduce Dev-Primitives (Development Primitives), a modular and executable abstraction that transforms repository components from passive software artifacts into active participants in software engineering.
  • Each Dev-Primitive pairs a repository artifact with a resident LLM, which gives the artifact an agent-native interface grounded in its own implementation and dependencies, enabling natural-language reasoning, inter-component communication, and localized self-modification.

Sources (1)

  • [1]Harness Engineering for Software Engineering via Modular Executable Dev-Primitives
    arXiv (AI, ML, NLP, CV, robotics, multi-agent) · Oct 6, 06:32 AM
    Building on Dev-Primitives, we propose HERMES, a Harness Engineering framework for software engineeRing via Modular Executable Dev-PrimitiveS, which instantiates these primitives at repository scale through a dependency-aware dynamic activation mechanism and a bug diagnosis mechanism that maps execution evidence back to the components that must be revised.
    Large language models (LLMs) equipped with terminal access have demonstrated strong capabilities in automating software engineering tasks.

Extractive summary: sentences quoted from the sources.

Before this

  1. Oct 4, 2026nerkyor/Qwen3.8-27B-Coder390-EfficientThink-Opus5.5-GPT6Astra-Grok4.7-DSV4Pro-K3-SFT-RLOO-MTP-DFlash2
  2. Oct 1, 2026Inherit-MAS: Test-Time Evolution of Multi-Agent Systems through Workflow and Execution Inheritance
  3. Sep 30, 2026huggingface/transformers v5.18.0: Release 5.18.0
  4. Sep 29, 2026NVIDIA/TensorRT-LLM v1.3.0rc29
  5. Sep 28, 2026Holo4: powering generalist computer-use agents
  6. Sep 9, 2026vllm-project/vllm v0.29.0

Related