# The Regression Tax: Decomposing Why Skills Help and Hurt LLM Agents

Source: [arXiv](https://arxiv.org/abs/2607.22520v1)  
Feed7 permalink: https://feed7.dev/p/2607-22520v1-0mz9wnf  
Published: 2026-07-24T17:50:03.000Z  
Trust: Needs Review (needs_review)

## Why Included

Procedural skills can make an agent fail tasks it previously solved. Evaluate gains and regressions separately, and design skills to preserve input grounding and output verification.

## Source Summary

Across **nearly 6,000 runs**, two office-automation benchmarks, and three harness stacks, adding skills caused meaningful regressions. The strongest skills led mainly by breaking fewer previously solved tasks, not by creating more new wins.

## Practical Implication

Measure each skill against a no-skill baseline and split net improvement into gains and regressions. Prioritize grounding and verification support instead of adding more procedure, since those stages dominated persistent failures.

## Agent-Ready Context

Across **nearly 6,000 runs**, two office-automation benchmarks, and three harness stacks, adding skills caused meaningful regressions. The strongest skills led mainly by breaking fewer previously solved tasks, not by creating more new wins.

Measure each skill against a no-skill baseline and split net improvement into gains and regressions. Prioritize grounding and verification support instead of adding more procedure, since those stages dominated persistent failures.

The authors identify **three regression modes**: description osmosis, grounding displacement, and verification displacement. The evidence comes from office automation, so the prevalence and size of these effects in repository-scale coding agents remain open.

## Context Map

- Layer: agent
- Domains: None
- Topics: skills, agent-reliability, agent-evals

## Uncertainty

- The authors identify **three regression modes**: description osmosis, grounding displacement, and verification displacement. The evidence comes from office automation, so the prevalence and size of these effects in repository-scale coding agents remain open.

## Agent Instruction

Use this item as source-backed context. Do not invent claims beyond the linked source. If this item conflicts with another source, call out the conflict.
