Skip to content

Topic

Context Engineering

Curating what information enters an LLM's context window at inference time, including retrieval, compression, memory and tool-result filtering.

Current stories

build1 publisher

About forty agent tools now read Anthropic's SKILL.md workflow format

Anthropic's open SKILL.md format is read by roughly forty agent tools, among them Codex, Cursor, Copilot and Gemini CLI. Each installed skill costs about 100 tokens a session until a task matches its description and the full workflow loads.

Publishers:dev.to

Reality

Evidence45
Adoption45
Hype gap+20
Incentives
Insufficient
Confidence35
build1 publisher

LangChain drops about 4,000 base input tokens from every default Deep Agents turn

The harness lost its hidden system prompt, 43% of its builtin tool descriptions and its todo list middleware. LangChain's own footnote says reward confidence intervals span zero for every model tested, so the evals settle the token saving more firmly than the quality.

Publishers:langchain.com

Reality

Evidence58
Adoption30
Hype gap+18
Incentives82
Confidence46
build1 publisher

Writing the task text before the research encodes a guess

Anton Brilliantov spent eight parts specifying agent handoffs down to a single acceptance command. His ninth names the task shapes where those facts are still unknown, and the reading that has to happen first.

Publishers:dev.to

Reality

Evidence32
Adoption10
Hype gap−10
Incentives25
Confidence45

Earlier coverage

  1. TOON's 49% character saving falls to 33% against JSON that was already minified

    Build · August 31, 2026 · 1 publisher

  2. Testing a skill means running the scenario again on the next model version

    Build · August 31, 2026 · 1 publisher

  3. After five months behind main, classifying 312 conflict hunks helped turn a two-week rebase estimate into 11 hours

    Build · August 30, 2026 · 1 publisher

  4. An unsupervised agent loop billed $38 before anything in the system said stop

    Build · August 30, 2026 · 1 publisher

  5. A memory layer beat CLAUDE.md by 22.2 points, and 54 of 72 test pairs never moved

    Build · August 24, 2026 · 1 publisher

  6. Agents denied a fact do not stop, and read traces cannot tell you they lied

    Build · August 24, 2026 · 1 publisher

  7. Unity's AI problem is not the prompt: the load-bearing context lives in the prefabs

    Build · August 24, 2026 · 1 publisher

  8. Coding agents cost $4,125 a month because 73% of it is context you already sent

    Build · August 23, 2026 · 1 publisher

  9. Anthropic cut 80% of Claude Code's system prompt and the evals did not move

    Invest · August 23, 2026 · 1 publisher

  10. The fix for a confused coding agent is a smaller input, not a bigger window

    Build · August 23, 2026 · 1 publisher

  11. OpenClaw makes the channel the architecture, and the reasoning loop a lodger

    Build · August 23, 2026 · 1 publisher

  12. Agent memory that learns from wins is grading the user, not the context

    Build · August 23, 2026 · 1 publisher

  13. The personal agent is a folder, not a model: four files and less memory than you thought

    Product · August 22, 2026 · 1 publisher

  14. Open Knowledge Format: when the fact already has a name, the chunker is the bug

    Build · August 22, 2026 · 1 publisher

  15. Model choice is becoming a line item, and the differentiator moved up the stack

    Product · August 22, 2026 · 1 publisher

  16. Netflix's plain-text recommender won on 40x fewer labels, and the bill moved rather than vanished

    Build · August 22, 2026 · 1 publisher

  17. Stage Gates Got Cheap Again, And That Is The Whole Argument For "Waterfall 2.0"

    Product · August 21, 2026 · 1 publisher

  18. Three tools, three spellings of the same glob: agent rules do not port

    Build · August 20, 2026 · 1 publisher

  19. Pocock's /wayfinder bets that the bottleneck in overnight agents is planning, not code

    Build · August 20, 2026 · 1 publisher

  20. JFrog measured 847 log lines to find 9, and that ratio is now a budget line

    Build · August 20, 2026 · 1 publisher

  21. Adronite's Codistry makes token count, not context window, the axis of competition

    Product · August 19, 2026 · 2 publishers

  22. Agent memory has a dose-response curve, and the cheapest dose won the biggest gain

    Build · August 18, 2026 · 1 publisher

  23. The reason your agent gets worse after an hour is that nothing ever leaves the context window

    Build · August 18, 2026 · 1 publisher

  24. Your inference bill is an architecture defect: declare the task before you call the model

    Build · August 18, 2026 · 1 publisher

  25. The payload is rebuilt every turn, so stop treating your prompt as a shipped artifact

    Build · August 15, 2026 · 1 publisher

  26. Context rot at 15 iterations: two toolkits that move the spec into Git

    Build · August 15, 2026 · 1 publisher

  27. Agent reliability is a harness problem, not a prompt problem

    Build · August 15, 2026 · 1 publisher