Andreas Belitz
Blog Notes Projects About

All Posts

All AI Engineering Tools & Workflow 3D Printing & Making LLM Patterns Career & Thinking Infrastructure
Measuring my agent harness against plain Claude Code
Sep 29, 2026 · 8 min read · AI Engineering

Measuring my agent harness against plain Claude Code

I built a multi-agent harness and ran it head to head against a plain control on a real benchmark. On these tasks it cost about five times more and did not score better. What it bu...

agents evaluation benchmarking swe-bench harness
Your coding agent isn't faster on Linux. Your environment is just less of a variable.
Sep 13, 2026 · 7 min read · Tools & Workflow

Your coding agent isn't faster on Linux. Your environment is just less of a variable.

Coding agents really do run better on Linux, but mostly for contingent reasons. The lasting point is that determinism matters more than raw speed once you're running agents at volu...

coding-agents developer-experience docker determinism
An agent picked the artwork, and it didn't pick mine
Aug 23, 2026 · 11 min read · AI Engineering

An agent picked the artwork, and it didn't pick mine

The AI Art Magazine is running an open call where agents submit the work and agents judge it. I registered one and let it choose. It didn't choose the piece I wanted, or the one it...

agents verification art making claude
A rule nobody checks is a decoration
Aug 11, 2026 · 15 min read · AI Engineering

A rule nobody checks is a decoration

I've been running my context-management discipline on a real client programme where most of the execution is done by agents. The principles held up. The rules that were only writte...

context-management agents governance verification delivery
The Boiler Project, and the context engineering that made it work
Jul 20, 2026 · 13 min read · 3D Printing & Making

The Boiler Project, and the context engineering that made it work

A Home Assistant optimisation that turns a hot-water tank into the most valuable battery in the house, and a worked example of how deliberate context management turned a rumour and...

home-assistant making context-management self-hosting energy automation
The context operating model I use for the projects that don't fit in a chat
Jul 7, 2026 · 9 min read · LLM Patterns

The context operating model I use for the projects that don't fit in a chat

My .md-file memory system was built for code. This is the version I run for consulting work (RfPs, discoveries, research) out of one shared folder. It uses the same hierarchical co...

context-management cowork knowledge-work workflow subagents
Context rot has a name now: here's what four months of living with it taught me
Jun 18, 2026 · 9 min read · AI Engineering

Context rot has a name now: here's what four months of living with it taught me

A new piece lays out the mechanisms behind context rot and gives concrete token thresholds. I have been hitting those thresholds by hand since February. This post covers where the...

context-management context-rot attention agents claude
Claude Fable 5: what a new tier above Opus means for context management and agent teams
Jun 10, 2026 · 10 min read · AI Engineering

Claude Fable 5: what a new tier above Opus means for context management and agent teams

Anthropic launched Claude Fable 5 yesterday, its first Mythos-class model, priced at double Opus. The context window did not grow. What changed is the management layer around it.

claude fable-5 context-management agents agent-teams anthropic
Bigger context, worse agents
May 22, 2026 · 5 min read · AI Engineering

Bigger context, worse agents

I've been experimenting with feeding agents larger volumes of context. The pattern is steady: the bigger I make it, the worse it looks. Meanwhile, the rest of the field is racing t...

context-management context-rot knowledge-graphs retrieval agents
I built an AI to watch a bird. She left.
May 17, 2026 · 6 min read · 3D Printing & Making

I built an AI to watch a bird. She left.

I built an AI to watch a Common Redstart nesting in my garden. Two weeks later she left the box and never came back. The AI didn't notice — and that was the lesson.

side-project claude gemini prompting vision ai-architecture sweden
← Newer Page 1 of 3 Older →

© 2026 Andreas Belitz v:1ed2117

RSS GitHub LinkedIn