Durable agent work
Team coding agents are being framed as task systems with memory, ownership, and validation records. Charlie’s post argues that a Slack thread, GitHub comment, Linear issue, scheduled wake, or review request should become a durable task that can create child tasks, track branches and pull requests, preserve test output, and handle follow-up questions. The post claims cost gains with small models, but it does not provide a controlled benchmark.
Two builder reports show why that task shape matters. Bearhug’s founder claims a non-programmer managed seven coding agents for 21 days, spent about $5,000, and produced more than 75,000 lines of production code for an executive talent marketplace. A YIMBY civic-data post describes three public local-housing data projects built with Claude Code in a few hours each. These are production anecdotes. They show speed and scope, while leaving quality, maintenance cost, and reproducibility largely unmeasured.