List from agenticcoding.com
AI for Developers
Tools, techniques and ideas for building software with AI.
Containarium v0.101.2
What changed Added report whether a run’s model traffic is scanned (#2367) (#2465) wire the inbound response scan into the daemon’s model gateway (#2367) (#2464) cache the inbound policy read; exp
The Non-Compassionate Case for Model Welfare
Anthropic is updating their usage policy to prohibit “sustained and needless abusive or cruel behavior toward [their] models.”

Release 0.57.1
JSONL sessions now report final usage, normalize token accounting across sessions, validate saved ChatGPT sessions before reuse, and preserve usage-limit error handling. Background-job reconciliation and snapshot handling, supervisor chunking and question/restriction distinctions, gate scope and historical evidence, and spinner behavior around tracing output have been corrected.
Release Candidate v1.6.13-rc.99
Automated release candidate build from main.\n\nnpm: npm install gitnexus@rc\nVersion: 1.6.13-rc.99\nTarget base: 1.6.13 (rc #99)\nSource commit (main): 22aeeeb\nRelease commit (versioned tree): 446e52c\n\nRelease candidates are pre-stable builds intended for early testing. Stable releases remain on the latest dist-tag. What's Changed 🚨 Security Improve MCP startup compatibility and lazy-load CLI commands by
Claude Code vs local Qwen on a Mac in 2026: which should you go for?
The switch posts are half right. They skip the prompt Claude Code sends every turn, and how slowly a local model on a Mac reads it. Continue reading on Towards AI »


How to turn AI production feedback into better agents
Your AI agent’s service is healthy. Its answers might still be getting worse.

Much more than you wanted to know about wombats
Epistemic status: infected. Amateur Wombat enthusiast. Facts checked and footnoted. AI use: research (fact-check and extra facts), review, grammar and light phrasing. I remember very vividly how my unhealthy interest in wombats started. It was an ordinary day and I was not anticipating anything special. I was browsing YouTube and stumbled upon a video called Wombat attack inside tunnel.[1]Back then I didn't even know what a wombat was. In the video a guy went into an abandoned pipeline in Australia and a wombat entered it right after him.
What the [cyber] am I paying Anthropic for?
I was trying to set up Matrix calls for my LAN homeserver, tunneled through an external gateway using Chisel. I asked Claude to help me set up a secure configuration.
V0.211.3: chore(release): 0.211.3 (#973)
chore(release): 0.211.3 Release-time preparation from merged main. Feature PRs do not carry version or changelog edits. A maintainer must approve the bot-created Actions workflow run before its checks can run.
How to build a trusted intelligence layer for reliable AI agents
From context engineering to trusted intelligence: a practical architecture for relevant evidence, governed access and human control.

V0.203.4: fix(deps): publish the 0.203 line on agent-core 0.10.2 (0.203.4) (#971)
fix(deps): publish the 0.203 line on agent-core 0.10.2 (0.203.4) v0.203.3 (#970) did not publish. Its packed-consumer check found two agent-interface copies, because agent-core 0.10.3 (2026-10-06) moved to agent-interface 3 in a patch release, and the 0.203 line's ^0.10.2 now resolves it beside the consumer's interface 2.
The Agents in Production Aren’t Mine. Here’s What Their Server Sees
What our MCP server logs about agents we don’t control, and what my own scheduled agents really costSomewhere between June 3 and September 2, 2026, an agent asked our production MCP server for a tool called GBContent.getItems(sectionId, opts, onOk, onErr). Parentheses, parameter names, callbacks, the whole JavaScript signature, sent as the name of a tool. The server has never had a tool by that name.

So, am i in? Fully in or half in? (Startup)
submitted by /u/IT-BAER [link] [comments]

V0.203.3: feat(meta-eval): backport the judge gate to 0.203 (0.203.3) (#970)
feat(meta-eval): backport the judge gate to 0.203 (0.203.3) GTM pins agent-runtime 0.291.0 and agent-knowledge 18.0.0, whose agent-eval peer windows end below 0.204. The judge gate shipped in 0.210/0.211 (#961, #963), so GTM's adoption (#1445) broke its peer-floor ship check and was reverted (#1448). No agent-runtime or agent-knowledge release on GTM's interface-2 cohort admits agent-eval 0.211.

Two new POTENTIAL planets?
got the idea from u/This_Cell_1829's post last week about his TESS planet candidate and wanted to see if it could be done again. im not an astronomer, was just curious abt how many undiscovered planets there are potentially. I used opus 5.5 for this we went through about 2,900 nearby small stars (all within ~200 light years). picked ones tess saw in a few diffrent years but only in its wide field images so nobodys pipeline ever searched all those years together.

Why Your Agent Loop Costs More Than Its Model Price Tag
Opus costs $4 per million input tokens. Haiku costs $1 [1]. Swap an agent from one to the other and you’d expect the bill to drop by roughly 75%. In a tool-heavy agent it often drops by a fraction of that, because the per-token price was never the biggest variable in the equation. The loop around the model is.

Github-v1.2.26: chore: promote v2 to dev
Takes the v2 tree, with packages/console, packages/web, packages/stats, packages/function, infra and sst.config.ts merged up to latest dev, and retargets V2 publishing and deploys from the v2 branch to dev.
Utilitarianism and Autism
This is a crosspost from my substack. Effective altruism has received a lot of media attention since the Department of War CTO of the United States declared that there follow “Americanism, NOT Effective Altruism.” One of the main critiques of effective altruists is to accuse them of being utilitarians. It’s an especially tiring criticism, especially from other academics, who should know better than to spread disinformation. As one of the founders of the movement put it:
Showing 20 of 100 articles