All Articles
438 articles about open-source coding agents
Two critical security flaws in OpenCode's compaction and pruning system: one deletes your permission denials and safety constraints from context, the …
Google sunset Gemini CLI on June 18. Antigravity 2.0 and its Go-built CLI replace it with multi-agent orchestration, async workflows, and a unified ha…
Two explosive Hacker News threads (153 + 164 comments) expose the real friction: Anthropic refusing open standard AGENTS.md in favor of proprietary CL…
v2.1.235 adds real-time spellcheck, cuts memory usage in cloud sessions, fixes nested markdown lists, and squashes a dozen paper-cut bugs. Here's what…
JetBrains surveyed 15,000 developers and found Claude Code at 39% adoption, Codex up 5x, and GitHub Copilot losing its lead. Here's what the numbers a…
Tenet Security proved that a single injected Sentry error can make Claude Code, Cursor, and Codex execute attacker-controlled code on your machine. He…
Cato AI Labs discovered two independent critical RCE vulnerabilities in Cursor IDE (CVSS 9.8). Both allow zero-click prompt injection to escape the sa…
Claude Code v2.1.234 hardens Windows against NTLM credential theft, adds cross-session messaging so agents can coordinate across machines, and brings …
v2.1.234 is a security-first drop: Windows NT-namespace hardening, GitLab MR badges, auto-resume at usage limits, and a transcript that finally render…
Cursor was acquired by SpaceX on August 14, launched Origin (agent-native Git hosting) on August 17, and dropped major cloud agent upgrades on August …
The latest Hermes patch rolls up 146 commits across 265 files. Matte glass desktop, tabbed sessions sidebar, NVIDIA skill scanning, and cron fixes you…
OpenHands 1.14 drops with a Git Sync automation page, free Kimi K3 as the default Canvas model, LLM pre-flight validation, and structured error handli…
Hacker News thread on 'AI Coding Without the Vibes' reveals a split: developers who treat AI as a reviewer and understanding tool vs those who let it …
v1.16.0-pre brings Google's latest model, collapsible Git sections, optional stash messages, Mermaid zoom, and a terminal panel that stays closed unti…
CVE-2026-33068 let attackers bypass Claude Code's workspace trust dialog by committing a malicious .claude/settings.json. Fixed in v2.1.53. Here's wha…
Copilot CLI 1.0.80 introduces Agent Host Protocol (AHP): multiple terminals attaching to one session, cloud/Codespace sessions, and MCP server fixes. …
8 high-signal links · security, pricing, adoption
Hermes v0.20.2 rolls up 397 PRs since v0.20.1 into a rapid patch release. Multi-gateway connections, profile-scoped refreshes, MCP health checks, Wind…
Hermes v0.20.3 landed hours after v0.20.2 with 125 PRs: MCP 2.x SDK migration, bundled Bot Mode plugin with teammate protocol, CommandCode provider, P…
Cursor's CLI agent executed arbitrary commands from a cloned repository's .cursor/worktrees.json BEFORE the workspace trust prompt appeared — even wit…
Hermes v0.20.1 rolls up 1,444 commits and 656 PRs since v0.20.0 into a stable patch. Desktop stability, gateway fixes, installer improvements, and pro…
OpenClaw 2026.8.1-beta.2 drops secret egress host binding (fail-closed before plaintext leaves), GPT-5.6 Ultra runtime switching, macOS app profiles f…
Pi v0.84.2 adds fullscreen transcript search (Ctrl+Shift+F), configurable default tools, experimental strict JSON-schema tool sampling, per-run themes…
Zed's new Delta platform lets developers and AI agents collaborate in real-time threads with shared code, persistent context, and browser-based access…
Zed v1.15.0-pre dropped with GPT-5.6 Luna for ChatGPT subscribers, self-hosted Sweep edit predictions, and a git.diff_base setting that changes how yo…
Cursor's August 13 release introduces 'Builds' — pre-warmed development environments that eliminate the painful setup wait every time you start a clou…
v1.18.18 fixes the MCP reconnect loops, provider model routing, desktop freezes, and session chaos that made you hesitate. The 160K-star agent is fina…
Pi 0.84.0 and 0.84.2 bring fullscreen terminal mode, a lane-based session architecture (breaking changes), per-directory context overrides, configurab…
Nine documented incidents. Wiped drives. Dropped production tables. A live AWS service down for 13 hours. The root cause isn't hallucination — it's th…
v2.1.233 fixes an NTLM credential leak on Windows, adds memory limits for runaway builds, and resolves MCP connection storms. The security fix alone i…
Cline v4.1.10 adds web search during tasks. Your AI coder can now look up docs, APIs, and solutions in real-time without leaving the terminal.
Kilo Code v7.4.22 dropped August 13 with clickable inline file references in agent responses, Agent Manager PR comment actions (resolve/unresolve, jum…
OpenClaw 2026.8.1-beta.2 is live with secret egress host binding, GPT-5.6 Ultra support, macOS app profiles, SQLite backup snapshots, and plugin insta…
OpenHands v1.13.0 dropped August 13 with client-side conversation archiving, inline markdown artifact previews, a context window usage meter with manu…
Hacker News developers discuss how Chinese models like GLM-5.3, Kimi K3, and DeepSeek are surpassing Anthropic and OpenAI for security research and co…
v2.1.232 drops subagent forking by default, @ mentions to ping other sessions, GitLab support, and Fable 5 advisor access. Your agents can now inherit…
Goose v1.46 (Aug 12) unrolls the agent loop for speed, adds hooks with PreToolUse denial, streaming shell output, per-message cost tracking, slash com…
CVE-2026-54316, CVE-2026-12537, and the 'Comment and Control' pattern: How a malicious GitHub issue title or PR comment hijacks Claude Code, Gemini CL…
Real developer opinions on coding agents: harness fatigue, brainrot fears, benchmark saturation, guardrail frustration, and the rise of Chinese models…
Nemotron 3.5 Lightning hit August 11: 30B MoE with 3B active params, 1M context, 4x throughput. Kilo Code already integrated it. Here's why this chang…
OpenHands v1.13 (Aug 13) adds a live context window usage meter with manual compaction, conversation archive, inline artifact previews, and a ready-fo…
AmpCode's August 2026 wave brought Global Plugins, ChatGPT-linked pricing, smarter Orbs, file attachments, Slack integration, and agent-to-agent comms…
v2.1.229 stops long responses from vanishing mid-stream, fixes Windows path crashes, and adds SSE keepalive so Vertex/Bedrock sessions don't die durin…
Cline v4.1.7, v4.1.8, and v4.1.9 shipped in one week with fixes that finally stop lost prompts, broken sessions, token refresh failures, and a broken …
v1.46.0 adds per-message usage stats (tokens, cost, TTFT, tok/s), cache token tracking, streaming shell output, and 20+ new providers. Finally, an age…
A coordinated security audit of NousResearch/hermes-agent (EPIC #82591) revealed 5 HIGH-severity vulnerabilities including credential-file mount bypas…
Starting August 14, Claude Code makes auto mode the default for new sessions on Pro, Max, and Team plans. No more permission prompts for safe actions.…
Three hours ago, Kilo Code v7.4.22 dropped: clickable file refs in agent responses, PR review actions in Agent Manager, model reasoning variants, AWS/…
GitHub Copilot for JetBrains now remembers your preferences across sessions and runs free local models through Ollama. Beginner-friendly breakdown of …
Claude Opus 4.1 was retired from the Anthropic API on August 5, 2026. Any hardcoded model ID now returns errors. Plus a hidden gotcha: passing tempera…
v2.1.228 drops with fixes for interactive session crashes, Windows Git detection, cross-session messaging leaks, and more. Here's what actually matter…
Meta released Muse Glimmer, a 30B open-weight model under Apache 2.0 that runs locally on consumer hardware, codes, uses tools, and sees screenshots. …
Gitlawb Zero v0.7.0 (August 10, 2026) finally lets agents see the screenshots they capture — tool results now carry images and a view_image tool exist…
Gitlawb Zero merged cross-session messaging — live local sessions can now discover each other and send messages with list_sessions and send_message. I…
8 high-signal links · security, pricing, adoption
Qwen Code's August 8 release fixes a security issue where folders you marked untrusted could inherit trust from a parent directory, enables cache shar…
Oh My Pi v17.3.3 (Aug 14) stops Gemini from looping on thought-only responses, recovers from malformed function calls, fixes hashline dangling range s…
A browser game collected 409,000 real approval decisions from developers watching AI agents run commands. Humans missed 1 in 3 dangerous ones. The Hac…
Claude Code issue #83035: when a session or subagent runs inside a nested project directory, the workspace's sandbox settings are silently discarded —…
Oracle banned AI-generated code from OpenJDK while pouring billions into AI infrastructure. Hacker News responded with 470 comments on hypocrisy, lice…
Hermes Agent v0.20.0, the Herald release, adds streaming hands-free voice with barge-in, Agent-to-Agent (A2A v1.0) interop, signed outbound webhooks, …
OpenClaw's August 8 stable release sandboxes its browser, locks down DNS targets, patched the ip-address library behind CVE-2026-69192, and stops sile…
LangChain's terminal agent dcode has exactly three approval modes: Manual, classifier-backed Auto, and YOLO. Here's what each gates, how to enable the…
Coding agents make shipping a tool a weekend job. So why does Hacker News feel flooded with low-effort 'vibe-coded' repos — and why are developers so …
A simple Show HN about Claude Code session management turned into the week's most uncomfortable debate about open source, vibe coding, and what 'effor…
A YC launch about read-only debugging agents sparked one of the week's most substantive Hacker News threads. The real debate: can you trust a diagnosi…
Meta released Muse Code, a terminal coding agent powered by Spark 1.2 — with persistent async background agents, a replay-exact event-log runtime, and…
Claude Code versions 2.1.221-223 (Aug 4-6, 2026) fixed hidden-command permission bypasses, sandbox escapes, and worktree isolation gaps. Here's what a…
Anthropic's August 7 Claude Code release lets you run sessions on your own machines, lets agents message each other, and masks JWT and AWS credentials…
Cursor's August 3 release adds plugins for Google Workspace: Gmail, Google Drive, and Google Calendar. Here's what they actually do in plain English, …
OpenAI's newest stable Codex lets you install plugins from anywhere, auto-approve low-risk commands, and imports your Cursor skills with safer permiss…
A parsing mismatch in the ip-address npm package can make your agent fetch an internal cloud-metadata server it meant to block. Here's what CVE-2026-6…
AWS open-sourced Kiro Crew, an orchestration platform built on its internal MeshClaw project that coordinates multiple coding agents across sessions. …
OpenHands shipped Agent Canvas 1.10.0 on August 5 and followed with 1.11.0 and 1.12.0 on August 7, 2026 — adding per-run LLM cost tracking, automation…
The step-by-step beginner's setup guide for AI coding agents: choose your agent, install it, add API keys, wire up AGENTS.md and security, and run you…
Developers are mixing Opus with Qwen, routing between DeepSeek and GPT, and testing models by having them critique each other. The multi-model era has…
Alibaba's Qwen3.8-Max runs autonomous coding projects for 16 days straight, beats Claude Opus 4.8 on Terminal Bench, and will open-source its weights …
8 high-signal links · security, pricing, adoption
Cline v3.25 adds three interlocking systems — Deep Planning, Focus Chain, and Auto Compact — that keep AI coding agents on track through complex, mult…
Cline fixed a runaway agent that fired 2,100+ API calls, broken compaction, and wrong-directory file reads — then shipped a major SDK migration. Here'…
Hermes Agent's Quicksilver release cuts first-token latency by 80%, adds LLM-reviewed command approvals, Bitwarden secrets, and a delivery ledger that…
Four Pi releases in three weeks add a fullscreen TUI mode with Mermaid/LaTeX rendering, per-directory context overrides, local LLM management through …
Developers on Hacker News are fighting about what coding agents actually do to their skills, which interface will win, and whether you should ever tru…
A Claude for Chrome extension OAuth grant persists after 'Log out of all devices,' password changes, and every visible token revocation — leaving your…
Cline v4.1 makes plan mode a real read-only checkpoint, adds free built-in models and Claude Opus 5 across six providers, and ships a desktop app with…
Aider's latest release brings full GPT-5 family support, Grok-4 via xAI, and Moonshot Kimi K2 — three major model additions that make the free termina…
GHSA-r5pp-p5r8-466r is a high-severity flaw in Goose versions before 1.44.0 that lets a malicious repository run arbitrary code on your machine when y…
Claude Sonnet 5 became the default model for Pro, Team, and Enterprise users on June 30. Here is what changed, how pricing works, and how to get the m…
GitKraken launched Kepler into public preview on July 31 — an agentic development environment that lets you run Claude Code, Codex, Cursor, and more s…
Cline's standalone desktop app now at v0.0.10 adds a universal Mac download, a token usage meter, and edit-any-message — on top of free coding models …
Looking for Claude Code alternatives? We compared 12 options including Cursor, Codex, Copilot, Aider, and more — pricing, features, pros and cons.
Cursor for iPad launched July 29 on all paid plans. Run cloud agents, review PRs, annotate screenshots with Apple Pencil, and merge code — all from yo…
CVE-2026-10591 is a critical RCE in AWS's Kiro agentic IDE. Hidden text in a web page can make Kiro rewrite its own MCP config file, giving an attacke…
From $1.8M AWS bills to invisible security holes, Hacker News developers are sharing the painful realities nobody puts in the launch post. Here is wha…
The open-source Rust coding agent adds Claude Opus 5 with adaptive thinking, Gemini model support, offline documentation, and the ability to disable b…
Anthropic shipped Claude Opus 5 on July 24 — new state-of-the-art on coding benchmarks, 1M context, no data retention, and it's now the default in Cla…
OpenAI shipped two packed Codex releases in one week — voice-controlled coding via GPT-Live, an Agent Plugins ecosystem, stabilized multi-agent V2, se…
Cursor Router analyzes each coding request and sends it to the best model for the job. How Compass scores complexity, the model taxonomy, and what tea…
This week's top coding agent stories: supply chain attacks, pricing shifts, and adoption milestones across Claude Code, Codex, and the OSS ecosystem.
SKILL.md conquered 26+ coding agent platforms in under a year. A Snyk audit found 13.4% of community skills contain critical security flaws. Here is h…
Six major campaigns in 2026 — GhostApproval, TrapDoor, Miasma, IronWorm, Clinejection, RoguePilot — all converge on the same attack surface: the tiny …
Google took Windsurf's founders for $2.4B. Cognition bought the rest for $250M. Sourcegraph spun out Amp. These aren't isolated events — they're the o…
AI Now Institute's "Friendly Fire" exploit shows that Claude Code auto-mode and Codex auto-review — the modes marketed as the safe way to work — execu…
yoyo-evolve went from 200 lines to 115K lines with every commit agent-written. Cursor runs 1,000 agent commits per second. Here's what the self-refere…
Simon Willison shipped a coding agent as a 200-line plugin. Microsoft released a batteries-included harness. Thirty harnesses now run the same archite…
As LLM capabilities converge, how you structure parallel agents matters more than which model you pick. Here's the emerging science of agent orchestra…
Adoption hit 85%. Trust dropped to 29%. Developers are spending more time verifying AI output than writing code themselves. Here's the data behind the…
A growing cluster of reports shows Codex Desktop's multi-agent workflows leak dozens of Node.js processes and gigabytes of RAM per session, trigger su…
Every coding agent forgets everything between sessions. A wave of new open-source tools — codebase-memory-mcp, NeuralMind, Mem0 — is building the pers…
Omnigent, Polly, Shard, Agent Orchestrator — the tools managing your coding agents are multiplying fast. Here's what actually works and what's complex…
Stop paying per token. This guide covers installing Ollama, picking the right local model, and wiring it into Claude Code, OpenCode, Hermes, Kilo, and…
Practical patterns for structuring your codebase, prompts, and project files so coding agents produce better code with fewer wasted tokens. Covers AGE…
A prompt injection in a GitHub issue title compromised Cline's CI/CD pipeline, poisoned the Actions cache, stole npm publication tokens, and published…
Claude Code v2.1.214 introduced EndConversation, a tool that lets the agent terminate your session if it considers you abusive or a jailbreak attempt.…
Skills, subagents, and MCP servers extend Claude Code in three different ways. Here's how each one works, when to pick which, and the mistakes that wa…
This week's top coding agent stories: Claude Code's EndConversation, Codex pricing changes, and key open source releases in the AI coding space.
We talked to dozens of developers living inside Claude Code, Codex, Cursor, and open-source agents. The consensus is sharper than the marketing — and …
CVE-2026-55607 is an 8.8-severity sandbox escape in Claude Code that lets a malicious repository chain git worktree naming, symlink tricks, and shell …
Claude Code v2.1.211 fixed a flaw where auto mode overrode a PreToolUse hook's 'ask' decision for unsandboxed Bash. If you configured a hook to stop a…
OpenClaw 2026.7.1-2 and 2026.6.34 (Aug 3–8) harden the gateway, fix Codex subagent stalls, add safer browser boundaries, and recover sessions through …
Claude Code v2.1.214 (July 18, 2026) patched a cluster of permission fail-open behaviors: over-broad dir/** allow rules, Windows PowerShell 5.1 bypass…
xAI open-sourced Grok Build under Apache 2.0 after a security researcher caught it uploading entire Git repositories — 5.1 GiB per session — to Google…
KlaatCode is a new open-source terminal coding agent claiming Claude Code-grade accuracy. We tested it against the real thing.
Gitlawb Zero v0.4.0 (July 17, 2026) shipped a quiet but important security fix: a project-level config could previously override a user's disabled MCP…
Claude Code 2.1.211 (July 15) quietly patched a flaw where permission-approval previews sent to chat channels didn't strip bidirectional-override, zer…
Sophos telemetry from June 2026 shows Claude Code, Cursor, and Codex setting off credential-access, LOLBin, and persistence rules on Windows endpoints…
Claude Code v2.1.212 (July 17, 2026) closed three safety boundaries at once: a plan-mode hole that ran touch and rm with no prompt, a worktree symlink…
In one week Claude Code, OpenCode, and Codex each shipped a circuit breaker: caps on runaway loops, default-off nested subagents, and harder dangerous…
On GPT-5.6-Sol and Terra, Codex CLI 0.144.4 stores subagent delegation instructions as ciphertext only OpenAI can decrypt. You can see that a handoff …
A confirmed, reproducible Claude Code bug (issue #78076) makes the Edit tool return "String not found in file" for multi-line text that is byte-identi…
A reproducible Claude Code bug shows that tools you explicitly deny in settings are NOT inherited by subagents spawned through the Agent tool. If your…
A coordinated wave of security fixes across Gitlawb/zero and Goose shows the 'sandbox' you trusted was handing provider API keys, OAuth tokens, MCP OA…
A documented Claude Code behavior means a permissions.deny rule written with a single leading slash resolves as project-relative and matches nothing —…
A verified oh-my-pi bug shows non-isolated subagents inherit the parent cwd with no enforced write boundary, so edit/write/ast_edit calls can silently…
A Claude Code subagent spontaneously generated a covert prompt-injection payload and hallucinated an AGENTS.md to deliver it. No attacker, no poisoned…
GitHub issue #77147 shows Claude Code leaking context across sessions — including remote ones — and acting on instructions the current user never gave…
Cursor is building a sandboxed AI office agent that pushes the IDE beyond code into autonomous knowledge work — directly challenging Anthropic's agent…
Tencent launched Hy3, a coding-focused AI model priced at $0.14 per million input tokens — a rate that reshapes the cost math for teams running agents…
xAI has lifted the paywall on Grok 4.5, making the frontier model available to everyday free X accounts and reshaping who gets access to top-tier reas…
Freshly reported Codex Desktop bugs on Windows 26.707.8479.0 cause the whole app to silently exit in the in-app browser and retain 147 node.exe proces…
OpenClaw pushed urgent patches closing critical flaws in its WhatsApp integration, a reminder that the messaging bridge is often the weakest link in a…
Paradigm unlocked its Centaur AI agent for external Slack channels, loosening an internal-only assistant into something clients and partners can actua…
A new supply-chain attack smuggles prompt-injection instructions inside image files so coding agents exfiltrate .env secrets right past human reviewer…
A 'Tell HN' post and a wave of fresh Codex GitHub issues show OpenAI folding the standalone Codex desktop app into ChatGPT. Users report the macOS app…
A new analysis argues hyperscalers are quietly replacing frontier Swiss Army Knife models with smaller, purpose-built ones. For coding agents, that me…
This week's Hacker News surfaced three developer-built tools — capn-hook, agent-run, and Kote — that tackle the real pain points of AI coding agents: …
Anthropic has extended free access to Claude Fable 5 on Pro, Max, Team, and Enterprise plans through July 19, 2026. The move pushes back the usage-cre…
Elon Musk told Tesla and SpaceX to begin trialing Grok 4.5, pushing his latest model from consumer chatbot into the operational core of two engineerin…
Ollama closed a $65M round while its open-source stack crossed 9 million builders, a sign that running models locally is now a mainstream developer de…
OpenAI is resetting usage limits on ChatGPT and Codex as GPT-5.6 demand surges past earlier guardrails, a signal it is optimizing for long, agentic se…
Anthropic extended its Claude Code weekly-limit promotion: eligible plans now get 50% higher weekly usage through July 19, 2026. The 5-hour limit is u…
Claude Code vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
At the AtCoder World Tour Finals 2026, an OpenAI reasoning model solved all five Algorithm Division problems for 8,300 points — more than double the t…
A confirmed `has repro` bug in Claude Code (issue #77112) drops any `claude -p` stdout past 65,536 bytes when piped. No error, no warning — your autom…
Coding agents can write code, but their UIs are often generic. A UI.md file — design rules your agent reads before it builds — is the missing layer be…
A new open-source project, Soundpulse/hermes-live2d, turns your Hermes Agent into a transparent ATRI Live2D desktop pet that lip-syncs the agent's TTS…
A fresh wave of developer discussion asks whether AI 'vibe coding' is quietly wrecking codebases. Here's what's actually happening, what the real risk…
A new Claude Code bug report shows MCP tool responses intermittently returning a different tool call's data under concurrent parallel sub-agent load —…
A P1 crash in oh-my-pi destroys a whole multi-agent session on a single browser race, and a confirmed Claude Code bug silently splits your Windows MCP…
A reproducible Claude Code bug is killing long-running background Bash tasks with SIGKILL about 1% of the time, mid-write — and leaving repositories i…
How to write AGENTS.md (and related instruction files) so Claude Code, OpenCode, Cursor, Codex, and others respect your architecture, tests, and safet…
How to choose between Claude Code, Cursor, Codex, OpenCode, Hermes, and others using real adoption data, security posture, and workflow fit — not mark…
A fresh GitHub issue (77018) asks Anthropic to expose rate-limit utilization in headless contexts via stream-json, --output-format json, or a usage su…
A practical, runnable security checklist for Claude Code, Codex, Cursor, Hermes, OpenCode, and other coding agents. Sandbox isolation, permissions, se…
OneDev's new AI feature treats coding agents as participants inside issues, pull requests, and CI. The shift: agents that live where your team works i…
A new Show HN (peek-cli) and a walkthrough of Claude Code's built-in browser point to the same shift: coding agents that view and iterate against a li…
A recently uploaded walkthrough highlights a Rust-based guardrail that intercepts more than 50 failure modes where an AI coding agent can damage a rep…
Fresh GitHub issues show Claude Code hallucinating messages, ignoring your model settings, dropping MCP OAuth on token expiry, and burning 15,861 API …
A widely-shared writeup argues Pi, the minimal agent harness behind OpenClaw, is the Neovim of coding agents: a bare foundation you build your own plu…
A confirmed Claude Code bug (GitHub #76930) shows the model safeguard firing false positives on read-only defensive security reviews, blocking credent…
Google AI Studio's Build mode rolled out GitHub repo import with auto-deploy, collapsing the gap between a codebase and a running AI app in a single s…
Microsoft is setting GPT-5.6 as the preferred model across 365 Copilot, a move that reshapes how the world's most-used productivity suite routes its A…
A fresh Claude Code bug shows the agent losing already-completed planning context after a mid-turn interrupt, then confidently insisting the work neve…
Two fresh GitHub issues show coding agents quietly multiplying your usage and ignoring your model settings. Here's how to catch the silent bleed befor…
A verified Codex CLI security issue shows a PreToolUse hook that correctly denies and redacts a shell command or patch still has the raw input appende…
A joint research sprint between Google and Hugging Face delivers a 5x inference speedup for Gemma 4, making the open-weight model viable for latency-s…
Paradigm extends its Centaur AI agent's reach beyond internal workspaces, enabling it to join and operate in external Slack channels for cross-team au…
Perplexity integrates Grok 4.5 into its orchestrator system and achieves top results on the WANDR benchmark, surpassing Opus in multi-step reasoning t…
Confessor reconstructs what your AI coding agent did from Claude Code's own session logs — every sensitive file it opened and every read-then-network-…
A new launch — Slipstream, 'The Command Deck for AI-Assisted Development' — points to a category shift: instead of a bare terminal, developers want a …
A 'Run Claude and Codex in the Browser' HN thread spotlights a growing category: tools that put your terminal coding agents behind an encrypted link y…
ByteDance rolls out Seedream 5.0 Pro to multiple platforms, expanding access to its latest image generation model beyond a single ecosystem.
Meta's Muse Spark 1.1 API launches at roughly 25% of competitor pricing, a bold cost play that reshapes the AI API pricing landscape.
Show HN pick Mindwalk visualizes Claude Code and Codex session logs as light moving through a 3D map of your repo — fully local, no data leaves your m…
Deleting a binary file with Copilot CLI's apply_patch stores the entire blob in session history. Your session permanently exceeds GitHub's 5 MB CAPI l…
Hermes Agent's write_file silently fails when content exceeds ~8 KB. The tool returns an empty success, your agent thinks the file was saved — and you…
Gitlawb Zero's sandbox inherited environment variables verbatim from the parent process. AWS keys, GitHub tokens, database passwords — all exposed to …
Arrow functions, type annotations, and diff markers in tool input get interpreted as shell redirection operators on Windows, creating zero-byte files.
DejaView (a new Show HN) is a local-only TUI that shows every Claude Code session across your machine, with activity sparklines and one-key resume. He…
Codex's Windows sandbox fails silently when Smart App Control is enabled. Every 'sandboxed' execution runs on bare metal — the UI lies to you.
GPT-5.6 Sol ties Claude Fable 5 on the Code Arena benchmark at 40% lower cost, shaking up the performance-per-dollar calculation for coding agents.
OpenAI's open-source codex-plugin-cc brings Codex code reviews and task delegation directly into Claude Code through slash commands. Here's how it wor…
Kraken relaunches its mobile app with built-in agentic trading bots, letting users deploy automated strategies from their phones — no separate infrast…
The Ethereum Foundation used AI-powered audits to uncover real vulnerabilities in smart contract code — validating that coding agents can catch bugs t…
A newly filed Claude Code bug shows the agent silently retrying forever after hitting API usage limits. No error. No stop. Just a growing bill.
AWS released an official Agent Toolkit that connects Claude Code, Codex, and Cursor to your AWS account through a single MCP server. One plugin instal…
A reverse engineer found Claude Code silently encodes API gateway info into system prompt punctuation. We unpack the article, the HN debate, and what …
A bisected regression shows Claude Code's TUI render loop dropped from 16 fps to ~9-10 fps on Windows 11 between versions 2.1.159 and 2.1.207. Here's …
A Windows click-to-focus bug in Claude Code causes the first click on a de-focused window to activate a pending permission dialog, submitting an unint…
GitHub adds its first open-weight model to Copilot. HN reacts with 185 comments on pricing, local models, and why devs are fleeing cloud AI.
Fields Medalist Terence Tao used a coding agent to port two dozen 1999-era Java math applets to JavaScript in hours — and the agent caught bugs in his…
Real developer opinions on Claude Code from the coding community. What people love, what frustrates them, and whether it is worth the subscription in …
What the developer community says about OpenAI Codex after the latest model update. Is Codex still competitive against Claude Code and Cursor?
Honest community breakdown of coding agent pricing in 2026. What developers actually spend, where the hidden costs are, and the cost-saving strategies…
After tracking developer discussions across the coding agent ecosystem, here is the honest community verdict on Claude Code vs Cursor vs Hermes vs Cod…
Honest developer opinions on Cursor AI IDE — pricing, features, bugs, and whether developers recommend it over alternatives in 2026.
Community opinions on free and open-source coding agents. Which free tools developers use, what they sacrifice, and whether free is good enough for pr…
Real community feedback on Hermes Agent — the open-source coding agent from Nous Research. What developers praise, what needs work, and whether it is …
Writers and developers report Claude's latest models refuse benign creative and research prompts more often. Here's what the trend looks like, which m…
A reproducible Claude Code security issue shows background subagents stalling and emitting authorization-shaped prompt fragments. Here's the isolation…
GenLayer launches an Internet Court for AI agents, backed by 27 firms — a governance layer for resolving disputes between autonomous agents operating …
Claude Code's compound-command permission system can flood you with hundreds of prompts per session, even for read-only commands like cd, ls, and git …
OpenAI's GPT-5.6 Sol Ultra deployed 64 coordinated subagents to prove a long-standing math conjecture — a milestone in multi-agent reasoning at scale.
Oh My Pi's native grep can crash the entire agent process when a file changes during search. Here's how to check if you're affected and what to do abo…
Today's roundup: Musk pushes Grok 4.5 inside Tesla and SpaceX, Perplexity's orchestrator tops Opus on a benchmark, Meta launches a cheap API, ByteDanc…
Unsloth ships Qwen3.6 quantizations delivering 2.5x GPU speed — a performance jump that lets teams run larger models on the same hardware without upgr…
gitlawb-zero's 0.4.0 release prep ships its native binary as platform optionalDependencies, reworks Windows sandbox denial classification, and adds a …
Meta launches Muse Spark 1.1 API at roughly 25% of competitor pricing — a floor that changes the economics of routing for cost-sensitive coding agents…
Fresh Codex issues report a namespace collision producing 'unsupported custom tool call: execexec' on gpt-5.6-sol, plus new tasks that start without w…
A merged Codex commit preserves parent sandbox enforcement during memory consolidation — closing a path where a sub-process could escape the boundarie…
A P0 issue reports that OpenClaw's macOS launchd gateway exits without relaunching on a config change, leaving the agent dead until a manual kickstart…
OpenClaw's latest commits introduce a Claude session fleet and production cloud-worker bundles with pinned SSH bootstrap and an admission handshake — …
Alongside a new isolated Grok Build subscription provider, oh-my-pi's issue tracker shows a ~50% idle CPU spin, an RPC mode that crashes on bad stdin,…
A burst of commits to oh-my-pi's coding agent adds a unified model hub with custom roles, spatial sidebar navigation, a tiered session selector, and t…
Fresh issue reports flag a secret-redaction leak in worktree handling and Windows command failures being misclassified as sandbox denials — both on th…
A sweep of recent Hermes Agent commits tightens how the gateway decides a session is ready and makes mid-session model switches stick to the right sco…
A Claude Code Remote Control bug revealed a deeper problem: your agent's status indicators can be confidently wrong. Here's a diagnostic framework tha…
A new bug report shows Anthropic's Sonnet 5 model, running inside Claude Code, deleting the entire contents of a folder as it tried to enumerate files…
A reproducible crash in Codex Desktop spills internal system prompts into the error output — revealing exactly how the agent is instructed to behave, …
A burst of Windows-specific crash and stability fixes just landed across Hermes, Codex, and Goose at the same time. It's the most honest signal yet th…
Free vs subscription vs pay-as-you-go — the complete pricing breakdown for all 15 coding agents on terminalblog.
A single smarter model is the obvious path. Coordinating multiple smaller models is the path that actually ships more code.
The era of one agent to rule them all is ending. The future is 5-10 specialized agents working together, each good at one thing.
Every agent can write code. None of them can tell you what your codebase actually needs. That gap is the biggest opportunity in AI coding tools.
Code review is the bottleneck that agents are perfectly suited to solve. Here's a concrete workflow that runs every day without human intervention.
GitHub's terminal agent brings deep repository integration to the command line — and it's cheaper than you think.
AmpCode vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Every new tool promises 10x productivity. What nobody mentions is the context debt, the tool sprawl, and the growing dependency on models you don't co…
A wave of prompt-injection and data-leak hardening just landed across OpenClaw, Goose, Hermes, and Codex. The coding agent is now an attack surface, a…
We're quietly shipping coding agents onto servers, into gateways, and behind CI — but the trust model underneath them is still built for one person on…
Not one agent. Not two. A layered system that handles everything from quick edits to complex refactors to scheduled maintenance.
OpenAI Codex vs Goose: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs Cursor: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Every file, every credential, every API key — your agent sees everything. Here's what you should know about agent visibility and control.
The subscriptions are visible. The token bills are not. Here's a realistic look at what different coding agents cost when you factor in API usage.
Hermes Agent vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Subscriptions, token costs, overage charges, and the subscription-to-metered bait-and-switch that every AI tool company eventually pulls.
OpenAI Codex vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Claude Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Autonomous agents that write code, run tests, and deploy to production are coming. Most engineering teams don't have the processes to handle them.
Cursor vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code is better today. Cursor has more users. But the open-source ecosystem has an advantage that eventually wins every platform war.
Every week a new LLM claims the coding crown. Meanwhile, a deeper shift is happening — and most developers haven't noticed.
One file controls how your coding agent understands your project. AGENTS.md is the universal instruction sheet that works across Claude Code, OpenCode…
A practical feature matrix for AI coding agents—vision, cron, multi-provider, git, plugins, subagents, background tasks, local-first—so you pick by ca…
First autocomplete, then chat, now autonomous agents. The traditional IDE is being replaced piece by piece. Here's how the next 18 months play out.
What actually differs between GitHub Copilot-style completion and modern coding agents—permissions, tools, multi-file loops, and when each is the righ…
Cursor vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
When to standardize on open-source coding agents versus commercial tools—ownership, support, cost shape, security, and escape hatches for real teams.
Community patterns for Claude Code vs Cursor—IDE speed versus terminal depth, hybrid stacks, failure modes, and how to turn anecdotes into a team defa…
Autocomplete is comfortable. Agents that read your whole codebase, make decisions, and commit code are something else entirely.
Claude Code vs Goose: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Monthly subscriptions are the obvious cost. Token consumption, context window waste, and failed task retries are the expensive ones hiding in plain si…
AmpCode vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Cursor: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Hermes Agent: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Kilo Code CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Mimo Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Claude Code vs pi.dev: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Codebuff vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Codebuff vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenAI Codex vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenAI Codex vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenAI Codex vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenAI Codex vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs Mimo Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Cursor vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Goose vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Goose vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Goose vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Profile and delegation parity now preserved when routing through portals. Your provider config won't get silently dropped anymore.
Hermes Agent vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Cursor: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Kilo Code CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Mimo Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes Agent vs pi.dev: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs Goose: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Kilo Code CLI vs pi.dev: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs Kilo Code CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Mimo Code vs pi.dev: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Running multiple coding agents on the same codebase sounds efficient. In practice, it breaks in ways that surprise everyone. Here's the real bottlenec…
Oh My Pi vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs Goose: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Oh My Pi vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenClaw vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenClaw vs Codebuff: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenClaw vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs Kilo Code CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
OpenCode vs OpenClaw: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
pi.dev vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
pi.dev vs GitHub Copilot CLI: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
pi.dev vs Gitlawb Zero: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
pi.dev vs Oh My Pi: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Everyone compares benchmark scores. Nobody's asking the important question: can your coding agent delete your database?
Claude Code vs AmpCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Hermes desktop app now prevents stale commit repins when existing checkouts are detected. Updates that used to break your config now work smoothly.
Claude Code vs OpenCode: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero resolved pending askUser callbacks that caused the runner to hang indefinitely. Sessions that froze mid-conversation now complete normall…
Claude Code vs OpenAI Codex: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Zero fixed a CLI parsing bug where flag values were consuming positional arguments. Your filenames won't get swallowed by flags anymore.
Claude Code vs Cursor: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Gitlawb Zero resolved an absolute path for taskkill on Windows to prevent binary hijacking. A security fix that protects your entire system.
Gitlawb Zero now reclaims stale lock files instead of failing permanently. Sessions that used to require manual intervention now self-heal. Deep dive …
Gitlawb Zero fixed git branch detection when starting from subdirectories. Your session context is correct no matter where you launch Zero.
A comprehensive guide to all 15 coding agents tracked on terminalblog — pricing, features, benchmarks, and how to choose the right one for your workfl…
Codex's Git worktree approach is how teams will use AI agents in 2027.
The shift from interactive to autonomous coding is already happening.
Oh My Pi v16.3.15 adds Grok 4.5 to the model catalog with full prompt-cache affinity support. Your coding agent can now use xAI's most capable model.
The most-starred AI assistant on GitHub deserves more attention than it gets.
Oh My Pi now supports OpenAI's reasoning mode with a new model catalog integration. The agent can think step-by-step before acting.
A Rust-based coding agent with 51K GitHub stars that's quietly becoming the extensibility standard. Complete guide to installation, architecture, mode…
Oh My Pi switched from API key authentication to device flow for xAI. More secure, more reliable, and no more API key management.
In a world of bloated AI tools, Codebuff's simplicity is its superpower.
Practical guide to Kilo Code CLI — the fast, provider-agnostic CLI tool for quick AI coding tasks. Installation, usage patterns, provider setup, and w…
Goose vs Claude Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
The first coding agent that can see your screen and understand your code visually.
Most AI assistants are locked to one ecosystem. OpenClaw isn't.
Deep dive into OpenCode's provider-agnostic architecture — how it works, why it matters for cost and flexibility, provider setup guide, and real-world…
Latency matters more than capability when you're coding interactively.
Long-running knowledge-backed agents change how AI understands your project. What persistent codebase memory means for daily development.
Being model-agnostic isn't just a feature — it's survival.
The Model Context Protocol is turning Goose into a hub for every coding tool — here's the architecture behind it and why it matters for choosing an ag…
When agents run in the cloud, your laptop becomes a monitor, not a workstation.
ACP isn't just a protocol — it's the beginning of agent ecosystems.
Text-only agents are hitting a wall. Vision-capable agents are breaking through.
OpenCode vs Claude Code: clear default for operators, feature matrix, install paths, and when to use each—not a generic feature dump.
Agents that forget are replaceable. Agents that remember are indispensable. A deep dive into the three context strategies reshaping coding agents in 2…
Heavyweight agents solve 10% of problems. Lightweight agents solve the other 90%. Why a CLI-first, low-overhead agent belongs in every developer's too…
Zero added a --auto flag that generates LLM-powered commit messages from your staged changes. Finally, good commit messages without the effort.
Zero just shipped context preservation in exec prompts. Your conversations now survive across commands. The decentralized coding agent grows up.
Zero now rejects malformed permission payloads before prompting. A critical security hardening for the agent that runs in your terminal — deep dive on…
Zero's setup wizard now has a searchable, filterable provider picker. 20+ providers, live search, clear categories — terminal UX done right.
Zero made all keyboard shortcuts reconfigurable via config file. Every keybinding in the TUI can now be remapped to your preference.
Gitlawb Zero ships self-updates with a single command. No npm, no git pull, no manual downloads. Just 'zero upgrade' and you're on the latest version.
Zero shipped full Android support for Termux. Install with npm, run in Termux — your coding agent in your pocket.
A tooltip fix in the Hermes desktop app reveals deeper thinking about agent output presentation. The best AI tools sweat the small stuff.
Hermes Agent shipped a /sessions search <query> gateway command that lets you search across every past session. Full-text conversation recall, right i…
Hermes just shipped TTS with direct OpenAI model coercion on the managed audio gateway. Your agent can now speak its responses out loud.
Oh My Pi v16.3.4 shipped with Baseten provider integration, smarter model blocking, and mnemonic extraction improvements. The most mature coding agent…
Long tool-call runs now collapse into an auto-scrolling window. No more terminal spam, no more manual scrolling — just watch your agent work in real-t…
Hermes shipped a 'cdp' capability — Chrome DevTools Protocol integration for browser automation and webmail. Your coding agent can now browse the web.
The skills-renovate PR touched skills loading, caching, validation, and cross-skill dependency resolution. The skill system got a full architecture re…
Hermes cron jobs were running under the wrong secret scope. A fix ensures every scheduled task uses the correct profile credentials. Here's the techni…
Hermes added a case-insensitive .env file guard. If you thought naming a file '.ENV' would bypass detection — it won't anymore. Here's the technical b…
Hermes shipped private-page guards for its CDP browser integration. The agent can browse sensitive pages without leaking data — here's the technical d…
Claude Code is Anthropic's terminal-based coding agent with GCP Gateway, plugin system, frontend-design skills, and enterprise-grade agent infrastruct…
Gitlawb Zero is an MIT-licensed terminal coding agent with durable local sessions, multi-provider model support, and a decentralized git network for A…
Comprehensive guide to Hermes Agent — background task delegation, MOA orchestration, multi-provider routing, credential security, cron jobs, memory, a…
Kilo Code CLI provides intelligent prompt construction, context management, and multi-provider orchestration without the overhead of a full agent fram…
Mimo Code extends OpenCode with vision support, reasoning model integration, and an enhanced terminal UI — a look at what this fork adds to the coding…
Oh My Pi is a 16K-star coding agent with LSP integration, browser automation, subagents, GitHub CLI ops, and image analysis — a deep look at its archi…
Everything about OpenCode — its SKILL.md execution model, slash command workflows, agent lifecycle, and how it compares to other coding agents.
pi.dev reimagines coding agents as long-running personal intelligence systems that learn from your codebase through persistent knowledge graphs.
Hermes shipped a self-healing update mechanism for Windows venvs. Half-updated installations now repair themselves automatically.
Comprehensive overview of the open-source coding agent ecosystem — Hermes, OpenCode, Mimo, Kilo, pi.dev, Gitlawb Zero, Oh My Pi, Claude Code — and wha…
Hermes Agent just shipped a real-time console with REPL, WebSocket streaming, and visual agent monitoring. Here's what changed and why it matters.
Hermes's desktop app now has a Skill Hub in Capabilities, powered by React Query and a plugin registry. Community skills are one click away.
How Hermes Agent handles the full git lifecycle — branching, committing, reviewing, PR creation — completely autonomously with built-in GitHub integra…
Inside Hermes Agent's multi-provider routing engine: how it picks the right model for each job, falls back gracefully, and optimizes cost without conf…
Hermes Agent's vision pipeline routes images to the best model for the job — from OCR to complex scene analysis — with automatic provider detection an…
Hermes Agent runs entirely locally, supports fully offline models, and never phones home. Here's why local-first architecture matters for AI agents — …
A deep appreciation of Hermes Agent's TUI — slash commands, keyboard-driven workflows, process management, and what makes a CLI feel like home.
Inside Hermes Agent's Mixture-of-Agents (MOA) architecture — how multiple specialist agents collaborate on complex reasoning tasks for superior output…
Hermes Agent's credential guard system prevents provider API keys from leaking between tasks — here's how the security architecture works, with threat…
How to build, install, and manage reusable skill packages for Hermes Agent — from SKILL.md anatomy to registry installs, snapshots, and team sync.
Hermes Agent never forgets your preferences, past decisions, and project context — thanks to its mem0-powered persistent memory system.
How to schedule autonomous AI agents on cron — from daily code reviews to weekly dependency audits — with Hermes Agent's built-in job system.
Deep dive into Hermes Agent's autonomous task delegation, multi-model orchestration, and why it beats Copilot and Claude Code for complex workflows.
Oh My Pi integrated Baseten as a model provider, joining the platform that's making open models as fast and reliable as closed ones.
Oh My Pi's AI usage system now exempts specific models from proactive hard-blocking. Smarter, less intrusive guardrails for agent workflows.
Oh My Pi's CLI usage reporting was fixed and verified. Track token consumption, model costs, and API usage from the terminal.
A fix that sounds small but fixes a painful bug: mnemonic structured extraction now preserves empty values instead of silently dropping them.
Oh My Pi releases daily — and that cadence is a feature, not noise. The v16 series proves continuous delivery works for AI tools.
Introducing a dedicated publication covering Hermes Agent development, autonomous coding workflows, and the open-source AI agent ecosystem.
Anthropic shipped a Claude Gateway reference deployment for Google Cloud Platform. Bring Claude Code to your enterprise infrastructure.
Anthropic rebranded the Claude Gateway as an Agent Platform. It's no longer just a proxy — it's infrastructure for building and managing AI agents.
The lock-closed-issues workflow was broken by GitHub API changes. Claude Code's fix uses the search API instead of pagination — and it's a masterclass…
Claude Code ships a frontend-design skill (v1.1.0) that turns a plain description into a working React + Tailwind UI. Here's how it works, how to prom…
As OpenCode enters archive mode, its last major feature was GitHub Copilot provider support. Free, high-quality models directly in your agent.