search
agentic coding
Trends
- 1
Developer JuliusBrussee has released an open-source project called caveman, a skill and proxy for coding agents written in Go. It prompts AI coding assistants to write in a stripped-down, 'caveman' style, which the project says cuts token consumption by around 65 percent while keeping the code output intact. The project is drawing attention among developers looking to reduce costs on popular coding agents.
- 2
Developer obra has released Superpowers, an open-source agentic skills framework and software development methodology designed to make AI coding agents work more effectively. The project, written in Shell, describes itself as a methodology that 'works,' and is drawing attention as developers look for practical ways to structure and improve how autonomous coding agents handle real software projects.
- 3
A JavaScript project called Ponytail by developer Dietrich Gebert is circulating among coders. It is designed to make AI coding agents behave like 'the laziest senior dev in the room', pushing them to write as little code as possible. The pitch is that the best code is the code you never wrote, a philosophy that is drawing attention and engagement from developers tracking new AI coding tools.
- 4
TypeScript educator Matt Pocock has published a repository called 'skills' on GitHub, described as 'Skills for Real Engineers', drawn from his personal .agents directory. The collection packages configurations he uses with AI coding agents. It quickly climbed GitHub's trending ranks, drawing attention from developers interested in how experienced engineers structure AI-assisted workflows.
- 5Open-source model routing targets top coding-agent performance▼Show HN: Open-source model routing for coding agents at Astra-level performance
A developer has released an open-source tool that routes requests between AI models for coding agents, claiming performance on par with Astra-level systems. The launch invites developers to test whether smart routing across cheaper models can match premium single-model results. Early discussion centers on benchmarks, reliability, and whether routing adds real gains for coding workflows.
- 6
A new open-source project called context-mode, written in TypeScript, aims to optimise how AI coding agents use their context windows. It claims to sandbox tool output with a 98% size reduction, persist session memory between tasks, and enforce routing across 17 platforms using MCP and hooks. Developers are sharing it as a practical response to the common problem of agents running out of usable context during long sessions.
- 7
A discussion thread on Hacker News asks whether developers are actually producing good code with AI coding agents. The question invites engineers to share firsthand experiences on whether these tools deliver production-quality results or mostly generate code that needs heavy review and rework.
- 8
Developer Corey Haines published a JavaScript toolkit called marketingskills, which adds marketing capabilities to Claude Code and other AI agents. The collection covers conversion rate optimisation, copywriting, SEO, analytics and growth engineering. It is gaining traction on GitHub, where early users see it as a way to turn coding assistants into marketing assistants.
- 9
A developer-published open-source project called CodeGraph is drawing attention. It builds a pre-indexed knowledge graph of a codebase that automatically syncs as code changes, and works with AI coding assistants including Claude Code, Codex, Gemini, Cursor, OpenCode, AntiGravity, Kiro, Copilot and Hermes Agent. The pitch is fewer tokens and tool calls, with everything running locally.
- 10New agent development environment ships with coding agent orchestrator●Agent Development Environment (ADE)and orchestrator shipping with coding agents
Developer Cezar, known as cezar.run, has released an Agent Development Environment (ADE) together with an orchestrator designed to work alongside coding agents. The tool bundles an environment for building and running agents with a layer that coordinates multiple coding agents. Discussion is centred on developers evaluating whether this setup improves how they run and manage AI coding agents day to day.
- 11Developers Delete Unit Tests as AI Coding Agents Spread●Developers Delete Unit Tests as AI Agents Reshape Coding
Software developers are reportedly removing unit tests from their codebases as AI coding agents take on more of the programming work. The practice is drawing debate, with some arguing tests become unnecessary when AI generates and verifies code, while others warn deleting tests undermines safety nets and long-term code quality.
- 12Corral tool kills every command an AI agent starts▼Show HN: Corral – Kill every command your agent starts
A developer has released Corral, an open-source tool shared on Hacker News that can kill every command started by an AI coding agent. The tool is aimed at giving users a way to stop runaway processes or unwanted commands launched by autonomous agents. It is available on GitHub, and early engagement on Hacker News suggests interest in agent safety and control tooling.
- 13Graphene launches as data analysis toolkit for coding agents●Show HN: Graphene – Data analysis toolkit for your coding agent
A new open-source project called Graphene has been released, offering a data analysis toolkit designed to work with AI coding agents such as coding assistants that write and run code. The toolkit is available on GitHub, where developers can try it and contribute. Early reactions on Hacker News are modest but curious, with the community discussing whether agent-focused data tools will become a standard part of developer workflows.
- 14Television launches as open source GUI for AI agent harnesses●Show HN: Television – an open source GUI for your agent harness
A developer has released Television, an open source graphical interface for agent harnesses, on the community discussion site Hacker News under its Show HN format. The tool, available at television.run, gives developers a visual front end for managing AI coding agents. Early reaction is small but positive, with modest engagement as users evaluate whether a GUI layer adds value to terminal-based agent workflows.
- 15SwiftFairy update speeds up AI code reviews on macOS●Just pushed an update to SwiftFairy 🧚, our native macOS MCP server that reviews your agent’s code locally for correctnes
Developer hishnash has released version 2026.10.1 of SwiftFairy, a native macOS MCP server that reviews AI agents' code locally for correctness, performance and maintainability. The update lets agents send file paths instead of full source code, making large reviews faster. It is a small but notable release for developers running AI coding agents on Macs, reflecting growing interest in local, privacy-friendly tooling.
- 16AI Agents Can Code, But Do They Understand Business?▼Your AI Agents Can Write Code. Can They Understand Your Business?
A new commentary asks whether AI agents that can write code actually grasp the business contexts they are deployed in. The piece highlights a growing concern among companies adopting agentic AI: technical capability alone does not guarantee that automated systems understand organizational goals, workflows, or domain nuances, raising questions about how firms should evaluate and supervise AI agents before trusting them with real business decisions.
- 17TIRx Harness Released for Agentic GPU Programming●TIRx Harness: An Open Compiler Harness for Agentic GPU Programming
The MLC team has introduced TIRx Harness, an open compiler harness aimed at agentic GPU programming, allowing AI agents to write and optimize GPU code through a compiler-driven workflow. The announcement, published on the MLC blog, is drawing attention from developers interested in combining large language model agents with low-level performance engineering and GPU kernel development.
- 18Anthropic says AGI could arrive within a couple of years●AGI in a couple of years — Anthropic's RL leads on coding agents, Frontier Math and biology
Anthropic is claiming its reinforcement learning systems lead on coding agents, Frontier Math benchmarks and biology tasks, with the company suggesting artificial general intelligence could arrive within a couple of years. The claim has drawn attention because of its unusually concrete timeline and its focus on measurable performance in coding, mathematics and scientific research capabilities.
- 19oh-my-agent adds test reruns when workflows stop●oh-my-agent (OMA) can rerun a configured test script when an active workflow tries to stop. This... # ai # testing # ope
oh-my-agent (OMA), an open-source AI agent tool, now reruns a configured test script whenever an active workflow attempts to stop, returning details on what failed when its test gate blocks completion. The feature targets developers using AI coding agents, aiming to prevent unfinished or broken changes from slipping through by forcing automated checks before a workflow ends.
- 20Tanuki open-source tool brings protocol-first Linux AD triage to AI agents●Excited to share a new open-source project: Tanuki 🦝 Protocol-first Linux Active Directory triage for AI coding agents &
A security researcher has released Tanuki, a new open-source tool described as protocol-first Linux Active Directory triage built for AI coding agents and operators. The developer says most coding agents hallucinate or propose noisy, dangerous tactics when handling Linux AD environments, such as recommending weak crypto settings. The project is being shared with the infosec community.
- 21AI coding agent Pi reaches version 1.0●Coding-Agent Pi: Mehr Platz im Prompt, weniger Ballast Pi ist ein erweiterbarer KI-Coding-Agent fürs Terminal. Version 1
Pi, an extensible AI coding agent that runs in the terminal, has released version 1.0.0. The update streamlines the tool's code mode, adds image generation, and launches in fullscreen mode, giving developers more prompt space and less overhead. The German tech press coverage highlights the release as a step toward leaner AI-assisted programming directly from the command line, drawing attention from developer communities tracking terminal-based AI tools.
- 22Five lessons from running Claude Code as an hourly agent●5 lessons from running an hourly Claude Code cloud agent on a large production monorepo: proving review comments, loop p
An engineer has shared five lessons from running Anthropic's Claude Code as a cloud agent every hour on a large production monorepo. The write-up covers validating review comments, preventing the agent from getting stuck in loops, fixing a 403 error, and reducing token consumption. The post is drawing attention from developers interested in using AI coding agents for automated code review in real production environments.
- 23AI agents ignore brand guidelines written as documents●Brand guidelines PDFs are not runtime. Agents do not reliably obey a paragraph that says "keep the... # branding # ai #
Developers are discussing a recurring problem with AI coding agents: brand rules stored in PDFs or prose guidelines are not reliably followed by automated systems. The argument circulating is that branding should be encoded as compile-time rules and template parameterization, so constraints are enforced structurally rather than left to an agent to interpret and obey.
- 24The Four Horsemen of Agentic Coding●The Four Horsemen of Agentic Coding Article URL: https:// distantprovince.substack.com/p /the-four-horsemen-of-agentic-c
A Substack essay titled 'The Four Horsemen of Agentic Coding' is circulating among developers and startup circles. The piece examines the leading AI coding agents now shaping how software is written, framing them as dominant forces in the shift toward automated, agentic programming workflows. Discussion is modest so far, with a small thread on Hacker News, but the topic fits the fast-growing debate over which AI coding tools will define the next phase of software development.
- 25
A widely shared essay argues that in the era of AI agents, the harness—the scaffolding of tools, prompts, evaluation and workflow code wrapped around a model—is where a company's actual value sits, not the underlying model itself. As models become commoditised and interchangeable, the author contends the harness is the durable product, and effectively the company's true identity.
- 26MIT and Sakana AI unveil cheaper evaluation for self-improving coding agents●New MIT and Sakana AI framework uses an LLM judge to cut evaluation costs for self-improving coding agents
MIT and Sakana AI have introduced a new framework that uses a large language model as an automated judge to evaluate the output of self-improving coding agents. The approach is designed to significantly reduce evaluation costs, which typically require expensive human review or heavyweight testing as AI coding systems iterate and improve themselves.
- 27
MIT researchers have introduced SIFT, a new approach that significantly lowers the cost of evaluating coding AI agents. The method addresses the expense of running benchmark tests on agentic coding systems, which can require substantial compute. Details of how the technique works and how much it saves remain limited, but the news is drawing attention in the AI research community as demand grows for cheaper, faster agent evaluation.
- 28AI code reviewers miss subtle cheating in tests●The software factory assumes agents reviewing agents catches what tests miss. I gave 77 cheating diffs to three reviewer
An experiment tested whether AI reviewer models can catch cheating in code changes when agents review agents, an assumption behind automated software pipelines. Across 77 diffs containing deliberately planted cheats, three reviewer models caught every exotic trick but approved one case where an assertion was quietly made unfalsifiable, meaning the test could never fail. The finding raises doubts about relying on AI review alone to guarantee code quality where automated testing falls short.
- 29The Four Horsemen of Agentic Coding●The Four Horsemen of Agentic Coding https://distantprovince.substack.com/p/the-four-horsemen-of-agentic-coding # AI # Pr
A new essay explores what it calls the four horsemen of agentic coding, examining the key forces or figures shaping the rise of AI agents that write software autonomously. Programmers and AI observers are sharing and debating the piece as discussion grows over how agentic tools are changing development work and the future of the programming profession.
Repos
- anteloc/ldraw-nova Agent tooling for generative LEGO models building, built with Astra and Opus 5.5, powered by Jev
- Louis-CFM/coucou A tiny friend that lives in your notch (macOS) or at the top of your screen (Windows, Linux) and keeps an eye on your co
- DietrichGebert/ponytail Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
- mvschwarz/openrig Build your own network of agents from Claude Code, Codex and Pi: persistent teams with roles, shared context and owned w
- pbakaus/impeccable The design language that makes your AI harness better at design.
- yetone/magpie Every agent's model. One place. Codex on DeepSeek, Claude Code on Kimi, from the menu bar.
- paperclipai/paperclip The open-source app everyone uses to manage agents at work
- heygen-com/hyperframes Write HTML. Render video. Built for agents.
- JuliusBrussee/caveman 🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking l
- coreyhaines31/marketingskills Marketing skills for Claude Code and AI agents. CRO, copywriting, SEO, analytics, and growth engineering.
- earendil-works/pi AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
- mksglu/context-mode Context window optimization for AI coding agents. Sandboxes tool output (98% reduction), persists session memory, and
- colbymchenry/codegraph Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode, AntiGrav
- mattpocock/skills Skills for Real Engineers. Straight from my .agents directory.
- cursor/plugins Cursor plugin specification and official plugins
- obra/superpowers An agentic skills framework & software development methodology that works.
- google/skills Agent Skills for Google products and technologies
- Panniantong/Agent-Reach Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili, XiaoHo
- kaankiziltug/logo-design-skill A comprehensive logo-design skill for Claude, Gemini CLI, Codex and other AI agents: principles, process, SVG craft, tes
- nanaism/yomiyasu AI生成の日本語を自然な日本語へ推敲するAgent Skill / Agent Skill for Refining AI-Generated Japanese into Natural Japanese