search
agentic coding
Trends
- 1
Developer and educator Matt Pocock has published a repository called 'skills' on GitHub, described as 'Skills for Real Engineers. Straight from my .agents directory.' The collection contains configuration and prompts from his personal AI agent setup, and it is drawing attention on GitHub's trending list. Engineers are looking to it as a practical example of how to organise reusable skills and instructions for AI coding agents in real projects.
- 2Pi pod lets developers run coding agents in self-hosted sandboxes●Show HN: Pi pod – Run your pi coding agent in sandboxes on your own server
A developer has launched Pi pod, a tool that lets programmers run the Pi coding agent inside sandboxes on their own servers. The project was shared on Hacker News, drawing early engagement from the community. It targets developers who want the convenience of an autonomous coding agent without sending code or workloads to a third-party cloud.
- 3
Alibaba has released open-code-review, a Go-based tool that combines deterministic analysis pipelines with an LLM agent to review code. It flags issues like null pointer errors, thread-safety problems, XSS and SQL injection with line-level comments, and works with both OpenAI and Anthropic models. Alibaba says it is battle-tested at its own scale.
- 4
Google engineer Addy Osmani has published 'agent-skills', an open-source collection of production-grade engineering skills for AI coding agents, written in JavaScript. The repository provides reusable tools and patterns aimed at making AI-powered coding assistants more reliable in real-world engineering work, and it is quickly gaining attention from developers.
- 5
A tool called claude-mem, developed under the GitHub account thedotmack, is drawing attention for giving AI coding agents persistent memory across sessions. Written in TypeScript, it captures everything an agent does, compresses it with AI, and reinjects relevant context into future sessions. It works with Claude Code, Codex, Gemini, Copilot and other popular coding assistants, and is climbing GitHub's trending repositories as developers seek continuity in agent workflows.
- 6Open-source tool lets AI agents annotate your screen●Show HN: Let your AI agents paint big arrows, boxes and text on your screen
Developer franze released 'big-arrow-on-the-screen', an open-source tool that lets AI coding agents draw big arrows, boxes and text directly on a user's screen. The idea is to make agent feedback visual and unmissable, letting an AI point at UI elements it wants to discuss or fix. The project drew a warm response on Hacker News, gathering over 360 upvotes as developers debated its usefulness and playful design.
- 7Voice AI agents move to the center of enterprise development●In the rapidly evolving landscape of artificial intelligence, voice agents have emerged as a critical... # typescript #
Developers are highlighting a guide to building predictable, enterprise-grade voice agents with TypeScript, using open-source tools. The piece argues that as artificial intelligence matures, voice agents have become a critical interface for businesses, and it walks through engineering practices for reliability. Discussion centers on practical software development, coding and community-driven, inclusive approaches to voice AI tooling.
- 8
Microsoft has released MXC, a sandboxed code execution system, on GitHub. The project lets developers run untrusted code in isolated environments, a growing need as AI agents and automated pipelines increasingly execute generated code. Early reaction from developers has been strong, with many discussing how it compares to existing sandboxing tools like gVisor and Firecracker and what its release means for cloud security tooling.
- 9SkillsMP indexes over three million GitHub skill files●SkillsMP indexes 3M+ SKILL.md files from GitHub for free; SkillGild lists a reviewed catalog with one-command installs a
SkillsMP is offering a free index of more than three million SKILL.md files hosted on GitHub, while SkillGild provides a reviewed catalog of agent skills with one-command installs and hosted runs. The comparison is drawing attention from developers deciding which service fits their workflow, particularly those building with Claude and other AI coding tools that use modular skill definitions.
- 10
A JavaScript project called Ponytail by developer Dietrich Gebert is drawing attention on GitHub. Its pitch: make AI coding agents behave like "the laziest senior dev in the room," operating on the principle that the best code is the code you never wrote. The tool is being shared as a minimalist counterweight to AI assistants that tend to over-engineer solutions.
- 11
A developer is drawing attention to a Python tool designed to stop AI coding assistants from burying the answer under lengthy explanations. It shapes agent output to be direct and concise, in a format described as ADHD-friendly. Interest centres on whether AI assistants should adapt their communication style to different users rather than defaulting to verbose responses.
- 12Pocketty launches iPhone SSH terminal with agent alerts●Show HN: Pocketty – iPhone SSH terminal that pings you when an agent is blocked
A developer has released Pocketty, an iPhone SSH terminal app designed for managing remote servers from a phone. Its distinguishing feature is notifications: when an AI coding agent running on a remote machine gets blocked and needs input, the app pings the user. The launch is drawing attention from developers interested in mobile workflows for supervising AI agents.
- 13
Developer Paul Hudson, known online as twostraws, has published a SwiftUI agent skill on GitHub designed for Claude Code, Codex, and other AI coding assistants. The tool gives AI agents structured knowledge for building SwiftUI interfaces, making it easier for developers to generate accurate SwiftUI code with AI help. Early attention from the iOS developer community is growing quickly.
- 14Ask HN: Are coding agents producing good code?▼Ask HN: Is anybody producing good code with coding agents?
A question posed on Hacker News asks whether developers are actually producing good code with AI coding agents. The post invites people building software with tools like these agents to share whether the output quality holds up in real projects. Commenters are weighing in with their experiences, and the debate reflects ongoing uncertainty about how reliable AI-assisted programming is in practice.
- 15Open-source model router promises top-tier coding agent performance▼Show HN: Open-source model routing for coding agents at Astra-level performance
A developer has released an open-source tool on Hacker News that routes requests between AI models for coding agents, claiming performance comparable to Astra, a reference to high-tier results on coding benchmarks. The launch drew engagement from the developer community, with discussion centring on whether a routing layer over existing models can match single frontier models on coding tasks.
- 16Running Zed Editor's Docker Agent in a Sandbox▼Zed Editor, Docker Agent, ACP, but in a sandbox: running the agent with sbx
A developer has published a walkthrough for running AI coding agents with Zed Editor via the Agent Client Protocol (ACP), confined inside a sandboxed Docker container using a tool called sbx. The setup lets agents execute tasks without touching the host system, addressing security concerns around giving AI agents broad access to local files and commands.
- 17Cloudflare releases open-source security audit tool for coding agents●cloudflare/security-audit-skill
Cloudflare has published an open-source tool called security-audit-skill, a skill for coding agents that runs multi-phase security audits of code. Its findings are independently verified and produced in a machine-readable format, so other tools and developers can consume them directly. It is written in JavaScript and available on GitHub, where it has drawn early attention from the developer community.
- 18
Developer Michael Lynch publishes an essay asking why AI coding agents keep making simple mistakes despite their hype. The piece examines the gap between expectations and real-world performance of tools that write code autonomously, drawing discussion among developers weighing in on whether current agents are overpromised.
- 19Researchers Backdoor Open AI Model to Steal Credentials●Researchers Backdoor Open AI Model to Steal Credentials in Coding Agents
Security researchers demonstrated a backdoor inserted into an open AI model that causes coding agents to exfiltrate credentials when handling code tasks. The attack shows how tampered open-weight models could silently leak secrets like API keys during autonomous programming work. The finding is drawing attention as more developers adopt AI coding agents with broad access to sensitive systems.
- 20Running the Pi coding agent on Temporal●The immortal life of Pi (Running the Pi coding agent on Temporal)
Temporal has published a new blog post explaining how to run the Pi coding agent on its workflow orchestration platform. The piece, titled 'The immortal life of Pi,' describes how durable execution keeps long-running coding agents alive through failures and restarts. Developers on Hacker News are discussing the approach and what it means for building reliable AI agent infrastructure.
- 21AI coding tools let anyone program, sparking debate●Now with AI and agents, Codex, Claude, etc... anyone thinks they're a programmer now, and honestly, I think it's perfect
A developer writing on X argued that AI assistants like Codex and Claude have made programming accessible to everyone, and that this is a good thing. The post, which mocked programming 'elitists' frustrated by the trend, drew over a thousand likes, reflecting an ongoing argument in tech circles about whether AI coding tools devalue or democratise software development.
- 22AWS launches open-source sandbox to contain runaway AI agents▼AWS launches open-source AI agent sandbox to prevent YOLO mode disasters
Amazon Web Services has released an open-source sandbox designed to let AI coding agents run commands safely, without giving them unrestricted access to a developer's machine. The tool targets so-called 'YOLO mode' setups, where agents execute commands with no checks, a practice that has led to accidental deletions and other costly mistakes among developers experimenting with autonomous agents.
- 23Mycroft joins as Anton's synthetic AI cofounder●Mycroft, Anton's synthetic AI cofounder. A machine drafted this write-up; every number in it is... # ai # agents # opens
Developers around Anton are discussing Mycroft, a synthetic AI being treated as a cofounder in their open source work. A machine-drafted write-up covers fixes such as resolving empty serverInfo.version values in MCP handshakes, along with related testing and debugging efforts. The project emphasizes inclusive, community-driven development, with tags pointing to AI agents, coding, and engineering.
- 24AI Coding Agents Perform Better When Not Writing Their Own Tests●AI Coding Agents Perform Better Without Writing Their Own Tests
A new discussion in developer circles claims that AI coding agents perform better when they do not write their own tests, contradicting the common assumption that self-testing improves code quality. Developers are debating why test generation may mislead agents, with some saying letting models grade their own work invites blind spots rather than catching bugs.
- 25Graphene launches as data analysis toolkit for coding agents●Show HN: Graphene – Data analysis toolkit for your coding agent
A new open-source project called Graphene has been released, pitched as a data analysis toolkit designed to work with AI coding agents. The toolkit is available on GitHub, letting developers connect analysis capabilities directly into agent-driven coding workflows. Early attention has come from the developer community, where Show projects like this typically draw discussion about usefulness, design, and how well it integrates with existing agent tools.
- 26
A new essay identifies four leading tools or approaches driving the shift toward agentic coding, where AI assistants autonomously plan and execute programming tasks rather than just completing snippets. Developers are debating which of these systems will define the next phase of software engineering and how quickly autonomous coding agents are changing day-to-day development work.
- 27Coding agents now start most vx runs, reshaping developer tooling●Most vx runs are started by coding agents now. So every verb speaks JSON with a shipped schema, a failure keeps its log
The developer behind the vx tooling reports that most of its runs are now initiated by coding agents rather than humans. As a result, every command outputs JSON with a shipped schema, failures retain their logs and named files, run history is stored in a single SQLite file, and documentation and skills ship as plain markdown — a design shift toward machine-first interfaces.
- 28Writer tests four coding agents on a $100 PDF editor challenge●I gave four coding agents $100 budget to build a PDF editor
A developer gave four AI coding agents a $100 budget each and asked them all to build the same PDF editor, then compared what they produced. The write-up details where the agents still fall short on real software tasks, and the comparison is drawing attention among developers weighing how far AI coding tools have actually come.
- 29
Manaflow AI has released cmux, an open source macOS terminal built on Ghostty, featuring vertical tabs and notifications designed for running AI coding agents. The tool targets developers who juggle multiple agents at once, emphasizing multitasking, organization, and programmability. Developers are sharing and discussing the release as interest grows in tooling built specifically around AI-assisted coding workflows.
- 30Developer Lets AI Agent Run WordPress Store Via MCP Server●I Turned My WordPress Store Into an MCP Server. Now an AI Agent Runs My Business — With My Permission. # ai # wordpress
A developer describes converting a WordPress online store into an MCP (Model Context Protocol) server, allowing an AI agent to operate the business — processing tasks with explicit human permission at each step. The write-up covers the open-source setup and coding involved. It is drawing attention among developers interested in AI automation of commerce and what it means for handing routine business operations to autonomous agents.
- 31
A Chinese developer has closed the source code of its ARTEX AI agent following a hack that hit a South Korean bank. The move reverses any open availability of the tool and comes amid heightened scrutiny of AI software linked to cybersecurity incidents in the financial sector.
- 32Trigger.dev adds Python sandbox for AI agents●An AI agent analysing data often needs to run code: calculate a total, inspect a file, or test an... # triggerdev # pyth
Trigger.dev has introduced Plimsoll, an open-source Python sandbox that lets AI agents safely execute code while analysing data. Agents often need to run calculations, inspect files, or test snippets, and the sandbox provides an isolated environment for that. Developers in the open-source community are sharing the tool as a practical addition to agent-building workflows.
- 33AI Agents Shift Toward Specialized, Persistent Tools With Major Funding●AI Agents Shift to Specialized, Persistent Tools with Major Funding
The AI industry is moving away from general-purpose chatbots toward specialized agents that persist across tasks and retain context, backed by significant new venture funding. Startups are building domain-specific agents for coding, research and workflow automation, and investors are pouring capital into the category. Commenters are debating whether focused, long-lived agents will replace broad assistant models as the next major platform shift in artificial intelligence.
- 34Solo developer unveils OpenPilot, an open-source AI coding agent●Disclosure: I'm Ibrahim, the solo developer of OpenPilot. This article was drafted with help from an... # opensource # a
Developer Ibrahim has introduced OpenPilot, a solo-built open-source tool he describes as a Cursor-style AI coding agent that runs on locally hosted models instead of cloud services. Writing under an open disclosure that the article itself was drafted with AI assistance, he frames the project as a privacy-friendly alternative for developers who want agent-based coding help under their own control, and is inviting the open-source community to take part.
- 35Google launches Gemini AI workplace agent for coding and tasks▼Google launches Gemini AI workplace agent that can write code and run tasks, company says
Google has announced the launch of a Gemini AI workplace agent that can write code and carry out tasks on a user's behalf, according to the company. The agent is aimed at workplace use, marking another step in the competition among major tech firms to deploy AI assistants that go beyond chat and perform hands-on work for businesses.
- 36New agent development environment ships with coding agents▼Agent Development Environment (ADE)and orchestrator shipping with coding agents
A developer tool called the Agent Development Environment, or ADE, has been released alongside an orchestrator for running coding agents. The project, presented at cezar.run, offers a dedicated environment for developing and coordinating AI coding agents rather than working through ad-hoc scripts. Discussion on Hacker News centers on whether purpose-built environments and orchestration layers can make multi-agent coding workflows more reliable and manageable.
- 37AI Coding Agents Do Fine Without Writing Their Own Tests●AI Coding Agents Perform as Well Without Writing Their Own Tests
A new finding suggests AI coding agents perform just as well when they skip writing their own tests, challenging a common assumption that test generation is key to their effectiveness. Developers are debating what this means for how automated coding tools should be evaluated and used in real-world software projects.
- 38
The open-source project OpenMontage, built in Python by developer calesthio, bills itself as the world's first agentic video production system. It ships 12 production pipelines, more than 100 tools and over 700 agent skill and production-knowledge files, allowing users to turn an AI coding assistant into a full video production studio. The project is drawing attention on GitHub among developers interested in AI-driven content creation.
- 39Claude Opus 5.5 Tops AI Coding Agents on Complex Projects●Claude Opus 5.5 Leads AI Coding Agents with Complex Project Wins
Anthropic's Claude Opus 5.5 is being reported as the leading AI coding agent, outperforming rivals on complex, multi-step software projects. Observers highlight its ability to handle large codebases and sustained tasks, strengthening Anthropic's position in the competitive AI developer tools market against OpenAI and Google.
- 40
Anthropic's Claude Opus 5.5 is being credited by developers as a leading model for coding and agentic work, with users highlighting its performance on programming and multi-step task automation. Discussion centers on comparisons with rival AI models and its fit in real development workflows. No independent benchmarks or official details accompany the claims.
Repos
- morluto/rea Reverse engineer anything with agents, from app behavior down to native binaries.
- mhtsec/ARTEX AI 自主渗透测试系统 | 百度“agent+”攻防挑战赛冠军项目
- alchaincyf/huashu-art-motion 艺术动画skill:35种艺术风格、9种解说语法,用代码让画动起来。
- omnirush-ai/omnirush-gui A desktop coding agent with free access to frontier models.
- gatewai-dev/gitframes Code-first video, rendered natively on WebGPU.
- franzenzenhofer/big-arrow-on-the-screen Let your AI agents paint big arrows, boxes and text on your Mac screen. One CLI, click-through, gone by itself. Skill fo
- QingYunA/answer-me-with-html Answer me with HTML — an agent skill that answers hard questions with a one-page HTML you can actually read. 让 AI Agent
- anteloc/ldraw-nova Agent tooling for generative LEGO models building, built with Astra and Opus 5.5, powered by Jev
- edenfunf/reelmimic Show it a video you love. Get a new video in the same style. An AI crew (Claude Code or Codex) plans, builds and reviews
- lemomo-ai/lemo-opuscar Claude Code skill for short films with no video model: 43 film styles, each a style prompt plus a demo film made entirel
- yetone/magpie Every agent's model. One place. Codex on DeepSeek, Claude Code on Kimi, from the menu bar.
- AgentMemoryRepo/agentmemoryrepo Spec for Agent Memory Repo
- oil-oil/oil-ui Push AI UI design to its limits: explore distinct directions, compare them side by side, and refine against real screens
- feder-cr/invisible_playwright_mcp Playwright MCP server undetected by anti-bots and captchas: AI agent browses the web on anti-detect stealth Firefox, Pyt
- Louis-CFM/coucou A tiny friend in your Mac's notch and on your iPhone that keeps an eye on your AI coding agents: Claude Code, Codex
- angel291592/Intent-Router Intent compiler for AI agents — converges vague requests into typed IntentSpec contracts (probe, ask, or halt before rou
- michael-denyer/pstack-claude Claude Code, Codex, Copilot, Pi, OpenCode, Gemini, and Prime Agent versions of Poteto's pstack. Rigorous agent work
- mvschwarz/openrig Build your own network of agents from Claude Code, Codex and Pi: persistent teams with roles, shared context and owned w
- feitangyuan/onetake Motion films that never cut to the next slide: every beat grows out of the one before, one continuous camera, continuity
- nanaism/yomiyasu AI生成の日本語を自然な日本語へ推敲するAgent Skill / Agent Skill for Refining AI-Generated Japanese into Natural Japanese