Coding agents lower the cost of reverse-engineering home devices
Coding agents reduce the cost of writing and maintaining code for reverse-engineering tasks.
Coding agents reduce the cost of writing and maintaining code for reverse-engineering tasks.
SkillCorpus filters ~821,000 open agent skills into 96,401 high-quality skills organized by a 16-class taxonomy and three quality facets.
Amazon Quick is a new agentic AI assistant from AWS designed to automate administrative tasks in sales workflows.
Lila Sciences operates a 24/7 automated lab with robotics and AI to generate scientific data at scale.
Anthropic’s Claude is the primary agent orchestration platform for 40% of enterprises, more than double any rival, followed by Microsoft (18%) and OpenAI (13%).
Shippy is a maritime AI agent built by Allen Institute for AI (AI2) for high-stakes ocean monitoring decisions
A new arXiv study introduces a six-hop relay testbed to measure how message formats affect information fidelity across agent tiers.
GATS eliminates LLM calls during planning by using a three-layer world model and UCB1-based tree search.
A new arXiv paper introduces DeepSearch-World, a deterministic and verifiable environment for training web agents with reproducible tools for search and page reading.
xAI launched Grok 4.5, its first Opus-class model post-Cursor partnership, targeting coding and agents workflows.
Modal closed a $355M Series C to build infrastructure tailored to AI agents, not human developers.
Lilian Weng published a synthesis of 35 papers on harness engineering for recursive self-improvement (RSI).
Claude Cowork, Anthropic’s agent for general knowledge work, is now available on web and mobile for Max subscribers, expanding beyond its desktop origins.
Anthropic’s Fable 5 leads a new 657-task agent benchmark with 48.6% success, narrowly ahead of Opus 4.8 at 48.5%.
A software developer describes a prompt pattern that instructs an agent (Fable) to use its own judgment about when to write tests and which subagent model to invoke.
llm-coding-agent 0.1a0 is a new Python library that implements a coding agent using the LLM framework.
A debate at the AI Engineer World’s Fair questioned the viability of autonomous software factories and agentic loops.
Adobe is prototyping "agentic sites" that assemble web pages in real time based on individual user intent using an LLM and existing content as a grounding corpus.
Vercel’s Chief of Software says agents are a new type of software that demands different primitives for context, tools, resumability, and long-running work compared to web applications.
Impeccable introduces a design ‘skill’ system that translates designer terms like ‘bolder’ or ‘quieter’ into operational guidance for coding agents.