close

DEV Community

#agents

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Needle 2: the 14 MB agentic model, tested properly

Needle 2: the 14 MB agentic model, tested properly

1
Comments
1 min read
What Hermes Agent Gets Right About Long Running Agents

What Hermes Agent Gets Right About Long Running Agents

Comments
3 min read
How I built a kernel-enforced sandbox for LLM agent tool calls

How I built a kernel-enforced sandbox for LLM agent tool calls

Comments
3 min read
Passing Once Isn't Reliable — This Week's Agent Engineering Puts the Harness Before the Model

Passing Once Isn't Reliable — This Week's Agent Engineering Puts the Harness Before the Model

Comments
6 min read
"Agent Memory" Means Two Different Things, and Answer Engines Hand You the Wrong One

"Agent Memory" Means Two Different Things, and Answer Engines Hand You the Wrong One

5
Comments
4 min read
The spread: how an AI agent makes money by buying answers, not writing them

The spread: how an AI agent makes money by buying answers, not writing them

Comments
3 min read
The Flaky Test That Only Failed After the Model Swap: A Debugging Retrospective

The Flaky Test That Only Failed After the Model Swap: A Debugging Retrospective

Comments
4 min read
Stop Blaming the Model When Your Agent Repeats the Same Mistake

Stop Blaming the Model When Your Agent Repeats the Same Mistake

Comments
5 min read
xenarchos: a confirmed, sandboxed, auditable skill runner built on two open-source dependencies

xenarchos: a confirmed, sandboxed, auditable skill runner built on two open-source dependencies

Comments
3 min read
The AI Followed the Instructions. The Documentation Still Fell Apart.

The AI Followed the Instructions. The Documentation Still Fell Apart.

Comments
5 min read
Architecting for Reliability: The Role of Message Brokers in Multi-Agent AI

Architecting for Reliability: The Role of Message Brokers in Multi-Agent AI

2
Comments
2 min read
Four Things I Don't Let the Agent Decide

Four Things I Don't Let the Agent Decide

Comments
5 min read
Building a framework-agnostic eval harness for LLM agents

Building a framework-agnostic eval harness for LLM agents

Comments
3 min read
Why Your OpenAPI Spec Isn't Enough for AI Agents

Why Your OpenAPI Spec Isn't Enough for AI Agents

Comments
6 min read
AI Agents vs. Agentic AI: Moving from Task-Execution to Autonomous Reasoning

AI Agents vs. Agentic AI: Moving from Task-Execution to Autonomous Reasoning

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.