close

DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Passing Once Isn't Reliable — This Week's Agent Engineering Puts the Harness Before the Model

Passing Once Isn't Reliable — This Week's Agent Engineering Puts the Harness Before the Model

Comments
6 min read
"Agent Memory" Means Two Different Things, and Answer Engines Hand You the Wrong One

"Agent Memory" Means Two Different Things, and Answer Engines Hand You the Wrong One

5
Comments
4 min read
My Local LLM Was Running at 1.6% of Its Context. Here's the Setting That Fixed It

My Local LLM Was Running at 1.6% of Its Context. Here's the Setting That Fixed It

Comments
2 min read
Stop Blaming the Model When Your Agent Repeats the Same Mistake

Stop Blaming the Model When Your Agent Repeats the Same Mistake

Comments
5 min read
Fine-Tuning IBM Granite 4.1 8B for Banking-Specific AI Guardrails

Fine-Tuning IBM Granite 4.1 8B for Banking-Specific AI Guardrails

Comments
5 min read
Your AI Eval Has a Blind Spot. You Built It.

Your AI Eval Has a Blind Spot. You Built It.

3
Comments 1
3 min read
An AI pentest agent that structurally can't hallucinate a vulnerability — and runs offline

An AI pentest agent that structurally can't hallucinate a vulnerability — and runs offline

Comments
3 min read
Spring AI Retries and Embeddings: Failed Answers and Full Re-indexes — LLM Cost Control 4/4

Spring AI Retries and Embeddings: Failed Answers and Full Re-indexes — LLM Cost Control 4/4

Comments
9 min read
Nightly Drift Checks: Catch a Free Model's Behavior Change Before Your Users Do

Nightly Drift Checks: Catch a Free Model's Behavior Change Before Your Users Do

Comments
6 min read
Claude Code Surpasses GitHub Copilot: 90% of Developers Adopt AI Coding

Claude Code Surpasses GitHub Copilot: 90% of Developers Adopt AI Coding

Comments
6 min read
The Replay Bundle That Remembers What Happened

The Replay Bundle That Remembers What Happened

4
Comments
6 min read
MAESTRO: threat-modeling AI agents in seven layers

MAESTRO: threat-modeling AI agents in seven layers

2
Comments
3 min read
MCP's Token Overhead: Why Agent Tool Protocols Burn 4-32x More Tokens Than CLI and How to Fix It

MCP's Token Overhead: Why Agent Tool Protocols Burn 4-32x More Tokens Than CLI and How to Fix It

1
Comments
5 min read
Benchmarks can't tell you if agent memory helps your team. A paired control can

Benchmarks can't tell you if agent memory helps your team. A paired control can

Comments
4 min read
How to Implement Retrieval-Augmented Generation (RAG) with Re-Ranking

How to Implement Retrieval-Augmented Generation (RAG) with Re-Ranking

Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.