Sponsored Content

DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
TIL - Choosing Between Code, an LLM Call, and an AI Agent

TIL - Choosing Between Code, an LLM Call, and an AI Agent

Comments
12 min read
Fully Autonomous. Except for the Human Doing the Hard Part.

Fully Autonomous. Except for the Human Doing the Hard Part.

1
Comments
2 min read
Adaptive Cognitive AI (ACAI): Chapter 1 - Introduction & System Vision

Adaptive Cognitive AI (ACAI): Chapter 1 - Introduction & System Vision

Comments
3 min read
How I Shrunk My Agent's Core from 15 to 9: Why Hit Rate Alone Can't Tell You Which Rules to Retire

How I Shrunk My Agent's Core from 15 to 9: Why Hit Rate Alone Can't Tell You Which Rules to Retire

1
Comments
10 min read
I Told the AI "A Scanner Flagged This" — and It Agreed With Everything

Frontier models fall hard for confirmation bias

I Told the AI "A Scanner Flagged This" — and It Agreed With Everything

13
Comments 15
10 min read
The Geometry Measures, the AI Teaches: An Agent That Verifies Descriptive-Geometry Constructions, with the Google GenAI SDK, Vertex AI & OpenCV

The Geometry Measures, the AI Teaches: An Agent That Verifies Descriptive-Geometry Constructions, with the Google GenAI SDK, Vertex AI & OpenCV

Comments
5 min read
My AI agent tried to delete my secrets. It couldn't.

My AI agent tried to delete my secrets. It couldn't.

1
Comments
9 min read
Your Error Messages Are an API Now

Your Error Messages Are an API Now

Comments 2
4 min read
The Model Writes, the Judge Measures: Anatomy of an LLM Judge

The Model Writes, the Judge Measures: Anatomy of an LLM Judge

Comments
12 min read
Why RL Training Collapses on Long-Horizon Agents

Why RL Training Collapses on Long-Horizon Agents

Comments 2
10 min read
What 78K attack samples taught me about catching prompt injection

What 78K attack samples taught me about catching prompt injection

Comments
2 min read
I built a spend cap for LLM calls. It failed by 4.2x under parallel load.

I built a spend cap for LLM calls. It failed by 4.2x under parallel load.

1
Comments 2
5 min read
Your Knowledge Graph Is Wasting 70% of Its Tokens

Your Knowledge Graph Is Wasting 70% of Its Tokens

1
Comments
2 min read
A 2-Token Prompt and a 39,966-Token Bill: Measuring What My Agent Actually Costs

A 2-Token Prompt and a 39,966-Token Bill: Measuring What My Agent Actually Costs

2
Comments 1
5 min read
Is Agent Memory Just RAG With Extra Steps? We Opened the Source Code to Find Out

Vendor codebase audit settles the RAG debate

Is Agent Memory Just RAG With Extra Steps? We Opened the Source Code to Find Out

6
Comments 8
6 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.