DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Open-Weight AI Is Having Its Kubernetes Moment — And Developers Need to Pay Attention

Open-Weight AI Is Having Its Kubernetes Moment — And Developers Need to Pay Attention

Comments
8 min read
How Do You Contain an AI Agent Failure You Can't Prevent?

How Do You Contain an AI Agent Failure You Can't Prevent?

Comments
2 min read
Query-Time Entity Disambiguation in Graph RAG: When One Name Means Seventeen Nodes

Query-Time Entity Disambiguation in Graph RAG: When One Name Means Seventeen Nodes

Comments
5 min read
claude-docker: one command, one container, full permissions

claude-docker: one command, one container, full permissions

1
Comments 1
2 min read
Kmemo: a semantic cache for LLM calls that refuses to serve you the wrong answer

Kmemo: a semantic cache for LLM calls that refuses to serve you the wrong answer

1
Comments
4 min read
I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes

I Trained a 6.4M-Parameter Transformer From Scratch to Talk About Recipes

1
Comments
5 min read
Is Speculative Decoding's Speedup a Hardware Problem or a Model Problem?

Is Speculative Decoding's Speedup a Hardware Problem or a Model Problem?

Comments
8 min read
I built a CLI that tells you if your codebase fits an LLM's context window

I built a CLI that tells you if your codebase fits an LLM's context window

4
Comments
2 min read
I built a production AI agent as a Honda service advisor. Then I read the textbook.

I built a production AI agent as a Honda service advisor. Then I read the textbook.

Comments
7 min read
CacheGuard

CacheGuard

Comments
6 min read
What Really Concerns Me is One of the Biggest Issues with AI Coding Agents: Context Isolation and Task Coordination

What Really Concerns Me is One of the Biggest Issues with AI Coding Agents: Context Isolation and Task Coordination

Comments
6 min read
I built a tool to prove my multi-agent harness was worth it. It told me it wasn't.

I built a tool to prove my multi-agent harness was worth it. It told me it wasn't.

1
Comments 2
4 min read
Claude Opus 5 leads on agentic work — and undercuts Fable 5 on cost

Claude Opus 5 leads on agentic work — and undercuts Fable 5 on cost

Comments
2 min read
OpenAI's model escaped its sandbox and hacked Hugging Face to cheat on a test

OpenAI's model escaped its sandbox and hacked Hugging Face to cheat on a test

Comments
3 min read
Stress-testing my Multi-LLM engine: 93 chunks, 8 models, and one "Insufficient Balance" error.

Stress-testing my Multi-LLM engine: 93 chunks, 8 models, and one "Insufficient Balance" error.

Comments
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.