DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Your AI Agent's Logs Are Lying to You: A 4-Field Schema That Actually Works

Your AI Agent's Logs Are Lying to You: A 4-Field Schema That Actually Works

Comments
5 min read
Why MCP Hookups, Not Just Models, enable Agent Potential in Production

Why MCP Hookups, Not Just Models, enable Agent Potential in Production

Comments
5 min read
Why your AI agent needs more than just an OpenAI API key

Why your AI agent needs more than just an OpenAI API key

2
Comments
5 min read
The Looping Principle: A Simple Mental Model for Understanding AI Agents

The Looping Principle: A Simple Mental Model for Understanding AI Agents

Comments
4 min read
Seven Real Failures of an LLM Agent Operating a CAD Kernel (and How the Architecture Contained Them)

Seven Real Failures of an LLM Agent Operating a CAD Kernel (and How the Architecture Contained Them)

Comments 7
6 min read
This Smart-Home Agent Treats Its Own 1B Model as Untrusted Input

This Smart-Home Agent Treats Its Own 1B Model as Untrusted Input

Comments
3 min read
The RAG Bug That Isn't an Error: Bad Retrieval

The RAG Bug That Isn't an Error: Bad Retrieval

11
Comments 3
1 min read
CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

1
Comments
4 min read
737x faster LangGraph checkpoints, and the case where Rust lost

737x faster LangGraph checkpoints, and the case where Rust lost

2
Comments 2
5 min read
EchoLeak: zero-click data theft from an AI assistant

EchoLeak: zero-click data theft from an AI assistant

1
Comments
3 min read
LLM Policy as Code: Version-Controlled Governance for Model and Agent Access

LLM Policy as Code: Version-Controlled Governance for Model and Agent Access

Comments
4 min read
How I Cut 30% LLM Costs: RAG Context Pruning Cost Reduction

How I Cut 30% LLM Costs: RAG Context Pruning Cost Reduction

Comments
10 min read
Why I'm writing about making AI agents actually reliable

Why I'm writing about making AI agents actually reliable

Comments
1 min read
Taming LLM Tail Latency: Dynamic Request Hedging with Java's JEP 480 Structured Concurrency

Taming LLM Tail Latency: Dynamic Request Hedging with Java's JEP 480 Structured Concurrency

Comments
2 min read
How to Build a Streaming Chatbot API in Python with FastAPI and SSE

How to Build a Streaming Chatbot API in Python with FastAPI and SSE

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.