DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Tool Definition Drift: When Your Agent's Toolset Outgrows Its Prompt

Tool Definition Drift: When Your Agent's Toolset Outgrows Its Prompt

Comments
8 min read
5 Prompt Mistakes That Make AI Generate Worse Code (With Fixes)

5 Prompt Mistakes That Make AI Generate Worse Code (With Fixes)

Comments
2 min read
Why Small LLMs Fail at Tool Calling: The Shocking Discovery from Our Llama 3B Benchmark

Why Small LLMs Fail at Tool Calling: The Shocking Discovery from Our Llama 3B Benchmark

Comments
11 min read
Day 2 - RAG - What is Vector DB ?

Day 2 - RAG - What is Vector DB ?

Comments
3 min read
Type-Guided Constrained Decoding: How to Stop LLMs from Hallucinating Code

Type-Guided Constrained Decoding: How to Stop LLMs from Hallucinating Code

Comments
7 min read
LLM Routing: How to cut AI Infrastructure costs by 70% Without losing quality

LLM Routing: How to cut AI Infrastructure costs by 70% Without losing quality

1
Comments 1
5 min read
Indeed Data API: Extract Structured JSON in 2026

Indeed Data API: Extract Structured JSON in 2026

Comments
8 min read
A Production Readiness Checklist for Remote MCP Servers

A Production Readiness Checklist for Remote MCP Servers

Comments
6 min read
Prompt Caching Is Quietly Becoming the Operating System of AI Agents

Prompt Caching Is Quietly Becoming the Operating System of AI Agents

5
Comments
1 min read
Multi-Tenant Token Budgets: Quota Patterns That Don't Starve Your Best Customers

Multi-Tenant Token Budgets: Quota Patterns That Don't Starve Your Best Customers

Comments 1
9 min read
I Built My Own LLM Observability Tool — Here’s Why and How

I Built My Own LLM Observability Tool — Here’s Why and How

1
Comments 1
4 min read
Few-Shot Selection at Runtime: Why Static Examples Hurt Edge Cases

Few-Shot Selection at Runtime: Why Static Examples Hurt Edge Cases

Comments
8 min read
Cost-Aware LLM Routing: Sending 30% of Traffic to a Cheaper Model Without Quality Loss

Cost-Aware LLM Routing: Sending 30% of Traffic to a Cheaper Model Without Quality Loss

Comments
10 min read
Why Everyone's Talking About AI Agents (And Why You Should Be Too)

Why Everyone's Talking About AI Agents (And Why You Should Be Too)

1
Comments 1
3 min read
Anthropic Stop Reasons in Production: 5 Cases That Need Different Reactions

Anthropic Stop Reasons in Production: 5 Cases That Need Different Reactions

Comments
8 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.