DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The LLM-shaped hole in your XGBoost pipeline

The LLM-shaped hole in your XGBoost pipeline

Comments
1 min read
How I cut my multi-turn LLM API costs by 90% (O(N ) O(N))

How I cut my multi-turn LLM API costs by 90% (O(N ) O(N))

Comments
2 min read
Six Principles in Practice: How an Agentic E2E Found 11 Production Bugs in 8 Runs

Six Principles in Practice: How an Agentic E2E Found 11 Production Bugs in 8 Runs

Comments
13 min read
Chunking in RAG: why your splitter matters more than your embedding model

Chunking in RAG: why your splitter matters more than your embedding model

2
Comments
5 min read
What MCP Really Is — A Demo You Can Run on Your Laptop in 5 Minutes

What MCP Really Is — A Demo You Can Run on Your Laptop in 5 Minutes

Comments
10 min read
Why Your AI Agent Keeps Overreaching — And How to Fix It with a Boundary Contract

Why Your AI Agent Keeps Overreaching — And How to Fix It with a Boundary Contract

1
Comments
4 min read
GeekNews AI Weekly Deep Dive - 2026-06-08

GeekNews AI Weekly Deep Dive - 2026-06-08

Comments
1 min read
Token Consumption Optimization in LLM Applications

Token Consumption Optimization in LLM Applications

1
Comments 1
2 min read
Simple A2A implementation with Strands

Simple A2A implementation with Strands

6
Comments
4 min read
The $47K agent loop: why logging, monitoring, and max_tokens all failed to stop it

The $47K agent loop: why logging, monitoring, and max_tokens all failed to stop it

3
Comments 1
6 min read
LLM integration with Vercel AI SDK

LLM integration with Vercel AI SDK

Comments
4 min read
Hearth: scale-to-zero LLM serving on Kubernetes — and you can hack on it without a GPU

Hearth: scale-to-zero LLM serving on Kubernetes — and you can hack on it without a GPU

2
Comments 1
3 min read
Why your quantized LLM loses its MTP heads and how to keep them

Why your quantized LLM loses its MTP heads and how to keep them

1
Comments
5 min read
Is Vibe Coding Over? The Free Lunch Is Ending, Not the Movement

Is Vibe Coding Over? The Free Lunch Is Ending, Not the Movement

Comments
4 min read
About Sharing Local Inference: A Marketplace for Renting Idle GPUs with an OpenAI-Compatible Backend

About Sharing Local Inference: A Marketplace for Renting Idle GPUs with an OpenAI-Compatible Backend

Comments
8 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.