DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
What VAKRA Reveals About Why Agents Actually Fail

What VAKRA Reveals About Why Agents Actually Fail

Comments
3 min read
Docker with AI: A Practical Guide to Running LLMs, Agents and MCP

Docker with AI: A Practical Guide to Running LLMs, Agents and MCP

1
Comments 3
7 min read
Writing High-Performance Kernels in TileLang, from GEMM to MLA

Writing High-Performance Kernels in TileLang, from GEMM to MLA

1
Comments 1
12 min read
How to Actually Benchmark Open-Source LLMs Before Ditching Your API Provider

How to Actually Benchmark Open-Source LLMs Before Ditching Your API Provider

Comments
5 min read
What is RAG? A Beginner's Guide to Retrieval-Augmented Generation (For Engineers Who Actually Build It)

What is RAG? A Beginner's Guide to Retrieval-Augmented Generation (For Engineers Who Actually Build It)

1
Comments
5 min read
Cache-Aware Spawning: What Changed in llm-cli-gateway, a Week On

Cache-Aware Spawning: What Changed in llm-cli-gateway, a Week On

Comments
12 min read
Prefix caching in vLLM under multi-tenant agent traffic

Prefix caching in vLLM under multi-tenant agent traffic

Comments 2
4 min read
What to Expect from GTK Cyber at Black Hat USA 2026

What to Expect from GTK Cyber at Black Hat USA 2026

Comments
3 min read
Why most AI tools fail at infrastructure troubleshooting

Why most AI tools fail at infrastructure troubleshooting

3
Comments
2 min read
The Hidden Costs of AI in Production (And How Developers Can Reduce Them)

The Hidden Costs of AI in Production (And How Developers Can Reduce Them)

Comments
4 min read
AI guardrails are not security boundaries

AI guardrails are not security boundaries

6
Comments 2
4 min read
Agents of Chaos: a field study of 16 agent failures (and refusals)

Agents of Chaos: a field study of 16 agent failures (and refusals)

Comments 1
4 min read
16 constitutional AI models built on a Chromebook

16 constitutional AI models built on a Chromebook

Comments 1
1 min read
I Raised Gemma 4's Token Cap. The Dense Model Stopped Refusing.

Gemma 4 Challenge: Write about Gemma 4 Submission

I Raised Gemma 4's Token Cap. The Dense Model Stopped Refusing.

24
Comments 9
7 min read
I built a local-first movie recommender with Corrective-RAG (cited explanations, hybrid retrieval, runs entirely on Ollama)

I built a local-first movie recommender with Corrective-RAG (cited explanations, hybrid retrieval, runs entirely on Ollama)

Comments 1
1 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.