DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
The gap between detecting hallucinations and handling them

The gap between detecting hallucinations and handling them

2
Comments
2 min read
How We Solved the Hidden Problem of Cheap LLMs

How We Solved the Hidden Problem of Cheap LLMs

3
Comments 2
8 min read
When AI Meets Reality: Why “Hello World” Isn’t Enough for LLM Systems

When AI Meets Reality: Why “Hello World” Isn’t Enough for LLM Systems

Comments
2 min read
Sharing your prompts is the new telling people your dreams

Sharing your prompts is the new telling people your dreams

Comments
3 min read
Sonnet 4.6 vs Haiku 4.5 vs Opus 4.6: I Tested 3 Claude Models on 10 Real Tasks

Sonnet 4.6 vs Haiku 4.5 vs Opus 4.6: I Tested 3 Claude Models on 10 Real Tasks

Comments
3 min read
What's new in llm-cli-gateway

What's new in llm-cli-gateway

1
Comments
5 min read
Why I Built TokenBar: AI Spend Should Not Be a Monthly Surprise

Why I Built TokenBar: AI Spend Should Not Be a Monthly Surprise

Comments
1 min read
How Top Companies Are Shipping AI Agents Today (Apr 15)

How Top Companies Are Shipping AI Agents Today (Apr 15)

Comments
3 min read
I built an open-source LLM eval framework as a BCA student — hallucination detection, red-teaming, regression tracking

I built an open-source LLM eval framework as a BCA student — hallucination detection, red-teaming, regression tracking

Comments
1 min read
RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

Comments
6 min read
The day I realized AI costs need a warning light

The day I realized AI costs need a warning light

Comments
2 min read
Your model speed benchmark is measuring the wrong thing

Your model speed benchmark is measuring the wrong thing

Comments
3 min read
Do Androids Dream of Your Electric Life?

Do Androids Dream of Your Electric Life?

1
Comments
16 min read
Your AI speed benchmark is measuring the one workload you don't run

Your AI speed benchmark is measuring the one workload you don't run

Comments
3 min read
Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.