DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
How We Solved the Hidden Problem of Cheap LLMs

How We Solved the Hidden Problem of Cheap LLMs

3
Comments 2
8 min read
When AI Meets Reality: Why “Hello World” Isn’t Enough for LLM Systems

When AI Meets Reality: Why “Hello World” Isn’t Enough for LLM Systems

Comments
2 min read
Sharing your prompts is the new telling people your dreams

Sharing your prompts is the new telling people your dreams

Comments
3 min read
Sonnet 4.6 vs Haiku 4.5 vs Opus 4.6: I Tested 3 Claude Models on 10 Real Tasks

Sonnet 4.6 vs Haiku 4.5 vs Opus 4.6: I Tested 3 Claude Models on 10 Real Tasks

Comments
3 min read
What's new in llm-cli-gateway

What's new in llm-cli-gateway

1
Comments
5 min read
Why I Built TokenBar: AI Spend Should Not Be a Monthly Surprise

Why I Built TokenBar: AI Spend Should Not Be a Monthly Surprise

Comments
1 min read
How Top Companies Are Shipping AI Agents Today (Apr 15)

How Top Companies Are Shipping AI Agents Today (Apr 15)

Comments
3 min read
I built an open-source LLM eval framework as a BCA student — hallucination detection, red-teaming, regression tracking

I built an open-source LLM eval framework as a BCA student — hallucination detection, red-teaming, regression tracking

Comments
1 min read
The day I realized AI costs need a warning light

The day I realized AI costs need a warning light

Comments
2 min read
RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

Comments
6 min read
Your model speed benchmark is measuring the wrong thing

Your model speed benchmark is measuring the wrong thing

Comments
3 min read
Moving Beyond My Go-To LLM: Is Anyone Using Anthropic's Mythos? 🤖

Moving Beyond My Go-To LLM: Is Anyone Using Anthropic's Mythos? 🤖

5
Comments 1
1 min read
Do Androids Dream of Your Electric Life?

Do Androids Dream of Your Electric Life?

1
Comments
16 min read
Your AI speed benchmark is measuring the one workload you don't run

Your AI speed benchmark is measuring the one workload you don't run

Comments
3 min read
Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.