DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The day I realized AI costs need a warning light

The day I realized AI costs need a warning light

Comments
2 min read
RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

RAG Series (22): Long Context vs RAG — Do We Even Need RAG?

Comments
6 min read
Your model speed benchmark is measuring the wrong thing

Your model speed benchmark is measuring the wrong thing

Comments
3 min read
Moving Beyond My Go-To LLM: Is Anyone Using Anthropic's Mythos? 🤖

Moving Beyond My Go-To LLM: Is Anyone Using Anthropic's Mythos? 🤖

5
Comments 1
1 min read
Do Androids Dream of Your Electric Life?

Do Androids Dream of Your Electric Life?

1
Comments
16 min read
Your AI speed benchmark is measuring the one workload you don't run

Your AI speed benchmark is measuring the one workload you don't run

Comments
3 min read
Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Boosting llama.cpp with Auto-Tuning, Qwen Quantization Benchmarks, & Mobile Ollama AI Servers

Comments
3 min read
SafePaths: How We Reduced Token Consumption by 85% — The Benchmark Story

SafePaths: How We Reduced Token Consumption by 85% — The Benchmark Story

Comments
2 min read
AI boyfriends are 10x bigger than AI girlfriends

Vanishing nightlife and risk-free AI bonds

AI boyfriends are 10x bigger than AI girlfriends

17
Comments 10
15 min read
I Built an AI Agent to Do My Pre-Refinement. It Turned Into a Mirror of How We Wrote Tickets.

I Built an AI Agent to Do My Pre-Refinement. It Turned Into a Mirror of How We Wrote Tickets.

1
Comments
10 min read
340% and Climbing: What the CIS Prompt Injection Report Means for Enterprise AI Agents

340% and Climbing: What the CIS Prompt Injection Report Means for Enterprise AI Agents

Comments
10 min read
Your Tech Stack Has an AI Problem: How to Audit and Fix It in 2026

Your Tech Stack Has an AI Problem: How to Audit and Fix It in 2026

Comments
8 min read
RAG Series (21): Performance Optimization — Faster and Cheaper

RAG Series (21): Performance Optimization — Faster and Cheaper

Comments
7 min read
How the itrstats tax assistant works: one query, every layer

How the itrstats tax assistant works: one query, every layer

1
Comments
10 min read
Voice-Controlled Local AI Agent

Voice-Controlled Local AI Agent

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.