DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
We Gave AI the Keys. Nobody Asked If It Knows How to Drive.

We Gave AI the Keys. Nobody Asked If It Knows How to Drive.

Comments
4 min read
RAG vs Fine-Tuning: Which Approach Should You Choose?

RAG vs Fine-Tuning: Which Approach Should You Choose?

Comments
3 min read
Quantization formats compared: GGUF vs GPTQ vs AWQ vs NF4

Quantization formats compared: GGUF vs GPTQ vs AWQ vs NF4

Comments
7 min read
The 'Own Hardware for AI' Myth

The 'Own Hardware for AI' Myth

Comments
11 min read
What Are Tokens and Why Do They Matter in LLMs?

What Are Tokens and Why Do They Matter in LLMs?

2
Comments
3 min read
I Ditched Vector Search for My Coding Agent's Memory. FTS5 Won.

I Ditched Vector Search for My Coding Agent's Memory. FTS5 Won.

Comments 1
4 min read
エージェントAPI代、月数万円になってない?マルチモデルルーティングでコストを10分の1にする実践ガイド

エージェントAPI代、月数万円になってない?マルチモデルルーティングでコストを10分の1にする実践ガイド

Comments
2 min read
Mixture of Experts (MoE) Explained Simply: How Modern AI Models Get Bigger Without Getting Slower

Mixture of Experts (MoE) Explained Simply: How Modern AI Models Get Bigger Without Getting Slower

5
Comments 1
5 min read
The hard part of an AI quiz generator isn't the questions — it's the wrong answers

The hard part of an AI quiz generator isn't the questions — it's the wrong answers

1
Comments
2 min read
Best AI Client for Mac (2026): Elvean vs Jan vs Msty vs LM Studio

Best AI Client for Mac (2026): Elvean vs Jan vs Msty vs LM Studio

Comments
6 min read
LLM token budgeting for startups: the playbook before you have a finance function

LLM token budgeting for startups: the playbook before you have a finance function

1
Comments
12 min read
Our AI support agent doesn't use RAG - here's the math

Our AI support agent doesn't use RAG - here's the math

9
Comments 6
10 min read
Tokens per Word: GPT-5 vs Claude vs GPT-4, Measured Across 7 Languages

Tokens per Word: GPT-5 vs Claude vs GPT-4, Measured Across 7 Languages

Comments
3 min read
We Tracked 1M LLM API Calls — 60% Were Wasting Money on the Wrong Model

We Tracked 1M LLM API Calls — 60% Were Wasting Money on the Wrong Model

Comments
3 min read
Cohere's North Mini Code, LLM Token Optimization & OpenMed Healthcare AI Highlight Local AI Advancements

Cohere's North Mini Code, LLM Token Optimization & OpenMed Healthcare AI Highlight Local AI Advancements

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.