DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
The KV Cache Is a Market, Not a Cache

The KV Cache Is a Market, Not a Cache

Comments
5 min read
How to batch moderate existing posts and comments with an LLM classification API

How to batch moderate existing posts and comments with an LLM classification API

Comments 1
6 min read
Ollama 0.32.0の「ローカルAIエージェント」を本番機で試せなかった話 — 共有GPUインフラのバージョンを上げない判断と、リリースノートと現物のドリフト

Ollama 0.32.0の「ローカルAIエージェント」を本番機で試せなかった話 — 共有GPUインフラのバージョンを上げない判断と、リリースノートと現物のドリフト

Comments
1 min read
Close your editor before heavy jobs? The heavy job lives inside my editor

Close your editor before heavy jobs? The heavy job lives inside my editor

Comments
4 min read
How I Ran Gemma 4 26B on M-Series Mac: 2GB RAM, 1.8 tok/s

How I Ran Gemma 4 26B on M-Series Mac: 2GB RAM, 1.8 tok/s

Comments
9 min read
The Biggest AI Stories Weren’t Features. They Were Dependency

The Biggest AI Stories Weren’t Features. They Were Dependency

Comments
12 min read
AI Coding Tip 031 - Stop Over-Prompting Reasoning Models

AI Coding Tip 031 - Stop Over-Prompting Reasoning Models

2
Comments
8 min read
Spring AI: Bringing Generative AI into Spring Boot Applications

Spring AI: Bringing Generative AI into Spring Boot Applications

1
Comments 1
9 min read
Build a Local LLM Chatbot with Ollama and Python

Build a Local LLM Chatbot with Ollama and Python

Comments
4 min read
RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother

RAG Beyond the Demo: Pipeline, Citations, Evaluation, and When Not to Bother

Comments 2
7 min read
Four Models Cited My Numbers Perfectly. One Still Misread Them.

Four Models Cited My Numbers Perfectly. One Still Misread Them.

Comments
6 min read
OpenEval: Why LLM Evaluation Needs a Standard Format

OpenEval: Why LLM Evaluation Needs a Standard Format

Comments
1 min read
Your multi-agent system isn't hitting prompt cache. Your system prompt is the reason.

Your multi-agent system isn't hitting prompt cache. Your system prompt is the reason.

4
Comments 6
4 min read
Show HN: Telechat – Self-hosted Claude on Telegram/WhatsApp/Slack, no cloud relay

Show HN: Telechat – Self-hosted Claude on Telegram/WhatsApp/Slack, no cloud relay

2
Comments
2 min read
Deploying DeepSeek V3 (LLM) Using SGLang

Deploying DeepSeek V3 (LLM) Using SGLang

6
Comments 1
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.