DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Run and Compare AI Evaluations with a CLI for Developers and Coding Agents

Run and Compare AI Evaluations with a CLI for Developers and Coding Agents

4
Comments
9 min read
claude-real-video: feeding an LLM the frames that actually matter

claude-real-video: feeding an LLM the frames that actually matter

Comments
3 min read
I Almost Hand-Wrote a FHIR Schema. Then I Found Out I Didn't Have To.

I Almost Hand-Wrote a FHIR Schema. Then I Found Out I Didn't Have To.

1
Comments
3 min read
AI Memory Is Not Just a Bigger Context Window

AI Memory Is Not Just a Bigger Context Window

Comments
1 min read
Your reasoning model isn't dumb. Your parser is throwing away its best answers.

Your reasoning model isn't dumb. Your parser is throwing away its best answers.

1
Comments 2
4 min read
How I Cut LLM Token Usage by 99% — 3 Production Engineering Steps

How I Cut LLM Token Usage by 99% — 3 Production Engineering Steps

2
Comments 2
2 min read
Building Local AI Agents in Java with Tools4AI and Ollama: An Insurance Claims Use Case

Building Local AI Agents in Java with Tools4AI and Ollama: An Insurance Claims Use Case

2
Comments
10 min read
Mistral's Move into Physical AI with Robostral Navigate

Mistral's Move into Physical AI with Robostral Navigate

Comments
3 min read
What Actually Happens When You Ask ChatGPT a Question? A Step-by-Step Journey Inside an LLM

What Actually Happens When You Ask ChatGPT a Question? A Step-by-Step Journey Inside an LLM

1
Comments
7 min read
Your AI Agent's Bill Tripled Overnight. The Prompt Cache Broke, Not the Model.

Your AI Agent's Bill Tripled Overnight. The Prompt Cache Broke, Not the Model.

Comments
3 min read
Ship a Production RAG Chatbot in a Weekend with Claude, pgvector, and FastAPI

Ship a Production RAG Chatbot in a Weekend with Claude, pgvector, and FastAPI

Comments
6 min read
Helicone is now in maintenance mode. Here's how to switch to OpenObserve in 5 minutes.

Helicone is now in maintenance mode. Here's how to switch to OpenObserve in 5 minutes.

Comments
3 min read
Bonsai-27B: A 1-Bit LLM for On-Device Inference with Llama.cpp and MLX

Bonsai-27B: A 1-Bit LLM for On-Device Inference with Llama.cpp and MLX

Comments
3 min read
MiniCPM5-1B-Claude-Opus-Fable5-Thinking: A Compact LLM for Enhanced Coding and Instruction Following

MiniCPM5-1B-Claude-Opus-Fable5-Thinking: A Compact LLM for Enhanced Coding and Instruction Following

Comments
3 min read
Human-in-the-Loop for the OpenAI Agents SDK

Human-in-the-Loop for the OpenAI Agents SDK

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.