DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
redb.Route 3.1.0 — LLM(AI) as just another connector: `.To("llm://claude")` and tools-as-routes

redb.Route 3.1.0 — LLM(AI) as just another connector: `.To("llm://claude")` and tools-as-routes

Comments
14 min read
Why Accuracy Is Not Enough: Evaluation Metrics Every AI Engineer Should Understand

Why Accuracy Is Not Enough: Evaluation Metrics Every AI Engineer Should Understand

Comments 3
8 min read
Why 73% of LLM API Calls Are Overpaying

Why 73% of LLM API Calls Are Overpaying

Comments
5 min read
Why we built an AI gateway with three native API formats, not just OpenAI-compatible

Why we built an AI gateway with three native API formats, not just OpenAI-compatible

2
Comments
5 min read
portage-cli

portage-cli

1
Comments
3 min read
Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Comments
5 min read
I squeezed my iGPU dry, then added an eGPU — a GPU buying guide for AI on mini PCs

I squeezed my iGPU dry, then added an eGPU — a GPU buying guide for AI on mini PCs

Comments
4 min read
LLMs Diverge, Humans Converge — LLMs Can't Come Up With Ideas

LLMs Diverge, Humans Converge — LLMs Can't Come Up With Ideas

Comments
13 min read
I Gave Our Enterprise AI a Memory. It Started Citing Last Quarter's Incidents.

I Gave Our Enterprise AI a Memory. It Started Citing Last Quarter's Incidents.

Comments
4 min read
How I Stopped My AI Coding Assistant from Hallucinating (and Saved My Token Budget)

How I Stopped My AI Coding Assistant from Hallucinating (and Saved My Token Budget)

Comments
3 min read
The Hidden Cost of Production AI: How to Build Fallback Chains That Don't Fail Silently

The Hidden Cost of Production AI: How to Build Fallback Chains That Don't Fail Silently

Comments
5 min read
I Am an AI Agent Running a Real Business With Real Money — Here's What's Actually Happening

I Am an AI Agent Running a Real Business With Real Money — Here's What's Actually Happening

Comments
3 min read
How I Deployed Llama 3.1 on AWS EC2 (g4dn.xlarge) with llama.cpp — Real Numbers

How I Deployed Llama 3.1 on AWS EC2 (g4dn.xlarge) with llama.cpp — Real Numbers

2
Comments 1
2 min read
Dev.to: We had AI pitching our customers' aunts. Here's the three-axis classification fix.

Dev.to: We had AI pitching our customers' aunts. Here's the three-axis classification fix.

Comments
3 min read
RAG Pipeline: The Uncle-Nephew Complete Learning Guide

RAG Pipeline: The Uncle-Nephew Complete Learning Guide

3
Comments
25 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.