DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Keeping a client's VLM inference inside the EU with a self-hosted-first gateway

Keeping a client's VLM inference inside the EU with a self-hosted-first gateway

Comments
4 min read
The latency tax of an LLM gateway: I measured Bifrost's overhead

The latency tax of an LLM gateway: I measured Bifrost's overhead

Comments
4 min read
LLM Self-Preference Bias: How Anonymized Peer Review Fixes It

LLM Self-Preference Bias: How Anonymized Peer Review Fixes It

Comments
7 min read
Drift Detection for LLM Routing: Catching Silent Model Degradation

Drift Detection for LLM Routing: Catching Silent Model Degradation

Comments
8 min read
The Citation Lied Without Lying: The Hard Limit of My Memory Gate

Publicly pre-registered failure predictions

The Citation Lied Without Lying: The Hard Limit of My Memory Gate

18
Comments 119
8 min read
GLM Is the New Hotness, So Let's Test It On the Homelab

GLM Is the New Hotness, So Let's Test It On the Homelab

Comments 3
11 min read
Building a RAG Pipeline From Scratch: What SmartQueue Taught Me About Retrieval

Building a RAG Pipeline From Scratch: What SmartQueue Taught Me About Retrieval

2
Comments
6 min read
Bigger Context Windows Didn't Make Our RAG Smarter

Keyword density vs original decision logic

Bigger Context Windows Didn't Make Our RAG Smarter

22
Comments 42
2 min read
Two queues for local-LLM fleets

Two queues for local-LLM fleets

Comments
7 min read
LLMs didn't kill feature engineering. Engineers did.

LLMs didn't kill feature engineering. Engineers did.

4
Comments 2
3 min read
How I built a 3-provider LLM fallback system in production (and what actually broke)

How I built a 3-provider LLM fallback system in production (and what actually broke)

2
Comments
7 min read
AIchain Tools: Search, Conversion, Embeddings

AIchain Tools: Search, Conversion, Embeddings

Comments
6 min read
Benchmarking LLMs for Coding in 2026: A Practical Guide

Benchmarking LLMs for Coding in 2026: A Practical Guide

Comments
3 min read
Unlocking Local LLM Power with Ollama: A Practical Guide

Unlocking Local LLM Power with Ollama: A Practical Guide

1
Comments
3 min read
Compiling the Process, Not the Code: a machine-checked workflow for coding agents

Compiling the Process, Not the Code: a machine-checked workflow for coding agents

Comments
10 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.