DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

1
Comments
5 min read
X's Feed Ranking Algorithm: How Grok Ranks 500M Posts in 200ms

X's Feed Ranking Algorithm: How Grok Ranks 500M Posts in 200ms

Comments
8 min read
The Request Is the Wrong Unit of Scale for LLMs on Kubernetes

The Request Is the Wrong Unit of Scale for LLMs on Kubernetes

Comments
12 min read
Redacting PII in LLM Traces Without Losing Debuggability

Redacting PII in LLM Traces Without Losing Debuggability

Comments
6 min read
Stop Using Raw Vector Search: Implement GraphRAG with Spring AI and Neo4j

Stop Using Raw Vector Search: Implement GraphRAG with Spring AI and Neo4j

Comments
2 min read
How to Access DeepSeek API from Outside China (2026 Guide)

How to Access DeepSeek API from Outside China (2026 Guide)

Comments 2
3 min read
Lenovo's AI Host P7: 190 TOPS, 30W, 122B Models — Too Good to Be True?

Lenovo's AI Host P7: 190 TOPS, 30W, 122B Models — Too Good to Be True?

Comments
3 min read
Lookspan: local-first observability for AI agents

Lookspan: local-first observability for AI agents

Comments
1 min read
The Physical Laws of AI Migrations: Architecting an LLM Orchestrator that Survives Reality

The Physical Laws of AI Migrations: Architecting an LLM Orchestrator that Survives Reality

3
Comments 1
4 min read
I Thought I Was Cataloging Ways AI Agents Fail. I Was Describing Cross-Layer Coherence.

State drift across four agent layers

I Thought I Was Cataloging Ways AI Agents Fail. I Was Describing Cross-Layer Coherence.

16
Comments 62
7 min read
How Markus Builds AI Teams That Actually Ship — Not Just Chat

How Markus Builds AI Teams That Actually Ship — Not Just Chat

Comments
5 min read
Meet Deliberation: 400+ models is easy, knowing which ones earn a place is hard.

Meet Deliberation: 400+ models is easy, knowing which ones earn a place is hard.

4
Comments
11 min read
Bootstrap confidence intervals for your LLM eval metrics

Bootstrap confidence intervals for your LLM eval metrics

Comments 2
4 min read
One API Key for GPT, Claude, Gemini, and Qwen: A Practical Guide to OpenAI-Compatible Model Routing

One API Key for GPT, Claude, Gemini, and Qwen: A Practical Guide to OpenAI-Compatible Model Routing

Comments
5 min read
[Open-Source LLM Agent #2] Streaming a LangGraph Agent as OpenAI-Compatible SSE (with a Thinking Panel)

[Open-Source LLM Agent #2] Streaming a LangGraph Agent as OpenAI-Compatible SSE (with a Thinking Panel)

Comments 2
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.