DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Tame the AI Black Box: How to Continuously Monitor Your Brand’s LLM Visibility

Tame the AI Black Box: How to Continuously Monitor Your Brand’s LLM Visibility

Comments
4 min read
NVIDIA RTX Spark Superchip: Unified CPU–GPU Memory

NVIDIA RTX Spark Superchip: Unified CPU–GPU Memory

Comments 1
8 min read
I published pip install ajah-sdk and npm install ajah-sdk — here's what they do

I published pip install ajah-sdk and npm install ajah-sdk — here's what they do

1
Comments
2 min read
Replacing Fragile CSS Selectors with LLM-Powered Zero-Shot JSON Extraction

Replacing Fragile CSS Selectors with LLM-Powered Zero-Shot JSON Extraction

Comments
8 min read
I built a CLI tool that explains any error in plain English — just pipe it

I built a CLI tool that explains any error in plain English — just pipe it

Comments
1 min read
Hot-swapping GGUF models kernel-panicked my M4 Mac: wired memory, llama.cpp, and why we restart the server instead

Hot-swapping GGUF models kernel-panicked my M4 Mac: wired memory, llama.cpp, and why we restart the server instead

Comments
5 min read
When Gatekeepers Panic: The Encyclopédie, Open AI Models, and the Politics of Accessible Knowledge

When Gatekeepers Panic: The Encyclopédie, Open AI Models, and the Politics of Accessible Knowledge

Comments
22 min read
What Is an MCP Proxy - And When Do You Actually Need a Gateway Instead?

What Is an MCP Proxy - And When Do You Actually Need a Gateway Instead?

1
Comments
7 min read
Retrieval-Augmented Self-Recall — Part 3: Teaching RAG to Say \"I Don't Know\

Retrieval-Augmented Self-Recall — Part 3: Teaching RAG to Say \"I Don't Know\

1
Comments
5 min read
Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Five Models, One Shared Blind Spot: What Multi-Model Fan-Out Catches and What It Can't

Comments 2
7 min read
Stop guessing your AI bill: one endpoint for GPT-5.5, Claude & Gemini at a flat per-call price

Stop guessing your AI bill: one endpoint for GPT-5.5, Claude & Gemini at a flat per-call price

Comments 2
2 min read
Your RAG System Is Lying To You About That Table

Your RAG System Is Lying To You About That Table

13
Comments 2
4 min read
Structured Output Gives You Syntax. It Doesn't Give You Semantics

Structured Output Gives You Syntax. It Doesn't Give You Semantics

Comments 6
5 min read
A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

A Chinese 8B model beat the Western 8B models at Japanese RAG. I still wouldn't put it in the default deployment — and that distinction is the point.

Comments
4 min read
Part 2 — Why Does One System Need Three Chunking Strategies? And One Document Type Shouldn't Be Chunked At All

Part 2 — Why Does One System Need Three Chunking Strategies? And One Document Type Shouldn't Be Chunked At All

6
Comments
9 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.