DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Unlocking the Power of RAG Systems with LangChain and Vector Databases

Unlocking the Power of RAG Systems with LangChain and Vector Databases

Comments
3 min read
AirLLM Shrinks 70B LLMs to 4GB VRAM; DPO & Supermemory Boost Open Models

AirLLM Shrinks 70B LLMs to 4GB VRAM; DPO & Supermemory Boost Open Models

Comments
3 min read
Switching our LLM-as-judge from 5-class to binary in CI: the patterns we kept

Switching our LLM-as-judge from 5-class to binary in CI: the patterns we kept

Comments
3 min read
From Chatbots to Personal AI Agents: The Infrastructure Developers Actually Need

From Chatbots to Personal AI Agents: The Infrastructure Developers Actually Need

1
Comments
18 min read
RAG pilots fail when the sources are not ready

RAG pilots fail when the sources are not ready

Comments
2 min read
The most expensive bug in an AI agent is the one it's confident about

The most expensive bug in an AI agent is the one it's confident about

Comments
3 min read
AWS Optimizes Starts, Adaptive Worms Rise, and LLM Memory Gets Local

AWS Optimizes Starts, Adaptive Worms Rise, and LLM Memory Gets Local

Comments
2 min read
# Enterprise RAG’s Biggest Risk: Answers That Look Correct but Aren’t

# Enterprise RAG’s Biggest Risk: Answers That Look Correct but Aren’t

Comments
7 min read
The Prompt Quality Report: What 1,018 Scored Prompts Reveal

The Prompt Quality Report: What 1,018 Scored Prompts Reveal

1
Comments
2 min read
Everyone is hardening the structure. Nobody is passing down the why.

Everyone is hardening the structure. Nobody is passing down the why.

Comments 13
3 min read
I built a circuit breaker for LLM agents after seeing someone lose $200 overnight

I built a circuit breaker for LLM agents after seeing someone lose $200 overnight

1
Comments
6 min read
I Built an LLM Filter That Prefers Silence Over Slop, and the Eval Harness That Keeps It Honest

I Built an LLM Filter That Prefers Silence Over Slop, and the Eval Harness That Keeps It Honest

Comments
4 min read
Can You Build an Alternative to LLMs? 8 Months, ~200 Failed Experiments, One Wall. 2

The wall of transferable causal transition

Can You Build an Alternative to LLMs? 8 Months, ~200 Failed Experiments, One Wall. 2

11
Comments 8
9 min read
From Commerce to E-Commerce to MCP-Commerce: The Third Wave

From Commerce to E-Commerce to MCP-Commerce: The Third Wave

Comments
3 min read
Asynchronous Telemetry Blindness in AI Streaming Clients: A PoC Where Text Renders, Billing Stays at Zero

Asynchronous Telemetry Blindness in AI Streaming Clients: A PoC Where Text Renders, Billing Stays at Zero

1
Comments 1
7 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.