DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
OpenClaw Skills Ecosystem and Practical Production Picks

OpenClaw Skills Ecosystem and Practical Production Picks

Comments
11 min read
Architecting for Speed and Precision: My Blueprint for a Production-Ready RAG System

Architecting for Speed and Precision: My Blueprint for a Production-Ready RAG System

1
Comments
4 min read
7 Production RAG Mistakes I Made (And How to Fix Them)

7 Production RAG Mistakes I Made (And How to Fix Them)

1
Comments
5 min read
Two engines for AI slide decks: HTML output vs gpt-image-2 (and how we solved CJK rendering)

Two engines for AI slide decks: HTML output vs gpt-image-2 (and how we solved CJK rendering)

3
Comments
3 min read
I Built an Offline AI Career Advisor Using Gemma 4 — Here's Exactly How It Works

I Built an Offline AI Career Advisor Using Gemma 4 — Here's Exactly How It Works

Comments
6 min read
Lakera Guard Was Acquired for $300M. Here Is the Free Alternative We Built for Developers.

Lakera Guard Was Acquired for $300M. Here Is the Free Alternative We Built for Developers.

Comments
4 min read
Why RAG Pipelines Silently Hallucinate — And The Decay Score That Catches It Before The LLM Does

Why RAG Pipelines Silently Hallucinate — And The Decay Score That Catches It Before The LLM Does

Comments
2 min read
MCP Security in 2026: How to Protect Your AI Agents from Prompt Injection

MCP Security in 2026: How to Protect Your AI Agents from Prompt Injection

Comments
7 min read
Is That Really 'a'? How Homoglyph Attacks Bypass LLM Security Filters (with Python examples)

Is That Really 'a'? How Homoglyph Attacks Bypass LLM Security Filters (with Python examples)

Comments
6 min read
Two architectures that didn't help small-model agent memory on a free T4

Two architectures that didn't help small-model agent memory on a free T4

Comments
12 min read
How Thinking Machines built interactivity into the model

How Thinking Machines built interactivity into the model

Comments
4 min read
qwen2.5-coder is too slow for Claude Code on a Mac. Here's the fix.

qwen2.5-coder is too slow for Claude Code on a Mac. Here's the fix.

Comments 7
8 min read
When Your LLM Provider Pulls the Rug: Lessons from Anthropic's OAuth Shutdown

When Your LLM Provider Pulls the Rug: Lessons from Anthropic's OAuth Shutdown

Comments
2 min read
How ChatGPT Works (Simple Explanation for Beginners)

How ChatGPT Works (Simple Explanation for Beginners)

Comments
2 min read
LLM Gateway Explained — Build One With LiteLLM + LangChain

LLM Gateway Explained — Build One With LiteLLM + LangChain

1
Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.