DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
When your agent does something bad, can you tell which agent did it?

When your agent does something bad, can you tell which agent did it?

Comments
4 min read
Two Patterns for Reducing LLM Costs in Data-Heavy RAG Apps

Two Patterns for Reducing LLM Costs in Data-Heavy RAG Apps

1
Comments
6 min read
My LLM drift tracker flagged four regressions this week. All four were wrong.

My LLM drift tracker flagged four regressions this week. All four were wrong.

4
Comments 21
4 min read
Cornell Notes on Context Layer

Cornell Notes on Context Layer

Comments
2 min read
How I generate LLM test cases that actually catch bugs

How I generate LLM test cases that actually catch bugs

1
Comments 1
6 min read
73% of AI-agent credential leaks trace back to one mundane thing: debug logging

73% of AI-agent credential leaks trace back to one mundane thing: debug logging

Comments
2 min read
Your AI agent is only as secure as the tools and agents it calls

Your AI agent is only as secure as the tools and agents it calls

Comments
4 min read
I Grepped My Own Claude Code Logs and Found the Hidden Tag Anthropic Never Shows You

I Grepped My Own Claude Code Logs and Found the Hidden Tag Anthropic Never Shows You

1
Comments 2
6 min read
Beyond the Hype: Testing Gemma-4-12B Agentic GGUFs in the Wild

Beyond the Hype: Testing Gemma-4-12B Agentic GGUFs in the Wild

Comments
2 min read
How Transformer Decoders Generate Text — From Causal Masking to Decoding

How Transformer Decoders Generate Text — From Causal Masking to Decoding

Comments
5 min read
How I Built a Production WhatsApp AI Assistant for Mexican SMBs with Claude and n8n

How I Built a Production WhatsApp AI Assistant for Mexican SMBs with Claude and n8n

1
Comments
3 min read
Building a Free AI PDF Assistant: How I Solved Parsing Issues and Minimized LLM Costs

Building a Free AI PDF Assistant: How I Solved Parsing Issues and Minimized LLM Costs

Comments
3 min read
DiffusionGemma: 4x faster text generation

DiffusionGemma: 4x faster text generation

18
Comments
5 min read
I Enabled MCP on My AI Coding Agent and My Token Bill Tripled: Here's the Math

I Enabled MCP on My AI Coding Agent and My Token Bill Tripled: Here's the Math

Comments
4 min read
How to Rank Local LLMs by Cost per Correct Answer (Measured GPU Energy, 8 Ollama Models)

How to Rank Local LLMs by Cost per Correct Answer (Measured GPU Energy, 8 Ollama Models)

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.