DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Your AI Agent Can Be Socially Engineered. Here Are 3 Attacks That Prove It.

Your AI Agent Can Be Socially Engineered. Here Are 3 Attacks That Prove It.

4
Comments
4 min read
Building a Context-Aware AI Chat Without a Vector Database

Building a Context-Aware AI Chat Without a Vector Database

Comments
6 min read
Multi-Model LLM Orchestration with OpenRouter

Multi-Model LLM Orchestration with OpenRouter

Comments
6 min read
SimCore: I built a social simulation engine where LLM agents live on a real map of your city

SimCore: I built a social simulation engine where LLM agents live on a real map of your city

Comments
1 min read
I Tried Speculative Decoding on RTX 4060 8GB — Every Config Was Slower Than Baseline

I Tried Speculative Decoding on RTX 4060 8GB — Every Config Was Slower Than Baseline

1
Comments
8 min read
How We Used 5 LLM APIs and 25 AI Agents to Write a 60-Page Book in One Session

How We Used 5 LLM APIs and 25 AI Agents to Write a 60-Page Book in One Session

Comments
12 min read
MEMORY.md Every Turn? That’s Noise, Not Memory.

MEMORY.md Every Turn? That’s Noise, Not Memory.

8
Comments 2
5 min read
AI Memory Architectures Compared: Long Context vs RAG vs Vector Store vs Hybrid (With Benchmarks)

AI Memory Architectures Compared: Long Context vs RAG vs Vector Store vs Hybrid (With Benchmarks)

Comments
10 min read
Fine-Tuning DeepSeek V4 vs GPT-5 vs Claude for Legal AI — Cost, Accuracy & Real Benchmarks

Fine-Tuning DeepSeek V4 vs GPT-5 vs Claude for Legal AI — Cost, Accuracy & Real Benchmarks

Comments
8 min read
The Claude Code Team Declares Emergencies When This One Metric Drops.

The Claude Code Team Declares Emergencies When This One Metric Drops.

Comments
7 min read
I cut Claude API costs by 90% with prompt caching. Here's what I learned before I had to shut it down.

I cut Claude API costs by 90% with prompt caching. Here's what I learned before I had to shut it down.

1
Comments
10 min read
The AI Scaffolding Tax đź’°: The Hidden 70% Nobody Warns You About When Building with LLMs

The AI Scaffolding Tax đź’°: The Hidden 70% Nobody Warns You About When Building with LLMs

Comments
8 min read
When agent trace metrics lie: the span tree double-counting problem

When agent trace metrics lie: the span tree double-counting problem

Comments
9 min read
EVAL #008: NVIDIA Just Open-Sourced an Inference Engine. Now What?

EVAL #008: NVIDIA Just Open-Sourced an Inference Engine. Now What?

1
Comments
10 min read
Local LLMs & Edge AI: Hardware Boost, Security Fixes, and Extreme Compression

Local LLMs & Edge AI: Hardware Boost, Security Fixes, and Extreme Compression

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.