DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
RAG SOTA: I Tested 7 Pipelines and Built SEQUOIA (Open Source)

RAG SOTA: I Tested 7 Pipelines and Built SEQUOIA (Open Source)

Comments 2
2 min read
Claude Code and Codex are logging your token usage locally. Here is how to read it.

Claude Code and Codex are logging your token usage locally. Here is how to read it.

Comments
3 min read
Is AI Getting Quietly Dumber? A 24/7 Benchmark That Catches LLM Degradation

Is AI Getting Quietly Dumber? A 24/7 Benchmark That Catches LLM Degradation

Comments 2
6 min read
Your LLM bill is not your capacity plan. Here's the math that pages you at 2am.

Your LLM bill is not your capacity plan. Here's the math that pages you at 2am.

Comments
4 min read
Your AI agent talks to one EU user on Aug 2. Can you prove it disclosed it was AI?

Your AI agent talks to one EU user on Aug 2. Can you prove it disclosed it was AI?

Comments 1
3 min read
Beyond Pay-Per-Token: How Enterprises Barter Architecture for AI Access

Beyond Pay-Per-Token: How Enterprises Barter Architecture for AI Access

Comments
3 min read
LLM Prompt Injection & Guardrail Security

LLM Prompt Injection & Guardrail Security

1
Comments 1
5 min read
Why RAG Fails in Enterprise R&D (And What Actually Works)

Why RAG Fails in Enterprise R&D (And What Actually Works)

Comments 1
5 min read
LLM Structured Output Validation in Python That Holds Up

LLM Structured Output Validation in Python That Holds Up

Comments
14 min read
Agents need a black box recorder, not more memory

Agents need a black box recorder, not more memory

Comments
3 min read
The HTTP Code Your AI Agent Doesn't Handle Yet: 402

The HTTP Code Your AI Agent Doesn't Handle Yet: 402

2
Comments 12
12 min read
AI Reliability: What It Is, Why It Matters, and How to Fix It

AI Reliability: What It Is, Why It Matters, and How to Fix It

Comments
9 min read
What is Agent Memory and why does it matter?

What is Agent Memory and why does it matter?

Comments
7 min read
LLaMA.cpp Gets Qwen MTP Boost, Ring-2.6-1T for Ollama, AMD GPU Fixes

LLaMA.cpp Gets Qwen MTP Boost, Ring-2.6-1T for Ollama, AMD GPU Fixes

Comments
3 min read
Determinism as a feature: when to let your agent call a math API instead of reasoning

Determinism as a feature: when to let your agent call a math API instead of reasoning

1
Comments 5
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.