DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Tokens, Context, and Why Small AI Tasks Aren't Cheap

Tokens, Context, and Why Small AI Tasks Aren't Cheap

Comments 1
4 min read
Nvidia H100 and GPU Pricing 2026: Buy, Rent, and Cloud Costs Explained

Nvidia H100 and GPU Pricing 2026: Buy, Rent, and Cloud Costs Explained

1
Comments
5 min read
Comparing AI Agents Python Library Options for Production

Comparing AI Agents Python Library Options for Production

Comments
7 min read
We stopped writing eval cases by hand. Now every prod incident becomes one.

We stopped writing eval cases by hand. Now every prod incident becomes one.

Comments
2 min read
PII Redaction Built Entirely in the Browser

PII Redaction Built Entirely in the Browser

4
Comments 4
2 min read
Clioloop — The Open-Source AI Agent with Agentic Fusion

Clioloop — The Open-Source AI Agent with Agentic Fusion

2
Comments
2 min read
AI Agentic Workflow Explained: A Quick Tour of Harness, Tools, Skills, MCP, and Memory

AI Agentic Workflow Explained: A Quick Tour of Harness, Tools, Skills, MCP, and Memory

Comments
5 min read
Is Your AI Agent Production-Ready? Define the Bar First

Is Your AI Agent Production-Ready? Define the Bar First

Comments
3 min read
Distinguishing wrong from absent

Distinguishing wrong from absent

Comments 2
1 min read
Eval-Gated AI Releases: Treating Retrieval Quality Like Unit Tests

Eval-Gated AI Releases: Treating Retrieval Quality Like Unit Tests

1
Comments 1
2 min read
Keeping a client's VLM inference inside the EU with a self-hosted-first gateway

Keeping a client's VLM inference inside the EU with a self-hosted-first gateway

Comments
4 min read
The latency tax of an LLM gateway: I measured Bifrost's overhead

The latency tax of an LLM gateway: I measured Bifrost's overhead

Comments
4 min read
Autonomy Is the Bug: Why Self-Driving Agents Hallucinate When the Model Barely Does

Autonomy Is the Bug: Why Self-Driving Agents Hallucinate When the Model Barely Does

6
Comments 6
8 min read
GLM Is the New Hotness, So Let's Test It On the Homelab

GLM Is the New Hotness, So Let's Test It On the Homelab

Comments 3
11 min read
We got tired of stuffing our AI agent's entire chat history into every prompt, so we built an API that doesn't

We got tired of stuffing our AI agent's entire chat history into every prompt, so we built an API that doesn't

Comments 6
1 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.