DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
ArchGuard: Detect Architecture Drift Before It Becomes Technical Debt

ArchGuard: Detect Architecture Drift Before It Becomes Technical Debt

2
Comments
6 min read
Our LiDAR detector spent 40% of its time in voxelization, not convs

Our LiDAR detector spent 40% of its time in voxelization, not convs

1
Comments
4 min read
The SLM-First Agent: Why 2026's Best Agentic Systems Run on Small Models

The SLM-First Agent: Why 2026's Best Agentic Systems Run on Small Models

Comments
6 min read
AI Builder Notes - May 2026

AI Builder Notes - May 2026

Comments
7 min read
The Budget Guide to Prompt Engineering: Save Money with Every Token

The Budget Guide to Prompt Engineering: Save Money with Every Token

Comments
13 min read
Designing a 3-Tier LLM Fallback Router with Cooldown Locking

Designing a 3-Tier LLM Fallback Router with Cooldown Locking

Comments
5 min read
How we optimized our LLM pipeline to cut token usage by 70%

How we optimized our LLM pipeline to cut token usage by 70%

1
Comments
3 min read
Mellum2 MoE, Heretic Censorship Removal, & NVIDIA Cosmos 3 Omni-model for Local AI

Mellum2 MoE, Heretic Censorship Removal, & NVIDIA Cosmos 3 Omni-model for Local AI

Comments
3 min read
A weekly tool-use test is the only signal

A weekly tool-use test is the only signal

Comments
3 min read
I built an llms.txt for Salesforce — so AI stops writing deprecated Apex

I built an llms.txt for Salesforce — so AI stops writing deprecated Apex

Comments
4 min read
Provider drift broke our regression evals. We pinned versions through Bifrost.

Provider drift broke our regression evals. We pinned versions through Bifrost.

Comments
4 min read
What every LLM call in your Flutter app actually costs

What every LLM call in your Flutter app actually costs

Comments
1 min read
Centralising tool access for our prompt-assembly agent with Bifrost MCP gateway

Centralising tool access for our prompt-assembly agent with Bifrost MCP gateway

Comments
4 min read
How We Cut AI Infrastructure Costs by 94% Without Sacrificing Quality (And How You Can Too)

How We Cut AI Infrastructure Costs by 94% Without Sacrificing Quality (And How You Can Too)

Comments
13 min read
Cache-hit dispersion is the 7th vendor-risk axis — and the one your invoice can't see

Cache-hit dispersion is the 7th vendor-risk axis — and the one your invoice can't see

Comments
8 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.