DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Provider drift broke our regression evals. We pinned versions through Bifrost.

Provider drift broke our regression evals. We pinned versions through Bifrost.

Comments
4 min read
What every LLM call in your Flutter app actually costs

What every LLM call in your Flutter app actually costs

Comments
1 min read
Centralising tool access for our prompt-assembly agent with Bifrost MCP gateway

Centralising tool access for our prompt-assembly agent with Bifrost MCP gateway

Comments
4 min read
How We Cut AI Infrastructure Costs by 94% Without Sacrificing Quality (And How You Can Too)

How We Cut AI Infrastructure Costs by 94% Without Sacrificing Quality (And How You Can Too)

Comments
13 min read
Cache-hit dispersion is the 7th vendor-risk axis — and the one your invoice can't see

Cache-hit dispersion is the 7th vendor-risk axis — and the one your invoice can't see

Comments
8 min read
What I Learned Building a Local RAG Agent

What I Learned Building a Local RAG Agent

Comments
5 min read
I let GPT-4o and a cheaper model fight over my inbox. GPT-4o lost.

I let GPT-4o and a cheaper model fight over my inbox. GPT-4o lost.

8
Comments 5
4 min read
One image schema for four VLM providers: we stopped reformatting payloads

One image schema for four VLM providers: we stopped reformatting payloads

Comments
4 min read
Introducing Noesis: An AI-Native Runtime for Epistemic Honesty

Introducing Noesis: An AI-Native Runtime for Epistemic Honesty

Comments
2 min read
Rails, Not Rules: enforcing a coding agent's domain vocabulary with checks

Rails, Not Rules: enforcing a coding agent's domain vocabulary with checks

Comments
13 min read
#3 setting env

#3 setting env

Comments
2 min read
Prompting styles - Basic

Prompting styles - Basic

Comments
2 min read
Why Chunking Matters in RAG: The Hidden Key to Better Retrieval

Why Chunking Matters in RAG: The Hidden Key to Better Retrieval

Comments
2 min read
Free contextual chunk headers: heading-aware chunking for hybrid retrieval

Free contextual chunk headers: heading-aware chunking for hybrid retrieval

Comments
4 min read
Why JSON is Becoming a Bottleneck for AI Agents

Why JSON is Becoming a Bottleneck for AI Agents

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.