DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Why Your AI Feels Dumb (And How MCP Fixes It)

Why Your AI Feels Dumb (And How MCP Fixes It)

6
Comments 2
3 min read
The Orphan Axiom Problem in Ontology-Based RAG

The Orphan Axiom Problem in Ontology-Based RAG

Comments
6 min read
All Data and AI Weekly #224-12 Jan 2026

All Data and AI Weekly #224-12 Jan 2026

5
Comments
4 min read
The LLM Control Stack: From Words to Weights

The LLM Control Stack: From Words to Weights

Comments
4 min read
LLMs Can Now Write GPU Kernels That Beat torch.compile

LLMs Can Now Write GPU Kernels That Beat torch.compile

1
Comments
7 min read
The Squeezing Effect: Why Your Aligned AI Model Gets Worse

The Squeezing Effect: Why Your Aligned AI Model Gets Worse

Comments
3 min read
Developers Love Tools. AI Needs Better Instructions.

Developers Love Tools. AI Needs Better Instructions.

8
Comments
3 min read
Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem

Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem

Comments
12 min read
How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications

How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications

5
Comments
8 min read
Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%

Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%

Comments
7 min read
Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

Comments
5 min read
The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember

The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember

3
Comments
9 min read
📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌

📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌

1
Comments
1 min read
Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents

Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents

Comments
5 min read
Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis

Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.