DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
All Data and AI Weekly #224-12 Jan 2026

All Data and AI Weekly #224-12 Jan 2026

5
Comments
4 min read
The LLM Control Stack: From Words to Weights

The LLM Control Stack: From Words to Weights

Comments
4 min read
LLMs Can Now Write GPU Kernels That Beat torch.compile

LLMs Can Now Write GPU Kernels That Beat torch.compile

1
Comments
7 min read
The Squeezing Effect: Why Your Aligned AI Model Gets Worse

The Squeezing Effect: Why Your Aligned AI Model Gets Worse

Comments
3 min read
Developers Love Tools. AI Needs Better Instructions.

Developers Love Tools. AI Needs Better Instructions.

8
Comments
3 min read
Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem

Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem

Comments
12 min read
How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications

How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications

5
Comments
8 min read
Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%

Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%

Comments
7 min read
The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember

The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember

3
Comments
9 min read
Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

Comments
5 min read
📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌

📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌

1
Comments
1 min read
Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents

Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents

Comments
5 min read
Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis

Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis

Comments
5 min read
A Deep Dive into Deep Agent Architecture for AI Coding Assistants

A Deep Dive into Deep Agent Architecture for AI Coding Assistants

4
Comments 1
16 min read
Part 1: Why Transformers Still Forget

Part 1: Why Transformers Still Forget

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.