DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Text-to-SQL Finally Gets Real: DySQL-Bench, BibSQL, DLBench Fix the 'Perfect Query' Myth

Text-to-SQL Finally Gets Real: DySQL-Bench, BibSQL, DLBench Fix the 'Perfect Query' Myth

1
Comments
3 min read
OpenAI just raised $122B. Frontier inference pricing hasn't moved in 9 weeks

OpenAI just raised $122B. Frontier inference pricing hasn't moved in 9 weeks

1
Comments
1 min read
Two-Pass LLM Processing: When Single-Pass Classification Isn't Enough

Two-Pass LLM Processing: When Single-Pass Classification Isn't Enough

Comments
5 min read
AI Integration Without AI Researchers: What Engineering Teams Actually Need in 2026

AI Integration Without AI Researchers: What Engineering Teams Actually Need in 2026

2
Comments 1
4 min read
The 5 Types of AI Agent Memory Every Developer Needs to Know (Part 1)

The 5 Types of AI Agent Memory Every Developer Needs to Know (Part 1)

7
Comments 4
8 min read
Helix AI Studio v2.1.0 — 7 AI Providers, CLI Integration, gemma4 Default

Helix AI Studio v2.1.0 — 7 AI Providers, CLI Integration, gemma4 Default

Comments 1
1 min read
AI Doesn't Replace Thinking. It Replaces Forgetting.

AI Doesn't Replace Thinking. It Replaces Forgetting.

Comments
5 min read
Claude Code Token Crisis: Why I Built a Local Agent Instead of Switching to Codex

Claude Code Token Crisis: Why I Built a Local Agent Instead of Switching to Codex

Comments 3
3 min read
Gemma 4 VRAM Requirements: The hardware guide I wish I had

Gemma 4 VRAM Requirements: The hardware guide I wish I had

2
Comments
1 min read
# 🚀 How Large Language Models (LLMs) Actually Work (With Diagrams + Code)

# 🚀 How Large Language Models (LLMs) Actually Work (With Diagrams + Code)

2
Comments
2 min read
5 Context Window Tricks That Cut My Token Usage in Half

5 Context Window Tricks That Cut My Token Usage in Half

Comments 1
3 min read
Sub-Agent Architectures: Patterns, Trade-offs, and a Kotlin Implementation

Sub-Agent Architectures: Patterns, Trade-offs, and a Kotlin Implementation

Comments
11 min read
The Anatomy of an AI Agent: Memory, Tools, Planning, and Execution Explained

The Anatomy of an AI Agent: Memory, Tools, Planning, and Execution Explained

3
Comments
6 min read
Why Faster First Tokens Matter More Than Total Response Time

Why Faster First Tokens Matter More Than Total Response Time

Comments
10 min read
We Spent a Week Evaluating a Context Compression Tool, Then Killed It

We Spent a Week Evaluating a Context Compression Tool, Then Killed It

Comments 1
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.