DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Cutting LLM API Cost Without Rewriting Your OpenAI SDK Integration

Cutting LLM API Cost Without Rewriting Your OpenAI SDK Integration

Comments
5 min read
When scraping orchestration is the wrong abstraction for LLM workflows

When scraping orchestration is the wrong abstraction for LLM workflows

Comments
4 min read
Make your site citable by AI: a technical GEO checklist (with code)

Make your site citable by AI: a technical GEO checklist (with code)

Comments
4 min read
What I Learned Building a Real-Time AI Voice Agent

What I Learned Building a Real-Time AI Voice Agent

2
Comments
2 min read
Computer Use Agents Go Local: A Deep Technical Dive into On-Device GUI Automation, Quantized Inference & Holo3.1

Computer Use Agents Go Local: A Deep Technical Dive into On-Device GUI Automation, Quantized Inference & Holo3.1

Comments
18 min read
Small Models, Great Tools: The Engineering Behind a Local AI Agent in Production

Small Models, Great Tools: The Engineering Behind a Local AI Agent in Production

2
Comments 2
7 min read
Bedrock Codex, Robust MILP, Multi‑Model Deliberation, Tree‑Based Molecule Ops, and MoE Quantization

Bedrock Codex, Robust MILP, Multi‑Model Deliberation, Tree‑Based Molecule Ops, and MoE Quantization

Comments
2 min read
AI.Insaf (@ai_tablet) — Полный архив постов канала

AI.Insaf (@ai_tablet) — Полный архив постов канала

Comments
4 min read
The Rise of Agentic Engineering — Part 1: The Prompt Engineering Era

The Rise of Agentic Engineering — Part 1: The Prompt Engineering Era

1
Comments
8 min read
Tokenmaxxing Is a 2026 Anti-Pattern: Why Your Team's Token Bill Is Up 10x and What

Tokenmaxxing Is a 2026 Anti-Pattern: Why Your Team's Token Bill Is Up 10x and What

1
Comments
6 min read
Agent Architecture Is a Compute Allocation Problem: The Advisor Strategy, Cost-Curve Frame Recursed

Agent Architecture Is a Compute Allocation Problem: The Advisor Strategy, Cost-Curve Frame Recursed

1
Comments
17 min read
Self-hosted video creation is coming

Self-hosted video creation is coming

Comments
1 min read
We ran an AI 'peer organization' (Claude + Codex + Gemini) for 7 weeks. Here is the operational record.

We ran an AI 'peer organization' (Claude + Codex + Gemini) for 7 weeks. Here is the operational record.

3
Comments 58
5 min read
LLM SQL Guard Architecture: Parser, Catalog, Policy Engine, Audit Log

LLM SQL Guard Architecture: Parser, Catalog, Policy Engine, Audit Log

Comments
1 min read
The Infrastructure Rule That Prevents AI Automation Disasters

The Infrastructure Rule That Prevents AI Automation Disasters

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.