DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Introducing Batch Processing for ZeroGPU

Introducing Batch Processing for ZeroGPU

1
Comments
3 min read
I Replaced $800/mo in API Costs with a Local Llama 4 Setup for E-Commerce

I Replaced $800/mo in API Costs with a Local Llama 4 Setup for E-Commerce

Comments
4 min read
GPU cloud servers for AI workloads: how to choose the right instance and deploy without waste

GPU cloud servers for AI workloads: how to choose the right instance and deploy without waste

1
Comments
15 min read
What is an Agentic AI Developer? (And Why It's the Most In-Demand Role of 2026)

What is an Agentic AI Developer? (And Why It's the Most In-Demand Role of 2026)

1
Comments
3 min read
Qwen 3.6, llama.cpp Speculative Decoding, Deepseek TileKernels for Local AI on Consumer GPUs

Qwen 3.6, llama.cpp Speculative Decoding, Deepseek TileKernels for Local AI on Consumer GPUs

Comments
3 min read
Doby: How I Cut Claude Code's Navigation Tokens by 95% with a Spec-First Workflow

Doby: How I Cut Claude Code's Navigation Tokens by 95% with a Spec-First Workflow

Comments
1 min read
I built a new file format to cut AI token costs by 70% — here's how it works

I built a new file format to cut AI token costs by 70% — here's how it works

1
Comments
5 min read
Best MCP Server Directories for Developers

Best MCP Server Directories for Developers

2
Comments 1
17 min read
I open-sourced a 4-agent blood-panel triage workflow on heym, with a deterministic Python safety gate that runs BEFORE any LLM token

I open-sourced a 4-agent blood-panel triage workflow on heym, with a deterministic Python safety gate that runs BEFORE any LLM token

5
Comments 1
5 min read
LocalForge: I built a self-hosted LLM control plane with intelligent routing and LoRA finetuning

LocalForge: I built a self-hosted LLM control plane with intelligent routing and LoRA finetuning

Comments
2 min read
Most RAG Problems Are R(etrieval) Problems

Most RAG Problems Are R(etrieval) Problems

4
Comments 5
3 min read
Your LLM Is Wrong. Your Codebase Is Why.

Your LLM Is Wrong. Your Codebase Is Why.

1
Comments 10
5 min read
Federico@Cursor,Dimma@Fireworks深入探讨Composer2技术

Federico@Cursor,Dimma@Fireworks深入探讨Composer2技术

Comments
2 min read
5 gotchas I hit moving LLM logs from Postgres to ClickHouse

5 gotchas I hit moving LLM logs from Postgres to ClickHouse

1
Comments 2
8 min read
The Actual Cost of Self-Hosting Your LLM (Nobody Does This Math First)

The Actual Cost of Self-Hosting Your LLM (Nobody Does This Math First)

Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.