DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
A local model opened 41 of our pull requests in five weeks. The model is the least interesting part.

A local model opened 41 of our pull requests in five weeks. The model is the least interesting part.

Comments
10 min read
Your First LLM API on Kubernetes: From Model to Curl Request

Your First LLM API on Kubernetes: From Model to Curl Request

Comments
10 min read
Prompt Caching vs Fine-Tuning: Cost-Effective LLM Strategies

Prompt Caching vs Fine-Tuning: Cost-Effective LLM Strategies

1
Comments
3 min read
Async LLM inference in CI: stop build workers blocking on slow jobs

Async LLM inference in CI: stop build workers blocking on slow jobs

Comments
4 min read
GLM-5.2 open agent benchmark: 22% Less Tool Failure

GLM-5.2 open agent benchmark: 22% Less Tool Failure

1
Comments
8 min read
When AI-Generated SQL Becomes Untrustworthy: How to Restore Confidence in Our Data

When AI-Generated SQL Becomes Untrustworthy: How to Restore Confidence in Our Data

5
Comments
8 min read
My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

Comments
10 min read
Vibecoding: When the AI writes the code and you manage the intent

Vibecoding: When the AI writes the code and you manage the intent

Comments 1
3 min read
Usage-Based AI Coding Needs Runtime Budgets, Not Just Billing Dashboards

Usage-Based AI Coding Needs Runtime Budgets, Not Just Billing Dashboards

1
Comments
4 min read
One Go interface, ten LLMs, three transport classes

One Go interface, ten LLMs, three transport classes

Comments
6 min read
OpenAI-Compatible APIs Are Great Until Streaming Breaks: What I Check Before Switching Providers

OpenAI-Compatible APIs Are Great Until Streaming Breaks: What I Check Before Switching Providers

Comments
7 min read
How to Test Structured Outputs Across Multiple AI Models

How to Test Structured Outputs Across Multiple AI Models

Comments 1
3 min read
Building an MCP Server with Flama

Building an MCP Server with Flama

11
Comments
10 min read
How much does context cost an AI coding agent? grep vs graph vs LSP, measured across 936 runs

How much does context cost an AI coding agent? grep vs graph vs LSP, measured across 936 runs

Comments
12 min read
Where Does the LLM Fit?

Where Does the LLM Fit?

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.