DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Claude's default teaching shape has no return: the 5-node loop that fixes it

Claude's default teaching shape has no return: the 5-node loop that fixes it

1
Comments
6 min read
Cola DLM — Text Generation That Plans Before It Writes

Cola DLM — Text Generation That Plans Before It Writes

Comments
4 min read
Your "Claude Opus" API Might Not Be Claude Opus

Your "Claude Opus" API Might Not Be Claude Opus

Comments
4 min read
I built a reasoning harness for LLM agents. Here's what an agent receives when it calls it.

I built a reasoning harness for LLM agents. Here's what an agent receives when it calls it.

1
Comments 2
4 min read
Rethinking Open Source Contribution in the Age of AI Agents, featuring vLLM Core Maintainer Roger Wang at MLSys'26

Rethinking Open Source Contribution in the Age of AI Agents, featuring vLLM Core Maintainer Roger Wang at MLSys'26

8
Comments 6
3 min read
How to choose the right AIOps platform

How to choose the right AIOps platform

Comments
4 min read
AI-generated accessibility, an update — frontier models still fail, but skills change the game

AI-generated accessibility, an update — frontier models still fail, but skills change the game

4
Comments 4
6 min read
I Built a Private AI Assistant That Queries My Git History and Project Management Data — Using Only Local LLMs

I Built a Private AI Assistant That Queries My Git History and Project Management Data — Using Only Local LLMs

Comments 2
5 min read
Qwen3.6 GGUF Benchmarks, Ternary Bonsai 1.58-bit Models, & Ollama Code Explainer Tool

Qwen3.6 GGUF Benchmarks, Ternary Bonsai 1.58-bit Models, & Ollama Code Explainer Tool

Comments
3 min read
How to Run LLMs Locally When Cloud AI Gets Too Invasive

How to Run LLMs Locally When Cloud AI Gets Too Invasive

Comments
5 min read
Most document AI questions aren't retrieval problems

Most document AI questions aren't retrieval problems

4
Comments
4 min read
How I got 80% code retrieval accuracy without vectors, embeddings, or any ML

How I got 80% code retrieval accuracy without vectors, embeddings, or any ML

Comments
2 min read
Agentic AI's Infrastructure Boom Meets Its Reliability Problem

Agentic AI's Infrastructure Boom Meets Its Reliability Problem

Comments
3 min read
Why I built ragwise: pip-installable RAG with hybrid search, streaming, and agent tools by default

Why I built ragwise: pip-installable RAG with hybrid search, streaming, and agent tools by default

Comments
4 min read
The Real Problems Start After Your MCP Server Works

The Real Problems Start After Your MCP Server Works

4
Comments 9
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.