DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
I built an LLM eval framework from scratch. Here is what I wish I had bought instead.

I built an LLM eval framework from scratch. Here is what I wish I had bought instead.

1
Comments
6 min read
A practical checklist for evaluating an OpenAI-compatible AI API gateway

A practical checklist for evaluating an OpenAI-compatible AI API gateway

Comments
2 min read
Running a 27B Model with 16 GB: Bonsai 2-Bit on M1

Running a 27B Model with 16 GB: Bonsai 2-Bit on M1

Comments
4 min read
I Built an AI Liar's Dice Opponent That Remembers How You Play

I Built an AI Liar's Dice Opponent That Remembers How You Play

5
Comments 5
6 min read
2026년 7월 14일 AI·LLM 이슈 다이제스트 — 에이전트가 사무실로 들어오는데, 문단속은 누가 하나

2026년 7월 14일 AI·LLM 이슈 다이제스트 — 에이전트가 사무실로 들어오는데, 문단속은 누가 하나

Comments
1 min read
How I Built a Lightweight Local AI Assistant (Python + PyQt6 + Ollama)

How I Built a Lightweight Local AI Assistant (Python + PyQt6 + Ollama)

Comments
1 min read
NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM

NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM

Comments
4 min read
Stop Re-Prefilling: Cross-Model KV Cache Transfer Makes LLM Swaps 25x Faster

Stop Re-Prefilling: Cross-Model KV Cache Transfer Makes LLM Swaps 25x Faster

1
Comments
4 min read
What I learned building a long-lived AI agent (the boring version)

Real-world architecture beats benchmark hype

What I learned building a long-lived AI agent (the boring version)

15
Comments 31
6 min read
最近整理了几个 API 中转站,高效安全又便捷!

最近整理了几个 API 中转站,高效安全又便捷!

Comments
1 min read
The Most Dangerous Bias of Your AI Assistant Is That It Agrees with You – Part 2: Why We Also Need to Remove Rules Again

The Most Dangerous Bias of Your AI Assistant Is That It Agrees with You – Part 2: Why We Also Need to Remove Rules Again

5
Comments 4
7 min read
Agentic tool-use eval on a local 35B (Q8): trap-tool avoidance is solid, but I can't tell if my failures are the model or my harness

Agentic tool-use eval on a local 35B (Q8): trap-tool avoidance is solid, but I can't tell if my failures are the model or my harness

1
Comments 3
3 min read
Fine-Tuning vs RAG: The Decision Framework I Actually Use

Fine-Tuning vs RAG: The Decision Framework I Actually Use

Comments
4 min read
AI dragons: capable of everything, governed by nothing

AI dragons: capable of everything, governed by nothing

Comments
4 min read
How to Build a Good Human-in-the-Loop for AI Coding Agents

How to Build a Good Human-in-the-Loop for AI Coding Agents

1
Comments
7 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.