DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Human-in-the-Loop Approval for LangChain Agents

Human-in-the-Loop Approval for LangChain Agents

Comments
4 min read
Agentic Loops: Quick Guide

Agentic Loops: Quick Guide

Comments
2 min read
AI Agent Memory Explained: How to Build Systems Your Agents Actually Remember

AI Agent Memory Explained: How to Build Systems Your Agents Actually Remember

Comments
6 min read
How Much Does KV Cache Actually Save? A Real Look at Chinese LLM Cache Pricing

How Much Does KV Cache Actually Save? A Real Look at Chinese LLM Cache Pricing

Comments
1 min read
TurboQuant, Four Months Later: Chasing Google's 6x VRAM Claim Into the Wild

TurboQuant, Four Months Later: Chasing Google's 6x VRAM Claim Into the Wild

Comments
6 min read
I built an LLM eval framework from scratch. Here is what I wish I had bought instead.

I built an LLM eval framework from scratch. Here is what I wish I had bought instead.

1
Comments
6 min read
A practical checklist for evaluating an OpenAI-compatible AI API gateway

A practical checklist for evaluating an OpenAI-compatible AI API gateway

Comments
2 min read
Running a 27B Model with 16 GB: Bonsai 2-Bit on M1

Running a 27B Model with 16 GB: Bonsai 2-Bit on M1

Comments
4 min read
2026년 7월 14일 AI·LLM 이슈 다이제스트 — 에이전트가 사무실로 들어오는데, 문단속은 누가 하나

2026년 7월 14일 AI·LLM 이슈 다이제스트 — 에이전트가 사무실로 들어오는데, 문단속은 누가 하나

Comments
1 min read
How I Built a Lightweight Local AI Assistant (Python + PyQt6 + Ollama)

How I Built a Lightweight Local AI Assistant (Python + PyQt6 + Ollama)

Comments
1 min read
NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM

NVIDIA Unveils Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4: A Deployment-Optimized Hybrid MoE LLM

Comments
4 min read
Stop Re-Prefilling: Cross-Model KV Cache Transfer Makes LLM Swaps 25x Faster

Stop Re-Prefilling: Cross-Model KV Cache Transfer Makes LLM Swaps 25x Faster

1
Comments
4 min read
The Most Dangerous Bias of Your AI Assistant Is That It Agrees with You – Part 2: Why We Also Need to Remove Rules Again

The Most Dangerous Bias of Your AI Assistant Is That It Agrees with You – Part 2: Why We Also Need to Remove Rules Again

5
Comments 4
7 min read
最近整理了几个 API 中转站,高效安全又便捷!

最近整理了几个 API 中转站,高效安全又便捷!

Comments
1 min read
Agentic tool-use eval on a local 35B (Q8): trap-tool avoidance is solid, but I can't tell if my failures are the model or my harness

Agentic tool-use eval on a local 35B (Q8): trap-tool avoidance is solid, but I can't tell if my failures are the model or my harness

1
Comments 3
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.