DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Chronicle: Rethinking Codebase Context for AI Coding Agents

Chronicle: Rethinking Codebase Context for AI Coding Agents

3
Comments
4 min read
Why MTP doesn't speed up your llama.cpp inference (and how to actually fix it)

Why MTP doesn't speed up your llama.cpp inference (and how to actually fix it)

1
Comments
5 min read
Building a Voice-Controlled AI Agent using AssemblyAI and Groq

Building a Voice-Controlled AI Agent using AssemblyAI and Groq

1
Comments
3 min read
LLM Benchmark Rankings 2026: 15 Models Tested on 38 Real Coding Tasks

LLM Benchmark Rankings 2026: 15 Models Tested on 38 Real Coding Tasks

1
Comments
28 min read
I Built a Debugger for LLM Agents — Here's Why "Observability" Wasn't Enough

I Built a Debugger for LLM Agents — Here's Why "Observability" Wasn't Enough

3
Comments 1
2 min read
The Shai-Hulud Worm Is Now Open Source — Here's How to Stop Self-Replicating Prompts Before They Reach Your LLM

The Shai-Hulud Worm Is Now Open Source — Here's How to Stop Self-Replicating Prompts Before They Reach Your LLM

1
Comments
4 min read
Building KernelMind Part 2: Hybrid Retrieval, Reranking, and Actually Retrieving Useful Code

Building KernelMind Part 2: Hybrid Retrieval, Reranking, and Actually Retrieving Useful Code

2
Comments 3
5 min read
Brazilian Lawyers Fined R$84,000 for Prompt Injection in Court — Here's What Caught Them (and What Didn't)

Brazilian Lawyers Fined R$84,000 for Prompt Injection in Court — Here's What Caught Them (and What Didn't)

2
Comments
5 min read
A CLI tool to score fine-tuning dataset quality before training starts

A CLI tool to score fine-tuning dataset quality before training starts

2
Comments
3 min read
You Probably Don't Need a Custom Agent

You Probably Don't Need a Custom Agent

Comments
3 min read
Stop Blaming Your Prompts. It’s the Architecture, Stup1d!

Stop Blaming Your Prompts. It’s the Architecture, Stup1d!

1
Comments 1
2 min read
20260324_snn_vs_gpu_en

20260324_snn_vs_gpu_en

Comments
6 min read
Open-Weight AI Model Licenses Compared: What MiniMax's Controversy Means for You

Open-Weight AI Model Licenses Compared: What MiniMax's Controversy Means for You

1
Comments
5 min read
VRAMを増やせば解決する、は物理的に間違っている — HBM・CXL・Unified Memoryが取れなかったもの

VRAMを増やせば解決する、は物理的に間違っている — HBM・CXL・Unified Memoryが取れなかったもの

Comments
4 min read
llama.cppの設定で8GBの性能が5倍変わる — 主要オプションの最適値を出した

llama.cppの設定で8GBの性能が5倍変わる — 主要オプションの最適値を出した

Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.