DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
KVarN, Cost.dev, headroom — the week the agent runtime bill got itemized

KVarN, Cost.dev, headroom — the week the agent runtime bill got itemized

1
Comments
4 min read
No, AI Is Not "Just a Token Predictor"

No, AI Is Not "Just a Token Predictor"

2
Comments
13 min read
I'm 13, Building a CLI Tool for LLM Cost Tracking, and Shipping It in 10 Days

Privacy-first local logging with no account

I'm 13, Building a CLI Tool for LLM Cost Tracking, and Shipping It in 10 Days

8
Comments 9
3 min read
GPU autoscaling on Kubernetes with KEDA: building an external scaler with NVML

GPU autoscaling on Kubernetes with KEDA: building an external scaler with NVML

Comments
3 min read
New `llama.cpp` Updates, AI Agents for Any LLM, and Quantized Vector Index for Local Inference

New `llama.cpp` Updates, AI Agents for Any LLM, and Quantized Vector Index for Local Inference

Comments
3 min read
I wired 908 creator dossiers into my Substack commenter. Here is what changed.

I wired 908 creator dossiers into my Substack commenter. Here is what changed.

Comments
3 min read
The 20% of your AI agent's tool schemas that's pure cruft (and the one-liner to strip it)

The 20% of your AI agent's tool schemas that's pure cruft (and the one-liner to strip it)

Comments
2 min read
Personal Context vs. Shared Context: A Deep Dive Into How Humans and Organizations Should Feed Their AI Agents

Personal Context vs. Shared Context: A Deep Dive Into How Humans and Organizations Should Feed Their AI Agents

1
Comments
22 min read
DeepSeek V4 Pro vs MiMo V2.5 Pro - Debugging Benchmark

DeepSeek V4 Pro vs MiMo V2.5 Pro - Debugging Benchmark

Comments 1
6 min read
We stopped Googling and started Prompting

We stopped Googling and started Prompting

Comments
4 min read
Prompt injection and LLM security for SaaS

Prompt injection and LLM security for SaaS

1
Comments
10 min read
Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Compass v1.1.0 · we shipped a memory plugin that catches its own consumption drift

Comments
5 min read
From Code Completion to Autonomous Reasoning: What the Oceanus Leak Tells Us About the Future of AI Software Engineering

From Code Completion to Autonomous Reasoning: What the Oceanus Leak Tells Us About the Future of AI Software Engineering

Comments
7 min read
How to Handle LLM API Errors & Rate Limits in Node.js

How to Handle LLM API Errors & Rate Limits in Node.js

Comments
4 min read
I measured the token cost of 13 real AI agents (GitHub's MCP server alone is 3,546 tokens/turn)

I measured the token cost of 13 real AI agents (GitHub's MCP server alone is 3,546 tokens/turn)

Comments 1
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.