DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I Built a Local LLM Rig to Escape API Bills. Then I Paid OpenAI Again.

I Built a Local LLM Rig to Escape API Bills. Then I Paid OpenAI Again.

Comments 2
1 min read
BeeLlama.cpp enhances llama.cpp, Qwen 35B hits 128K context, iOS local LLMs with Ollama

BeeLlama.cpp enhances llama.cpp, Qwen 35B hits 128K context, iOS local LLMs with Ollama

Comments
3 min read
The expensive part of an AI agent failure is usually the retry loop

The expensive part of an AI agent failure is usually the retry loop

Comments 1
3 min read
AI Evals, Part 2: Error Analysis The Unglamorous Superpower Behind Good Evals

AI Evals, Part 2: Error Analysis The Unglamorous Superpower Behind Good Evals

Comments
5 min read
Deterministic reliability stack for LLM pipelines

Deterministic reliability stack for LLM pipelines

Comments
1 min read
LLM Token Counting and Cost Optimization: A Practical Guide

LLM Token Counting and Cost Optimization: A Practical Guide

1
Comments
5 min read
Generation 1 — Standalone Models (2018–2022)

Generation 1 — Standalone Models (2018–2022)

Comments
5 min read
Why Most WordPress SEO Plugins Are Not Ready for AI Search Yet

Why Most WordPress SEO Plugins Are Not Ready for AI Search Yet

Comments
5 min read
A Survey of LLM-based Deep Search Agents Adaptive Path Planning via Weighted A* and Heuristic Rewards

A Survey of LLM-based Deep Search Agents Adaptive Path Planning via Weighted A* and Heuristic Rewards

Comments
4 min read
How Stripe, Shopify, and Airbnb Build AI Harnesses

How Stripe, Shopify, and Airbnb Build AI Harnesses

Comments
3 min read
I Trained an LLM on 75K of My Own Messages So It Would Stop Writing Like a Chatbot

I Trained an LLM on 75K of My Own Messages So It Would Stop Writing Like a Chatbot

Comments
8 min read
Claude Code Chose a Stock Ticker Over Someone's Life. We Investigated.

Claude Code Chose a Stock Ticker Over Someone's Life. We Investigated.

Comments 1
10 min read
I ran local LLMs on my phone for a week, and now my desktop setup feels like overkill

I ran local LLMs on my phone for a week, and now my desktop setup feels like overkill

9
Comments
1 min read
Structuring Raw Interaction Data in AI Agents using Weaviate Engram

Structuring Raw Interaction Data in AI Agents using Weaviate Engram

2
Comments
3 min read
tierKV: A Distributed KV Cache That Makes Evicted Blocks Faster to Restore Than GPU Cache Hits

tierKV: A Distributed KV Cache That Makes Evicted Blocks Faster to Restore Than GPU Cache Hits

1
Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.