DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Parameter Count Is the Worst Way to Pick a Model on 8GB VRAM

Parameter Count Is the Worst Way to Pick a Model on 8GB VRAM

Comments
5 min read
Seven principles of real memory for AI agents

Seven principles of real memory for AI agents

Comments
12 min read
I tried every major LLM observability platform. Traceport changed how I think about AI gateways.

I tried every major LLM observability platform. Traceport changed how I think about AI gateways.

Comments
4 min read
I Added Self-Hosted GPU Training to MetaClaw — Here's How to Train Your AI Agent on Your Own A100s

I Added Self-Hosted GPU Training to MetaClaw — Here's How to Train Your AI Agent on Your Own A100s

Comments
2 min read
Inference Engines - A visual deep dive into the layers of an LLM

Inference Engines - A visual deep dive into the layers of an LLM

Comments
1 min read
12 million tokens, linear cost: Subquadratic's bet against the attention tax

12 million tokens, linear cost: Subquadratic's bet against the attention tax

Comments
3 min read
How I stopped Claude Code from hallucinating 42% of my React Code

How I stopped Claude Code from hallucinating 42% of my React Code

Comments
14 min read
Your AI Agent’s Prompt Is the Bottleneck, Not the Model

Your AI Agent’s Prompt Is the Bottleneck, Not the Model

Comments
7 min read
MCP Explained Simply: How AI Talk to Your Data

MCP Explained Simply: How AI Talk to Your Data

3
Comments 2
28 min read
Step 3.5 Flash Series New Release! Available to All Step Plan Users!

Step 3.5 Flash Series New Release! Available to All Step Plan Users!

5
Comments
2 min read
Forget Your RAG: Build Your Own LLM Wiki in C# with Ollama + Kimi (Step‑by‑Step Guide)

Forget Your RAG: Build Your Own LLM Wiki in C# with Ollama + Kimi (Step‑by‑Step Guide)

6
Comments
10 min read
Here's what stopped breaking, when you make LLM agents author in two formats

Here's what stopped breaking, when you make LLM agents author in two formats

3
Comments
7 min read
Running AI on a Budget: 11 Tactics for Enterprise-Scale Efficiency

Running AI on a Budget: 11 Tactics for Enterprise-Scale Efficiency

1
Comments
8 min read
I Analyzed AI Coding Mistakes and Built an ESLint Plugin to Catch Them

I Analyzed AI Coding Mistakes and Built an ESLint Plugin to Catch Them

1
Comments
4 min read
We Gave an AI Agent a Long Context Caching Idea. Here's what happened next!

We Gave an AI Agent a Long Context Caching Idea. Here's what happened next!

2
Comments
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.