DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Slaying the Gemma Beast: How We Fixed Local AI and Shipped Search

Slaying the Gemma Beast: How We Fixed Local AI and Shipped Search

Comments
13 min read
Putting the GPU to Work: Running Local LLMs on a Home Lab

Putting the GPU to Work: Running Local LLMs on a Home Lab

Comments
10 min read
Model Showdown Round 2: Adding Gemma, Kimi, and 579 GB of Stubborn Optimism

Model Showdown Round 2: Adding Gemma, Kimi, and 579 GB of Stubborn Optimism

Comments
11 min read
Mamba/SSM Basics

Mamba/SSM Basics

4
Comments
3 min read
Why Eddie Oz's 'LLMs Under Siege' Is the Defensive Wake-Up Call AI Security Needed

Why Eddie Oz's 'LLMs Under Siege' Is the Defensive Wake-Up Call AI Security Needed

7
Comments
5 min read
Query The Quantum

Query The Quantum

1
Comments
4 min read
MTP Isn't Always a Win: 1.95x on My 3090, but Speculative Decoding Is Hardware-Dependent

MTP Isn't Always a Win: 1.95x on My 3090, but Speculative Decoding Is Hardware-Dependent

1
Comments
3 min read
What does a missing description on an MCP tool actually do? Four failure modes I traced from real MCP servers

What does a missing description on an MCP tool actually do? Four failure modes I traced from real MCP servers

Comments 2
7 min read
Paper opinion: Execution Lineage vs Agent Loops (arXiv 2605.06365)

Paper opinion: Execution Lineage vs Agent Loops (arXiv 2605.06365)

Comments
4 min read
TrueFoundry vs Bifrost: Which AI Gateway Actually Scales With Your Team?

TrueFoundry vs Bifrost: Which AI Gateway Actually Scales With Your Team?

2
Comments 1
7 min read
Our voice agent passed every test and still woke me up at 3am

Our voice agent passed every test and still woke me up at 3am

Comments
4 min read
WWDC 2026 - Apple Just Opened the Foundation Models Framework to Any LLM Provider

WWDC 2026 - Apple Just Opened the Foundation Models Framework to Any LLM Provider

8
Comments
6 min read
The problem wasn't that the AI wrote bad code — weak specs caused unstable implementations

The problem wasn't that the AI wrote bad code — weak specs caused unstable implementations

Comments
2 min read
What Is RAG? Why LLM Memory Alone Is Never Enough

What Is RAG? Why LLM Memory Alone Is Never Enough

3
Comments
5 min read
Why RAG is Like Playing Space Invaders. The Higher the Level the More Difficult it Becomes to Win.

Why RAG is Like Playing Space Invaders. The Higher the Level the More Difficult it Becomes to Win.

Comments
15 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.