DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How to Run an 80B Qwen Model in 4.3GB of RAM: The Edge AI Revolution Explained

How to Run an 80B Qwen Model in 4.3GB of RAM: The Edge AI Revolution Explained

Comments
6 min read
The AI Feature Is Cheap to Build and Expensive to Run

The AI Feature Is Cheap to Build and Expensive to Run

Comments
3 min read
Why I Built a Local AI Assistant Instead of Yet Another Chatbot

Why I Built a Local AI Assistant Instead of Yet Another Chatbot

Comments
5 min read
5 Coding Models You Can Actually Run on a Laptop in 2026

5 Coding Models You Can Actually Run on a Laptop in 2026

Comments
5 min read
DeepSeek V4 Is Now the 'Kill Line' of AI Models — Here's What That Means

DeepSeek V4 Is Now the 'Kill Line' of AI Models — Here's What That Means

Comments
3 min read
Building a Ride-Share Zone-Balancing Agent with LangGraph — Part 3: Giving the Agent Memory

Building a Ride-Share Zone-Balancing Agent with LangGraph — Part 3: Giving the Agent Memory

1
Comments 3
7 min read
The server rejected the norepinephrine — and that was the best thing that happened

The server rejected the norepinephrine — and that was the best thing that happened

Comments
2 min read
OpenAI divise par cinq le prix de Luna, pas ses coûts

OpenAI divise par cinq le prix de Luna, pas ses coûts

Comments
6 min read
From "token" to "MoE": the LLM glossary in dependency order

From "token" to "MoE": the LLM glossary in dependency order

Comments
17 min read
Mnemara updated with the past couple of months of changes.

Mnemara updated with the past couple of months of changes.

1
Comments
2 min read
CoMem Explained — From Paper to Working Code in 10 Minutes

CoMem Explained — From Paper to Working Code in 10 Minutes

Comments
6 min read
CLAUDE.md vs. a memory MCP: what actually goes where

CLAUDE.md vs. a memory MCP: what actually goes where

Comments 2
6 min read
Building low-latency semantic memory for coding agents with LanceDB

Building low-latency semantic memory for coding agents with LanceDB

Comments 1
9 min read
What Actually Happens During a Single LLM Call?

What Actually Happens During a Single LLM Call?

Comments
1 min read
The Promptfoo Acquisition Made Me Realize I Was Evaluating LLMs on Easy Mode

The Promptfoo Acquisition Made Me Realize I Was Evaluating LLMs on Easy Mode

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.