Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Local LLM Acceleration: Quantization, TTS, and 1M Tokens/Sec
soy
soy
soy
Follow
Mar 26
Local LLM Acceleration: Quantization, TTS, and 1M Tokens/Sec
#
ai
#
machinelearning
#
llm
Comments
Add Comment
4 min read
Designing Agent Fleets That Survive Rate Limits: A Production Architecture Guide
Rhumb
Rhumb
Rhumb
Follow
Mar 31
Designing Agent Fleets That Survive Rate Limits: A Production Architecture Guide
#
ai
#
agents
#
programming
#
llm
Comments
Add Comment
6 min read
I Replaced Cloud AI APIs With a $600 Mac Mini — Here's What Actually Works
MaxxMini
MaxxMini
MaxxMini
Follow
Mar 27
I Replaced Cloud AI APIs With a $600 Mac Mini — Here's What Actually Works
#
ai
#
machinelearning
#
llm
#
programming
1
 reaction
Comments
Add Comment
4 min read
Why Your Agent's Eval Suite Won't Catch Production Failures
Devon
Devon
Devon
Follow
Mar 27
Why Your Agent's Eval Suite Won't Catch Production Failures
#
python
#
ai
#
agents
#
llm
Comments
Add Comment
6 min read
Multi-Agent Systems Break Differently Than Single Agents
Devon
Devon
Devon
Follow
Mar 27
Multi-Agent Systems Break Differently Than Single Agents
#
python
#
ai
#
agents
#
llm
Comments
Add Comment
7 min read
The Prompt Tax Most LLM Teams Are Silently Paying
Parag Darade
Parag Darade
Parag Darade
Follow
Apr 29
The Prompt Tax Most LLM Teams Are Silently Paying
#
ai
#
llm
#
rag
#
machinelearning
2
 reactions
Comments
Add Comment
4 min read
Lemonade v10.3: Run Local LLMs, Image Gen, and Speech on Your Own GPU for Free
ArshTechPro
ArshTechPro
ArshTechPro
Follow
Apr 29
Lemonade v10.3: Run Local LLMs, Image Gen, and Speech on Your Own GPU for Free
#
ai
#
opensource
#
llm
#
programming
10
 reactions
Comments
Add Comment
5 min read
How LLMs Memorize Phone Numbers (and How Labs Stop It)
Gabriel Anhaia
Gabriel Anhaia
Gabriel Anhaia
Follow
Apr 29
How LLMs Memorize Phone Numbers (and How Labs Stop It)
#
ai
#
llm
#
security
#
privacy
1
 reaction
Comments
Add Comment
7 min read
Engineering approach: Startup Mode v/s Big Tech Mode
Kirti Rathore
Kirti Rathore
Kirti Rathore
Follow
Apr 29
Engineering approach: Startup Mode v/s Big Tech Mode
#
architecture
#
llm
#
performance
#
startup
2
 reactions
Comments
Add Comment
6 min read
RAG and Vector Databases:
Sivakami Thangaraj
Sivakami Thangaraj
Sivakami Thangaraj
Follow
Apr 30
RAG and Vector Databases:
#
ai
#
database
#
llm
#
rag
1
 reaction
Comments
Add Comment
2 min read
Making OpenClaw Use the Right Model for Each Task
Devon
Devon
Devon
Follow
Mar 26
Making OpenClaw Use the Right Model for Each Task
#
python
#
ai
#
agents
#
llm
Comments
Add Comment
5 min read
7 AI Gateways That Actually Work in Production (2026 Guide)
Varshith V Hegde
Varshith V Hegde
Varshith V Hegde
Follow
Apr 29
7 AI Gateways That Actually Work in Production (2026 Guide)
#
ai
#
devops
#
programming
#
llm
38
 reactions
Comments
2
 comments
11 min read
Building an AI Agent That Owns Post-Call Execution: Architecture Decisions
SpurIQ Engineering
SpurIQ Engineering
SpurIQ Engineering
Follow
Apr 29
Building an AI Agent That Owns Post-Call Execution: Architecture Decisions
#
ai
#
architecture
#
llm
#
revenue
1
 reaction
Comments
Add Comment
6 min read
Why Strict JSON Mode Doesn't Stop Hallucinated Tool Calls
Gabriel Anhaia
Gabriel Anhaia
Gabriel Anhaia
Follow
Apr 29
Why Strict JSON Mode Doesn't Stop Hallucinated Tool Calls
#
ai
#
llm
#
agents
#
python
Comments
Add Comment
7 min read
Every LLM Eval Library Has the Same Bug: Stochastic Judges Used as Deterministic Oracles
Gabriel Anhaia
Gabriel Anhaia
Gabriel Anhaia
Follow
Apr 29
Every LLM Eval Library Has the Same Bug: Stochastic Judges Used as Deterministic Oracles
#
ai
#
testing
#
llm
#
observability
Comments
Add Comment
7 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account