Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
Python
import antigravity
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Why Your RAG System Returns Garbage (And How to Actually Fix It)
Alan West
Alan West
Alan West
Follow
Mar 27
Why Your RAG System Returns Garbage (And How to Actually Fix It)
#
rag
#
llm
#
python
#
ai
Comments
Add Comment
5 min read
I Built a Semantic Cache That Cuts LLM API Costs by 72% - What Actually Worked and What Didn't
Vinay Kumar Reddy Budideti
Vinay Kumar Reddy Budideti
Vinay Kumar Reddy Budideti
Follow
Mar 4
I Built a Semantic Cache That Cuts LLM API Costs by 72% - What Actually Worked and What Didn't
#
ai
#
python
#
opensource
#
machinelearning
Comments
Add Comment
6 min read
How to deploy NexusQuant in production (and what's missing)
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
How to deploy NexusQuant in production (and what's missing)
#
machinelearning
#
llm
#
production
#
python
Comments
Add Comment
4 min read
Building Privacy-Preserving Machine Learning: A Practical Guide to Federated Learning
Dinesh Garikapati
Dinesh Garikapati
Dinesh Garikapati
Follow
Mar 27
Building Privacy-Preserving Machine Learning: A Practical Guide to Federated Learning
#
machinelearning
#
privacy
#
python
#
tutorial
2
reactions
Comments
Add Comment
4 min read
Cache semántico y FAQ matching: cómo reduje un 40% el costo de LLM en mi motor RAG
Martin Palopoli
Martin Palopoli
Martin Palopoli
Follow
Apr 7
Cache semántico y FAQ matching: cómo reduje un 40% el costo de LLM en mi motor RAG
#
rag
#
python
#
ai
#
postgres
Comments
Add Comment
8 min read
How I Implemented End-to-End SSE Streaming: From LLM to Browser Through Nginx
Martin Palopoli
Martin Palopoli
Martin Palopoli
Follow
Apr 7
How I Implemented End-to-End SSE Streaming: From LLM to Browser Through Nginx
#
sse
#
python
#
fastapi
#
javascript
Comments
Add Comment
13 min read
Compress your LLM's KV cache 33x with zero training
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Compress your LLM's KV cache 33x with zero training
#
python
#
machinelearning
#
llm
#
opensource
Comments
Add Comment
2 min read
Why E8 lattice quantization beats scalar quantization for KV caches
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Why E8 lattice quantization beats scalar quantization for KV caches
#
python
#
machinelearning
#
math
#
llm
Comments
Add Comment
2 min read
Longer contexts are easier to compress (not harder)
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Longer contexts are easier to compress (not harder)
#
python
#
machinelearning
#
llm
#
performance
Comments
Add Comment
2 min read
Why Your AI Agents Are Burning Cash and How to Fix It
Alan West
Alan West
Alan West
Follow
Mar 27
Why Your AI Agents Are Burning Cash and How to Fix It
#
ai
#
llm
#
agents
#
python
Comments
Add Comment
5 min read
Build and deploy a RAG pipeline as a REST API in under 5 minutes with RAGLight
Bessouat40
Bessouat40
Bessouat40
Follow
Mar 4
Build and deploy a RAG pipeline as a REST API in under 5 minutes with RAGLight
#
python
#
ai
#
rag
#
opensource
Comments
Add Comment
3 min read
Building the Trust Layer for AI Trading Agents
laguia
laguia
laguia
Follow
Mar 4
Building the Trust Layer for AI Trading Agents
#
ai
#
mcp
#
api
#
python
1
reaction
Comments
Add Comment
2 min read
Why I Run the Entire Pipeline Twice to Match Products
Kingsley Onoh
Kingsley Onoh
Kingsley Onoh
Follow
Mar 4
Why I Run the Entire Pipeline Twice to Match Products
#
python
#
ecommerce
#
shopify
#
backend
Comments
Add Comment
5 min read
I Built a Bot That Watches multiple Markets at Once and Finds Risk-Free Trades (arbitrage)
Claw Arbs
Claw Arbs
Claw Arbs
Follow
Apr 7
I Built a Bot That Watches multiple Markets at Once and Finds Risk-Free Trades (arbitrage)
#
python
#
asyncio
#
architecture
#
ai
Comments
1
comment
8 min read
Building a Multimodal Cross Cloud Live Agent with ADK, Azure ACA, and Gemini CLI
xbill
xbill
xbill
Follow
for
Google Developer Experts
Apr 5
Building a Multimodal Cross Cloud Live Agent with ADK, Azure ACA, and Gemini CLI
#
googleadk
#
python
#
gemini
#
azureaca
4
reactions
Comments
Add Comment
6 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account