Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
What I Learned Testing 12 Compression Approaches That Failed
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
What I Learned Testing 12 Compression Approaches That Failed
#
machinelearning
#
llm
#
research
#
python
Comments
Add Comment
6 min read
Why Your RAG System Returns Garbage (And How to Actually Fix It)
Alan West
Alan West
Alan West
Follow
Mar 27
Why Your RAG System Returns Garbage (And How to Actually Fix It)
#
rag
#
llm
#
python
#
ai
Comments
Add Comment
5 min read
Six Characters Fixed My AI's Personality: A Fine-Tuning Story
Meridian_AI
Meridian_AI
Meridian_AI
Follow
Mar 17
Six Characters Fixed My AI's Personality: A Fine-Tuning Story
#
ai
#
machinelearning
#
llm
#
engineering
Comments
Add Comment
4 min read
How to deploy NexusQuant in production (and what's missing)
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
How to deploy NexusQuant in production (and what's missing)
#
machinelearning
#
llm
#
production
#
python
Comments
Add Comment
4 min read
NexusQuant vs KVTC vs TurboQuant vs CommVQ — honest comparison
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
NexusQuant vs KVTC vs TurboQuant vs CommVQ — honest comparison
#
machinelearning
#
llm
#
performance
#
benchmark
Comments
Add Comment
4 min read
NexusQuant benchmarks: every number, honestly
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
NexusQuant benchmarks: every number, honestly
#
machinelearning
#
llm
#
performance
#
opensource
Comments
Add Comment
5 min read
Why Your AI Agents Are Burning Cash and How to Fix It
Alan West
Alan West
Alan West
Follow
Mar 27
Why Your AI Agents Are Burning Cash and How to Fix It
#
ai
#
llm
#
agents
#
python
Comments
Add Comment
5 min read
Longer contexts are easier to compress (not harder)
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Longer contexts are easier to compress (not harder)
#
python
#
machinelearning
#
llm
#
performance
Comments
Add Comment
2 min read
Why E8 lattice quantization beats scalar quantization for KV caches
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Why E8 lattice quantization beats scalar quantization for KV caches
#
python
#
machinelearning
#
math
#
llm
Comments
Add Comment
2 min read
Compress your LLM's KV cache 33x with zero training
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
Compress your LLM's KV cache 33x with zero training
#
python
#
machinelearning
#
llm
#
opensource
Comments
Add Comment
2 min read
Why Your AI Forgets Everything — and How MemPalace Fixes It
Jayde72
Jayde72
Jayde72
Follow
Apr 7
Why Your AI Forgets Everything — and How MemPalace Fixes It
#
ai
#
llm
#
github
#
opensource
1
reaction
Comments
Add Comment
2 min read
How to benchmark NexusQuant on your own model
João André Gomes Marques
João André Gomes Marques
João André Gomes Marques
Follow
Apr 7
How to benchmark NexusQuant on your own model
#
llm
#
python
#
machinelearning
#
opensource
Comments
Add Comment
3 min read
Introducing llm-lean-log: Token-Efficient Chat Logging for AI Agents
Luu Vinh Loc
Luu Vinh Loc
Luu Vinh Loc
Follow
Mar 4
Introducing llm-lean-log: Token-Efficient Chat Logging for AI Agents
#
ai
#
llm
#
logger
Comments
Add Comment
4 min read
Llama vs Mistral vs Phi: Complete Open-Source LLM Comparison for Enterprise (2026)
Jaipal Singh
Jaipal Singh
Jaipal Singh
Follow
Mar 4
Llama vs Mistral vs Phi: Complete Open-Source LLM Comparison for Enterprise (2026)
#
ai
#
llm
#
opensource
#
enterprise
Comments
Add Comment
16 min read
Why We Ditched Bedrock Agents for Nova Pro and Built a Custom Orchestrator
Alex Vega
Alex Vega
Alex Vega
Follow
Apr 5
Why We Ditched Bedrock Agents for Nova Pro and Built a Custom Orchestrator
#
agents
#
architecture
#
aws
#
llm
2
reactions
Comments
2
comments
7 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account