DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The Quantization Audit: Why Leaderboard Scores Lie About Local Agent Capabilities

The Quantization Audit: Why Leaderboard Scores Lie About Local Agent Capabilities

Comments 1
1 min read
Claude AI da Anthropic: Conheça os Diferenciais Que Destacam Este Modelo [PT-BR]

Claude AI da Anthropic: Conheça os Diferenciais Que Destacam Este Modelo [PT-BR]

Comments
4 min read
We Just Handed AI Agents the Keys to the Password Vault. What Could Go Wrong?

We Just Handed AI Agents the Keys to the Password Vault. What Could Go Wrong?

1
Comments
3 min read
Batch Processing vs Real-Time Inference: When to Use Each for Image Generation

Batch Processing vs Real-Time Inference: When to Use Each for Image Generation

Comments
6 min read
Engenharia de Prompts para Massa de Dados: Escalando Testes com Cobertura e Sem Duplicidade utilizando LLMs

Engenharia de Prompts para Massa de Dados: Escalando Testes com Cobertura e Sem Duplicidade utilizando LLMs

Comments
5 min read
GPT-5.6 Sol tops the frontend leaderboard

GPT-5.6 Sol tops the frontend leaderboard

1
Comments 2
1 min read
Ollama Complete Guide: Run LLMs Locally in 2026

Ollama Complete Guide: Run LLMs Locally in 2026

1
Comments
4 min read
Five Comments That Redesigned My LLM Verification Pipeline

Five Comments That Redesigned My LLM Verification Pipeline

5
Comments 44
26 min read
Grok 4.5 is Here, And It's All About Developer Efficiency

Grok 4.5 is Here, And It's All About Developer Efficiency

Comments 1
3 min read
Kimi K3 API Guide: Reasoning, Tool Calling, Structured Output, and Vision

Kimi K3 API Guide: Reasoning, Tool Calling, Structured Output, and Vision

Comments 1
6 min read
I built a readability test for my own compression format. It scored 0/24.

Column misalignment in compressed prompts

I built a readability test for my own compression format. It scored 0/24.

8
Comments 10
5 min read
The Retrieval Problem Nobody Talks About: When Your AI Knows Too Much Old Information

The Retrieval Problem Nobody Talks About: When Your AI Knows Too Much Old Information

1
Comments
3 min read
AI Agents That Live Inside a Dreamed-Up World

AI Agents That Live Inside a Dreamed-Up World

1
Comments 2
4 min read
How much VRAM do you actually need to run Llama 3 or Gemma locally?

How much VRAM do you actually need to run Llama 3 or Gemma locally?

Comments
4 min read
Stop Hiding the Chain of Thought: Stream Claude 4.5 Native Thinking Blocks with Spring AI and SSE

Stop Hiding the Chain of Thought: Stream Claude 4.5 Native Thinking Blocks with Spring AI and SSE

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.