DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
The Quantization Audit: Why Leaderboard Scores Lie About Local Agent Capabilities

The Quantization Audit: Why Leaderboard Scores Lie About Local Agent Capabilities

Comments 1
1 min read
Microsoft FastContext: a Repo-Explorer Subagent Cuts Coding-Agent Tokens 60%: Explorer-Subagent Context Offloading

Microsoft FastContext: a Repo-Explorer Subagent Cuts Coding-Agent Tokens 60%: Explorer-Subagent Context Offloading

Comments
7 min read
Claude AI da Anthropic: Conheça os Diferenciais Que Destacam Este Modelo [PT-BR]

Claude AI da Anthropic: Conheça os Diferenciais Que Destacam Este Modelo [PT-BR]

Comments
4 min read
Batch Processing vs Real-Time Inference: When to Use Each for Image Generation

Batch Processing vs Real-Time Inference: When to Use Each for Image Generation

Comments
6 min read
GPT-5.6 Sol tops the frontend leaderboard

GPT-5.6 Sol tops the frontend leaderboard

1
Comments 2
1 min read
Engenharia de Prompts para Massa de Dados: Escalando Testes com Cobertura e Sem Duplicidade utilizando LLMs

Engenharia de Prompts para Massa de Dados: Escalando Testes com Cobertura e Sem Duplicidade utilizando LLMs

Comments
5 min read
Harness Engineering 101 — สิ่งที่อยู่ใต้พรมของ Agentic AI

Harness Engineering 101 — สิ่งที่อยู่ใต้พรมของ Agentic AI

Comments
3 min read
Grok 4.5 is Here, And It's All About Developer Efficiency

Grok 4.5 is Here, And It's All About Developer Efficiency

Comments 1
3 min read
จาก chatbot ธรรมดา สู่ AI ที่ทำงานแทนเราได้ — เราเติมอะไรเข้าไปบ้าง?

จาก chatbot ธรรมดา สู่ AI ที่ทำงานแทนเราได้ — เราเติมอะไรเข้าไปบ้าง?

Comments
2 min read
The Retrieval Problem Nobody Talks About: When Your AI Knows Too Much Old Information

The Retrieval Problem Nobody Talks About: When Your AI Knows Too Much Old Information

1
Comments
3 min read
How much VRAM do you actually need to run Llama 3 or Gemma locally?

How much VRAM do you actually need to run Llama 3 or Gemma locally?

Comments
4 min read
Stop Hiding the Chain of Thought: Stream Claude 4.5 Native Thinking Blocks with Spring AI and SSE

Stop Hiding the Chain of Thought: Stream Claude 4.5 Native Thinking Blocks with Spring AI and SSE

Comments
2 min read
consent_url is not a governance layer — what Whire got right and what comes next

consent_url is not a governance layer — what Whire got right and what comes next

Comments
3 min read
How Retrieval‑Augmented Generation Is Revolutionizing Real‑Time, Personalized Career Coaching on AI‑Powered Talent Platforms

How Retrieval‑Augmented Generation Is Revolutionizing Real‑Time, Personalized Career Coaching on AI‑Powered Talent Platforms

Comments
7 min read
How Retrieval‑Augmented Generation Is Revolutionizing Real‑Time, Personalized Career Coaching on AI‑Powered Talent Platforms

How Retrieval‑Augmented Generation Is Revolutionizing Real‑Time, Personalized Career Coaching on AI‑Powered Talent Platforms

Comments
7 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.