DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
RAG LLM: Why Your AI Costs 10x More Than It Should (And How to Fix It)

RAG LLM: Why Your AI Costs 10x More Than It Should (And How to Fix It)

Comments
5 min read
Agentic AI: How LLMs Really Work Behind the Scenes

Agentic AI: How LLMs Really Work Behind the Scenes

8
Comments
4 min read
Nov7, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

Nov7, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab

1
Comments
3 min read
How To Run an Open-Source LLM on Your Personal Computer

How To Run an Open-Source LLM on Your Personal Computer

4
Comments 1
6 min read
TOON Benchmarks: A Critical Analysis of Different Results

TOON Benchmarks: A Critical Analysis of Different Results

2
Comments 1
7 min read
Prompt Caching Slashed My AI Bills by 90%. Here's What Nobody Tells You.

Prompt Caching Slashed My AI Bills by 90%. Here's What Nobody Tells You.

Comments
5 min read
LLPY-14: Evaluación y Métricas de Calidad - Midiendo el Éxito del RAG

LLPY-14: Evaluación y Métricas de Calidad - Midiendo el Éxito del RAG

Comments
12 min read
Understanding RAG: How AI Models Learn to Search Before They Speak

Understanding RAG: How AI Models Learn to Search Before They Speak

1
Comments
3 min read
Spring AI RAG, Demystified: From Toy Demos to Production-Grade Retrieval

Spring AI RAG, Demystified: From Toy Demos to Production-Grade Retrieval

Comments 1
12 min read
🧑‍🚀 Choosing the Right Engine to Launch Your LLM (LM Studio, Ollama, and vLLM)

🧑‍🚀 Choosing the Right Engine to Launch Your LLM (LM Studio, Ollama, and vLLM)

1
Comments 3
3 min read
The Orchestrator Pattern: Routing Conversations to Specialized AI Agents

The Orchestrator Pattern: Routing Conversations to Specialized AI Agents

2
Comments 1
11 min read
AI Security Tools Find Critical curl Vulnerabilities

AI Security Tools Find Critical curl Vulnerabilities

Comments
9 min read
Context Engineering: Giving AI Agents Memory Without Breaking the Token Budget

Context Engineering: Giving AI Agents Memory Without Breaking the Token Budget

3
Comments 3
11 min read
Why Claude Code's Unix Philosophy Beats Other AI Assistants

Why Claude Code's Unix Philosophy Beats Other AI Assistants

Comments
8 min read
🧑‍🚀 Mission Accomplished: How an Engineer-Astronaut Prepared Meta’s CRAG Benchmark for Launch in Docker

🧑‍🚀 Mission Accomplished: How an Engineer-Astronaut Prepared Meta’s CRAG Benchmark for Launch in Docker

1
Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.