DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Prompt Caching Explained: How to Cut LLM Costs by 30–99%

Prompt Caching Explained: How to Cut LLM Costs by 30–99%

Comments 1
5 min read
Prompt-Based vs. Native Tool-Calling: Navigating the Local LLM Implementation Minefield

Prompt-Based vs. Native Tool-Calling: Navigating the Local LLM Implementation Minefield

Comments
1 min read
Meet Kent 2.0 - Your Coding Accomplice

Meet Kent 2.0 - Your Coding Accomplice

Comments
4 min read
Grammarly costs $12/mo — a local LLM does it for free (Chrome + Ollama)

Grammarly costs $12/mo — a local LLM does it for free (Chrome + Ollama)

Comments
6 min read
Serverless GPU Inference: Deploy Any Hugging Face Model on Google Cloud Run

Serverless GPU Inference: Deploy Any Hugging Face Model on Google Cloud Run

Comments 2
4 min read
Most Teams Ask the Wrong Question About RAG vs Fine-Tuning

Most Teams Ask the Wrong Question About RAG vs Fine-Tuning

Comments
2 min read
RAG vs Fine-Tuning: Which One Should You Actually Choose?

RAG vs Fine-Tuning: Which One Should You Actually Choose?

1
Comments
6 min read
My RAG Benchmark is lying to me

My RAG Benchmark is lying to me

3
Comments 1
5 min read
Resolving CP949 Errors in Local LLM Benchmarking and Building an Automatic Model Recommendation System

Resolving CP949 Errors in Local LLM Benchmarking and Building an Automatic Model Recommendation System

5
Comments
4 min read
One API key, three terminals: driving many models from the shell with llm, mods and aichat

One API key, three terminals: driving many models from the shell with llm, mods and aichat

1
Comments
5 min read
trajectory-sentinel v0.1.0: correlación de señales de defensa de agentes

trajectory-sentinel v0.1.0: correlación de señales de defensa de agentes

3
Comments
2 min read
goal-anchor v0.1.0: integridad de objetivo para agentes multi-paso

goal-anchor v0.1.0: integridad de objetivo para agentes multi-paso

3
Comments
2 min read
wallet-guard v0.1.0: guardrails de bucle y presupuesto para agentes

wallet-guard v0.1.0: guardrails de bucle y presupuesto para agentes

3
Comments
2 min read
adi-shield v0.1.0: detección de inyección de prompt en 5 vectores

adi-shield v0.1.0: detección de inyección de prompt en 5 vectores

3
Comments
2 min read
scope-lib v0.1.0: evaluación de alcance para agentes de IA en 3 criterios

scope-lib v0.1.0: evaluación de alcance para agentes de IA en 3 criterios

3
Comments
2 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.