DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Local Inference Powers Browser Sign Language, Open-Source Agent Infra, & AI Engineering Guides

Local Inference Powers Browser Sign Language, Open-Source Agent Infra, & AI Engineering Guides

Comments
3 min read
🚀 Why Enterprise Knowledge Systems Are Still Broken (And How We Fixed It with an AI Copilot)

🚀 Why Enterprise Knowledge Systems Are Still Broken (And How We Fixed It with an AI Copilot)

Comments
3 min read
How I Tested 5 Small LLMs on a Weak PC (Intel i5, No GPU) – And Found a Winner

How I Tested 5 Small LLMs on a Weak PC (Intel i5, No GPU) – And Found a Winner

Comments
6 min read
Cursor's compression isn't a bug. It's how it works.

Cursor's compression isn't a bug. It's how it works.

Comments
9 min read
AI Is No Longer a Luxury. It's a Workspace Necessity.

AI Is No Longer a Luxury. It's a Workspace Necessity.

Comments 2
1 min read
Bounded retries for agent tool calls: the budget that stopped our infinite-loop incidents

Bounded retries for agent tool calls: the budget that stopped our infinite-loop incidents

Comments
2 min read
I measure how fast 42 LLMs actually answer. Here's the honest method.

I measure how fast 42 LLMs actually answer. Here's the honest method.

1
Comments 1
2 min read
I built Python and Node.js SDKs for my open-source LLM observability gateway — and I need a hosting sponsor

I built Python and Node.js SDKs for my open-source LLM observability gateway — and I need a hosting sponsor

Comments
1 min read
The SLM Advantage: Why Enterprises Are Choosing Small Language Models Over GPT-Scale AI

The SLM Advantage: Why Enterprises Are Choosing Small Language Models Over GPT-Scale AI

Comments
1 min read
Prompt Caching Explained: How to Cut LLM Costs by 30–99%

Prompt Caching Explained: How to Cut LLM Costs by 30–99%

Comments 1
5 min read
Prompt-Based vs. Native Tool-Calling: Navigating the Local LLM Implementation Minefield

Prompt-Based vs. Native Tool-Calling: Navigating the Local LLM Implementation Minefield

Comments
1 min read
Meet Kent 2.0 - Your Coding Accomplice

Meet Kent 2.0 - Your Coding Accomplice

Comments
4 min read
Grammarly costs $12/mo — a local LLM does it for free (Chrome + Ollama)

Grammarly costs $12/mo — a local LLM does it for free (Chrome + Ollama)

Comments
6 min read
Serverless GPU Inference: Deploy Any Hugging Face Model on Google Cloud Run

Serverless GPU Inference: Deploy Any Hugging Face Model on Google Cloud Run

Comments 2
4 min read
Most Teams Ask the Wrong Question About RAG vs Fine-Tuning

Most Teams Ask the Wrong Question About RAG vs Fine-Tuning

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.