DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
How RAGScope Knows Which Chunks Your LLM Actually Used

How RAGScope Knows Which Chunks Your LLM Actually Used

Comments 2
4 min read
Is Your Agent Skill Actually Good? Microsoft's Dual-Paper Deep Dive into Skill Evaluation and Self-Evolving Optimization

Is Your Agent Skill Actually Good? Microsoft's Dual-Paper Deep Dive into Skill Evaluation and Self-Evolving Optimization

Comments 1
14 min read
I Built a Production-Oriented Multi-Provider AI Chatbot in Rust — Here's How

I Built a Production-Oriented Multi-Provider AI Chatbot in Rust — Here's How

3
Comments 1
5 min read
Progressive Distillation

Progressive Distillation

2
Comments
4 min read
Llama-Server Router Mode - Dynamic Model Switching Without Restarts

Llama-Server Router Mode - Dynamic Model Switching Without Restarts

Comments
9 min read
AI for Knowledge Management: Real Workflows That Hold Up

AI for Knowledge Management: Real Workflows That Hold Up

Comments
8 min read
I Built a GPU Dataset for LLM Inference — Here’s What I Learned

I Built a GPU Dataset for LLM Inference — Here’s What I Learned

1
Comments
2 min read
Building Your First Real GPT Is Not a Prompting Exercise

Building Your First Real GPT Is Not a Prompting Exercise

1
Comments 11
6 min read
Your LLM Bill Is Too High. Here's How to Fix It (Part 1)

Your LLM Bill Is Too High. Here's How to Fix It (Part 1)

Comments
3 min read
I built an LLM eval rig in a weekend. Most of it was wrong.

I built an LLM eval rig in a weekend. Most of it was wrong.

1
Comments
4 min read
How to Detect Prompt Injection in Your LLM Agent — Python, 5 Minutes

How to Detect Prompt Injection in Your LLM Agent — Python, 5 Minutes

Comments
5 min read
Is Gemini 3.5 Flash Actually Better at Coding Than 3.1 Pro? I Tested It with Real Rust Code

Is Gemini 3.5 Flash Actually Better at Coding Than 3.1 Pro? I Tested It with Real Rust Code

Comments 6
5 min read
Deep Dive into Open Agent SDK (Part 6): Multi-LLM Providers and Runtime Controls

Deep Dive into Open Agent SDK (Part 6): Multi-LLM Providers and Runtime Controls

Comments
13 min read
ModelChain: Measurable LLM Router with Adaptive Model Selection, Real-Time Scoring, Budget Guards and Failover for Node.js, Edge and Browser

ModelChain: Measurable LLM Router with Adaptive Model Selection, Real-Time Scoring, Budget Guards and Failover for Node.js, Edge and Browser

1
Comments 1
3 min read
Skills for eval-driven agent optimization

Skills for eval-driven agent optimization

1
Comments
1 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.