DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
vLLM GGUF FAQ: Ten Search Questions, Answered

vLLM GGUF FAQ: Ten Search Questions, Answered

5
Comments 2
3 min read
Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

1
Comments 1
17 min read
A Prompt-Injection Firewall for AWS: What I'd Build, and What I Measured

A Prompt-Injection Firewall for AWS: What I'd Build, and What I Measured

Comments
11 min read
Gone for Good? I asked 10 models which deleted data could actually be restored

Kaggle Benchmarking Challenge Submission

Gone for Good? I asked 10 models which deleted data could actually be restored

Comments 1
10 min read
My Second Brain Has Two AIs and Works With My Laptop Shut

My Second Brain Has Two AIs and Works With My Laptop Shut

Comments
6 min read
RAG vs. Fine-Tuning: The Wrong Question to Ask When Building AI Systems

RAG vs. Fine-Tuning: The Wrong Question to Ask When Building AI Systems

Comments
25 min read
The Developer's Guide to AI Tokenomics: Why Your Coding Agent Can Burn Through Credits in Minutes

The Developer's Guide to AI Tokenomics: Why Your Coding Agent Can Burn Through Credits in Minutes

Comments
9 min read
Securing AI Agents: Enforce Boundaries Outside the Model, Not Inside the Prompt

Securing AI Agents: Enforce Boundaries Outside the Model, Not Inside the Prompt

1
Comments 1
4 min read
The cached-prefix crossover: when the cheaper LLM becomes the expensive one

The cached-prefix crossover: when the cheaper LLM becomes the expensive one

Comments
4 min read
Making Qwen Faster on an RTX 3090

Making Qwen Faster on an RTX 3090

1
Comments 1
8 min read
Measure Once, Distribute Everywhere: A Central Telemetry Engine for GenAI

Measure Once, Distribute Everywhere: A Central Telemetry Engine for GenAI

Comments 1
3 min read
The Average Latency Lie: What 100 LLM API Calls Taught Me About Benchmarking

The Average Latency Lie: What 100 LLM API Calls Taught Me About Benchmarking

Comments
5 min read
Private AI SOC Triage Lab: Can a Small Local LLM Triage Alerts Safely?

Private AI SOC Triage Lab: Can a Small Local LLM Triage Alerts Safely?

Comments
3 min read
Your AI Agent Has a Carbon Footprint. Nobody's Measuring It.

Your AI Agent Has a Carbon Footprint. Nobody's Measuring It.

Comments 1
3 min read
How to Build a Competitive Intelligence Agent for Public Companies

How to Build a Competitive Intelligence Agent for Public Companies

6
Comments 1
11 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.