DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Evaluate LLM code generation with LLM-as-judge evaluators

Evaluate LLM code generation with LLM-as-judge evaluators

7
Comments
12 min read
vLLM On-Demand Gateway: Zero-VRAM Standby for Local LLMs on Consumer GPUs

vLLM On-Demand Gateway: Zero-VRAM Standby for Local LLMs on Consumer GPUs

2
Comments 1
4 min read
Query Live AI Inference Pricing with the ATOM MCP Server

Query Live AI Inference Pricing with the ATOM MCP Server

2
Comments
3 min read
Add DeFi Superpowers to Claude with MCP Skills

Add DeFi Superpowers to Claude with MCP Skills

1
Comments
4 min read
Empirically Testing Skill Scanners Against Traditional Obfuscation

Empirically Testing Skill Scanners Against Traditional Obfuscation

Comments
8 min read
45 MCP Tools: Everything Your Claude Agent Can Do with a Wallet

45 MCP Tools: Everything Your Claude Agent Can Do with a Wallet

Comments
4 min read
Hybrid RAG System over SEC Filings

Hybrid RAG System over SEC Filings

Comments
19 min read
AI News Roundup: Claude Code Security, ggml.ai + Hugging Face, and 17K tok/s Silicon Llama

AI News Roundup: Claude Code Security, ggml.ai + Hugging Face, and 17K tok/s Silicon Llama

Comments
3 min read
MCP + Wallet: When AI Agents Can Actually Pay

MCP + Wallet: When AI Agents Can Actually Pay

Comments
4 min read
You Don’t Need a Bigger Model — You Need a Stable One

You Don’t Need a Bigger Model — You Need a Stable One

Comments
3 min read
When Code Becomes Cheap, Thinking Becomes Expensive

When Code Becomes Cheap, Thinking Becomes Expensive

1
Comments
4 min read
I Added Langfuse to My RAG App and It Immediately Caught Two Bugs

I Added Langfuse to My RAG App and It Immediately Caught Two Bugs

Comments
7 min read
When CLAUDE.md Stops Working: Adding Vector Memory to Claude Code

When CLAUDE.md Stops Working: Adding Vector Memory to Claude Code

1
Comments
10 min read
From expensive tokens to intelligent compression: how we optimize LLM costs in production

From expensive tokens to intelligent compression: how we optimize LLM costs in production

Comments
4 min read
Beyond SEO: Generative Engine Optimization (GEO). How to Implement `llms.txt` and RAG-Friendly Markup

Beyond SEO: Generative Engine Optimization (GEO). How to Implement `llms.txt` and RAG-Friendly Markup

1
Comments 2
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.