DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Setting Up a Local AI Coding Agent with Ollama and Aider (part 3)

Setting Up a Local AI Coding Agent with Ollama and Aider (part 3)

Comments
5 min read
When Should You /clear? A 1913 Inventory Formula Has the Answer

When Should You /clear? A 1913 Inventory Formula Has the Answer

3
Comments 1
9 min read
Robust-GAP: Achieving Zero-Hallucination Causal Summarization in Hierarchical RAG

Robust-GAP: Achieving Zero-Hallucination Causal Summarization in Hierarchical RAG

7
Comments
6 min read
A key-value store where the query language is Lua, and you can build RAG inside it

A key-value store where the query language is Lua, and you can build RAG inside it

1
Comments 1
5 min read
My AI System Logged 35,669 LLM Calls. It Still Couldn’t Tell Me What They Cost.

My AI System Logged 35,669 LLM Calls. It Still Couldn’t Tell Me What They Cost.

Comments
7 min read
I almost burned ₹4,000 on Claude API overnight — so I built llm-cost-guard

I almost burned ₹4,000 on Claude API overnight — so I built llm-cost-guard

Comments
3 min read
MCP na prática: Tools, Resources e quando usar cada um

MCP na prática: Tools, Resources e quando usar cada um

1
Comments
9 min read
Moonshot AI's Kimi K3 Is Here: A 2.8 Trillion Parameter Open MoE Model That Pushes Long-Context AI Forward

Moonshot AI's Kimi K3 Is Here: A 2.8 Trillion Parameter Open MoE Model That Pushes Long-Context AI Forward

1
Comments
3 min read
A required field made my AI fabricate statistics

A required field made my AI fabricate statistics

Comments 2
5 min read
Top 10 Open Source & Open-Weight AI Models in July 2026: Capabilities, Architecture, and Estimated Training Costs

Top 10 Open Source & Open-Weight AI Models in July 2026: Capabilities, Architecture, and Estimated Training Costs

Comments
19 min read
Getting an LLM to Actually Follow Your Output Format (Without Fighting It Every Request)

Getting an LLM to Actually Follow Your Output Format (Without Fighting It Every Request)

2
Comments 2
3 min read
Self-Hosting Your First LLM for Enterprise: What Nobody Tells You Before You Start

Self-Hosting Your First LLM for Enterprise: What Nobody Tells You Before You Start

2
Comments
3 min read
Rate limiting, email alerts, health checks, and Grafana — what we shipped to make Ajah production-ready

Rate limiting, email alerts, health checks, and Grafana — what we shipped to make Ajah production-ready

Comments
2 min read
Stop Wasting Your LLM Context Window: A Practical Strategy

Stop Wasting Your LLM Context Window: A Practical Strategy

Comments
3 min read
Stop Parsing LLM Junk: Zero-Latency JSON with Claude Prefill, Spring AI, and Java 26 Records

Stop Parsing LLM Junk: Zero-Latency JSON with Claude Prefill, Spring AI, and Java 26 Records

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.