DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Claude Sonnet 4.6: 1M Context, 300K Output, Agentic Coding

Claude Sonnet 4.6: 1M Context, 300K Output, Agentic Coding

Comments
10 min read
NyayAI: AI-Powered Legal Intelligence for India

NyayAI: AI-Powered Legal Intelligence for India

1
Comments
13 min read
I Built a Local AI Gateway That Talks to Claude, ChatGPT, DeepSeek and Gemini — Without a Single API Key

I Built a Local AI Gateway That Talks to Claude, ChatGPT, DeepSeek and Gemini — Without a Single API Key

Comments
3 min read
How AI Memory Actually Works: Context Windows and RAG

How AI Memory Actually Works: Context Windows and RAG

Comments
8 min read
The Hidden Tax: Why OpenAI Charges Up to 60% More for Spanish Prompts (and How to Fix It)

The Hidden Tax: Why OpenAI Charges Up to 60% More for Spanish Prompts (and How to Fix It)

Comments
7 min read
When AI Reads Blueprints: The Hidden Attack Surface of Multimodal Engineering Intelligence

When AI Reads Blueprints: The Hidden Attack Surface of Multimodal Engineering Intelligence

12
Comments 6
7 min read
Claude Opus 4.7 Hit 87.6% on SWE-bench. The Story Is What It Didn't Charge You.

Claude Opus 4.7 Hit 87.6% on SWE-bench. The Story Is What It Didn't Charge You.

Comments
7 min read
agent-consistency – a Python consistency layer for multi-agent workflows

agent-consistency – a Python consistency layer for multi-agent workflows

Comments
1 min read
Why Does AI Make Things Up? A Dev's Guide to Hallucination

Why Does AI Make Things Up? A Dev's Guide to Hallucination

2
Comments 2
4 min read
Production-Ready MCP Servers in 60 Seconds (Auth, Rate Limits, Audit Logs Included)

Production-Ready MCP Servers in 60 Seconds (Auth, Rate Limits, Audit Logs Included)

2
Comments 1
2 min read
Custom Silicon, Agentic Search, and Smarter Fine-Tuning

Custom Silicon, Agentic Search, and Smarter Fine-Tuning

Comments
2 min read
Why RAG Breaks in Real-World Systems (and How I’m Trying to Fix It)

Why RAG Breaks in Real-World Systems (and How I’m Trying to Fix It)

Comments
2 min read
Should we implicitly trust AI with our optimization problems? Like taxes?

Should we implicitly trust AI with our optimization problems? Like taxes?

Comments 2
10 min read
The open-weight licence trap: Apache 2.0 vs. the community-licence model

The open-weight licence trap: Apache 2.0 vs. the community-licence model

Comments
5 min read
Stop Benchmarking Embedding Models. 90% of Your Search Quality Lives Upstream.

Stop Benchmarking Embedding Models. 90% of Your Search Quality Lives Upstream.

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.