DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The AI writes the answer

The AI writes the answer

Comments
3 min read
An AI Agent Recommended a Malware Package. Here's the Failure Mode Behind It

An AI Agent Recommended a Malware Package. Here's the Failure Mode Behind It

Comments
4 min read
Note: Common Claude Architecture Challenges and Solutions

Note: Common Claude Architecture Challenges and Solutions

Comments
2 min read
I Tested GLM-5.3-Flash and Qwen3.8-Flash on 24 Real Tasks

I Tested GLM-5.3-Flash and Qwen3.8-Flash on 24 Real Tasks

Comments
6 min read
The Working Set That Never Saturated

The Working Set That Never Saturated

Comments 2
6 min read
A pre-flight checklist for shipping a Claude connector

A pre-flight checklist for shipping a Claude connector

Comments
5 min read
Indirect prompt injection in a RAG pipeline: one attack, step by step, and what actually stopped it

Indirect prompt injection in a RAG pipeline: one attack, step by step, and what actually stopped it

2
Comments 4
5 min read
When you need an Agent Gateway, not just another LLM proxy

When you need an Agent Gateway, not just another LLM proxy

Comments 1
2 min read
We benchmarked AI memory systems on LongMemEval — here is the honest data

We benchmarked AI memory systems on LongMemEval — here is the honest data

Comments
2 min read
One API Key, Multiple LLM Providers: Comparing Gateways for Python Text Classification

One API Key, Multiple LLM Providers: Comparing Gateways for Python Text Classification

Comments
5 min read
Designing a Multi-Model Answer Grid Without Hiding Uncertainty

Designing a Multi-Model Answer Grid Without Hiding Uncertainty

2
Comments
3 min read
Hermes Agent's Self-Improving Loop: What Recursive Learning Reveals About Agent Harness Architecture

Hermes Agent's Self-Improving Loop: What Recursive Learning Reveals About Agent Harness Architecture

Comments 1
5 min read
AuthLM: Why Your LLM Router Shouldn't Also Be Your Credential Manager

AuthLM: Why Your LLM Router Shouldn't Also Be Your Credential Manager

Comments
3 min read
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.

Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.

1
Comments 1
3 min read
From Guesswork to Evidence: How BotTalk Built a Systematic Approach to AI Model Selection

From Guesswork to Evidence: How BotTalk Built a Systematic Approach to AI Model Selection

1
Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.