DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
OpenEval: Why LLM Evaluation Needs a Standard Format

OpenEval: Why LLM Evaluation Needs a Standard Format

Comments
1 min read
Making Local AI Tool Calls More Reliable

Making Local AI Tool Calls More Reliable

2
Comments 6
1 min read
Show HN: Telechat – Self-hosted Claude on Telegram/WhatsApp/Slack, no cloud relay

Show HN: Telechat – Self-hosted Claude on Telegram/WhatsApp/Slack, no cloud relay

2
Comments
2 min read
Your agent returned 200 OK. Was it actually right?

Your agent returned 200 OK. Was it actually right?

Comments
2 min read
`finish_reason=length` Returned Empty Content — and the Error Message Lied to Me

`finish_reason=length` Returned Empty Content — and the Error Message Lied to Me

1
Comments
3 min read
Model + Harness = Agent: The Gap Isn’t Where You Think

Model + Harness = Agent: The Gap Isn’t Where You Think

Comments
4 min read
I Trust My AI Completely—Except When It Says “Done”

I Trust My AI Completely—Except When It Says “Done”

1
Comments 1
4 min read
Kmemo 2.0 is out, and the two gaps I admitted to in the first post are closed

Kmemo 2.0 is out, and the two gaps I admitted to in the first post are closed

4
Comments
5 min read
I built an AI observability platform with $0 – zero dependencies, zero ops, stateless

I built an AI observability platform with $0 – zero dependencies, zero ops, stateless

2
Comments 1
2 min read
The 22 Failure Modes Are Not 22 Problems. They Are Five.

The 22 Failure Modes Are Not 22 Problems. They Are Five.

1
Comments 1
4 min read
Context compaction is silently destroying your LLM agent's memory

Context compaction is silently destroying your LLM agent's memory

Comments 8
3 min read
Deploying a QAT Checkpoint Your Serving Stack Can't Load: Gemma 4 E2B in Pure JAX on One TPU

Deploying a QAT Checkpoint Your Serving Stack Can't Load: Gemma 4 E2B in Pure JAX on One TPU

8
Comments
10 min read
How to make your Next.js site appear in ChatGPT (and any LLM)

How to make your Next.js site appear in ChatGPT (and any LLM)

Comments
9 min read
The Same GraphRAG Comparison Wins and Loses. It Depends Which Instrument Judged It.

The Same GraphRAG Comparison Wins and Loses. It Depends Which Instrument Judged It.

10
Comments 23
3 min read
MCP moves context. It doesn't create it.

MCP moves context. It doesn't create it.

Comments 1
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.