DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Claude's invisible text watermarks: what practitioners need to know right now

Claude's invisible text watermarks: what practitioners need to know right now

Comments
3 min read
How LangChain extracted value out of their own agents

How LangChain extracted value out of their own agents

Comments
1 min read
Would You Let an AI Agent Move Your Money?

Would You Let an AI Agent Move Your Money?

Comments
5 min read
About Best in IT: Practical AI, Automation and Developer Tools

About Best in IT: Practical AI, Automation and Developer Tools

1
Comments
3 min read
Node.js API Key Text Classification: JSON Validation Before Multi-Provider Gateway Failover

Node.js API Key Text Classification: JSON Validation Before Multi-Provider Gateway Failover

Comments
6 min read
Adding Responses API to an Agent Framework

Adding Responses API to an Agent Framework

Comments
6 min read
A judge that agrees with your humans 92 percent of the time can be at 60 percent where the gate actually decides

A judge that agrees with your humans 92 percent of the time can be at 60 percent where the gate actually decides

Comments
8 min read
The Model Reading My Benchmark Mattered More Than the Memory System Did

The Model Reading My Benchmark Mattered More Than the Memory System Did

Comments
7 min read
Three layers of automated fact-checking for an LLM newsroom (and the bugs that forced each one)

Three layers of automated fact-checking for an LLM newsroom (and the bugs that forced each one)

Comments 1
3 min read
Guard Implementation Patterns to Stop AI Agent Runaway Behavior — 7 Types Extracted from Real-World Logs

Guard Implementation Patterns to Stop AI Agent Runaway Behavior — 7 Types Extracted from Real-World Logs

Comments
5 min read
A Better FP4 Gradient Quantizer That Training Couldn't Notice

A Better FP4 Gradient Quantizer That Training Couldn't Notice

Comments
7 min read
AgentCore Evaluations: How AWS Built a Framework-Agnostic Eval Layer Using OpenTelemetry as the Contract

AgentCore Evaluations: How AWS Built a Framework-Agnostic Eval Layer Using OpenTelemetry as the Contract

1
Comments
6 min read
Claude Opus 5 Is Too Good to Stay in a Browser—Put It in Discord, Slack, Telegram & LINE

Claude Opus 5 Is Too Good to Stay in a Browser—Put It in Discord, Slack, Telegram & LINE

1
Comments
3 min read
Needle 2: the 14 MB agentic model, tested properly

Needle 2: the 14 MB agentic model, tested properly

1
Comments
1 min read
Computer use leaves beta, request shape changes — the weekly AI engineering brief

Computer use leaves beta, request shape changes — the weekly AI engineering brief

Comments
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.