DEV Community

#agents

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Autonomous Beyond the Chatbox: Why AI Agents Require Dedicated Infrastructure to Work

Autonomous Beyond the Chatbox: Why AI Agents Require Dedicated Infrastructure to Work

Comments
3 min read
Who Checks the Work?

Who Checks the Work?

Comments
1 min read
Agent Leaderboards Mislead Under Distribution Shift (IBM): Predictive Validity

Agent Leaderboards Mislead Under Distribution Shift (IBM): Predictive Validity

Comments
7 min read
Building Multi-Agent Systems: When Your AI Should Spawn More AIs

Building Multi-Agent Systems: When Your AI Should Spawn More AIs

Comments
3 min read
Choosing the Right Model for Each Task in a Multi-Module AI Agent (Hermes Architecture)

Choosing the Right Model for Each Task in a Multi-Module AI Agent (Hermes Architecture)

Comments
10 min read
5 lessons from 5 interviews on AI agents and web scraping

5 lessons from 5 interviews on AI agents and web scraping

Comments 2
5 min read
Built a pay-per-call CVE fix remediation API for agents (x402, no account)

Built a pay-per-call CVE fix remediation API for agents (x402, no account)

1
Comments
1 min read
Building agents that track a hypothesis and alert when the evidence changes

Building agents that track a hypothesis and alert when the evidence changes

1
Comments 1
4 min read
Sakana AI's Fugu Explained: How the Multi-Agent Model Orchestrates Frontier LLMs

Sakana AI's Fugu Explained: How the Multi-Agent Model Orchestrates Frontier LLMs

Comments
7 min read
여러 에이전트가 협업하는 업무 자동화 시스템 설계 방법

여러 에이전트가 협업하는 업무 자동화 시스템 설계 방법

Comments
1 min read
What people get wrong about agent loops

What people get wrong about agent loops

Comments
1 min read
The OpenAI / Hugging Face Incident Was an Observability Failure First

The OpenAI / Hugging Face Incident Was an Observability Failure First

8
Comments 4
6 min read
Claude, Codex, Gemini, Grok: A Field Report on Agentic Memory Write Reliability

Claude, Codex, Gemini, Grok: A Field Report on Agentic Memory Write Reliability

Comments
5 min read
eval-harness - an evaluation framework for agentic cli tools

eval-harness - an evaluation framework for agentic cli tools

1
Comments 1
3 min read
Running Hermes Agent with Kokoro TTS: A Local-First AI Assistant Setup

Running Hermes Agent with Kokoro TTS: A Local-First AI Assistant Setup

6
Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.