Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Claude Code Now Runs Subagents in the Background by Default — What Actually Changed
Alvarito1983
Alvarito1983
Alvarito1983
Follow
Aug 7
Claude Code Now Runs Subagents in the Background by Default — What Actually Changed
#
agents
#
ai
#
claude
#
llm
1
 reaction
Comments
Add Comment
3 min read
Changing the AI engine moved 3 of 10 results. Changing the question moved 10 of 10.
Lana Plouffe
Lana Plouffe
Lana Plouffe
Follow
Aug 7
Changing the AI engine moved 3 of 10 results. Changing the question moved 10 of 10.
#
ai
#
llm
#
opensource
#
datascience
Comments
Add Comment
7 min read
Qwen3 235B Tops the Agentic Benchmark - What That Test Actually Measures
Basavaraj SH
Basavaraj SH
Basavaraj SH
Follow
Aug 7
Qwen3 235B Tops the Agentic Benchmark - What That Test Actually Measures
#
llm
#
agents
#
langgraph
#
qwen
Comments
Add Comment
2 min read
Buyer context moved 6 of the top 10 results on one market. I ran it again on an unrelated market and it moved 4.
Lana Plouffe
Lana Plouffe
Lana Plouffe
Follow
Aug 7
Buyer context moved 6 of the top 10 results on one market. I ran it again on an unrelated market and it moved 4.
#
ai
#
llm
#
datascience
#
machinelearning
Comments
Add Comment
7 min read
Your Sandbox Has a Hole in It, and the AI Agent Found It
Cor E
Cor E
Cor E
Follow
Aug 7
Your Sandbox Has a Hole in It, and the AI Agent Found It
#
ai
#
security
#
cybersecurity
#
llm
Comments
Add Comment
4 min read
You asked how I know what the model did. Here are the answers, with numbers.
Marc
Marc
Marc
Follow
Aug 6
You asked how I know what the model did. Here are the answers, with numbers.
#
ai
#
llm
#
monitoring
#
performance
Comments
Add Comment
6 min read
Human Oversight of AI Agents Failed 33% of the Time in Testing
Basavaraj SH
Basavaraj SH
Basavaraj SH
Follow
Aug 6
Human Oversight of AI Agents Failed 33% of the Time in Testing
#
agents
#
llm
#
aisafety
#
langchain
Comments
Add Comment
2 min read
Why AI Agents Say “Done” When the Task Actually Failed
Safiyev Marat
Safiyev Marat
Safiyev Marat
Follow
Aug 11
Why AI Agents Say “Done” When the Task Actually Failed
#
ai
#
opensource
#
llm
#
agents
6
 reactions
Comments
Add Comment
2 min read
Can a Mac mini run Kimi K3? I did the memory math
Tyson Cung
Tyson Cung
Tyson Cung
Follow
Aug 11
Can a Mac mini run Kimi K3? I did the memory math
#
kimi
#
kimik3
#
ai
#
llm
3
 reactions
Comments
Add Comment
3 min read
Teaching an Audio Model More About Barbados
Matt Hamilton
Matt Hamilton
Matt Hamilton
Follow
Aug 7
Teaching an Audio Model More About Barbados
#
ai
#
barbados
#
llm
#
machinelearning
Comments
Add Comment
15 min read
Upgrading the judge ends one score series and starts another
Maya Andersson
Maya Andersson
Maya Andersson
Follow
Aug 6
Upgrading the judge ends one score series and starts another
#
llm
#
statistics
#
datascience
#
testing
5
 reactions
Comments
Add Comment
9 min read
# I benchmarked AI agent memory in 2026 — and the numbers tell a different story than the marketing
Everest An
Everest An
Everest An
Follow
Aug 7
# I benchmarked AI agent memory in 2026 — and the numbers tell a different story than the marketing
#
agents
#
ai
#
llm
#
testing
Comments
Add Comment
4 min read
What Advisory Rules Actually Do in an Agent Loop
Lex
Lex
Lex
Follow
Aug 6
What Advisory Rules Actually Do in an Agent Loop
#
ai
#
llm
#
agents
#
python
Comments
Add Comment
7 min read
How Claude Marks AI-Generated Content?
Hassann
Hassann
Hassann
Follow
Aug 11
How Claude Marks AI-Generated Content?
#
ai
#
claude
#
llm
#
security
Comments
1
 comment
9 min read
Stop Building AI Wrappers. Start Building AI Systems
Shourya Kumar
Shourya Kumar
Shourya Kumar
Follow
Aug 6
Stop Building AI Wrappers. Start Building AI Systems
#
ai
#
api
#
llm
#
agents
Comments
Add Comment
2 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account