Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Claude Code faked its own work, then wrote me an unprompted confession
Jun
Jun
Jun
Follow
Jul 14
Claude Code faked its own work, then wrote me an unprompted confession
#
ai
#
llm
#
claudecode
#
agents
1
 reaction
Comments
1
 comment
9 min read
What Is an Agent Loop? How AI Agents Reason, Act, and Iterate
Rishabh Poddar
Rishabh Poddar
Rishabh Poddar
Follow
Jun 21
What Is an Agent Loop? How AI Agents Reason, Act, and Iterate
#
agents
#
ai
#
automation
#
llm
1
 reaction
Comments
Add Comment
5 min read
My Home AI's First Reply Took Four Minutes. Now It Takes Eleven Seconds.
Nova
Nova
Nova
Follow
Jul 14
My Home AI's First Reply Took Four Minutes. Now It Takes Eleven Seconds.
#
ai
#
llm
#
selfhosted
#
devops
2
 reactions
Comments
Add Comment
4 min read
From Raw Tickets to Verified Context: An AI-Driven Pipeline for GitHub Issues
lbobylev
lbobylev
lbobylev
Follow
Jul 25
From Raw Tickets to Verified Context: An AI-Driven Pipeline for GitHub Issues
#
ai
#
langchain
#
automation
#
llm
Comments
Add Comment
11 min read
Building Agentic Workflows in Java
Puneet Gupta
Puneet Gupta
Puneet Gupta
Follow
Jul 4
Building Agentic Workflows in Java
#
java
#
ai
#
llm
Comments
Add Comment
7 min read
Building Agentic Workflows in Python
Puneet Gupta
Puneet Gupta
Puneet Gupta
Follow
Jul 4
Building Agentic Workflows in Python
#
python
#
ai
#
llm
Comments
Add Comment
7 min read
Building Reliable LLM Applications in Java
Puneet Gupta
Puneet Gupta
Puneet Gupta
Follow
Jul 4
Building Reliable LLM Applications in Java
#
java
#
ai
#
llm
Comments
Add Comment
5 min read
You Recorded the Incident. Now Prove Your Fix Actually Works.
Tisha
Tisha
Tisha
Follow
Jul 22
You Recorded the Incident. Now Prove Your Fix Actually Works.
#
ai
#
llm
#
testing
#
machinelearning
21
 reactions
Comments
18
 comments
6 min read
I threw 750 autonomous LLM exploit attempts at a $10k sandbox bounty. Zero escapes.
Dipankar Sarkar
Dipankar Sarkar
Dipankar Sarkar
Follow
Jul 13
I threw 750 autonomous LLM exploit attempts at a $10k sandbox bounty. Zero escapes.
#
ai
#
agents
#
opensource
#
llm
2
 reactions
Comments
4
 comments
4 min read
I benchmarked Claude Code skills against a placebo — and half of mine failed
JinHyuk Sung
JinHyuk Sung
JinHyuk Sung
Follow
Jul 24
I benchmarked Claude Code skills against a placebo — and half of mine failed
#
ai
#
claude
#
llm
#
testing
1
 reaction
Comments
4
 comments
4 min read
An append-only audit log caught two accounting bugs in a 216-star usage tracker
Li Zhuojun
Li Zhuojun
Li Zhuojun
Follow
Jul 25
An append-only audit log caught two accounting bugs in a 216-star usage tracker
#
ai
#
llm
#
observability
#
opensource
Comments
Add Comment
6 min read
Building Reliable LLM Applications in Python
Puneet Gupta
Puneet Gupta
Puneet Gupta
Follow
Jul 4
Building Reliable LLM Applications in Python
#
python
#
ai
#
llm
Comments
Add Comment
5 min read
Before you run the AI debate 200 times, measure the die — temperature diversity vs. vendor diversity
Ryosuke Matsuzaki
Ryosuke Matsuzaki
Ryosuke Matsuzaki
Follow
Jul 24
Before you run the AI debate 200 times, measure the die — temperature diversity vs. vendor diversity
#
ai
#
llm
#
agents
#
softwareengineering
1
 reaction
Comments
2
 comments
5 min read
The Ruler Made of Itself
Nick Meinhold
Nick Meinhold
Nick Meinhold
Follow
Jul 4
The Ruler Made of Itself
#
llm
#
ai
#
machinelearning
#
evaluation
Comments
Add Comment
9 min read
I built a small library so my LLM agent stops double-charging people
vsangaraju
vsangaraju
vsangaraju
Follow
Jul 24
I built a small library so my LLM agent stops double-charging people
#
python
#
ai
#
opensource
#
llm
Comments
2
 comments
7 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account