Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Token-level eval harness for tool-calling agents: what we wired up
Marcus Chen
Marcus Chen
Marcus Chen
Follow
May 26
Token-level eval harness for tool-calling agents: what we wired up
#
machinelearning
#
llm
#
mlops
#
devops
Comments
1
 comment
4 min read
Agent as a Tool Call: Claude Code's Fork-Exec Pattern
eyesofish
eyesofish
eyesofish
Follow
May 26
Agent as a Tool Call: Claude Code's Fork-Exec Pattern
#
agents
#
architecture
#
claude
#
llm
2
 reactions
Comments
2
 comments
2 min read
I built a free LLM pricing tool that updates itself daily. here's how
CloudyBot
CloudyBot
CloudyBot
Follow
May 15
I built a free LLM pricing tool that updates itself daily. here's how
#
showdev
#
ai
#
llm
#
automation
3
 reactions
Comments
Add Comment
3 min read
LLM Cost Optimization for Agent Workflows: A Practical Guide
Omnithium
Omnithium
Omnithium
Follow
May 26
LLM Cost Optimization for Agent Workflows: A Practical Guide
#
costoptimization
#
llm
#
aiagents
#
engineering
Comments
2
 comments
13 min read
Evaluating Open-Weight LLMs for Phishing Simulation and Red Teaming
Jeff J. Bowie
Jeff J. Bowie
Jeff J. Bowie
Follow
Apr 22
Evaluating Open-Weight LLMs for Phishing Simulation and Red Teaming
#
ai
#
cybersecurity
#
llm
#
opensource
Comments
Add Comment
3 min read
Per-customer budget caps on our caption pipeline: 3 weeks with virtual keys
Elise Moreau
Elise Moreau
Elise Moreau
Follow
May 26
Per-customer budget caps on our caption pipeline: 3 weeks with virtual keys
#
llm
#
infrastructure
#
mlops
#
machinelearning
Comments
1
 comment
4 min read
Serving a Fleet of SLMs on One RTX 5080: Multi-Model on a Single Consumer GPU
Dharamendra Kumar
Dharamendra Kumar
Dharamendra Kumar
Follow
May 26
Serving a Fleet of SLMs on One RTX 5080: Multi-Model on a Single Consumer GPU
#
machinelearning
#
ai
#
llm
#
performance
1
 reaction
Comments
1
 comment
4 min read
I'm writing this down before I lose the thread
Gad Ofir
Gad Ofir
Gad Ofir
Follow
Apr 22
I'm writing this down before I lose the thread
#
agents
#
ai
#
architecture
#
llm
Comments
Add Comment
7 min read
Your LLM Forgets Everything. Give It a Wiki!
Artem M
Artem M
Artem M
Follow
May 26
Your LLM Forgets Everything. Give It a Wiki!
#
ai
#
productivity
#
opensource
#
llm
Comments
2
 comments
4 min read
CKP LLM: The Missing Layer Between Your AI Agent and Its Knowledge Base
Alessandro Marocchini
Alessandro Marocchini
Alessandro Marocchini
Follow
May 26
CKP LLM: The Missing Layer Between Your AI Agent and Its Knowledge Base
#
ai
#
llm
#
productivity
#
devtools
Comments
2
 comments
5 min read
Voice agent latency is a lie. The number you care about is barge-in interrupt rate.
Marcus Chen
Marcus Chen
Marcus Chen
Follow
May 26
Voice agent latency is a lie. The number you care about is barge-in interrupt rate.
#
ai
#
rust
#
performance
#
llm
1
 reaction
Comments
2
 comments
4 min read
The Pomodoro Timer Isn’t About Time, It’s About Engineering
YASHWANTH REDDY K
YASHWANTH REDDY K
YASHWANTH REDDY K
Follow
Apr 22
The Pomodoro Timer Isn’t About Time, It’s About Engineering
#
ai
#
vibecidearena
#
hackerearth
#
llm
4
 reactions
Comments
Add Comment
5 min read
Why Signatures Make Automatic Optimization Easier Than Writing Prompts Directly
Luhui Dev
Luhui Dev
Luhui Dev
Follow
Apr 22
Why Signatures Make Automatic Optimization Easier Than Writing Prompts Directly
#
ai
#
llm
#
promptengineering
#
softwareengineering
Comments
Add Comment
7 min read
Running a 70B LLM on Pure RISC-V: The MilkV Pioneer Deployment Journey
Bruno Verachten
Bruno Verachten
Bruno Verachten
Follow
Apr 22
Running a 70B LLM on Pure RISC-V: The MilkV Pioneer Deployment Journey
#
cpuinference
#
deepseekr1
#
llamacpp
#
llm
Comments
Add Comment
17 min read
Self-healing LLM routing: 13 providers, one fallback chain
Shiva
Shiva
Shiva
Follow
Apr 22
Self-healing LLM routing: 13 providers, one fallback chain
#
showdev
#
ai
#
architecture
#
llm
Comments
Add Comment
4 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account