Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
The Circuit Breaker Pattern for AI Agents
Brenn Hill
Brenn Hill
Brenn Hill
Follow
Aug 6
The Circuit Breaker Pattern for AI Agents
#
ai
#
machinelearning
#
llm
#
programming
7
reactions
Comments
2
comments
9 min read
How I reduced LLM context cost by 35% without changing code (Token Firewall)
miguel
miguel
miguel
Follow
Aug 6
How I reduced LLM context cost by 35% without changing code (Token Firewall)
#
ai
#
goolang
#
llm
#
mcp
1
reaction
Comments
1
comment
2 min read
Tokenization in AI: What Is It, Why Is It Used, and How Does It Work?
Prince Singh
Prince Singh
Prince Singh
Follow
Aug 2
Tokenization in AI: What Is It, Why Is It Used, and How Does It Work?
#
ai
#
llm
#
nlp
Comments
Add Comment
6 min read
LLM中如果一个问题容易验证 那么AI就容易学会解决!说说这个特性与P与NP问题的关联性
cognitalk
cognitalk
cognitalk
Follow
Aug 2
LLM中如果一个问题容易验证 那么AI就容易学会解决!说说这个特性与P与NP问题的关联性
#
ai
#
computerscience
#
llm
#
machinelearning
Comments
Add Comment
1 min read
The Open-Weight Inflection Point: Kimi K3, Claude Opus 5, and Microsoft MAI Signal a Market Shift
武乐丹
武乐丹
武乐丹
Follow
Aug 2
The Open-Weight Inflection Point: Kimi K3, Claude Opus 5, and Microsoft MAI Signal a Market Shift
#
ai
#
opensource
#
llm
#
technology
Comments
Add Comment
3 min read
I built a Laravel package for LLM workflows with hard token and cost limits
Anton Parmuzin
Anton Parmuzin
Anton Parmuzin
Follow
Aug 15
I built a Laravel package for LLM workflows with hard token and cost limits
#
showdev
#
laravel
#
llm
#
opensource
Comments
Add Comment
2 min read
Micro-compaction: amortizing context compression in agent loops
Michael Jordan
Michael Jordan
Michael Jordan
Follow
Aug 5
Micro-compaction: amortizing context compression in agent loops
#
llm
#
agents
#
ai
#
opensource
1
reaction
Comments
3
comments
5 min read
Kimi K3 is the largest open-weight model ever released — and you probably still can't run it
Alvarito1983
Alvarito1983
Alvarito1983
Follow
Aug 6
Kimi K3 is the largest open-weight model ever released — and you probably still can't run it
#
ai
#
llm
#
opensource
#
machinelearning
7
reactions
Comments
Add Comment
2 min read
Benchmarking GPT-4o, Claude 3.5 Sonnet, and Llama 3 for Automated Code Auditing & Vulnerability Detection
Saranyo Deyasi
Saranyo Deyasi
Saranyo Deyasi
Follow
Aug 1
Benchmarking GPT-4o, Claude 3.5 Sonnet, and Llama 3 for Automated Code Auditing & Vulnerability Detection
#
ai
#
cybersecurity
#
llm
#
vulnerabilities
Comments
Add Comment
2 min read
Same DeepSeek V4 Flash, Different Agent: Why the Runtime Changes the Result
OctoLab
OctoLab
OctoLab
Follow
Aug 1
Same DeepSeek V4 Flash, Different Agent: Why the Runtime Changes the Result
#
ai
#
agents
#
llm
#
programming
Comments
Add Comment
2 min read
How I Use DeepSeek V4 Flash: Reserve the Strongest Model for Uncertainty
OctoLab
OctoLab
OctoLab
Follow
Aug 1
How I Use DeepSeek V4 Flash: Reserve the Strongest Model for Uncertainty
#
ai
#
agentskills
#
programming
#
llm
Comments
Add Comment
2 min read
Building a podcast summarizer in 20 lines of Python
roberttomko
roberttomko
roberttomko
Follow
Aug 1
Building a podcast summarizer in 20 lines of Python
#
python
#
llm
#
rag
#
webdev
Comments
Add Comment
3 min read
Building Sluice: QoS-Aware Capacity Governance for Self-Hosted LLM Inference
Madhav M S
Madhav M S
Madhav M S
Follow
Aug 14
Building Sluice: QoS-Aware Capacity Governance for Self-Hosted LLM Inference
#
distributedsystems
#
llm
#
kubernetes
#
systemdesign
1
reaction
Comments
1
comment
19 min read
Stop Your AI Coding CLI From Wasting Tokens on "Hi" and "Thanks"
NaveenKumar Namachivayam ⚡
NaveenKumar Namachivayam ⚡
NaveenKumar Namachivayam ⚡
Follow
Aug 5
Stop Your AI Coding CLI From Wasting Tokens on "Hi" and "Thanks"
#
ai
#
beginners
#
cli
#
llm
5
reactions
Comments
5
comments
6 min read
Ingest-Time Compilation Takes On Query-Time RAG, and Agentic Retrieval Meets Its Limits
Felipe 0liveira
Felipe 0liveira
Felipe 0liveira
Follow
Aug 24
Ingest-Time Compilation Takes On Query-Time RAG, and Agentic Retrieval Meets Its Limits
#
ai
#
llm
#
machinelearning
#
rag
1
reaction
Comments
Add Comment
6 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account