DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I Added a Live Dashboard to My LLM Proxy. Zero Instrumentation. Just a URL Change.

I Added a Live Dashboard to My LLM Proxy. Zero Instrumentation. Just a URL Change.

Comments 2
3 min read
I built a Cyber-Cipher Text Decoder on Vibe Code Arena

I built a Cyber-Cipher Text Decoder on Vibe Code Arena

Comments
4 min read
I gave my AI agent a 2MB PDF. Here's what happened to my token count.

I gave my AI agent a 2MB PDF. Here's what happened to my token count.

1
Comments 1
5 min read
I got hit with a surprise AI bill, so I built TokenBar

I got hit with a surprise AI bill, so I built TokenBar

Comments
4 min read
Shipping FreeSay to 2GB-RAM Phones: What We Cut, What We Kept

Shipping FreeSay to 2GB-RAM Phones: What We Cut, What We Kept

Comments
2 min read
Coding Agents Are Breaking Containment

Coding Agents Are Breaking Containment

Comments
3 min read
OpenAI Agents SDK Tutorial: Build Multi-Agent AI Systems in Python (2025)

OpenAI Agents SDK Tutorial: Build Multi-Agent AI Systems in Python (2025)

1
Comments
9 min read
Improving AI with Situated Awareness and Practice Theory

Improving AI with Situated Awareness and Practice Theory

Comments 2
4 min read
I Built a Tool to Stop Guessing LLM API Costs. Here Is What I Learned.

I Built a Tool to Stop Guessing LLM API Costs. Here Is What I Learned.

1
Comments 3
3 min read
Morning notes: what I check before swapping LLM providers

Morning notes: what I check before swapping LLM providers

7
Comments
1 min read
I Reverse Engineered Claude's UI Widget — And It Changed How I Think About Building LLM Apps

I Reverse Engineered Claude's UI Widget — And It Changed How I Think About Building LLM Apps

Comments
4 min read
Why DDR5 Bandwidth Kills Dual-LLM Inference on APUs (Benchmarks Inside)

Why DDR5 Bandwidth Kills Dual-LLM Inference on APUs (Benchmarks Inside)

Comments
7 min read
We built a P2P AI inference network that runs on Android phones — here's what we learned

We built a P2P AI inference network that runs on Android phones — here's what we learned

Comments 1
4 min read
Null Bytes, Dead Streams, Last Chunk

Null Bytes, Dead Streams, Last Chunk

Comments
3 min read
2026 AI Engineer Interview Guide: RAG, LLMs, and Vector Databases

2026 AI Engineer Interview Guide: RAG, LLMs, and Vector Databases

3
Comments
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.