Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Achieving Maximum Throughput on vLLM with a Single RTX 3090: A Production Guide for 7B LLMs
ever9998
ever9998
ever9998
Follow
Apr 29
Achieving Maximum Throughput on vLLM with a Single RTX 3090: A Production Guide for 7B LLMs
#
llm
#
machinelearning
#
performance
#
tutorial
2
 reactions
Comments
1
 comment
4 min read
Seu agente de IA está desperdiçando 13.000 tokens antes de dizer "oi"
Rudson Kiyoshi Souza Carvalho
Rudson Kiyoshi Souza Carvalho
Rudson Kiyoshi Souza Carvalho
Follow
Apr 29
Seu agente de IA está desperdiçando 13.000 tokens antes de dizer "oi"
#
ai
#
agents
#
mcp
#
llm
Comments
Add Comment
4 min read
Why AI Hallucinates Even When It Knows the Answer
Sumanth Vallabhaneni
Sumanth Vallabhaneni
Sumanth Vallabhaneni
Follow
Mar 26
Why AI Hallucinates Even When It Knows the Answer
#
ai
#
deeplearning
#
llm
#
nlp
1
 reaction
Comments
Add Comment
5 min read
The Production Agent Checklist: What Every AI Agent Needs Before It Touches Real Users
Devon
Devon
Devon
Follow
Mar 26
The Production Agent Checklist: What Every AI Agent Needs Before It Touches Real Users
#
python
#
ai
#
agents
#
llm
Comments
Add Comment
9 min read
How I Built a Hallucination Detector for RAG Pipelines in Python
Devasish Banerjee
Devasish Banerjee
Devasish Banerjee
Follow
Mar 26
How I Built a Hallucination Detector for RAG Pipelines in Python
#
python
#
rag
#
llm
#
machinelearning
Comments
1
 comment
3 min read
Stop Hardcoding Model Fallbacks: Let Production Data Pick Your Paths
Devon
Devon
Devon
Follow
Mar 26
Stop Hardcoding Model Fallbacks: Let Production Data Pick Your Paths
#
python
#
ai
#
agents
#
llm
Comments
Add Comment
8 min read
SEO Is Dead? No. But the Game Changed.
Dmitry (Dee) Kargaev
Dmitry (Dee) Kargaev
Dmitry (Dee) Kargaev
Follow
Mar 27
SEO Is Dead? No. But the Game Changed.
#
ai
#
chatgpt
#
llm
#
marketing
Comments
Add Comment
11 min read
The AI Engineer's Toolkit: Moving Beyond Prompt Engineering to Build Robust AI Applications
Midas126
Midas126
Midas126
Follow
Mar 26
The AI Engineer's Toolkit: Moving Beyond Prompt Engineering to Build Robust AI Applications
#
ai
#
machinelearning
#
softwareengineering
#
llm
1
 reaction
Comments
Add Comment
5 min read
AI Agents in Production Are Flying Blind — AgentLens Fixes That
Farzan Hossan Shaikat
Farzan Hossan Shaikat
Farzan Hossan Shaikat
Follow
Apr 29
AI Agents in Production Are Flying Blind — AgentLens Fixes That
#
agents
#
ai
#
llm
#
monitoring
Comments
1
 comment
2 min read
Your AI agent wastes 13,000 tokens before saying "hello"
Rudson Kiyoshi Souza Carvalho
Rudson Kiyoshi Souza Carvalho
Rudson Kiyoshi Souza Carvalho
Follow
Apr 29
Your AI agent wastes 13,000 tokens before saying "hello"
#
ai
#
agents
#
mcp
#
llm
1
 reaction
Comments
Add Comment
4 min read
Your AI Agent Can Be Socially Engineered. Here Are 3 Attacks That Prove It.
Dishanth
Dishanth
Dishanth
Follow
Apr 28
Your AI Agent Can Be Socially Engineered. Here Are 3 Attacks That Prove It.
#
ai
#
cybersecurity
#
llm
#
security
4
 reactions
Comments
Add Comment
4 min read
Building a Context-Aware AI Chat Without a Vector Database
Ryan Carter
Ryan Carter
Ryan Carter
Follow
Apr 28
Building a Context-Aware AI Chat Without a Vector Database
#
ai
#
llm
#
webdev
#
tutorial
Comments
Add Comment
6 min read
SimCore: I built a social simulation engine where LLM agents live on a real map of your city
Elison Frankowski
Elison Frankowski
Elison Frankowski
Follow
Apr 28
SimCore: I built a social simulation engine where LLM agents live on a real map of your city
#
opensource
#
llm
#
python
#
simulation
Comments
Add Comment
1 min read
Multi-Model LLM Orchestration with OpenRouter
Ryan Carter
Ryan Carter
Ryan Carter
Follow
Apr 28
Multi-Model LLM Orchestration with OpenRouter
#
ai
#
llm
#
webdev
#
tutorial
Comments
Add Comment
6 min read
I Tried Speculative Decoding on RTX 4060 8GB — Every Config Was Slower Than Baseline
plasmon
plasmon
plasmon
Follow
Mar 25
I Tried Speculative Decoding on RTX 4060 8GB — Every Config Was Slower Than Baseline
#
llm
#
gpu
#
benchmark
#
ai
1
 reaction
Comments
Add Comment
8 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account