DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Zero-Idle Local LLMs: Running Llama 3 in AWS Lambda Containers

Zero-Idle Local LLMs: Running Llama 3 in AWS Lambda Containers

9
Comments 4
4 min read
Qwen 3 vs Llama 3: Configuring Local LLMs for Actual Performance

Qwen 3 vs Llama 3: Configuring Local LLMs for Actual Performance

Comments
5 min read
Diffusion Language Models: How NVIDIA Nemotron-Labs Diffusion Shatters the Autoregressive Speed Ceiling

Diffusion Language Models: How NVIDIA Nemotron-Labs Diffusion Shatters the Autoregressive Speed Ceiling

Comments
18 min read
Built a Predictive Incident Response Agent with LLMs and Vector Memory

Built a Predictive Incident Response Agent with LLMs and Vector Memory

Comments
6 min read
Aria: Building an AI Customer Support Agent with Persistent Memory

Aria: Building an AI Customer Support Agent with Persistent Memory

Comments
8 min read
Fixing blind spots in code reviews with Hindsight memory

Fixing blind spots in code reviews with Hindsight memory

Comments
2 min read
MCP Tool Poisoning: The AI Supply Chain Attack Nobody Is Talking About

MCP Tool Poisoning: The AI Supply Chain Attack Nobody Is Talking About

Comments 2
4 min read
Best AI Agent Security & Guardrails Tools in 2026: LLM Guard vs NeMo vs Guardrails AI

Best AI Agent Security & Guardrails Tools in 2026: LLM Guard vs NeMo vs Guardrails AI

1
Comments 1
3 min read
The 7-Layer Memory Architecture Behind Modern AI Agents

The 7-Layer Memory Architecture Behind Modern AI Agents

1
Comments
7 min read
BeeLlama v0.2.0: 164 tok/s on a 27B model, one RTX 3090

BeeLlama v0.2.0: 164 tok/s on a 27B model, one RTX 3090

Comments
3 min read
Anna's Archive llms.txt: a routing guide for LLM crawlers

Anna's Archive llms.txt: a routing guide for LLM crawlers

Comments
3 min read
Semantic caching thresholds and why they matter

Semantic caching thresholds and why they matter

Comments
12 min read
Five habits that separate the operator from the vibe-coder

Five habits that separate the operator from the vibe-coder

Comments
6 min read
Gate Zero: stop unfalsifiable prompts before they canonicalize as specs

Gate Zero: stop unfalsifiable prompts before they canonicalize as specs

Comments
5 min read
Eval-driven development for a local-LLM agent: how I shipped Lore 0.2.0 with confidence

Eval-driven development for a local-LLM agent: how I shipped Lore 0.2.0 with confidence

1
Comments
6 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.