DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Token Efficiency: 16 Algorithms, 5 Languages, Zero Guesswork

Token Efficiency: 16 Algorithms, 5 Languages, Zero Guesswork

Comments 1
4 min read
Building Persistent Memory for AI Agents: A 4-Layer File-Based Architecture

Building Persistent Memory for AI Agents: A 4-Layer File-Based Architecture

Comments
2 min read
OptiLeak: Efficient Prompt Reconstruction via Reinforcement Learning in Multi-tenant LLM Services

OptiLeak: Efficient Prompt Reconstruction via Reinforcement Learning in Multi-tenant LLM Services

1
Comments
4 min read
Mastering AI Agent Memory: Architecture for Power Users in 2024

Mastering AI Agent Memory: Architecture for Power Users in 2024

Comments 1
2 min read
One Billion Dollars Cannot Buy AI Visibility

One Billion Dollars Cannot Buy AI Visibility

1
Comments 1
3 min read
Build a Production-Ready RAG System Over Your Own Documents in 2026 – A Practical Tutorial

Build a Production-Ready RAG System Over Your Own Documents in 2026 – A Practical Tutorial

Comments 2
3 min read
AI's Infrastructure & Agents: From Chips to Code Automation

AI's Infrastructure & Agents: From Chips to Code Automation

Comments
4 min read
SentinelLM - A Proxy Middleware for Safer, Observable LLM Systems

SentinelLM - A Proxy Middleware for Safer, Observable LLM Systems

Comments
2 min read
Why AI-Generated Code is a Security Minefield (And What To Do About It)

Why AI-Generated Code is a Security Minefield (And What To Do About It)

Comments
4 min read
LLM Observability for Laravel - trace every AI call with Langfuse

LLM Observability for Laravel - trace every AI call with Langfuse

Comments 1
4 min read
The `prefer` parameter: routing AI requests by intent instead of model name

The `prefer` parameter: routing AI requests by intent instead of model name

1
Comments
3 min read
# Reading YC-Backed Code #1: claude-mem — Great Idea, Poor Implementation

# Reading YC-Backed Code #1: claude-mem — Great Idea, Poor Implementation

1
Comments
9 min read
16 GB VRAM LLM benchmarks with llama.cpp (speed and context)

16 GB VRAM LLM benchmarks with llama.cpp (speed and context)

Comments
4 min read
LangChain vs CrewAI vs AutoGen vs Dify: The Complete AI Agent Framework Comparison [2026]

LangChain vs CrewAI vs AutoGen vs Dify: The Complete AI Agent Framework Comparison [2026]

1
Comments
4 min read
MCP Gateways for Claude: Choosing the Right Infrastructure Layer

MCP Gateways for Claude: Choosing the Right Infrastructure Layer

2
Comments 2
8 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.