DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I built an open-source memory layer for LLMs — here's how it works

I built an open-source memory layer for LLMs — here's how it works

Comments
4 min read
I needed to know if the cheaper model was good enough. So I built an LLM-as-a-Judge pipeline

I needed to know if the cheaper model was good enough. So I built an LLM-as-a-Judge pipeline

Comments
2 min read
Why your LLM product hallucinates the one thing it shouldn't, and the architectural pattern that fixes it

Why your LLM product hallucinates the one thing it shouldn't, and the architectural pattern that fixes it

Comments
4 min read
OpenAI’s $1M API Credits, Holos’ Agentic Web, and Xpertbench’s Expert Tasks

OpenAI’s $1M API Credits, Holos’ Agentic Web, and Xpertbench’s Expert Tasks

Comments
2 min read
From Pydantic Model to AI Agent in 10 Lines of Python

From Pydantic Model to AI Agent in 10 Lines of Python

Comments
4 min read
Harness Engineering: The Architecture of Production-Grade AI Systems

Harness Engineering: The Architecture of Production-Grade AI Systems

1
Comments
4 min read
I tested speculative decoding on my home GPU cluster. Here's why it didn't help.

I tested speculative decoding on my home GPU cluster. Here's why it didn't help.

Comments
5 min read
KV Caching in LLMs

KV Caching in LLMs

Comments 3
4 min read
Letting AI Control RAG Search Improved Accuracy by 79%

Letting AI Control RAG Search Improved Accuracy by 79%

Comments
6 min read
Why Some AI Feels “Process-Obsessed” While Others Just Ship Code

Why Some AI Feels “Process-Obsessed” While Others Just Ship Code

Comments
1 min read
Cut AI Costs: Flutter On-Device LLM Integration Works

Cut AI Costs: Flutter On-Device LLM Integration Works

Comments
10 min read
Jetson Containers Quickstart on NVIDIA Jetson AGX Orin 64GB

Jetson Containers Quickstart on NVIDIA Jetson AGX Orin 64GB

Comments
8 min read
Two Kinds of AI Agents (And Why You Need Both)

Two Kinds of AI Agents (And Why You Need Both)

Comments
10 min read
How I Built Persistent Memory for AI Agents in Python

How I Built Persistent Memory for AI Agents in Python

Comments
2 min read
Why Your LLM App Fails in Production (and How to Debug It)

Why Your LLM App Fails in Production (and How to Debug It)

4
Comments 1
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.