DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
5 Ways Prompt Injection Can Silently Compromise Your AI App

5 Ways Prompt Injection Can Silently Compromise Your AI App

Comments
4 min read
I am 54 years old, 30 years in B2B sales. I tried to build my own AI sales coach. Here is what happened

I am 54 years old, 30 years in B2B sales. I tried to build my own AI sales coach. Here is what happened

Comments 1
2 min read
We let AI write the code. We just don't let it check its own work.

We let AI write the code. We just don't let it check its own work.

2
Comments 4
4 min read
LLM KV Cache Optimization, Open Model Evaluation, & Agent Engineering Skills for Local Deployment

LLM KV Cache Optimization, Open Model Evaluation, & Agent Engineering Skills for Local Deployment

Comments
3 min read
The two causes of your token bill

The two causes of your token bill

Comments
6 min read
I built a "boring" RAG demo over World Cup data — SQLite, sqlite-vec, and no framework

I built a "boring" RAG demo over World Cup data — SQLite, sqlite-vec, and no framework

Comments
4 min read
LLM Observability on Kubernetes: A Practical Guide

LLM Observability on Kubernetes: A Practical Guide

Comments
30 min read
Prompt cache, finally typed: shipping llm-ports 0.1.0-alpha.19

Prompt cache, finally typed: shipping llm-ports 0.1.0-alpha.19

Comments
6 min read
EvaluatingAgents, Securing AI, and Local LLMs Take Center Stage

EvaluatingAgents, Securing AI, and Local LLMs Take Center Stage

Comments
2 min read
oMLX 效能調校 KV Cache 與Concurrent Batching

oMLX 效能調校 KV Cache 與Concurrent Batching

Comments
9 min read
Memory Poisoning: The Silent Threat to AI Agents (and How to Defend Against It)

Memory Poisoning: The Silent Threat to AI Agents (and How to Defend Against It)

Comments
2 min read
LLM Fine-Tuning Guide: Full Fine-Tuning, LoRA, Learning Rate, and VRAM

LLM Fine-Tuning Guide: Full Fine-Tuning, LoRA, Learning Rate, and VRAM

Comments
11 min read
The Difference Between Search and Discovery

The Difference Between Search and Discovery

4
Comments 2
3 min read
Why You Need to Become a Neuro-Punk Right Now

Why You Need to Become a Neuro-Punk Right Now

Comments
6 min read
Token Cost Optimization: How to Cut LLM Inference Spend Without Cutting Quality

Token Cost Optimization: How to Cut LLM Inference Spend Without Cutting Quality

1
Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.