DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Mitigating Hallucinations in Theology AI: Implementing Groundedness Evaluation Pipelines

Mitigating Hallucinations in Theology AI: Implementing Groundedness Evaluation Pipelines

Comments 1
7 min read
Cli-Modelarium 0.1.4: 10 LLM providers now, with Qwen and GLM

Cli-Modelarium 0.1.4: 10 LLM providers now, with Qwen and GLM

Comments
2 min read
Designing Workflow-Level Guardrails on Top of Azure AI Foundry with elsai Guardrails

Designing Workflow-Level Guardrails on Top of Azure AI Foundry with elsai Guardrails

Comments
7 min read
【红杉播客】AI Neolab--Engram【主攻记忆与持续学习】--分享未来 AI 发展趋势的独特见解

【红杉播客】AI Neolab--Engram【主攻记忆与持续学习】--分享未来 AI 发展趋势的独特见解

Comments
2 min read
A Forced Dissent Slot Has a Floor: Read It by Convergence, Not Presence

A Forced Dissent Slot Has a Floor: Read It by Convergence, Not Presence

2
Comments
6 min read
Foreman 101: agentic coding as Kubernetes resources

Foreman 101: agentic coding as Kubernetes resources

3
Comments 1
9 min read
How I Let a Cloud AI Operate My Home Without Handing Over My Home

How I Let a Cloud AI Operate My Home Without Handing Over My Home

3
Comments
3 min read
LangGraph isn't cheaper than LangChain — unless you opt out of its defaults

LangGraph isn't cheaper than LangChain — unless you opt out of its defaults

2
Comments 2
7 min read
I edited a system prompt and had no way to prove it changed anything. So I built a measurement tool.

I edited a system prompt and had no way to prove it changed anything. So I built a measurement tool.

Comments
3 min read
How to Build a RAG Knowledge Base from Any Documentation Site in 5 Minutes

How to Build a RAG Knowledge Base from Any Documentation Site in 5 Minutes

Comments
2 min read
Reintroducing Neonmem or synopsis of what was before

Reintroducing Neonmem or synopsis of what was before

Comments
2 min read
Kmemo: a semantic cache for LLM calls that refuses to serve you the wrong answer

Exposes the hidden danger of similarity thresholds

Kmemo: a semantic cache for LLM calls that refuses to serve you the wrong answer

5
Comments 9
5 min read
Why Attention Becomes the Bottleneck — And How Efficient Attention Fixes It

Why Attention Becomes the Bottleneck — And How Efficient Attention Fixes It

Comments
3 min read
Construindo um assistente pessoal 100% local com Ollama, LangChain e voz — e as armadilhas que ninguém conta

Construindo um assistente pessoal 100% local com Ollama, LangChain e voz — e as armadilhas que ninguém conta

Comments
7 min read
I built a tool to prove my multi-agent harness was worth it. It told me it wasn't.

Stats prove simple prompts beat complex panels

I built a tool to prove my multi-agent harness was worth it. It told me it wasn't.

2
Comments 12
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.