DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Fable disappeared overnight. That's the best ad for open-weight AI anyone could have run.

Fable disappeared overnight. That's the best ad for open-weight AI anyone could have run.

1
Comments
3 min read
I thought Mnemara would save tokens for cloud based models, that was wrong.

I thought Mnemara would save tokens for cloud based models, that was wrong.

Comments
4 min read
We ported how brains manage the cost of thinking to LLM systems

We ported how brains manage the cost of thinking to LLM systems

Comments 2
9 min read
AI Agents write code that compiles, but they still lie to the user. Here is how to fix the pipeline

AI Agents write code that compiles, but they still lie to the user. Here is how to fix the pipeline

Comments
1 min read
I Benchmarked 47 LLM Providers Against Real Queries - Here's What I Found 📊

I Benchmarked 47 LLM Providers Against Real Queries - Here's What I Found 📊

Comments
8 min read
从 pip install 到生产部署:AI 自愈 Agent 10 分钟上线指南

从 pip install 到生产部署:AI 自愈 Agent 10 分钟上线指南

Comments
2 min read
Building and Running Llama.cpp on an Air-Gapped Mac

Building and Running Llama.cpp on an Air-Gapped Mac

Comments
3 min read
TitanCore Core-1 – Trillion-parameter LLM training infra in C++/CUDA with ZeRO-3

TitanCore Core-1 – Trillion-parameter LLM training infra in C++/CUDA with ZeRO-3

Comments
1 min read
AIMO: AI Mention Optimization — The Discipline of Being Recommended by AI Assistants

AIMO: AI Mention Optimization — The Discipline of Being Recommended by AI Assistants

Comments
6 min read
Multi-Agent Kill Switch: Why Stopping the Orchestrator Doesn't Stop the Swarm

Multi-Agent Kill Switch: Why Stopping the Orchestrator Doesn't Stop the Swarm

1
Comments 1
11 min read
GroundedQL: a semantic compiler for natural-language Postgres analytics

GroundedQL: a semantic compiler for natural-language Postgres analytics

Comments
2 min read
llama.cpp Optimizations & New Qwopus3.5-9B GGUF Model Boost Local AI Performance

llama.cpp Optimizations & New Qwopus3.5-9B GGUF Model Boost Local AI Performance

Comments
3 min read
Why I used three different critic roles instead of one (and what the eval taught me)

Why I used three different critic roles instead of one (and what the eval taught me)

Comments 2
6 min read
Fitting LLM Reply Suggestions Into Every Provider's Prompt Cache — Without Structured Output

Fitting LLM Reply Suggestions Into Every Provider's Prompt Cache — Without Structured Output

Comments 1
4 min read
Tackle High Token Usage with GraphRAG

Tackle High Token Usage with GraphRAG

1
Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.