DEV Community

#ollama

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I Built a Personal AI That Actually Knows My Projects (RAG + Ollama, Zero Cloud)

I Built a Personal AI That Actually Knows My Projects (RAG + Ollama, Zero Cloud)

Comments
5 min read
AI ERD Design With Ollama: Schema Help That Never Leaves Your Machine

AI ERD Design With Ollama: Schema Help That Never Leaves Your Machine

Comments
5 min read
Run Qwen Coder & DeepSeek Locally: The 2026 Free AI Pair-Programmer Setup

Run Qwen Coder & DeepSeek Locally: The 2026 Free AI Pair-Programmer Setup

Comments
7 min read
Ollama Model Loading RCE: Three Years of the Same Bug Class, One Self-Hosted LLM Runtime

Ollama Model Loading RCE: Three Years of the Same Bug Class, One Self-Hosted LLM Runtime

Comments
7 min read
It Fits and It Benchmarks Well. Will It Do Your Job?

It Fits and It Benchmarks Well. Will It Do Your Job?

2
Comments 1
4 min read
One MCP Server, Two Models: An Always-On Ops Agent That Costs $0

One MCP Server, Two Models: An Always-On Ops Agent That Costs $0

Comments
15 min read
I Built a Dead-Simple API Gateway for My Local LLMs in 50 Lines of Python

I Built a Dead-Simple API Gateway for My Local LLMs in 50 Lines of Python

Comments
4 min read
I built a 100% local invoice reader with Ollama + n8n — the real trick was teaching it NOT to guess

I built a 100% local invoice reader with Ollama + n8n — the real trick was teaching it NOT to guess

Comments
2 min read
I Built a Dead-Simple API Gateway for My Local LLMs in 50 Lines of Python

I Built a Dead-Simple API Gateway for My Local LLMs in 50 Lines of Python

1
Comments
4 min read
5 AI Tools You Can Run Locally for Free in 2026

5 AI Tools You Can Run Locally for Free in 2026

Comments
1 min read
Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Comments
3 min read
LLM Quantization Levels Compared: Q4_K_M vs Q8_0 vs FP16 [2026]

LLM Quantization Levels Compared: Q4_K_M vs Q8_0 vs FP16 [2026]

Comments
15 min read
Orchestrate Multiple Local LLMs: RTX 4090 Benchmarks

Orchestrate Multiple Local LLMs: RTX 4090 Benchmarks

1
Comments
9 min read
Eve: un modo “convention over configuration” per costruire e far girare AI agent in TypeScript (anche con modelli locali)

Eve: un modo “convention over configuration” per costruire e far girare AI agent in TypeScript (anche con modelli locali)

Comments
4 min read
I Ditched ChatGPT for Local LLMs and Saved $2,000 in a Year — The Real Numbers

I Ditched ChatGPT for Local LLMs and Saved $2,000 in a Year — The Real Numbers

1
Comments 1
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.