DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
DeepSeek pauses fundraise over Huawei deficit as Hugging Face demands $100M

DeepSeek pauses fundraise over Huawei deficit as Hugging Face demands $100M

7
Comments
9 min read
From a Sketch to a Working Multi‑Brain Dialogue Module — Design, Iteration, and Lessons Learned

From a Sketch to a Working Multi‑Brain Dialogue Module — Design, Iteration, and Lessons Learned

1
Comments 1
4 min read
GEO: How to Get Your Content Cited by AI Search Engines (With Data from the Princeton Study)

GEO: How to Get Your Content Cited by AI Search Engines (With Data from the Princeton Study)

Comments
2 min read
When Should an AI Agent Ask for Human Approval?

When Should an AI Agent Ask for Human Approval?

1
Comments 1
8 min read
Parsing documents for air-gapped RAG: no cloud, no JVM, no Python

Parsing documents for air-gapped RAG: no cloud, no JVM, no Python

Comments 5
5 min read
QLoRA: Fine-Tuning a 7B Model on a 16GB GPU (It Shrank to 5.4GB in Front of Me)

QLoRA: Fine-Tuning a 7B Model on a 16GB GPU (It Shrank to 5.4GB in Front of Me)

Comments
3 min read
If a 270M Model Already Worked, Why Did I Fine-Tune a 7B One?

If a 270M Model Already Worked, Why Did I Fine-Tune a 7B One?

Comments
3 min read
LoRA: I Trained <1% of a 1.5B Model and Matched a Full Fine-Tune

LoRA: I Trained <1% of a 1.5B Model and Matched a Full Fine-Tune

Comments
3 min read
Spring Boot MCP Server in 2026: The Transport Trap That Wastes Your Weekend

Spring Boot MCP Server in 2026: The Transport Trap That Wastes Your Weekend

2
Comments 2
3 min read
I Fine-Tuned a 270M Model on My Laptop (Full Fine-Tuning, From Scratch)

I Fine-Tuned a 270M Model on My Laptop (Full Fine-Tuning, From Scratch)

Comments
2 min read
I built an agent health checker, then it flunked itself — here's the audit

I built an agent health checker, then it flunked itself — here's the audit

Comments
9 min read
AMD ATOM + ATOMesh: Prefill/decode Disaggregation on ROCm

AMD ATOM + ATOMesh: Prefill/decode Disaggregation on ROCm

Comments
7 min read
Gemini 3.6 Flash & 3.5 Flash-Lite: Developer guide

Gemini 3.6 Flash & 3.5 Flash-Lite: Developer guide

22
Comments 2
6 min read
Transport, Surface, Skin: Building MCP Plugins That Survive the Spec

Transport, Surface, Skin: Building MCP Plugins That Survive the Spec

Comments
16 min read
Stop Wasting LLM Budgets: High-Performance Semantic Caching with Spring AI and pgvector

Stop Wasting LLM Budgets: High-Performance Semantic Caching with Spring AI and pgvector

Comments
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.