DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Base LLMs vs Instruction-Tuned LLMs: Understanding the Architecture Behind ChatGPT and Claude

Base LLMs vs Instruction-Tuned LLMs: Understanding the Architecture Behind ChatGPT and Claude

Comments
3 min read
Creating Personal AI Agents in Multiplayer Games with LoRA Adapters: An Efficient and Memory-Saving Solution

Creating Personal AI Agents in Multiplayer Games with LoRA Adapters: An Efficient and Memory-Saving Solution

8
Comments
4 min read
Build Better RAG Pipelines: Scraping Technical Docs to Clean Markdown

Build Better RAG Pipelines: Scraping Technical Docs to Clean Markdown

Comments 1
2 min read
The Prompting Trick That Fixed My AI Image Generation

The Prompting Trick That Fixed My AI Image Generation

11
Comments
7 min read
Reranking and Two-Stage Retrieval: Precision When It Matters Most

Reranking and Two-Stage Retrieval: Precision When It Matters Most

Comments
2 min read
Deploying NVIDIA Dynamo & LMCache for LLMs: Installation, Containers, and Integration

Deploying NVIDIA Dynamo & LMCache for LLMs: Installation, Containers, and Integration

4
Comments 2
2 min read
Bifrost: The Fastest Open Source LLM Gateway

Bifrost: The Fastest Open Source LLM Gateway

2
Comments
4 min read
🏠 Self-Hosted AI Code Generation: The Complete Guide to Building Your Private AI Coding Assistant

🏠 Self-Hosted AI Code Generation: The Complete Guide to Building Your Private AI Coding Assistant

15
Comments 1
6 min read
The Poetic Hack: Exploiting LLMs with Verse by Arvind Sundararajan

The Poetic Hack: Exploiting LLMs with Verse by Arvind Sundararajan

Comments
2 min read
TOON: Token-Oriented Object Notation – A Complete Guide for LLM Data Efficiency

TOON: Token-Oriented Object Notation – A Complete Guide for LLM Data Efficiency

Comments 1
3 min read
Finally Got My Dify Agent Working in Discord, Telegram and Slack

Finally Got My Dify Agent Working in Discord, Telegram and Slack

4
Comments
3 min read
From 16-bit to 4-bit: The Architecture for Scalable Personalized LLM Deployment

From 16-bit to 4-bit: The Architecture for Scalable Personalized LLM Deployment

5
Comments
6 min read
Dense vs Sparse Retrieval: Mastering FAISS, BM25, and Hybrid Search

Dense vs Sparse Retrieval: Mastering FAISS, BM25, and Hybrid Search

2
Comments
15 min read
Prompt‑Powered User Personas: From Messy Logs to Living Profiles

Prompt‑Powered User Personas: From Messy Logs to Living Profiles

Comments 1
14 min read
Why your AI assistant lies to you (and how to fix it)

Why your AI assistant lies to you (and how to fix it)

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.