DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
LLM Gateway Explained — Build One With LiteLLM + LangChain

LLM Gateway Explained — Build One With LiteLLM + LangChain

1
Comments
5 min read
When Your LLM Provider Pulls the Rug: Lessons from Anthropic's OAuth Shutdown

When Your LLM Provider Pulls the Rug: Lessons from Anthropic's OAuth Shutdown

Comments
2 min read
How ChatGPT Works (Simple Explanation for Beginners)

How ChatGPT Works (Simple Explanation for Beginners)

Comments
2 min read
From Tokens to Attention: My First Real Mental Model of LLMs

From Tokens to Attention: My First Real Mental Model of LLMs

1
Comments 2
5 min read
2026 Q1 is the year developers still build the agent harness. 2026 Q3 / 2027 is the year the LLM builds its own harness.

2026 Q1 is the year developers still build the agent harness. 2026 Q3 / 2027 is the year the LLM builds its own harness.

Comments 1
4 min read
Stop prompting "write me an API" — teach the LLM the shape first

Stop prompting "write me an API" — teach the LLM the shape first

Comments
2 min read
Why Much AI Memory Risks Rotting — The Exception Is the Memory of Being Wrong

Why Much AI Memory Risks Rotting — The Exception Is the Memory of Being Wrong

Comments
11 min read
What I Learned Building with Gemma 4

Gemma 4 Challenge: Write about Gemma 4 Submission

What I Learned Building with Gemma 4

3
Comments
4 min read
Reasoning Effort: Low, Medium, High: When Each Setting Actually Pays Off

Reasoning Effort: Low, Medium, High: When Each Setting Actually Pays Off

Comments
8 min read
Structured Output Validation: Pydantic/Zod vs In-Prompt Schema vs JSON Mode

Structured Output Validation: Pydantic/Zod vs In-Prompt Schema vs JSON Mode

Comments
9 min read
Prompt Diff Testing: A/B Your Prompts Without Changing the Model

Prompt Diff Testing: A/B Your Prompts Without Changing the Model

Comments
7 min read
Reranker Selection: Cross-Encoder vs LLM-as-Reranker vs ColBERT: Which Earns Its Latency

Reranker Selection: Cross-Encoder vs LLM-as-Reranker vs ColBERT: Which Earns Its Latency

Comments
9 min read
Chunk Overlap: The RAG Parameter Most Teams Pick Wrong

Chunk Overlap: The RAG Parameter Most Teams Pick Wrong

Comments
7 min read
Multi-Turn Agent Context Window: 4 Truncation Strategies That Don't Break the Agent

Multi-Turn Agent Context Window: 4 Truncation Strategies That Don't Break the Agent

Comments
9 min read
Streaming Tool Calls with Anthropic's API: The Buffer Pattern Nobody Documents

Streaming Tool Calls with Anthropic's API: The Buffer Pattern Nobody Documents

Comments
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.