Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.
#
llm
#
benchmarks
#
localllm
#
reproducibility
1
 reaction
Comments
1
 comment
3 min read
From Guesswork to Evidence: How BotTalk Built a Systematic Approach to AI Model Selection
Debapriya Dey
Debapriya Dey
Debapriya Dey
Follow
Aug 27
From Guesswork to Evidence: How BotTalk Built a Systematic Approach to AI Model Selection
#
ai
#
aws
#
llm
1
 reaction
Comments
Add Comment
4 min read
Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.
#
llm
#
benchmarks
#
localllm
#
privateai
2
 reactions
Comments
Add Comment
4 min read
Your RAG Isn't Broken. Your Retrieval Pipeline Is.
RAJSHREE
RAJSHREE
RAJSHREE
Follow
Aug 22
Your RAG Isn't Broken. Your Retrieval Pipeline Is.
#
rag
#
llm
#
ai
#
semanticsearch
Comments
Add Comment
16 min read
I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.
#
llm
#
prompting
#
debugging
#
localllm
1
 reaction
Comments
Add Comment
3 min read
Free AI Tokens Are a Trap: An Opinionated Cost Gate for Model Experiments
Finley Zhu
Finley Zhu
Finley Zhu
Follow
Aug 22
Free AI Tokens Are a Trap: An Opinionated Cost Gate for Model Experiments
#
ai
#
llm
#
python
#
devops
Comments
Add Comment
5 min read
Free vs Self-Hosted Models: A Break-Even Framework for Agent Workloads
Sam Sun
Sam Sun
Sam Sun
Follow
Aug 22
Free vs Self-Hosted Models: A Break-Even Framework for Agent Workloads
#
ai
#
llm
#
agents
#
opensource
Comments
Add Comment
5 min read
Batch LLM Jobs Without Breaking the Bank: A Queue-First Architecture for Free Tiers
Riley Wu
Riley Wu
Riley Wu
Follow
Aug 22
Batch LLM Jobs Without Breaking the Bank: A Queue-First Architecture for Free Tiers
#
python
#
llm
#
architecture
#
opensource
Comments
Add Comment
4 min read
A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number
#
llm
#
benchmarks
#
localllm
#
agents
1
 reaction
Comments
Add Comment
4 min read
Why is everyone so skeptical of AI memory tools? Fair question. Here are real answers.
Filippo Pilotta
Filippo Pilotta
Filippo Pilotta
Follow
Aug 22
Why is everyone so skeptical of AI memory tools? Fair question. Here are real answers.
#
ai
#
memory
#
mcp
#
llm
Comments
Add Comment
5 min read
How to route LLM requests by task difficulty (a practical guide to cutting API spend without losing quality)
Weio
Weio
Weio
Follow
Sep 5
How to route LLM requests by task difficulty (a practical guide to cutting API spend without losing quality)
#
ai
#
llm
#
webdev
#
programming
1
 reaction
Comments
Add Comment
5 min read
Opinion: Free AI Coding Tokens Are a Model Release, Not a Gift
Morgan Li
Morgan Li
Morgan Li
Follow
Aug 22
Opinion: Free AI Coding Tokens Are a Model Release, Not a Gift
#
ai
#
llm
#
testing
#
opinion
Comments
Add Comment
5 min read
From Massive to Miniature: How Small Language Models Are Engineered
Vishal Chandak
Vishal Chandak
Vishal Chandak
Follow
Aug 22
From Massive to Miniature: How Small Language Models Are Engineered
#
smalllanguagemodel
#
artificialintelligen
#
machinelearning
#
llm
Comments
Add Comment
12 min read
Quota Exhaustion Date: A Capacity Calculator for Free Model Quotas
Casey Sun
Casey Sun
Casey Sun
Follow
Aug 22
Quota Exhaustion Date: A Capacity Calculator for Free Model Quotas
#
ai
#
llm
#
devops
#
tutorial
Comments
Add Comment
4 min read
Why the Most Honest Coding-Agent Evaluation Runs on Free Infrastructure
Dakota Wu
Dakota Wu
Dakota Wu
Follow
Aug 22
Why the Most Honest Coding-Agent Evaluation Runs on Free Infrastructure
#
ai
#
agents
#
testing
#
llm
Comments
1
 comment
5 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account