DEV Community

#benchmarking

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
My context selector beat grep. An agent with grep beat it.

My context selector beat grep. An agent with grep beat it.

Comments
7 min read
Correctness Has a Price: We Benchmarked Fair Leaderboards

Correctness Has a Price: We Benchmarked Fair Leaderboards

5
Comments
6 min read
Benchmarking Bun and Node.js for a High-Throughput Video Metadata API

Benchmarking Bun and Node.js for a High-Throughput Video Metadata API

1
Comments
8 min read
We Put Our Router on an Academic Benchmark. Here Are the Numbers We'd Rather Hide.

We Put Our Router on an Academic Benchmark. Here Are the Numbers We'd Rather Hide.

1
Comments 1
5 min read
LiteSpeed vs Nginx for WordPress: 3 months of production benchmarks

LiteSpeed vs Nginx for WordPress: 3 months of production benchmarks

Comments
3 min read
Benchmarking Bun and Node for a High-Traffic Video Metadata API

Benchmarking Bun and Node for a High-Traffic Video Metadata API

Comments
9 min read
SQLAlchemy ORM Security: The Raw Query Escape Hatch

SQLAlchemy ORM Security: The Raw Query Escape Hatch

Comments
5 min read
A Benchmark Smelled Funny

A Benchmark Smelled Funny

Comments 1
9 min read
If 30% of Coding Tasks May Be Broken, Your Leaderboard Needs an Uncertainty Budget

If 30% of Coding Tasks May Be Broken, Your Leaderboard Needs an Uncertainty Budget

1
Comments
3 min read
Benchmarking Apple's SpeechAnalyzer API vs. Whisper: Performance, Accuracy, and Use Cases

Benchmarking Apple's SpeechAnalyzer API vs. Whisper: Performance, Accuracy, and Use Cases

Comments
2 min read
IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI

IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI

Comments
4 min read
How I Benchmarked an LLM Running Entirely on a Phone (No Cloud, No API)

How I Benchmarked an LLM Running Entirely on a Phone (No Cloud, No API)

Comments
16 min read
prima.cpp local llm benchmark: 15% Faster Than llama.cpp

prima.cpp local llm benchmark: 15% Faster Than llama.cpp

Comments
8 min read
My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

Comments
10 min read
Building an Official Performance Baseline for Vix.cpp Core v2.6.3

Building an Official Performance Baseline for Vix.cpp Core v2.6.3

Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.