Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
benchmark
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Glasshouse v0.1 Is Out: A Memory Benchmark for AI Systems
woochan
woochan
woochan
Follow
Sep 22
Glasshouse v0.1 Is Out: A Memory Benchmark for AI Systems
#
ai
#
memory
#
benchmark
#
wontopos
7
reactions
Comments
1
comment
2 min read
Onboarding benchmarks from real data across 464 SaaS products: median tour completion is 29%, 1-2 step tours complete at 73%, 9+ step tours at 8%
Olli
Olli
Olli
Follow
Sep 22
Onboarding benchmarks from real data across 464 SaaS products: median tour completion is 29%, 1-2 step tours complete at 73%, 9+ step tours at 8%
#
saas
#
onboarding
#
product
#
benchmark
Comments
1
comment
2 min read
How LLM Evaluation Actually Works: Inside the Satellite Geo QCM Leaderboard
RESK
RESK
RESK
Follow
Sep 22
How LLM Evaluation Actually Works: Inside the Satellite Geo QCM Leaderboard
#
llm
#
evaluation
#
benchmark
#
vision
1
reaction
Comments
Add Comment
4 min read
The 5 Walls Between a 3M req/s HTTP Benchmark and Production
Thomas Tartrau
Thomas Tartrau
Thomas Tartrau
Follow
Sep 21
The 5 Walls Between a 3M req/s HTTP Benchmark and Production
#
rust
#
performance
#
http
#
benchmark
Comments
Add Comment
7 min read
netcup VPS 1000 G12 benchmarked: how fast is it really?
serverkueche.de
serverkueche.de
serverkueche.de
Follow
Sep 20
netcup VPS 1000 G12 benchmarked: how fast is it really?
#
benchmark
#
vps
#
selfhosted
Comments
Add Comment
7 min read
A code review benchmark that isn't the vendor ranking itself
Tess Ainsley
Tess Ainsley
Tess Ainsley
Follow
Sep 20
A code review benchmark that isn't the vendor ranking itself
#
codereview
#
aicode
#
benchmark
#
reviewtools
Comments
Add Comment
4 min read
Enterprise Vector Database 2026: Qdrant vs Milvus vs pgvector vs Pinecone
devrudals
devrudals
devrudals
Follow
Sep 19
Enterprise Vector Database 2026: Qdrant vs Milvus vs pgvector vs Pinecone
#
ai
#
devops
#
benchmark
#
cloud
Comments
Add Comment
23 min read
You can read Claude Code's whole harness now. That's what every benchmark score throws away
Cole Halton
Cole Halton
Cole Halton
Follow
Sep 18
You can read Claude Code's whole harness now. That's what every benchmark score throws away
#
codingagents
#
harness
#
benchmark
#
eval
Comments
Add Comment
5 min read
TypeSafe’s JEV Model: Is It Really 193x Faster and 444x Cheaper?
Ariful Islam
Ariful Islam
Ariful Islam
Follow
Sep 22
TypeSafe’s JEV Model: Is It Really 193x Faster and 444x Cheaper?
#
jev
#
typesafe
#
ai
#
benchmark
2
reactions
Comments
1
comment
6 min read
JetBrains Ranked AI Agents on Real Kotlin Projects. The Token Column Is the Real Story.
jamilxt
jamilxt
jamilxt
Follow
Sep 15
JetBrains Ranked AI Agents on Real Kotlin Projects. The Token Column Is the Real Story.
#
ai
#
kotlin
#
java
#
benchmark
1
reaction
Comments
Add Comment
7 min read
Two "Codex CLI" models on the same benchmark: the harness hides the model
Cole Halton
Cole Halton
Cole Halton
Follow
Sep 14
Two "Codex CLI" models on the same benchmark: the harness hides the model
#
ai
#
codingagents
#
benchmark
#
eval
Comments
Add Comment
2 min read
How Fast Can Neovim Start? Benchmarking Popular Distros
glmlm
glmlm
glmlm
Follow
Sep 22
How Fast Can Neovim Start? Benchmarking Popular Distros
#
neovim
#
productivity
#
performance
#
benchmark
Comments
Add Comment
6 min read
Why I’m Building a New AI Memory Benchmark (And Why the Existing Ones Fall Short)
woochan
woochan
woochan
Follow
Sep 7
Why I’m Building a New AI Memory Benchmark (And Why the Existing Ones Fall Short)
#
ai
#
benchmark
#
startup
#
opensource
2
reactions
Comments
Add Comment
2 min read
Elysia 2 vs NestJS 12: Runtime +64.6%, Framework +10.6%
Davron Yuldashev
Davron Yuldashev
Davron Yuldashev
Follow
Sep 7
Elysia 2 vs NestJS 12: Runtime +64.6%, Framework +10.6%
#
benchmark
#
bunjs
#
elysia
#
performance
Comments
Add Comment
14 min read
เมื่อ Benchmark โกหกคุณ, SWE-Bench ProMax กับคะแนนจริงที่โมเดลเก่งสุดทำได้แค่ 41.2%
Nokka
Nokka
Nokka
Follow
Sep 5
เมื่อ Benchmark โกหกคุณ, SWE-Bench ProMax กับคะแนนจริงที่โมเดลเก่งสุดทำได้แค่ 41.2%
#
ai
#
benchmark
#
programming
#
machinelearning
Comments
Add Comment
2 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account