DEV Community

Machine Learning

A branch of artificial intelligence (AI) and computer science which focuses on the use of data and algorithms to imitate the way that humans learn, gradually improving its accuracy.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
AI agent benchmark: I gave 9 models a destroy button and a job that needed it

Kaggle Benchmarking Challenge Submission

AI agent benchmark: I gave 9 models a destroy button and a job that needed it

11
Comments 5
11 min read
12 of 13 AI models knew the new name and still wrote the old one

Kaggle Benchmarking Challenge Submission

12 of 13 AI models knew the new name and still wrote the old one

25
Comments 7
11 min read
Does your coding agent understand you in your native language?

Kaggle Benchmarking Challenge Submission

Does your coding agent understand you in your native language?

Comments 1
4 min read
We Built a Batch Annotation Platform for Dental X-rays Using SAM 2 — Here's What We Learned

We Built a Batch Annotation Platform for Dental X-rays Using SAM 2 — Here's What We Learned

Comments 1
4 min read
vLLM GGUF FAQ: Ten Search Questions, Answered

vLLM GGUF FAQ: Ten Search Questions, Answered

5
Comments 2
3 min read
RL 6: Pleasure-Seeking Neurons and the Birth of Backpropagation (1972–1980)

RL 6: Pleasure-Seeking Neurons and the Birth of Backpropagation (1972–1980)

Comments
13 min read
Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

1
Comments 1
17 min read
"The Gene Went Up. Could AI Tell What We Actually Know?"

Kaggle Benchmarking Challenge Submission

"The Gene Went Up. Could AI Tell What We Actually Know?"

Comments
7 min read
AI's Blind Spot for African Biodiversity

Kaggle Benchmarking Challenge Submission

AI's Blind Spot for African Biodiversity

Comments
2 min read
Qwen-Image 2.1 on 16 GB of VRAM, quantized to NF4: a real benchmark against FLUX

Qwen-Image 2.1 on 16 GB of VRAM, quantized to NF4: a real benchmark against FLUX

Comments 1
2 min read
I Set a Trap and Even Frontier Models Fell For It

Kaggle Benchmarking Challenge Submission

I Set a Trap and Even Frontier Models Fell For It

3
Comments 2
8 min read
100,000 Videos in Your Pocket: How Douyin's Recommender Reads Your Whole Life in O(1)

100,000 Videos in Your Pocket: How Douyin's Recommender Reads Your Whole Life in O(1)

Comments
7 min read
TeluguMixBench: Evaluating LLMs on Telugu and Telugu-English Code-Mixed Tasks

TeluguMixBench: Evaluating LLMs on Telugu and Telugu-English Code-Mixed Tasks

Comments 1
2 min read
SchemaShift: Can LLMs Reliably Debug Data Pipelines?

Kaggle Benchmarking Challenge Submission

SchemaShift: Can LLMs Reliably Debug Data Pipelines?

Comments
3 min read
I Distilled a 568M Multilingual Model Into a 37M Japanese-English Encoder — Here's What Survived

I Distilled a 568M Multilingual Model Into a 37M Japanese-English Encoder — Here's What Survived

1
Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.