DEV Community

Machine Learning

A branch of artificial intelligence (AI) and computer science which focuses on the use of data and algorithms to imitate the way that humans learn, gradually improving its accuracy.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
[Meta-RL] We told an AI agent 'you can fail 3 times.' Accuracy went up 19%.

[Meta-RL] We told an AI agent 'you can fail 3 times.' Accuracy went up 19%.

4
Comments
4 min read
I Built a Feedback Loop That Coaches LLMs at Runtime Using NumPy

I Built a Feedback Loop That Coaches LLMs at Runtime Using NumPy

Comments
3 min read
Wine classification - Vivino Qwasar

Wine classification - Vivino Qwasar

Comments
2 min read
Supprimer la censure des modèles LLM avec Heretic

Supprimer la censure des modèles LLM avec Heretic

1
Comments
8 min read
LLM 모델 검열 제거 방법: Heretic 활용

LLM 모델 검열 제거 방법: Heretic 활용

1
Comments
3 min read
When AI Stops to Think — And When It Shouldn't

When AI Stops to Think — And When It Shouldn't

Comments
3 min read
Why Your Website Needs a Robot Trust Certificate (And How to Get One for Free)

Why Your Website Needs a Robot Trust Certificate (And How to Get One for Free)

1
Comments 1
1 min read
The AI Agent Ecosystem in 2026: What's Actually Working (and What's Getting Canceled)

The AI Agent Ecosystem in 2026: What's Actually Working (and What's Getting Canceled)

1
Comments 2
3 min read
Solving "Use Machine Learning APIs on Google Cloud: Challenge Lab" — A Complete Guide

Solving "Use Machine Learning APIs on Google Cloud: Challenge Lab" — A Complete Guide

4
Comments
12 min read
6 Best Reinforcement Learning (RL) Tools in 2026

6 Best Reinforcement Learning (RL) Tools in 2026

5
Comments
16 min read
MiniMax M2.7 คืออะไร โมเดล AI พัฒนาตัวเองได้

MiniMax M2.7 คืออะไร โมเดล AI พัฒนาตัวเองได้

2
Comments
5 min read
How Did AI Learn to Be Nice? The Humans Behind the Curtain

How Did AI Learn to Be Nice? The Humans Behind the Curtain

7
Comments
5 min read
Detecting LLM Agent Contradictions Using NLI and Total Variance — A Python Implementation

Detecting LLM Agent Contradictions Using NLI and Total Variance — A Python Implementation

Comments 1
7 min read
Why Merging AI Models Fails (And How a 'Gossip Handshake' Fixed It)

Why Merging AI Models Fails (And How a 'Gossip Handshake' Fixed It)

2
Comments 9
2 min read
Why AI Agents Fall Apart on Real Work

Why AI Agents Fall Apart on Real Work

1
Comments 2
9 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.