DEV Community

Deep Learning

This tag is for discussing, sharing articles, and asking questions primarily on deep learning - a subfield of machine learning.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Generative Modeling: Learning the Distribution Behind the Data

Generative Modeling: Learning the Distribution Behind the Data

Comments
5 min read
I Implemented the Algorithm Behind ChatGPT From Scratch - Day 8 (PPO).

I Implemented the Algorithm Behind ChatGPT From Scratch - Day 8 (PPO).

11
Comments
3 min read
RL 1: Biological foundations and the "Law of Effect" (1898– 1949)

RL 1: Biological foundations and the "Law of Effect" (1898– 1949)

1
Comments
12 min read
From Learning Machine Learning to Competing on Kaggle: My First End-to-End Playground Competition Journey

From Learning Machine Learning to Competing on Kaggle: My First End-to-End Playground Competition Journey

Comments
9 min read
Separar canciones en stems con HTDemucs v4 en un servidor con poca RAM

Separar canciones en stems con HTDemucs v4 en un servidor con poca RAM

Comments
2 min read
Generative Modeling: From Data Distributions to Deep Generative Models

Generative Modeling: From Data Distributions to Deep Generative Models

Comments
10 min read
Detecting Objects in Satellite Imagery with YOLOv8: xView + DOTA to YOLO in Practice

Detecting Objects in Satellite Imagery with YOLOv8: xView + DOTA to YOLO in Practice

1
Comments 3
5 min read
权重即数据:神经网络权重空间学习如何成为 AI 的下一类训练集

权重即数据:神经网络权重空间学习如何成为 AI 的下一类训练集

Comments
2 min read
How Vision-Language Models Learned to Reason About Space (10 Papers, One Thread)

How Vision-Language Models Learned to Reason About Space (10 Papers, One Thread)

Comments
3 min read
GPU Architecture for ML (CUDA basics)

GPU Architecture for ML (CUDA basics)

Comments 1
8 min read
MoE Capacity Factor: Why Mixture-of-Experts Drops Your Tokens

MoE Capacity Factor: Why Mixture-of-Experts Drops Your Tokens

Comments
8 min read
I Taught an Agent to Act Directly - No Q-Values Needed (Day 6: REINFORCE)

I Taught an Agent to Act Directly - No Q-Values Needed (Day 6: REINFORCE)

5
Comments
3 min read
Building Your Own Deep Learning Framework: A 3–5 Month Journey From Zero to Tiny Transformer

Building Your Own Deep Learning Framework: A 3–5 Month Journey From Zero to Tiny Transformer

Comments
28 min read
The Easiest Way to Understand Backpropagation

The Easiest Way to Understand Backpropagation

1
Comments
2 min read
GradCuit: Credit-Assigned Gradient Flow for Robust Test-Time Latent Reasoning in LLMs

GradCuit: Credit-Assigned Gradient Flow for Robust Test-Time Latent Reasoning in LLMs

Comments
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.