DEV Community

Deep Learning

This tag is for discussing, sharing articles, and asking questions primarily on deep learning - a subfield of machine learning.

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
What Comes After LLMs? Mamba, Diffusion & World Models

What Comes After LLMs? Mamba, Diffusion & World Models

Comments
11 min read
What Your Loss Function Actually Tells the Model: MSE, Cross-Entropy, and the softmax Bug That Trains Anyway

What Your Loss Function Actually Tells the Model: MSE, Cross-Entropy, and the softmax Bug That Trains Anyway

Comments
13 min read
Neural Networks: Inspired by the Brain, Built for Data

Neural Networks: Inspired by the Brain, Built for Data

Comments
10 min read
We measured sharpness during loss spikes. The instrument started returning negative numbers.

We measured sharpness during loss spikes. The instrument started returning negative numbers.

Comments
3 min read
What I Learned Building a Mini TensorRT

What I Learned Building a Mini TensorRT

Comments
3 min read
AI LinkedIn headshot explained, a developer teardown

AI LinkedIn headshot explained, a developer teardown

Comments
6 min read
Neural Networks: Weights, Activation, and Backpropagation

Neural Networks: Weights, Activation, and Backpropagation

Comments
5 min read
Four Mechanisms Were Blamed for Loss Spikes in 2026. We Tested All Four at Once. None of Them Alone Causes Spikes.

Four Mechanisms Were Blamed for Loss Spikes in 2026. We Tested All Four at Once. None of Them Alone Causes Spikes.

Comments 1
3 min read
Inverse Problems: Why Predicting Backward Is Harder Than It Looks

Inverse Problems: Why Predicting Backward Is Harder Than It Looks

Comments
6 min read
The Use of AI in Computer Vision

The Use of AI in Computer Vision

Comments 1
3 min read
Every Greedy Metric Said the Model Was Improving. Then pass@64 Fell From 0.83 to 0.19

Every Greedy Metric Said the Model Was Improving. Then pass@64 Fell From 0.83 to 0.19

1
Comments
4 min read
AI Safety Researcher 工具箱:数据投毒防御、可解释性、进度追踪与判断力

AI Safety Researcher 工具箱:数据投毒防御、可解释性、进度追踪与判断力

Comments
2 min read
The Next Phase of Vision-Language Navigation

The Next Phase of Vision-Language Navigation

Comments
9 min read
Imag-Eval

Imag-Eval

Comments
1 min read
Neve - Towards a Unified Programming Model for the Complete Deep Learning Stack

Neve - Towards a Unified Programming Model for the Complete Deep Learning Stack

Comments
8 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.