d2l-zh

《动手学深度学习》:面向中文读者、能运行、可讨论。中英文版被70多个国家的500多所大学用于教学。

23 articles 75.9k View on GitHub ↗
23 articles
How d2l-zh Explains Learning Rate Optimization for Model Convergence

d2l-zh explains learning rate optimization for model convergence detailing optimal ranges and dynamic schedules like cosine decay with warm-up for stable, fast deep network training.

tutorial
Mar 1, 2026
Strategies for Handling Imbalanced Datasets in Deep Learning: A Practical Guide from d2l-zh

Discover d2l-zh strategies for imbalanced datasets in deep learning: reweighting loss, focal loss, and resampling. Improve your model accuracy now.

how-to-guide
Mar 1, 2026
How d2l-zh Addresses the Challenges of Training Very Deep Neural Networks

Learn how d2l-zh tackles deep neural network training challenges using residual connections, batch normalization, and more for stable, multi-layer network training.

deep-dive
Mar 1, 2026
Activation Functions in d2l-zh: Types, Implementation, and Usage Across Deep Learning Frameworks

Explore activation functions in d2l-zh: ReLU, tanh, Sigmoid, and masked Softmax. Understand their implementation and usage across PyTorch TensorFlow MXNet and Paddle for deep learning.

deep-dive
Mar 1, 2026
How d2l-zh Explains Gradient Descent and Its Variants: From First Principles to Adaptive Optimization

Explore how d2l-zh explains gradient descent and its variants using Taylor expansion, stochastic sampling, momentum, and adaptive learning rates for efficient optimization.

deep-dive
Mar 1, 2026
Hyperparameter Tuning in d2l-zh: Best Practices from Dive into Deep Learning

Master hyperparameter tuning in d2l-zh Learn best practices like validation sets and K-fold cross-validation to optimize your deep learning models.

best-practices
Mar 1, 2026
How d2l-zh Explains the Encoder-Decoder Architecture for Sequence-to-Sequence Tasks

Understand the encoder-decoder architecture for sequence-to-sequence tasks with d2l-zh. Learn how GRU-based recurrent networks implement this powerful framework for machine translation.

deep-dive
Mar 1, 2026
The Role and Impact of Batch Normalization in Model Training According to d2l-zh

Discover how batch normalization stabilizes hidden activations and accelerates deep neural network training according to d2l-zh. Learn its impact and role.

deep-dive
Mar 1, 2026
Preventing Overfitting in Deep Learning Models: Techniques from d2l-zh

Learn how to prevent overfitting in deep learning models with d2l-zh. Explore techniques like controlling model capacity, L₂ regularization, and dropout. Improve your model's generalization.

tutorial
Mar 1, 2026
How Word Embeddings Are Implemented and Utilized in d2l-zh

Discover how word embeddings are implemented and utilized in d2l-zh. Explore static pre-trained vectors like GloVe and fastText, and learn about learned embeddings within models like BERT.

how-to-guide
Mar 1, 2026
Differences Between LSTM and GRU in d2l-zh: Architectural and Performance Comparison

Explore LSTM vs GRU differences in d2l-zh. Understand how LSTM's three gates and cell state contrast with GRU's two gates for efficient long-term dependency management.

deep-dive
Mar 1, 2026
How d2l-zh Explains the Backpropagation Algorithm for Training Neural Networks

Learn how d2l-zh explains backpropagation using computational graphs and the chain rule for efficient neural network training. Understand gradient computation from output to input.

deep-dive
Mar 1, 2026

Have a question about this repo?

These articles cover the highlights, but your codebase questions are specific. Give your agent direct access to the source. Share this with your agent to get started:

Share the following with your agent to get started:
curl -s "https://instagit.com/install.md"

Works with
Claude Codex Cursor VS Code OpenClaw Any MCP Client

Maintain an open-source project? Get it listed too →