Skip to main navigation Skip to search Skip to main content

Fundamentals of Deep Learning

  • School of Physics, Harbin Institute of Technology

Research output: Chapter in Book/Report/Conference proceedingChapterpeer-review

Abstract

This chapter offers a comprehensive exposition of the fundamental theories and key techniques of deep learning, encompassing its core concepts, foundational models, learning objectives, optimization strategies, and representative architectures. It begins by introducing the historical context and the central issue of contribution allocation, using artificial neural networks to illustrate the information flow and their advantages in representation learning over traditional methods. The training pipeline is then examined in detail, including model formulation, loss function design, gradient-based optimization, regularization, and early stopping, providing a unified view of deep learning from both theoretical and practical perspectives. Structurally, the chapter highlights the principles of local connectivity, weight sharing, and pooling in convolutional neural networks, with classic architectures used to illustrate their evolution. Finally, this chapter introduces variational autoencoders and generative adversarial networks, discussing their modeling frameworks and application scenarios.

Original languageEnglish
Title of host publicationStudies in Computational Intelligence
PublisherSpringer Science and Business Media Deutschland GmbH
Pages1-31
Number of pages31
DOIs
StatePublished - 2026
Externally publishedYes

Publication series

NameStudies in Computational Intelligence
Volume1246
ISSN (Print)1860-949X
ISSN (Electronic)1860-9503

Keywords

  • Convolutional neural networks
  • Deep learning
  • Generative models
  • Neural networks
  • Optimization algorithms

Fingerprint

Dive into the research topics of 'Fundamentals of Deep Learning'. Together they form a unique fingerprint.

Cite this