Deep Learning: Advanced Backbones and Efficient GPU Training

External: Coursera Courses ↗ · Coursera

Open Course on External: Coursera

Free to audit · Opens on External: Coursera

Deep Learning: Advanced Backbones and Efficient GPU Training

Coursera · Advanced ·🧬 Deep Learning ·3mo ago
Skills: ML Pipelines80%

Key Takeaways

Master advanced deep learning architectures and efficient GPU training

Original Description

Master advanced deep learning architectures and efficient training techniques using PyTorch Lightning, timm, ConvNeXt, Vision Transformers, RoPE, SwiGLU, RMSNorm, and Weights & Biases. This course equips you to design, train, and benchmark modern backbones on limited GPU hardware for real-world production use. Module 1 introduces modern backbone architectures, tracing the evolution from ResNets to ConvNeXt and Vision Transformers, covering patch embeddings, multi-head self-attention, and position encodings. Module 2 dives into training dynamics and stabilization techniques including RMSNorm, SwiGLU activations, and Rotary Position Embeddings (RoPE) for stable, scalable training. Module 3 focuses on efficient training on limited GPUs using mixed precision (FP16/BF16), gradient accumulation, efficient data pipelines, and distributed training with DDP/FSDP in Lightning. Module 4 covers experiment tracking with TensorBoard and W&B, profiling FLOPs and throughput, and a hands-on ViT vs. CNN Showdown project with fine-tuning in timm. By the end of this course, you will: - Build and fine-tune ConvNeXt and Vision Transformer backbones using PyTorch Lightning and timm - Apply RMSNorm, SwiGLU, and RoPE to stabilize and scale deep transformer training - Implement mixed precision, gradient accumulation, and DDP/FSDP for efficient multi-GPU training - Design controlled CNN vs. ViT experiments with W&B tracking and PyTorch profiling Disclaimer: This is an independent educational resource created by Board Infinity for informational and educational purposes only. This course is not affiliated with, endorsed by, sponsored by, or officially associated with any company, organization, or certification body unless explicitly stated. The content provided is based on industry knowledge and best practices but does not constitute official training material for any specific employer or certification program. All company names, trademarks, service marks, and logos referenced are the pro
AI explanation not available for this lesson yet
This lesson is still being prepared for the AI tutor. In the meantime, explore lessons that are ready.
Browse explainer-ready lessons →

Related Reads

📰
Trained a neural net to reconstruct Bad Apple in real-time.
Reconstruct Bad Apple in real-time using a trained neural network and learn how to apply deep learning to video processing
Reddit r/deeplearning
📰
AI/ML Under the Hood — Part 29: CNN Breaking News: Proximity Matters
Learn how proximity affects CNNs with kernels, feature maps, padding, and strides
Medium · Deep Learning
📰
Deep Learning Scientists — Claude Cowork: The Deep Learning Scientist’s New Lab Partner
Meet Claude Cowork, a new tool for deep learning scientists to optimize their workflow and reduce the scarcity of compute and attention resources
Medium · Data Science
📰
Why Qwen3.8 27B Looked Brilliant in Testing but Failed to Ship My AI Newspaper
Learn why a high-performing AI model like Qwen3.8 27B failed to deliver in real-world application and how to avoid similar pitfalls
Medium · Deep Learning
Up next
Machine Learning Rust Candle Hugging Face Part 4
Stephen Blum
Watch →