Stochastic Gradient Descent (SGD)
الانحدار التدريجي العشوائي
نسخة سريعة من الانحدار التدريجي تحدّث الأوزان بالاعتماد على عينة عشوائية واحدة في كل خطوة.
Stochastic Gradient Descent (SGD)
Also translated asالهبوط الاحتمالي مع التدرج، تحسين الميل العشوائي
First appears in this corpus in: Méthode Générale pour la Résolution des Systèmes d'Équations Simultanées (1847)
Appears in these papers
- Adaptive Subgradient Methods for Online Learning and Stochastic Optimization2011in the sky ✦
- Adam: A Method for Stochastic Optimization2014in the sky ✦
- Decoupled Weight Decay Regularization2019in the sky ✦
- ImageNet Classification with Deep Convolutional Neural Networks2012in the sky ✦
- Neural Networks and the Bias/Variance Dilemma1992in the sky ✦
- BPR: Bayesian Personalized Ranking from Implicit Feedback2009in the sky ✦
- Unsupervised Visual Representation Learning by Context Prediction2015in the sky ✦
- A ConvNet for the 2020s2022in the sky ✦
- Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data2001in the sky ✦
- Cyclical Learning Rates for Training Neural Networks2017in the sky ✦
- Cyclical Learning Rates for Training Neural Networks2017in the sky ✦
- Deep Learning2015in the sky ✦
- Deep Learning2015in the sky ✦
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin2015in the sky ✦
- DeepWalk: Online Learning of Social Representations2014in the sky ✦
- Deep Interest Network for Click-Through Rate Prediction2018in the sky ✦
- FaceNet: A Unified Embedding for Face Recognition and Clustering2015in the sky ✦
- Enriching Word Vectors with Subword Information2017in the sky ✦
- Fully Convolutional Networks for Semantic Segmentation2015in the sky ✦
- 3D Gaussian Splatting for Real-Time Radiance Field Rendering2023in the sky ✦
- GloVe: Global Vectors for Word Representation2014in the sky ✦
- Greedy Function Approximation: A Gradient Boosting Machine2001in the sky ✦
- On the Difficulty of Training Recurrent Neural Networks2013in the sky ✦
- Méthode Générale pour la Résolution des Systèmes d'Équations Simultanées1847in the sky ✦
- Optimizing Neural Networks with Kronecker-Factored Approximate Curvature2015in the sky ✦
- Large Batch Optimization for Deep Learning: Training BERT in 76 Minutes2019in the sky ✦
- Large Batch Optimization for Deep Learning: Training BERT in 76 Minutes2019in the sky ✦
- Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour2017in the sky ✦
- Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour2017in the sky ✦
- Visualizing the Loss Landscape of Neural Nets2018in the sky ✦
- Visualizing the Loss Landscape of Neural Nets2018in the sky ✦
- Matrix Factorization Techniques for Recommender Systems2009in the sky ✦
- Some Methods of Speeding Up the Convergence of Iteration Methods1964in the sky ✦
- Some Methods of Speeding Up the Convergence of Iteration Methods1964in the sky ✦
- Natural Gradient Works Efficiently in Learning1998in the sky ✦
- Neural Collaborative Filtering2017in the sky ✦
- A Method for Solving the Convex Programming Problem with Convergence Rate O(1/k²)1983in the sky ✦
- A Neural Probabilistic Language Model2003in the sky ✦
- node2vec: Scalable Feature Learning for Networks2016in the sky ✦
- Policy Gradient Methods for Reinforcement Learning with Function Approximation1999in the sky ✦
- Proximal Policy Optimization Algorithms2017in the sky ✦
- QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation2018in the sky ✦
- Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation2014in the sky ✦
- RMSProp: Divide the Gradient by a Running Average of Its Recent Magnitude2012in the sky ✦
- RMSProp: Divide the Gradient by a Running Average of Its Recent Magnitude2012in the sky ✦
- A Stochastic Approximation Method1951in the sky ✦
- SGDR: Stochastic Gradient Descent with Warm Restarts2017in the sky ✦
- SGDR: Stochastic Gradient Descent with Warm Restarts2017in the sky ✦
- Shampoo: Preconditioned Stochastic Tensor Optimization2018in the sky ✦
- Sharpness-Aware Minimization for Efficiently Improving Generalization2021in the sky ✦
- Sharpness-Aware Minimization for Efficiently Improving Generalization2021in the sky ✦
- Exploring Simple Siamese Representation Learning2021in the sky ✦
- Visualizing Data Using t-SNE2008in the sky ✦
- UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction2018in the sky ✦
- UMAP: Uniform Manifold Approximation and Projection for Dimension Reduction2018in the sky ✦
- Auto-Encoding Variational Bayes2013in the sky ✦
- End-to-End Training of Deep Visuomotor Policies2016in the sky ✦
- Distributed Representations of Words and Phrases and Their Compositionality2013in the sky ✦