Loss
الفقد
القيمة الرقمية التي تحدد مدى انحراف وفارق تنبؤات النموذج عن القيم الحقيقية المرجعية.
Loss
Also translated asقيمة الخطأ الإجمالي، الخسارة الرقمية
First appears in this corpus in: Nouvelles méthodes pour la détermination des orbites des comètes (Method of Least Squares) (1805)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Adaptive Subgradient Methods for Online Learning and Stochastic Optimization2011in the sky ✦
- Adaptive Subgradient Methods for Online Learning and Stochastic Optimization2011in the sky ✦
- Adam: A Method for Stochastic Optimization2014in the sky ✦
- Highly Accurate Protein Structure Prediction with AlphaFold2021in the sky ✦
- Mastering the Game of Go Without Human Knowledge2017in the sky ✦
- beta-VAE: Learning Basic Visual Concepts with a Constrained Variational Framework2017in the sky ✦
- BPR: Bayesian Personalized Ranking from Implicit Feedback2009in the sky ✦
- Dynamic Routing Between Capsules2017in the sky ✦
- Training Compute-Optimal Large Language Models2022in the sky ✦
- Learning Transferable Visual Models from Natural Language Supervision2021in the sky ✦
- Representation Learning with Contrastive Predictive Coding2018in the sky ✦
- Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks2017in the sky ✦
- Cyclical Learning Rates for Training Neural Networks2017in the sky ✦
- Cyclical Learning Rates for Training Neural Networks2017in the sky ✦
- Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks2015in the sky ✦
- Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding2016in the sky ✦
- Deep Learning2015in the sky ✦
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin2015in the sky ✦
- Deformable Convolutional Networks2017in the sky ✦
- Extracting and Composing Robust Features with Denoising Autoencoders2008in the sky ✦
- Densely Connected Convolutional Networks2017in the sky ✦
- Deep Interest Network for Click-Through Rate Prediction2018in the sky ✦
- Dueling Network Architectures for Deep Reinforcement Learning2016in the sky ✦
- FaceNet: A Unified Embedding for Face Recognition and Clustering2015in the sky ✦
- Fast R-CNN2015in the sky ✦
- Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks2015in the sky ✦
- Fully Convolutional Networks for Semantic Segmentation2015in the sky ✦
- Explaining and Harnessing Adversarial Examples2015in the sky ✦
- Semi-Supervised Classification with Graph Convolutional Networks2017in the sky ✦
- Gaussian Error Linear Units (GELUs)2016in the sky ✦
- GloVe: Global Vectors for Word Representation2014in the sky ✦
- Glow: Generative Flow with Invertible 1×1 Convolutions2018in the sky ✦
- Google's Neural Machine Translation System: Bridging the Gap Between Human and Machine Translation2016in the sky ✦
- The Graph Neural Network Model2009in the sky ✦
- Going Deeper with Convolutions2014in the sky ✦
- Improving Language Understanding by Generative Pre-Training2018in the sky ✦
- Language Models Are Unsupervised Multitask Learners2019in the sky ✦
- Language Models Are Few-Shot Learners2020in the sky ✦
- GPT-4 Technical Report2023in the sky ✦
- On the Difficulty of Training Recurrent Neural Networks2013in the sky ✦
- Inductive Representation Learning on Large Graphs2017in the sky ✦
- Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets2022in the sky ✦
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling2014in the sky ✦
- Highway Networks2015in the sky ✦
- Hyperband: A Novel Bandit-Based Approach to Hyperparameter Optimization2018in the sky ✦
- Curiosity-Driven Exploration by Self-Supervised Prediction2017in the sky ✦
- Training Language Models to Follow Instructions with Human Feedback2022in the sky ✦
- On Information and Sufficiency1951in the sky ✦
- Distilling the Knowledge in a Neural Network2015in the sky ✦
- Regression Shrinkage and Selection via the Lasso1996in the sky ✦
- Layer Normalization2016in the sky ✦
- Nouvelles méthodes pour la détermination des orbites des comètes (Method of Least Squares)1805in the sky ✦
- LightGBM: A Highly Efficient Gradient Boosting Decision Tree2017in the sky ✦
- LLaMA: Open and Efficient Foundation Language Models2023in the sky ✦
- Visualizing the Loss Landscape of Neural Nets2018in the sky ✦
- Visualizing the Loss Landscape of Neural Nets2018in the sky ✦
- Masked Autoencoders Are Scalable Vision Learners2022in the sky ✦
- Mask R-CNN2017in the sky ✦
- Some Methods of Speeding Up the Convergence of Iteration Methods1964in the sky ✦
- Natural Gradient Works Efficiently in Learning1998in the sky ✦
- NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis2020in the sky ✦
- A Method for Solving the Convex Programming Problem with Convergence Rate O(1/k²)1983in the sky ✦
- A Neural Probabilistic Language Model2003in the sky ✦
- OPT: Open Pre-Trained Transformer Language Models2022in the sky ✦
- Optimal Brain Damage1989in the sky ✦
- Proximal Policy Optimization Algorithms2017in the sky ✦
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks2020in the sky ✦
- Ridge Regression: Biased Estimation for Nonorthogonal Problems1970in the sky ✦
- RMSProp: Divide the Gradient by a Running Average of Its Recent Magnitude2012in the sky ✦
- Scaling Laws for Neural Language Models2020in the sky ✦
- Segment Anything2023in the sky ✦
- A Stochastic Approximation Method1951in the sky ✦
- Sharpness-Aware Minimization for Efficiently Improving Generalization2021in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- Auto-Encoding Variational Bayes2013in the sky ✦
- Wasserstein GAN2017in the sky ✦