Parameter
المعلمة البنيوية
المتغيرات الداخلية (كالـلأوزان والانحيازات) التي يتعلمها النموذج ويعدلها تلقائياً أثناء التدريب.
Parameter
Also translated asالمتغير الداخلي للنموذج، الخاصية القابلة للتعلم، معلمة بنيوية
First appears in this corpus in: Nouvelles méthodes pour la détermination des orbites des comètes (Method of Least Squares) (1805)
Appears in these papers
- Adaptive Subgradient Methods for Online Learning and Stochastic Optimization2011in the sky ✦
- Adaptive Subgradient Methods for Online Learning and Stochastic Optimization2011in the sky ✦
- Adam: A Method for Stochastic Optimization2014in the sky ✦
- ImageNet Classification with Deep Convolutional Neural Networks2012in the sky ✦
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding2018in the sky ✦
- BPR: Bayesian Personalized Ranking from Implicit Feedback2009in the sky ✦
- Dynamic Routing Between Capsules2017in the sky ✦
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models2022in the sky ✦
- Training Compute-Optimal Large Language Models2022in the sky ✦
- Learning Transferable Visual Models from Natural Language Supervision2021in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Unsupervised Visual Representation Learning by Context Prediction2015in the sky ✦
- Gradient-Based Learning Applied to Document Recognition1998in the sky ✦
- Cyclical Learning Rates for Training Neural Networks2017in the sky ✦
- A Proposal for the Dartmouth Summer Research Project on Artificial Intelligence1956in the sky ✦
- Unsupervised Representation Learning with Deep Convolutional Generative Adversarial Networks2015in the sky ✦
- Continuous Control with Deep Reinforcement Learning2015in the sky ✦
- A Fast Learning Algorithm for Deep Belief Nets2006in the sky ✦
- Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding2016in the sky ✦
- Deep Learning2015in the sky ✦
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin2015in the sky ✦
- DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs2017in the sky ✦
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning2025in the sky ✦
- Deformable Convolutional Networks2017in the sky ✦
- Densely Connected Convolutional Networks2017in the sky ✦
- Deep Interest Network for Click-Through Rate Prediction2018in the sky ✦
- DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter2019in the sky ✦
- Are Transformers Effective for Time Series Forecasting?2023in the sky ✦
- Domain Randomization for Transferring Deep Neural Networks from Simulation to the Real World2017in the sky ✦
- Reconciling Modern Machine-Learning Practice and the Classical Bias–Variance Trade-Off2019in the sky ✦
- Dropout: A Simple Way to Prevent Neural Networks from Overfitting2014in the sky ✦
- ELIZA — A Computer Program for the Study of Natural Language Communication1966in the sky ✦
- Deep Contextualized Word Representations2018in the sky ✦
- Maximum Likelihood from Incomplete Data via the EM Algorithm1977in the sky ✦
- FaceNet: A Unified Embedding for Face Recognition and Clustering2015in the sky ✦
- Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks2015in the sky ✦
- Enriching Word Vectors with Subword Information2017in the sky ✦
- Finetuned Language Models Are Zero-Shot Learners2022in the sky ✦
- Graph Attention Networks2018in the sky ✦
- Gaussian Processes for Machine Learning2006in the sky ✦
- GloVe: Global Vectors for Word Representation2014in the sky ✦
- Glow: Generative Flow with Invertible 1×1 Convolutions2018in the sky ✦
- Google's Neural Machine Translation System: Bridging the Gap Between Human and Machine Translation2016in the sky ✦
- The Graph Neural Network Model2009in the sky ✦
- Going Deeper with Convolutions2014in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- Language Models Are Few-Shot Learners2020in the sky ✦
- GPTQ: Accurate Post-Training Quantization for Generative Pre-Trained Transformers2022in the sky ✦
- Grad-CAM: Visual Explanations from Deep Networks via Gradient-Based Localization2017in the sky ✦
- On the Difficulty of Training Recurrent Neural Networks2013in the sky ✦
- Méthode Générale pour la Résolution des Systèmes d'Équations Simultanées1847in the sky ✦
- Inductive Representation Learning on Large Graphs2017in the sky ✦
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling2014in the sky ✦
- Delving Deep into Rectifiers: Surpassing Human-Level Performance on ImageNet Classification2015in the sky ✦
- Highway Networks2015in the sky ✦
- A Tutorial on Hidden Markov Models and Selected Applications in Speech Recognition1989in the sky ✦
- Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset2017in the sky ✦
- The Mathematics of Statistical Machine Translation: Parameter Estimation1993in the sky ✦
- IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures2018in the sky ✦
- Training Language Models to Follow Instructions with Human Feedback2022in the sky ✦
- A New Approach to Linear Filtering and Prediction Problems1960in the sky ✦
- Optimizing Neural Networks with Kronecker-Factored Approximate Curvature2015in the sky ✦
- Distilling the Knowledge in a Neural Network2015in the sky ✦
- Regression Shrinkage and Selection via the Lasso1996in the sky ✦
- Latent Dirichlet Allocation2003in the sky ✦
- Nouvelles méthodes pour la détermination des orbites des comètes (Method of Least Squares)1805in the sky ✦
- LLaMA: Open and Efficient Foundation Language Models2023in the sky ✦
- Visual Instruction Tuning2023in the sky ✦
- LLM.int8(): 8-Bit Matrix Multiplication for Transformers at Scale2022in the sky ✦
- The Regression Analysis of Binary Sequences1958in the sky ✦
- LoRA: Low-Rank Adaptation of Large Language Models2021in the sky ✦
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces2023in the sky ✦
- A Logical Calculus of the Ideas Immanent in Nervous Activity1943in the sky ✦
- Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism2019in the sky ✦
- Mistral 7B2023in the sky ✦
- Mixtral of Experts2024in the sky ✦
- On the Mathematical Foundations of Theoretical Statistics1922in the sky ✦
- Some Methods of Speeding Up the Convergence of Iteration Methods1964in the sky ✦
- Natural Gradient Works Efficiently in Learning1998in the sky ✦
- Nearest Neighbor Pattern Classification1967in the sky ✦
- Neocognitron: A Self-Organizing Neural Network Model for Pattern Recognition1980in the sky ✦
- A Method for Solving the Convex Programming Problem with Convergence Rate O(1/k²)1983in the sky ✦
- A Neural Probabilistic Language Model2003in the sky ✦
- Neural Tangent Kernel: Convergence and Generalization in Neural Networks2018in the sky ✦
- Neural Tangent Kernel: Convergence and Generalization in Neural Networks2018in the sky ✦
- Optimal Brain Damage1989in the sky ✦
- PaLM: Scaling Language Modeling with Pathways2022in the sky ✦
- Policy Gradient Methods for Reinforcement Learning with Function Approximation1999in the sky ✦
- Proximal Policy Optimization Algorithms2017in the sky ✦
- QLoRA: Efficient Finetuning of Quantized LLMs2023in the sky ✦
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks2020in the sky ✦
- Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned2022in the sky ✦
- Deep Residual Learning for Image Recognition2015in the sky ✦
- RMSProp: Divide the Gradient by a Running Average of Its Recent Magnitude2012in the sky ✦
- Learning Phrase Representations Using RNN Encoder-Decoder for Statistical Machine Translation2014in the sky ✦
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control2023in the sky ✦
- Scaling Laws for Neural Language Models2020in the sky ✦
- Sequence to Sequence Learning with Neural Networks2014in the sky ✦
- A Stochastic Approximation Method1951in the sky ✦
- Sharpness-Aware Minimization for Efficiently Improving Generalization2021in the sky ✦
- Signature Verification Using a "Siamese" Time Delay Neural Network1993in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- Universal Language Model Fine-Tuning for Text Classification2018in the sky ✦
- On the Uniform Convergence of Relative Frequencies of Events to Their Probabilities1971in the sky ✦
- An Image Is Worth 16×16 Words: Transformers for Image Recognition at Scale2020in the sky ✦
- Robust Speech Recognition via Large-Scale Weak Supervision2022in the sky ✦