Fine-Tuning
الضبط الدقيق
تكييف نموذج مُدرَّب مسبقاً مع مهمة لاحقة محددة بتدريب (كل أو بعض) معاملاته على بيانات مُعنونة خاصة بالمهمة، مما يُتيح نقل المعرفة المُكتسبة من التدريب المسبق.
Adapting a pre-trained model to a specific downstream task by training (all or some of) its parameters on task-specific labeled data, enabling transfer of knowledge acquired during pre-training.
Also translated asإعادة التدريب الموجّه، التدريب التكميلي، التدريب الموجَّه اللاحق، التعديل التفصيلي للأوزان، التعديل الدقيق، التنغيم الدقيق، التنقيح، التنقيح الدقيق، التوليف الدقيق، الضبط التفصيلي، الضبط الموجَّه، الضبط المُوجَّه، المعايرة الدقيقة، المواءمة التخصصية
First appears in this corpus in: An Essay Towards Solving a Problem in the Doctrine of Chances (1763)
Appears in these papers
- Parameter-Efficient Transfer Learning for NLP2019in the sky ✦
- Parameter-Efficient Transfer Learning for NLP2019in the sky ✦
- ALIGN: Scaling Up Visual and Vision-Language Representation Learning with Noisy Text Supervision2021in the sky ✦
- ALIGN: Scaling Up Visual and Vision-Language Representation Learning with Noisy Text Supervision2021in the sky ✦
- Alpaca: A Strong, Replicable Instruction-Following Model2023in the sky ✦
- Alpaca: A Strong, Replicable Instruction-Following Model2023in the sky ✦
- Atlas: Few-shot Learning with Retrieval Augmented Language Models2023in the sky ✦
- Atlas: Few-shot Learning with Retrieval Augmented Language Models2023in the sky ✦
- BART: Denoising Sequence-to-Sequence Pre-Training for Natural Language Generation, Translation, and Comprehension2019in the sky ✦
- An Essay Towards Solving a Problem in the Doctrine of Chances1763in the sky ✦
- BEiT: BERT Pre-Training of Image Transformers2021in the sky ✦
- BEiT: BERT Pre-Training of Image Transformers2021in the sky ✦
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding2018in the sky ✦
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding2018in the sky ✦
- BLIP-2: Bootstrapping Language-Image Pre-Training with Frozen Image Encoders and Large Language Models2023in the sky ✦
- BLIP-2: Bootstrapping Language-Image Pre-Training with Frozen Image Encoders and Large Language Models2023in the sky ✦
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models2022in the sky ✦
- Chronos: Learning the Language of Time Series2024in the sky ✦
- Learning Transferable Visual Models from Natural Language Supervision2021in the sky ✦
- Evaluating Large Language Models Trained on Code2021in the sky ✦
- Evaluating Large Language Models Trained on Code2021in the sky ✦
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT2020in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Unsupervised Visual Representation Learning by Context Prediction2015in the sky ✦
- Unsupervised Visual Representation Learning by Context Prediction2015in the sky ✦
- Contriever: Unsupervised Dense Information Retrieval with Contrastive Learning2022in the sky ✦
- Contriever: Unsupervised Dense Information Retrieval with Contrastive Learning2022in the sky ✦
- A ConvNet for the 2020s2022in the sky ✦
- DeBERTa: Decoding-Enhanced BERT with Disentangled Attention2020in the sky ✦
- Reducing the Dimensionality of Data with Neural Networks2006in the sky ✦
- Reducing the Dimensionality of Data with Neural Networks2006in the sky ✦
- A Fast Learning Algorithm for Deep Belief Nets2006in the sky ✦
- A Fast Learning Algorithm for Deep Belief Nets2006in the sky ✦
- Deep Compression: Compressing Deep Neural Networks with Pruning, Trained Quantization and Huffman Coding2016in the sky ✦
- Deep Learning2015in the sky ✦
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning2025in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- Extracting and Composing Robust Features with Denoising Autoencoders2008in the sky ✦
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data2024in the sky ✦
- Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data2024in the sky ✦
- Emerging Properties in Self-Supervised Vision Transformers2021in the sky ✦
- DINOv2: Learning Robust Visual Features Without Supervision2023in the sky ✦
- DINOv2: Learning Robust Visual Features Without Supervision2023in the sky ✦
- DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter2019in the sky ✦
- DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter2019in the sky ✦
- ELECTRA: Pre-Training Text Encoders as Discriminators Rather Than Generators2020in the sky ✦
- Deep Contextualized Word Representations2018in the sky ✦
- Emergent Abilities of Large Language Models2022in the sky ✦
- The Falcon Series of Open Language Models2023in the sky ✦
- Fast R-CNN2015in the sky ✦
- Fast R-CNN2015in the sky ✦
- Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks2015in the sky ✦
- Fully Convolutional Networks for Semantic Segmentation2015in the sky ✦
- Fusion-in-Decoder: Leveraging Passage Retrieval with Generative Models for Open Domain Question Answering2021in the sky ✦
- Fusion-in-Decoder: Leveraging Passage Retrieval with Generative Models for Open Domain Question Answering2021in the sky ✦
- Flamingo: a Visual Language Model for Few-Shot Learning2022in the sky ✦
- Flamingo: a Visual Language Model for Few-Shot Learning2022in the sky ✦
- Finetuned Language Models Are Zero-Shot Learners2022in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- A Generalist Agent2022in the sky ✦
- Gemma: Open Models Based on Gemini Research and Technology2024in the sky ✦
- GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding2018in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- Improving Language Understanding by Generative Pre-Training2018in the sky ✦
- Language Models Are Unsupervised Multitask Learners2019in the sky ✦
- Language Models Are Few-Shot Learners2020in the sky ✦
- GPT-4 Technical Report2023in the sky ✦
- GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints2023in the sky ✦
- Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection2023in the sky ✦
- Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection2023in the sky ✦
- Group Normalization2018in the sky ✦
- Hindsight Experience Replay2017in the sky ✦
- HuggingGPT: Solving AI Tasks with ChatGPT and Its Friends in Hugging Face2023in the sky ✦
- HuggingGPT: Solving AI Tasks with ChatGPT and Its Friends in Hugging Face2023in the sky ✦
- HyDE: Precise Zero-Shot Dense Retrieval without Relevance Labels2022in the sky ✦
- Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset2017in the sky ✦
- Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset2017in the sky ✦
- Curiosity-Driven Exploration by Self-Supervised Prediction2017in the sky ✦
- ImageNet: A Large-Scale Hierarchical Image Database2009in the sky ✦
- Large Batch Optimization for Deep Learning: Training BERT in 76 Minutes2019in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- LIMA: Less Is More for Alignment2023in the sky ✦
- LIMA: Less Is More for Alignment2023in the sky ✦
- Symbolic Discovery of Optimization Algorithms2023in the sky ✦
- Symbolic Discovery of Optimization Algorithms2023in the sky ✦
- LLaMA: Open and Efficient Foundation Language Models2023in the sky ✦
- Visual Instruction Tuning2023in the sky ✦
- LLM.int8(): 8-Bit Matrix Multiplication for Transformers at Scale2022in the sky ✦
- The Lottery Ticket Hypothesis: Finding Sparse, Trainable Neural Networks2019in the sky ✦
- Masked Autoencoders Are Scalable Vision Learners2022in the sky ✦
- Mistral 7B2023in the sky ✦
- Mistral 7B2023in the sky ✦
- Momentum Contrast for Unsupervised Visual Representation Learning2020in the sky ✦
- Neural Collaborative Filtering2017in the sky ✦
- A Method for Solving the Convex Programming Problem with Convergence Rate O(1/k²)1983in the sky ✦
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2024in the sky ✦
- Open X-Embodiment: Robotic Learning Datasets and RT-X Models2024in the sky ✦
- OPT: Open Pre-Trained Transformer Language Models2022in the sky ✦
- OPT: Open Pre-Trained Transformer Language Models2022in the sky ✦
- PaLM: Scaling Language Modeling with Pathways2022in the sky ✦
- PaLM: Scaling Language Modeling with Pathways2022in the sky ✦
- A Time Series Is Worth 64 Words: Long-Term Forecasting with Transformers2023in the sky ✦
- A Time Series Is Worth 64 Words: Long-Term Forecasting with Transformers2023in the sky ✦
- Textbooks Are All You Need2023in the sky ✦
- π₀: A Vision-Language-Action Flow Model for General Robot Control2024in the sky ✦
- π₀: A Vision-Language-Action Flow Model for General Robot Control2024in the sky ✦
- The Power of Scale for Parameter-Efficient Prompt Tuning2021in the sky ✦
- QLoRA: Efficient Finetuning of Quantized LLMs2023in the sky ✦
- Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation2014in the sky ✦
- Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation2014in the sky ✦
- ReAct: Synergizing Reasoning and Acting in Language Models2023in the sky ✦
- ReAct: Synergizing Reasoning and Acting in Language Models2023in the sky ✦
- REALM: Retrieval-Augmented Language Model Pre-Training2020in the sky ✦
- REALM: Retrieval-Augmented Language Model Pre-Training2020in the sky ✦
- Reflexion: Language Agents with Verbal Reinforcement Learning2023in the sky ✦
- RETRO: Improving Language Models by Retrieving from Trillions of Tokens2022in the sky ✦
- RoBERTa: A Robustly Optimized BERT Pretraining Approach2019in the sky ✦
- RoBERTa: A Robustly Optimized BERT Pretraining Approach2019in the sky ✦
- Unsupervised Representation Learning by Predicting Image Rotations2018in the sky ✦
- Unsupervised Representation Learning by Predicting Image Rotations2018in the sky ✦
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control2023in the sky ✦
- Segment Anything2023in the sky ✦
- Self-Consistency Improves Chain of Thought Reasoning in Language Models2022in the sky ✦
- Self-Instruct: Aligning Language Models with Self-Generated Instructions2022in the sky ✦
- Self-Instruct: Aligning Language Models with Self-Generated Instructions2022in the sky ✦
- Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection2023in the sky ✦
- Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection2023in the sky ✦
- Sentence-BERT: Sentence Embeddings Using Siamese BERT-Networks2019in the sky ✦
- Sentence-BERT: Sentence Embeddings Using Siamese BERT-Networks2019in the sky ✦
- Sharpness-Aware Minimization for Efficiently Improving Generalization2021in the sky ✦
- Show and Tell: A Neural Image Caption Generator2015in the sky ✦
- A Simple Framework for Contrastive Learning of Visual Representations2020in the sky ✦
- SQuAD: 100,000+ Questions for Machine Comprehension of Text2016in the sky ✦
- STaR: Bootstrapping Reasoning With Reasoning2022in the sky ✦
- SWE-bench: Can Language Models Resolve Real-World GitHub Issues?2024in the sky ✦
- SWE-bench: Can Language Models Resolve Real-World GitHub Issues?2024in the sky ✦
- Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity2022in the sky ✦
- Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity2022in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- Toolformer: Language Models Can Teach Themselves to Use Tools2023in the sky ✦
- Toolformer: Language Models Can Teach Themselves to Use Tools2023in the sky ✦
- Universal Language Model Fine-Tuning for Text Classification2018in the sky ✦
- Universal Language Model Fine-Tuning for Text Classification2018in the sky ✦
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations2020in the sky ✦
- wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations2020in the sky ✦
- Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision2023in the sky ✦
- Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision2023in the sky ✦
- WebGPT: Browser-Assisted Question-Answering with Human Feedback2021in the sky ✦
- WebGPT: Browser-Assisted Question-Answering with Human Feedback2021in the sky ✦
- Robust Speech Recognition via Large-Scale Weak Supervision2022in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦
- You Only Look Once: Unified, Real-Time Object Detection2015in the sky ✦