Token
وحدة لغوية (رمز)
الجزء النصي الأصغر (كلمة أو جزء من كلمة) الناتج عن تفكيك العبارات لتسهيل الحوسبة.
Token
Also translated asعنصر نصي مجزأ، الـمفردة الرقمية
First appears in this corpus in: The Monte Carlo Method (1949)
Appears in these papers
- Alpaca: A Strong, Replicable Instruction-Following Model2023in the sky ✦
- BART: Denoising Sequence-to-Sequence Pre-Training for Natural Language Generation, Translation, and Comprehension2019in the sky ✦
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding2018in the sky ✦
- BLOOM: A 176B-Parameter Open-Access Multilingual Language Model2022in the sky ✦
- Neural Machine Translation of Rare Words with Subword Units2016in the sky ✦
- Training Compute-Optimal Large Language Models2022in the sky ✦
- Chronos: Learning the Language of Time Series2024in the sky ✦
- Evaluating Large Language Models Trained on Code2021in the sky ✦
- Evaluating Large Language Models Trained on Code2021in the sky ✦
- ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT2020in the sky ✦
- Contriever: Unsupervised Dense Information Retrieval with Contrastive Learning2022in the sky ✦
- DALL·E: Zero-Shot Text-to-Image Generation2021in the sky ✦
- DALL·E: Zero-Shot Text-to-Image Generation2021in the sky ✦
- A Proposal for the Dartmouth Summer Research Project on Artificial Intelligence1956in the sky ✦
- DeBERTa: Decoding-Enhanced BERT with Disentangled Attention2020in the sky ✦
- DeBERTa: Decoding-Enhanced BERT with Disentangled Attention2020in the sky ✦
- Decision Transformer: Reinforcement Learning via Sequence Modeling2021in the sky ✦
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning2025in the sky ✦
- DeepSeek-V3 Technical Report2024in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- Are Transformers Effective for Time Series Forecasting?2023in the sky ✦
- ELIZA — A Computer Program for the Study of Natural Language Communication1966in the sky ✦
- Deep Contextualized Word Representations2018in the sky ✦
- Are Emergent Abilities of Large Language Models a Mirage?2023in the sky ✦
- The Falcon Series of Open Language Models2023in the sky ✦
- Flamingo: a Visual Language Model for Few-Shot Learning2022in the sky ✦
- Finetuned Language Models Are Zero-Shot Learners2022in the sky ✦
- Graph Attention Networks2018in the sky ✦
- A Generalist Agent2022in the sky ✦
- Gaussian Processes for Machine Learning2006in the sky ✦
- Gemini: A Family of Highly Capable Multimodal Models2023in the sky ✦
- Gemma: Open Models Based on Gemini Research and Technology2024in the sky ✦
- Generative Agents: Interactive Simulacra of Human Behavior2023in the sky ✦
- GloVe: Global Vectors for Word Representation2014in the sky ✦
- Improving Language Understanding by Generative Pre-Training2018in the sky ✦
- Language Models Are Unsupervised Multitask Learners2019in the sky ✦
- Language Models Are Few-Shot Learners2020in the sky ✦
- GPT-4 Technical Report2023in the sky ✦
- Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection2023in the sky ✦
- Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection2023in the sky ✦
- Imagen: Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding2022in the sky ✦
- Jukebox: A Generative Model for Music2020in the sky ✦
- High-Resolution Image Synthesis with Latent Diffusion Models2022in the sky ✦
- Layer Normalization2016in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- LIMA: Less Is More for Alignment2023in the sky ✦
- Llama 2: Open Foundation and Fine-Tuned Chat Models2023in the sky ✦
- The Llama 3 Herd of Models2024in the sky ✦
- LLaMA: Open and Efficient Foundation Language Models2023in the sky ✦
- Visual Instruction Tuning2023in the sky ✦
- LLM.int8(): 8-Bit Matrix Multiplication for Transformers at Scale2022in the sky ✦
- LoRA: Low-Rank Adaptation of Large Language Models2021in the sky ✦
- Masked Autoencoders Are Scalable Vision Learners2022in the sky ✦
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces2023in the sky ✦
- Mistral 7B2023in the sky ✦
- Mixtral of Experts2024in the sky ✦
- Mixtral of Experts2024in the sky ✦
- The Monte Carlo Method1949in the sky ✦
- MusicLM: Generating Music From Text2023in the sky ✦
- Learning to Reason with LLMs2024in the sky ✦
- OPT: Open Pre-Trained Transformer Language Models2022in the sky ✦
- PaLM: Scaling Language Modeling with Pathways2022in the sky ✦
- A Time Series Is Worth 64 Words: Long-Term Forecasting with Transformers2023in the sky ✦
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks2020in the sky ✦
- REALM: Retrieval-Augmented Language Model Pre-Training2020in the sky ✦
- Learning Phrase Representations Using RNN Encoder-Decoder for Statistical Machine Translation2014in the sky ✦
- RoBERTa: A Robustly Optimized BERT Pretraining Approach2019in the sky ✦
- RT-1: Robotics Transformer for Real-World Control at Scale2022in the sky ✦
- RT-1: Robotics Transformer for Real-World Control at Scale2022in the sky ✦
- RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control2023in the sky ✦
- Scaling Laws for Neural Language Models2020in the sky ✦
- Self-Consistency Improves Chain of Thought Reasoning in Language Models2022in the sky ✦
- Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection2023in the sky ✦
- SGDR: Stochastic Gradient Descent with Warm Restarts2017in the sky ✦
- Video Generation Models as World Simulators2024in the sky ✦
- Sparks of Artificial General Intelligence: Early Experiments with GPT-42023in the sky ✦
- Fast Inference from Transformers via Speculative Decoding2023in the sky ✦
- Fast Inference from Transformers via Speculative Decoding2023in the sky ✦
- SQuAD: 100,000+ Questions for Machine Comprehension of Text2016in the sky ✦
- Swin Transformer: Hierarchical Vision Transformer Using Shifted Windows2021in the sky ✦
- Swin Transformer: Hierarchical Vision Transformer Using Shifted Windows2021in the sky ✦
- Switch Transformers: Scaling to Trillion Parameter Models with Simple and Efficient Sparsity2022in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- Toolformer: Language Models Can Teach Themselves to Use Tools2023in the sky ✦
- Toolformer: Language Models Can Teach Themselves to Use Tools2023in the sky ✦
- Tree of Thoughts: Deliberate Problem Solving with Large Language Models2023in the sky ✦
- An Image Is Worth 16×16 Words: Transformers for Image Recognition at Scale2020in the sky ✦
- WebGPT: Browser-Assisted Question-Answering with Human Feedback2021in the sky ✦
- Robust Speech Recognition via Large-Scale Weak Supervision2022in the sky ✦
- Distributed Representations of Words and Phrases and Their Compositionality2013in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦