Hidden State
الحالة المخفية
تمثيل داخلي يحتفظ به النموذج الارتجاعي عبر الخطوات الزمنية، يضغط فيه المعلومات من المُدخلات السابقة. في مامبا، الحالة المخفية ثابتة الحجم بصرف النظر عن طول التسلسل.
An internal representation maintained by a recurrent model across time steps, compressing information from past inputs. In Mamba, the hidden state is fixed-size regardless of sequence length.
Also translated asالحالة الكامنة، المتغير المستتر، المتغيّر الكامن، الوضع الداخلي الكامن
First appears in this corpus in: Extension of the Law of Large Numbers to Dependent Quantities (1906)
Appears in these papers
- Neural Machine Translation by Jointly Learning to Align and Translate2014in the sky ✦
- Backpropagation Through Time: What It Does and How to Do It1990in the sky ✦
- Learning Long-Term Dependencies with Gradient Descent is Difficult1994in the sky ✦
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding2018in the sky ✦
- A Fast Learning Algorithm for Deep Belief Nets2006in the sky ✦
- Deep Learning2015in the sky ✦
- Deep Speech 2: End-to-End Speech Recognition in English and Mandarin2015in the sky ✦
- DeepAR: Probabilistic Forecasting with Autoregressive Recurrent Networks2020in the sky ✦
- DeepAR: Probabilistic Forecasting with Autoregressive Recurrent Networks2020in the sky ✦
- Extracting and Composing Robust Features with Denoising Autoencoders2008in the sky ✦
- DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter2019in the sky ✦
- DistilBERT, a Distilled Version of BERT: Smaller, Faster, Cheaper and Lighter2019in the sky ✦
- Mastering Diverse Domains Through World Models2023in the sky ✦
- Mastering Diverse Domains Through World Models2023in the sky ✦
- Maximum Likelihood from Incomplete Data via the EM Algorithm1977in the sky ✦
- Google's Neural Machine Translation System: Bridging the Gap Between Human and Machine Translation2016in the sky ✦
- The Graph Neural Network Model2009in the sky ✦
- Improving Language Understanding by Generative Pre-Training2018in the sky ✦
- On the Difficulty of Training Recurrent Neural Networks2013in the sky ✦
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling2014in the sky ✦
- Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling2014in the sky ✦
- A Tutorial on Hidden Markov Models and Selected Applications in Speech Recognition1989in the sky ✦
- A New Approach to Linear Filtering and Prediction Problems1960in the sky ✦
- A New Approach to Linear Filtering and Prediction Problems1960in the sky ✦
- Layer Normalization2016in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- LLM.int8(): 8-Bit Matrix Multiplication for Transformers at Scale2022in the sky ✦
- Long Short-Term Memory1997in the sky ✦
- Mamba: Linear-Time Sequence Modeling with Selective State Spaces2023in the sky ✦
- Extension of the Law of Large Numbers to Dependent Quantities1906in the sky ✦
- Mixtral of Experts2024in the sky ✦
- Neural Message Passing for Quantum Chemistry2017in the sky ✦
- MuZero: Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model2020in the sky ✦
- MuZero: Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model2020in the sky ✦
- Neural Ordinary Differential Equations2018in the sky ✦
- Neural Turing Machines2014in the sky ✦
- Learning Phrase Representations Using RNN Encoder-Decoder for Statistical Machine Translation2014in the sky ✦
- Learning Phrase Representations Using RNN Encoder-Decoder for Statistical Machine Translation2014in the sky ✦
- Recurrent Neural Networks (RNNs): A Gentle Introduction and Overview2019in the sky ✦
- Recurrent Neural Networks (RNNs): A Gentle Introduction and Overview2019in the sky ✦
- Efficiently Modeling Long Sequences with Structured State Spaces2022in the sky ✦
- Efficiently Modeling Long Sequences with Structured State Spaces2022in the sky ✦
- Sequence to Sequence Learning with Neural Networks2014in the sky ✦
- Sequence to Sequence Learning with Neural Networks2014in the sky ✦
- Show and Tell: A Neural Image Caption Generator2015in the sky ✦
- Show, Attend and Tell: Neural Image Caption Generation with Visual Attention2015in the sky ✦
- Universal Language Model Fine-Tuning for Text Classification2018in the sky ✦
- VQA: Visual Question Answering2015in the sky ✦
- World Models2018in the sky ✦
- World Models2018in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦