-
Neural Machine Translation by Jointly Learning to Align and Translate
Paper • 1409.0473 • Published • 7 -
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 26 -
Hierarchical Reasoning Model
Paper • 2506.21734 • Published • 48
Collections
Discover the best community collections!
Collections including paper arxiv:1409.0473
-
Recurrent Neural Network Regularization
Paper • 1409.2329 • Published • 1 -
Pointer Networks
Paper • 1506.03134 • Published • 1 -
Order Matters: Sequence to sequence for sets
Paper • 1511.06391 • Published • 1 -
GPipe: Efficient Training of Giant Neural Networks using Pipeline Parallelism
Paper • 1811.06965 • Published • 1
-
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
You Only Look Once: Unified, Real-Time Object Detection
Paper • 1506.02640 • Published • 3 -
HEp-2 Cell Image Classification with Deep Convolutional Neural Networks
Paper • 1504.02531 • Published -
Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training
Paper • 2401.05566 • Published • 31
-
Recurrent Neural Network Regularization
Paper • 1409.2329 • Published • 1 -
Pointer Networks
Paper • 1506.03134 • Published • 1 -
Order Matters: Sequence to sequence for sets
Paper • 1511.06391 • Published • 1 -
GPipe: Efficient Training of Giant Neural Networks using Pipeline Parallelism
Paper • 1811.06965 • Published • 1
-
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
ImageNet Large Scale Visual Recognition Challenge
Paper • 1409.0575 • Published • 10 -
Sequence to Sequence Learning with Neural Networks
Paper • 1409.3215 • Published • 3 -
Language Models are Few-Shot Learners
Paper • 2005.14165 • Published • 19
-
SMOTE: Synthetic Minority Over-sampling Technique
Paper • 1106.1813 • Published • 1 -
Scikit-learn: Machine Learning in Python
Paper • 1201.0490 • Published • 1 -
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Paper • 1406.1078 • Published • 1 -
Distributed Representations of Sentences and Documents
Paper • 1405.4053 • Published
-
Neural Machine Translation by Jointly Learning to Align and Translate
Paper • 1409.0473 • Published • 7 -
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 26 -
Hierarchical Reasoning Model
Paper • 2506.21734 • Published • 48
-
Recurrent Neural Network Regularization
Paper • 1409.2329 • Published • 1 -
Pointer Networks
Paper • 1506.03134 • Published • 1 -
Order Matters: Sequence to sequence for sets
Paper • 1511.06391 • Published • 1 -
GPipe: Efficient Training of Giant Neural Networks using Pipeline Parallelism
Paper • 1811.06965 • Published • 1
-
Recurrent Neural Network Regularization
Paper • 1409.2329 • Published • 1 -
Pointer Networks
Paper • 1506.03134 • Published • 1 -
Order Matters: Sequence to sequence for sets
Paper • 1511.06391 • Published • 1 -
GPipe: Efficient Training of Giant Neural Networks using Pipeline Parallelism
Paper • 1811.06965 • Published • 1
-
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
ImageNet Large Scale Visual Recognition Challenge
Paper • 1409.0575 • Published • 10 -
Sequence to Sequence Learning with Neural Networks
Paper • 1409.3215 • Published • 3 -
Language Models are Few-Shot Learners
Paper • 2005.14165 • Published • 19
-
Attention Is All You Need
Paper • 1706.03762 • Published • 115 -
You Only Look Once: Unified, Real-Time Object Detection
Paper • 1506.02640 • Published • 3 -
HEp-2 Cell Image Classification with Deep Convolutional Neural Networks
Paper • 1504.02531 • Published -
Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training
Paper • 2401.05566 • Published • 31
-
SMOTE: Synthetic Minority Over-sampling Technique
Paper • 1106.1813 • Published • 1 -
Scikit-learn: Machine Learning in Python
Paper • 1201.0490 • Published • 1 -
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Paper • 1406.1078 • Published • 1 -
Distributed Representations of Sentences and Documents
Paper • 1405.4053 • Published