Papers tagged “pretraining”
17 papers · All papers →
-
Chameleon: Mixed-Modal Early-Fusion Foundation Models
-
LLaMA: Open and Efficient Foundation Language Models
-
Robust Speech Recognition via Large-Scale Weak Supervision
-
Emerging Properties in Self-Supervised Vision Transformers
-
Evaluating Large Language Models Trained on Code
-
Learning Transferable Visual Models from Natural Language Supervision
-
Masked Autoencoders Are Scalable Vision Learners
-
A Simple Framework for Contrastive Learning of Visual Representations
-
Bootstrap Your Own Latent: A New Approach to Self-Supervised Learning
-
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
-
Momentum Contrast for Unsupervised Visual Representation Learning
-
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
-
wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations
-
Language Models are Unsupervised Multitask Learners
-
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
-
Deep Contextualized Word Representations
-
Improving Language Understanding by Generative Pre-Training