-
Attention Is All You Need
Paper • 1706.03762 • Published • 141 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 33 -
Universal Language Model Fine-tuning for Text Classification
Paper • 1801.06146 • Published • 8 -
Language Models are Few-Shot Learners
Paper • 2005.14165 • Published • 21
Effi PRO
itseffi
AI & ML interests
None yet
Organizations
LLMs
-
meta-llama/Llama-2-70b-hf
Text Generation • 69B • Updated • 3.46k • 855 -
tiiuae/falcon-180B
Text Generation • 180B • Updated • 72 • 1.15k -
meta-llama/Meta-Llama-3-8B-Instruct
Text Generation • 8B • Updated • 1.26M • • 4.96k -
meta-llama/Meta-Llama-3-70B-Instruct
Text Generation • 71B • Updated • 43.7k • • 1.53k
Most influential papers
-
Attention Is All You Need
Paper • 1706.03762 • Published • 141 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 33 -
Universal Language Model Fine-tuning for Text Classification
Paper • 1801.06146 • Published • 8 -
Language Models are Few-Shot Learners
Paper • 2005.14165 • Published • 21
LLMs
-
meta-llama/Llama-2-70b-hf
Text Generation • 69B • Updated • 3.46k • 855 -
tiiuae/falcon-180B
Text Generation • 180B • Updated • 72 • 1.15k -
meta-llama/Meta-Llama-3-8B-Instruct
Text Generation • 8B • Updated • 1.26M • • 4.96k -
meta-llama/Meta-Llama-3-70B-Instruct
Text Generation • 71B • Updated • 43.7k • • 1.53k
models 0
None public yet