empero-ai/Qwythos-9B-Claude-Mythos-5-1M-GGUF Image-Text-to-Text • 9B • Updated about 1 month ago • 449k • 2.61k
view article Article Illustrating Reinforcement Learning from Human Feedback (RLHF) +2 natolambert, LouisCastricato, lvwerra, Dahoas • Dec 9, 2022 • 423
Running 3.98k The Ultra-Scale Playbook 🌌 3.98k The ultimate guide to training LLM on large GPU Clusters
HuggingFaceFW/fineweb-edu-classifier Text Classification • 0.1B • Updated Nov 17, 2024 • 14.2k • • 224