Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
V's picture

V

posteronthewall
2 1 4
aifeifei798's profile picture Hastagaras's profile picture 21world's profile picture
ยท

AI & ML interests

None yet

Recent Activity

reacted to GGUFGuy's post with ๐Ÿš€ about 12 hours ago
wait why can i post
reacted to GGUFGuy's post with ๐Ÿ”ฅ about 12 hours ago
wait why can i post
reacted to sergiopaniego's post with ๐Ÿ”ฅ 7 days ago
Can you do RL over taste? I've spent some time reproducing, in the open, Surya N's idea of training a model to paint with code. It's a coding model that learns to paint watercolours by writing JS code, trained with GRPO. I used TRL and OpenEnv for this, with the whole pipeline running on Hugging Face. The interesting part is that the reward has no correct answer, unlike a math problem. In this case it's based on the artistic preferences of the person who builds the dataset. Everything is published: the environment, the reference pool, the trained adapters, every painting of every run with the code that made it, and a write-up with all the decisions, including the ones that went wrong. Blog post: https://huggingface.co/blog/train-to-paint-with-code
View all activity

Organizations

None yet

posteronthewall 's models 2

posteronthewall/personal-loras

Updated Jul 30

posteronthewall/civvy

Updated Jul 3
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs