Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Huang's INTelligence lab

university
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

shrango  submitted a paper 2 days ago
On the Off-Policy Teacher in On-Policy Distillation
ChengsongHuang  authored a paper 10 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
ChengsongHuang  authored a paper 11 days ago
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling
View all activity

Papers

Process Rewards with Learned Reliability

View all Papers

Jiaxin Huang's profile picture Chengsong Huang's profile picture Langlin Huang's profile picture Jixuan Leng's profile picture Jinyuan Li's profile picture
HINT-lab 's papers 1
Submitted by
Jinyuan Li
51

Process Rewards with Learned Reliability

HINT-lab Huang's INTelligence lab
9 2
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs