Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

LAUNCH Lab

university
https://launch.eecs.umich.edu/
launchnlp
launchnlp
Activity Feed

AI & ML interests

Factuality, reasoning, alignment, LLM applications

Recent Activity

JieRuan  submitted a paper 6 days ago
SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
Ayoung01  authored a paper about 2 months ago
MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning
Ayoung01  authored a paper about 2 months ago
LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?
View all activity

Papers

SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning

View all Papers

Lu Wang's profile picture Yujian Liu's profile picture Shuyang Cao's profile picture Xinliang Frederick Zhang's profile picture xinyu hua's profile picture Yunxiang Zhang's profile picture Lechen Zhang's profile picture Farima Fatahi's profile picture sheza munir's profile picture KJ's profile picture Joe Peper's profile picture Ayoung Lee's profile picture Shitanshu Bhushan's profile picture Xin Liu's profile picture Muhammad Khalifa's profile picture Jie Ruan's profile picture Zohaib Khan's profile picture Jinyoung's profile picture
launch 's papers 3
Submitted by
Jie Ruan
10

SchemeArena: Factorized Stress Testing of Scheming in LLM Agents

launch LAUNCH Lab
1 2
Submitted by
Ayoung Lee
13

MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning

launch LAUNCH Lab
2
Submitted by
Muhammad Khalifa
2

Gaming the Judge: Unfaithful Chain-of-Thought Can Undermine Agent Evaluation

launch LAUNCH Lab
1
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs