Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

LAUNCH Lab

university
https://launch.eecs.umich.edu/
launchnlp
launchnlp
Activity Feed

AI & ML interests

Factuality, reasoning, alignment, LLM applications

Recent Activity

Ayoung01  authored a paper 10 days ago
MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning
Ayoung01  authored a paper 10 days ago
LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?
Ayoung01  authored a paper 10 days ago
Logit Arithmetic Elicits Long Reasoning Capabilities Without Training
View all activity

Papers

MET: Theory-Grounded and Culture-Aware Multilingual Moral Reasoning

Gaming the Judge: Unfaithful Chain-of-Thought Can Undermine Agent Evaluation

View all Papers

Lu Wang's profile picture Yujian Liu's profile picture Shuyang Cao's profile picture Xinliang Frederick Zhang's profile picture xinyu hua's profile picture Yunxiang Zhang's profile picture Lechen Zhang's profile picture Farima Fatahi's profile picture sheza munir's profile picture KJ's profile picture Joe Peper's profile picture Ayoung Lee's profile picture Shitanshu Bhushan's profile picture Xin Liu's profile picture Muhammad Khalifa's profile picture Jie Ruan's profile picture Zohaib Khan's profile picture Jinyoung's profile picture

launch 's Spaces 7

Running

LudoBench

🎲

Multimodal Game Reasoning Benchmark [ICLR 2026]

Apr 24
Sleeping
Agents

Answer Convergence Early Stopping

🛑

Demo for EMNLP Paper "Answer Convergence as a Signal..."

Jan 4
Runtime error

FactRBench

🏆

View and analyze long-form factuality leaderboard

Nov 3, 2025
Sleeping
3

ExpertLongBench

🚀

Leaderboard for ExpertLongBench

Sep 28, 2025
Sleeping
1

ManyICLBench

🚀

Leaderboard for ManyICLBench

Jun 20, 2025
Running

MLRC-BENCH

📊

Display model performance rankings

Apr 16, 2025
Running
3

Factbench

📈

View and compare language model factuality scores

Oct 30, 2024
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs