Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

PromptEval

https://github.com/felipemaiapolo/prompteval
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

borgr  authored a paper about 22 hours ago
LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content
borgr  authored a paper about 22 hours ago
Unforgettable Generalization in Language Models
borgr  authored a paper 18 days ago
Learning to combine Grammatical Error Corrections
View all activity

Felipe Maia Polo's profile picture Mikhail Yurochkin's profile picture Lucas Weber's profile picture Leshem Choshen's profile picture Ronald Xu's profile picture Mírian Silva's profile picture

PromptEval 's datasets 4

PromptEval/MMLU_multi_prompt

Viewer • Updated Dec 4, 2024 • 3.17M • 386 • 1

PromptEval/MMLU_multi_prompt_v0

Viewer • Updated Nov 26, 2024 • 3.17M • 464

PromptEval/PromptEval_MMLU_correctness

Viewer • Updated Jun 7, 2024 • 85.5k • 1.64k • 2

PromptEval/PromptEval_MMLU_full

Viewer • Updated Jun 7, 2024 • 21.1M • 4.72k • 3
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs