Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Marian Kannwischer's picture

Marian Kannwischer

canwiper
5 49
jasoncorkill's profile picture LinoGiger's profile picture Faypid's profile picture
ยท
  • kannwism
  • kannwischer

AI & ML interests

RLHF & Computer Vision

Recent Activity

reacted to jasoncorkill's post with ๐Ÿš€ about 3 hours ago
Most public benchmarks collapse model performance into one broad preference signal. That makes it hard to understand which capabilities differentiate between models. It's also almost impossible to inspect the evidence behind it. So @RapidataAI is releasing Benchmark.AI. We started with an SVG generation benchmark including 42 models, 500 prompts, 1.9M+ human judgements, 300K+ match-ups. We evaluate models separately on Preference, Alignment and Coherence, while making the prompts, outputs, match-ups and methodology public. Full dataset: https://huggingface.co/datasets/Rapidata/svg-benchmark Full benchmark: https://www.benchmark.ai/svg Methodology feedback and benchmark suggestions very welcome!
liked a dataset 2 months ago
Rapidata/svg-benchmark
liked a dataset 8 months ago
Rapidata/bananamark-dataset
View all activity

Organizations

mlo-data-cleaning's profile picture mlo-data-collab's profile picture Rapidata's profile picture

canwiper 's datasets

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs