Add Convence/ParseEmbed as an official benchmark on the Hub (If possible)

Hi Hugging Face team,

I’d like to register Convence/ParseEmbed as an official benchmark on the Hub.

Dataset: Convence/ParseEmbed · Datasets at Hugging Face

It includes a root eval.yaml with:

  • name: ParseEmbed
  • evaluation_framework: mteb
  • tasks: mean, text_formatting, table
  • config: parse-embed

ParseEmbed is a retrieval benchmark for embedding models. It tests whether models preserve parse-sensitive meaning under hard negatives, including semantic scope, formatting-sensitive text, and table grounding.

The dataset card documents the purpose, files, task IDs, and usage. The dataset loads with datasets using the parse-embed config and the task splits.

Could you please add it to the official benchmark allow-list?

Thanks! if something is wrong let me know

By any chance if a hf staff is reading this, respond please if you have time

Hmm… @lhoestq

Reminder for approval

BTW, already contacted via GitHub issue? (datasets,huggingface_hub, etc. )

Hey there, Yes i have right now

Tag a Staff member maybe

Reminder for approval

Might be true…

By the way, it looks like HF has a new feature. Apparently, we can now send feedback directly to the staff? Share your feedback with us

I’m not sure exactly what it’s for yet, but it might be worth trying out.