Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Leshem Choshen
borgr
66
28
23
Follow
jennm's profile picture
21world's profile picture
andrewtran117's profile picture
36 followers
·
25 following
https://ktilana.wixsite.com/leshem-choshen
LChoshen
borgr
leshemchoshen
LChoshen
AI & ML interests
Future of Humane AI, technology that matters to all of us CoLab PI - doing Science together
Recent Activity
new
activity
1 day ago
evaleval/EEE_datastore:
Add WILD aggregate evaluation logs (data/wild)
updated
a model
1 day ago
The-CoLab/llama3-7b-en-ru-v2
updated
a model
1 day ago
The-CoLab/llama3-7b-en-translated-ru-v2
View all activity
Organizations
borgr
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
evaleval/EEE_datastore
1 day ago
Add WILD aggregate evaluation logs (data/wild)
2
#174 opened 21 days ago by
borgr
New activity in
evaleval/EEE_datastore
4 days ago
[AlpacaEval] Replace v1/v2 records with the consolidated adapter output
2
#178 opened 5 days ago by
borgr
Vectara Hallucination Leaderboard: full 105-model roster
#181 opened 4 days ago by
borgr
New activity in
evaleval/EEE_datastore
5 days ago
[Submission] Add LEXam legal-reasoning leaderboard (36 models)
1
#179 opened 5 days ago by
borgr
New activity in
evaleval/EEE_datastore
10 days ago
Add Papers with Code adapter data (data/paperswithcode)
1
#177 opened 20 days ago by
borgr
New activity in
evaleval/EEE_datastore
21 days ago
Add Open Medical-LLM Leaderboard aggregate records (data/open-medical-llm)
#176 opened 21 days ago by
borgr
Add WILD instance-level sample + on-demand hosting proposal
#175 opened 21 days ago by
borgr
New activity in
evaleval/EEE_datastore
about 1 month ago
[Submission] Add long-context-code-retrieval eval
1
#165 opened about 1 month ago by
imhurl
[Submission] Add Kaggle Community Benchmarks results (1/14)
4
#158 opened about 2 months ago by
mrshu
Add BenchPress score-matrix data (data/benchpress)
1
#164 opened about 1 month ago by
borgr
New activity in
evaleval/EEE_datastore
about 2 months ago
[ACL SHARED TASK] Add OUP L2-Bench
11
#151 opened 2 months ago by
l2-bench
[Submission] DOVE
#143 opened 3 months ago by
eliyahabba
Remove converter script from utils (moved to evaleval/every_eval_ever)
1
#154 opened about 2 months ago by
borgr
Remove converter script from utils (moved to evaleval/every_eval_ever)
1
#155 opened about 2 months ago by
borgr
Remove converter script from utils (moved to evaleval/every_eval_ever)
#156 opened about 2 months ago by
borgr
[ACL Shared Task] Add BountyBench (DetectWorkflow) evaluation results
1
#67 opened 4 months ago by
mrpfisher
[ACL Shared Task] Add AlpacaEval 1.0 and 2.0 leaderboard data (324 models)
5
#69 opened 4 months ago by
karthikchundi
New activity in
evaleval/EEE_datastore
4 months ago
Add alphaXiv SOTA evaluations (27,976 records, 1,646 benchmarks)
10
#26 opened 6 months ago by
simpod
Add AlpacaEval 1.0 and 2.0 leaderboard data (324 models)
7
#65 opened 4 months ago by
karthikchundi
commented
a paper
5 months ago
General Agent Evaluation
Paper
•
2602.22953
•
Published
Feb 26
•
12
•
3
Load more