Babylm Leaderboard 2026
The BabyLM 2026 leaderboard
The BabyLM 2026 leaderboard
Track, rank and evaluate open Arabic LLMs and chatbots
Explore and compare multilingual LLM benchmarks
Please vote on TTS Arena V2 instead
Calculate GPU RAM requirements for ML models.
Compile and push a model graph to Hugging Face Hub
Performance of computer vision models on the ImageNet-1k
Comparing EEG models in an open and reproducible way
VRAM calculator โ can your GPU run any AI model?
Calculate GPU memory needed for training Hugging Face models
Explore and discover diffusion models from the Hugging Face hub
Explore and submit models for benchmarking
Merge AI models using a YAML configuration file
Track, rank and evaluate open LLMs in Portuguese
Request evaluation for a new model
View and submit LLM evaluations
Create a model card for Hugging Face Hub
Track, rank and evaluate open LLMs and chatbots
Explore and submit LLM benchmarks
Open Persian LLM Leaderboard
Export models to ONNX using Hugging Face
Explore and compare QA and long doc benchmarks
Measure BERT model performance using WebGPU and WASM
Explore and submit LLM benchmarks