MMMU-Benchmark/MMMU
This repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast.
$ git clone https://github.com/tatsu-lab/alpaca_eval.gitThis repo contains evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
γε€§θ―θ¨ζ¨‘εγδ½θ οΌθ΅΅ι«οΌζεζ― οΌε¨ζοΌε倩δΈοΌζη»§θ£
An unnecessarily tiny implementation of GPT-2 in NumPy.
MindSpore + π€Huggingface: Run any Transformers/Diffusers model on MindSpore with seamless compatibility and acceleration.
A simulation framework for RLHF and alternatives. Develop your RLHF method without collecting human data.
Data from GitHub Β· snapshot Sep 24, 2026