Language Models
PKU-Alignment/align-anything
Align Anything: Training All-modality Model with Feedback
PythonApache-2.0updated Nov 27, 2025
The official repo of Qwen-VL (通义千问-VL) chat & pretrained large vision language model proposed by Alibaba Cloud.
$ git clone https://github.com/QwenLM/Qwen-VL.gitAlign Anything: Training All-modality Model with Feedback
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
The official repo of MiniMax-Text-01 and MiniMax-VL-01, large-language-model & vision-language-model based on Linear Attention
Official repo for "Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models"
[CVPR 2024 Highlight🔥] Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
Data from GitHub · snapshot Sep 24, 2026