Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
PythonEvolvingLMMs-Lab/LLaVA-OneVision-2

LLaVA-OneVision-2

Fully Open Framework for Democratized Multimodal Training

74.4/100
1.1KForks: 77
View on GitHubHomepage →
Loading report...

Similar Projects

mlx-vlm

84

MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.

Python5.2K

VLMEvalKit

82

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

Python4.3K

Eagle

67

Eagle: Frontier Vision-Language Models with Data-Centric Strategies

Python3.2K

WhisperJAV

68

ASR/STT subtitle generator. Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD. Noise-robust for JAV

Python2.0K
Back to List