Building
Open research on multimodal models, evaluation, benchmarks, and lmms-eval tooling - shared as we discover.
LMMS-LAB // BREACH ACTIVENEURAL WEIGHT EXTRACTION
1/7LIVE
_
thinking:
_ LLaVA-OneVision-2 :
Towards Next-Generation Perceptual Intelligence
OneVision Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence
OneVision Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence
modelsLLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training
LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training
modelsLongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling
modelsOpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe
