LMMS Lab Logo
HomeBlog
About

Building the way to multimodal intelligence.

Open research on multimodal models, evaluation, benchmarks, and lmms-eval tooling - shared as we discover.

Explore ResearchAbout the Lab
LMMS-LAB // BREACH ACTIVENEURAL WEIGHT EXTRACTION
1/7LIVE
_
thinking:
_
Featured Research
APR 20, 2026

LLaVA-OneVision-2 :
Towards Next-Generation Perceptual Intelligence

models
Read Paper
Latest Publications View Archive
OneVision Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence
JAN 2026

OneVision Encoder: Codec-Aligned Sparsity as a Foundational Principle for Multimodal Intelligence

models
LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training
SEP 2025

LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training

models
LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling
NOV 2025

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling

models
OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe
NOV 2025

OpenMMReasoner: Pushing the Frontiers for Multimodal Reasoning with an Open and General Recipe

models
2026 LMMs-Lab
GitHubTwitter