Ultra-efficient multimodal language model from OpenBMB built on SigLIP2-400M and Qwen3.5-0.8B (~1B parameters). Supports single-image, multi-image, and video understanding with mixed 4x/16x visual token compression. Designed for edge deployment on iOS, Android, and HarmonyOS.
256K tokens
0
Aún no hay datos de proveedores.
Las asociaciones con proveedores aparecerán aquí cuando estén disponibles.