Switch language한국어
Back to the list

SenseTime releases SenseNova U1 models on HuggingFace

TL;DR AI

Key summary

2 min read
  1. SenseTime open-sourced two SenseNova U1 multimodal models on HuggingFace with open weights: an 8B dense MoT model and a ~3B-activated MoT model.

  2. The models use the NEO-Unify architecture, which replaces separate visual encoders and VAEs with a unified token space for multimodal understanding, reasoning, and generation.

  3. Released under Apache 2.0, the models are commercially usable and available for developers to build on.

  4. They target tasks like VQA, interleaved text-image generation, and content creation such as PPTs, posters, and diagrams.

Read the original