Switch language한국어
Back to the list

SenseTime open-sources SenseNova-Vision unified vision model

TL;DR AI

Key summary

2 min read
  1. SenseTime has fully open-sourced SenseNova-Vision, a unified multimodal vision model.

  2. The model handles detection, OCR, segmentation, depth estimation, 3D reconstruction, surface-normal prediction, and multi-view geometry.

  3. SenseTime also released the SenseNova-Vision Corpus-50M training dataset and plans to integrate the model into the SenseNova U-series.

  4. The move could reduce reliance on separate task-specific computer vision systems and accelerate adoption of multimodal AI tools.

Read the original