Switch language한국어
Back to the list

XPENG releases TuringViT for smart driving and humanoid robots

TL;DR AI

Key summary

2 min read
  1. XPENG has released TuringViT, a vision encoder for smart driving, smart cockpit systems, and humanoid robots.

  2. The model targets vision-language and vision-language-action use cases and comes in two variants, TuringViT-18L and TuringViT-24L.

  3. XPENG says it was trained on 850 million image-text pairs and delivered strong throughput and benchmark results at 1536×1536 resolution.

  4. The launch signals XPENG is building core multimodal AI infrastructure to compete in autonomous driving and robotics.

Read the original