xAI releases Grok-1.5V, a multimodal AI model with vision support
xAI, an artificial intelligence company under Musk, announced the launch of its first multimodal AI model, Grok-1.5V. In addition to its powerful text processing capabilities, Grok can also process various visual information, including documents, charts, screenshots, and photos. In benchmark tests in multiple fields, Grok-1.5V's performance is comparable to existing cutting-edge multimodal models. Especially in xAI's newly launched RealWorldQA benchmark test, Grok surpassed similar models in its ability to understand the real-world space. The RealWorldQA dataset contains more than 700 images and aims to evaluate the basic understanding ability of multimodal models in the physical world. Grok-1.5 will soon be open to early testers and existing users.
Disclaimer: The content of this article solely reflects the author's opinion and does not represent the platform in any capacity. This article is not intended to serve as a reference for making investment decisions.
You may also like
Bitget Launches PLUME On-chain Earn With 4.5% APR
Bitget Trading Club Championship (Phase 2) – Grab a share of 50,000 BGB, up to 500 BGB per user!
Bitget Trading Club Championship (Phase 2) – Grab a share of 50,000 BGB, up to 500 BGB per user!
Subscribe to UNITE Savings and enjoy up to 15% APR
Trending news
MoreCrypto prices
More








