Seed1.5-VL is an AI model developed by ByteDance (China), first published in May 2025. It works in the vision, language, multimodal and video domain, on tasks such as visual question answering, video description, language modeling/generation and 2 more.
Training it took an estimated 1.4×10²⁴ FLOP of compute (estimation method: hardware). It was trained on roughly 3T datapoints. Training ran on NVIDIA H800 SXM5.
Access: API access. Its weights are not openly released. Epoch AI rates the confidence of this record as confident.