Alibaba's Qwen team released Qwen3.8-Omni-Flash on Friday, a model that reads text, images, audio and video. Audio input now costs over 98% less per hour than the previous Qwen omni model. It holds a million tokens of context, but Alibaba did not publish the weights.