Qwen3.8-Omni-Flash
ストックにはログインが必要です
Qwen's omni-modal model built around agentic capabilities
Artificial Intelligence
GitHub
API
Video
Qwen3.8-Omni-Flash understands text, image, audio, and video with a 1M-token context, then plans, calls tools, and completes real work: editing video, making music videos, dubbing/translating shows, summarizing meetings into action items, even coding off what it hears and sees.
投票数: 0