QwenLM/Qwen3-VL
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
Appears on
Quick read
Latest capture 2026-09-08 10:55
0 paths
Agent instructions and tool configuration found in this repository.
No config files detected.
1 observed capture since 2026-09-08. Observed captures are shown by default.
Stars from first capture 0
Observed captures only
All tracked data
Observed snapshots
Observed snapshots
Nearest indexed repositories by embedding similarity.
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.
InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions
The official repo of Qwen (通义千问) chat & pretrained large language model proposed by Alibaba Cloud.
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
[CVPR 2025] Magma: A Foundation Model for Multimodal AI Agents