VITA-MLLM
VITA Team
Pinned Loading
Repositories
Showing 9 of 9 repositories
- Omni-Diffusion Public
✨✨[ICML 2026] Omni-Diffusion: Unified Multimodal Understanding and Generation with Masked Discrete Diffusion
- Freeze-Omni Public
✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
- VITA-Audio Public
✨✨[NeurIPS 2025] VITA-Audio: Fast Interleaved Cross-Modal Token Generation for Efficient Large Speech-Language Model
- Long-VITA Public
✨✨Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy
Top languages
Loading…
Most used topics
Loading…