Skip to content

Popular repositories Loading

  1. VITA VITA Public

    ✨✨VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

    Python 2.2k 165

  2. Freeze-Omni Freeze-Omni Public

    ✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

    Python 300 19

  3. Long-VITA Long-VITA Public

    ✨✨Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy

    Python 267 29

  4. LUCY LUCY Public

    LUCY: Linguistic Understanding and Control Yielding Early Stage of Her

    Python 32 3

  5. Sparrow Sparrow Public

    Sparrow: Data-Efficient Video-LLM with Text-to-Image Augmentation

    Jupyter Notebook 26

Repositories

Showing 5 of 5 repositories
  • LUCY Public

    LUCY: Linguistic Understanding and Control Yielding Early Stage of Her

    VITA-MLLM/LUCY’s past year of commit activity
    Python 32 3 4 0 Updated Mar 31, 2025
  • Sparrow Public

    Sparrow: Data-Efficient Video-LLM with Text-to-Image Augmentation

    VITA-MLLM/Sparrow’s past year of commit activity
    Jupyter Notebook 26 Apache-2.0 0 0 0 Updated Mar 28, 2025
  • VITA Public

    ✨✨VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

    VITA-MLLM/VITA’s past year of commit activity
    Python 2,198 165 48 1 Updated Mar 28, 2025
  • Long-VITA Public

    ✨✨Long-VITA: Scaling Large Multi-modal Models to 1 Million Tokens with Leading Short-Context Accuracy

    VITA-MLLM/Long-VITA’s past year of commit activity
    Python 267 29 4 1 Updated Mar 20, 2025
  • Freeze-Omni Public

    ✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM

    VITA-MLLM/Freeze-Omni’s past year of commit activity
    Python 300 19 8 2 Updated Jan 2, 2025

Top languages

Loading…

Most used topics

Loading…