Efficient AI Computing,
Transforming the Future.

Projects

To choose projects, simply check the boxes of the categories, topics and techniques.
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.

Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

ICLR 2025
 (
)

A new family of high-spatial compression autoencoders for accelerating high-resolution diffusion models.

LongVILA: Scaling Long-Context Visual Language Models for Long Videos

ICLR 2025
 (
)

LongVILA is a full-stack solution for long VLM, which incorporates novel training strategies, dataset, and the Multi-Modal Sequence Parallelism (MM-SP) system to efficiently handle long video understanding, achieving significant scalability, accuracy, and speed improvements on multi-modal benchmarks.

VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

ICLR 2025
 (
)

VILA-U is a Unified foundation model that integrates Video, Image, Language understanding and generation.

HART: Efficient Visual Generation with Hybrid Autoregressive Transformer

ICLR 2025
 (
)

HART is an autoregressive transformer that generates high resolution images with comparable quality to diffusion models, while offering 4.5-7.7x higher throughput.