Video models, accelerated and customized for your production.

Production grounding and evaluation for FastH3: MiniMax H3, post-trained and optimized by FastVideo.

Pending

Coming soon. Join the waitlist for updates and customization for your needs.

What We Do

We take open video models from weight to production

01

Accelerate

Post TrainingModel Adaptation

Reduce latency and serving costs with model-specific optimization tuned to your target hardware and production workload.

FastH3 technical path

02

Customize

Post TrainingModel Adaptation

Adapt open models to your visual language, formats, and workflows while preserving quality and creative control.

FastH3 post-training

03

Deploy

InferenceProduction Systems

Move from weights to a dependable production service with scalable infrastructure, observability, and hands-on support.

FastH3 results

Open Weights

We believe open video models will remain a critical part of the production stack. The Nuva Lab team has contributed to open-source AI for years. We optimize open-weight models for production today and will continue supporting and contributing to the models that come next.

FastH3

Current Work

MiniMax H3, post-trained and optimized by FastVideo.

Sparsity Kernel

T2V Validated Omni Ref Pending

90% Sparse≈10× fewer target-video QK/PV block pairs

Quantization-Aware Distillation (QAD)

T2V Validated Omni Ref Pending

12.5× Speedup50 → 4 denoising steps

NVFP4 Native

Pending

Weights / Attention / Linear

Releases & News

Open model work, production integrations, and ecosystem releases from Nuva Lab and our collaborators.

Partner

Collaborators

NVIDIA vLLM FastVideo

Working with an open-weight video model?

Talk to us