Best larger-system route
Studio 96
The smallest system in the range that reaches this model's preferred working allowance.
Explore Studio 96
Wan AI model
Wan documents a minimum 24GB GPU for 720p generation with model offload, dtype conversion and the text encoder on CPU. Team 32 clears that floor. An 80GB-class GPU can remove those offload options for faster execution.
Wan-AI/Wan2.2-TI2V-5B921dbaf3f167Wan-AI/Wan2.2-TI2V-5B Buyer verdict
Wan2.2 TI2V 5B is a practical starting point for local video generation because it has an official 24GB minimum and a 720p workflow.
These figures apply to the named model version. Quantisation, fine-tuning, context length, image resolution, batch size and serving software can materially change the hardware needed.
Hardware requirements
This is the official 720p command profile.
Wan states this can speed execution; actual latency still needs measurement.
Technical specification
Specifications shown for source version
921dbaf3f1674a56f47e83fb80a34bac8a8f203e, updated
7 August 2025.
Product compatibility
A system is listed as fitting when its GPU memory meets the requirement shown above. Speed, usable context, batch size and concurrent users still need testing with the final model and software configuration.
Best larger-system route
The smallest system in the range that reaches this model's preferred working allowance.
Explore Studio 96sme workstation
Meets the lower memory screen, but not the recommended working allowance.
Available GPU memory clears the lower 24GB VRAM with documented CPU offload settings screen but not the 80GB or more to remove the documented offload options working allowance. It may load, but choose a larger system when context, batch, concurrency or response time matter.
sme workstation
Meets the lower memory screen, but not the recommended working allowance.
Available GPU memory clears the lower 24GB VRAM with documented CPU offload settings screen but not the 80GB or more to remove the documented offload options working allowance. It may load, but choose a larger system when context, batch, concurrency or response time matter.
sme workstation
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
sme workstation
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
pcie rack
Meets the lower memory screen, but not the recommended working allowance.
Available GPU memory clears the lower 24GB VRAM with documented CPU offload settings screen but not the 80GB or more to remove the documented offload options working allowance. It may load, but choose a larger system when context, batch, concurrency or response time matter.
pcie rack
Meets the lower memory screen, but not the recommended working allowance.
Available GPU memory clears the lower 24GB VRAM with documented CPU offload settings screen but not the 80GB or more to remove the documented offload options working allowance. It may load, but choose a larger system when context, batch, concurrency or response time matter.
pcie rack
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
pcie rack
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
pcie rack
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
frontier partner
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
frontier partner
Also meets this model page's recommended working allowance.
Available GPU memory exceeds the model developer's published requirement. Confirm speed, context, batch size and concurrent users before purchase.
Deployment reality
Serving software
Before installation
System requirements
Performance: speed, usable context and concurrency depend on the selected system, software and workload. No benchmark is quoted on this page.
Commercial and legal boundary
Commercial use: permitted
The official model card states Apache 2.0 and places responsibility for lawful use on the operator.
Always retain the applicable notices and recheck the live terms for the intended organisation, territory, use and distribution route. Obtain legal advice where required; the official licence governs use.
Read the official licenceTechnical questions
It clears Wan's 24GB minimum and should use the documented offload route. Generation time and quality still need testing.
Wan states that 80GB or more can remove several offload options, which may improve speed and simplify the execution path.
Model revision, prompt, resolution, frames, duration, steps, batch, latency, peak VRAM and output review.
GPU memory is only one part of the system. Storage, data access, serving software, monitoring and administrator handover also affect a reliable deployment.
Privacy & cookies. Google Analytics measures site use so we can improve it. Read our Cookie Notice and Privacy Notice. To object, disable cookies in your browser.