Model reference · open weights

minimax_h3_ref2va_patchin_hf102

minimax_h3_ref2va_patchin_hf102 is an open-weight video model from t8star, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Licence fee required Video t8star 1 variants 3k downloads/mo
Request a licence + hosting quote All served models Not on the shared API today — deployed on request.

About

What minimax_h3_ref2va_patchin_hf102 is

MiniMax H3 Ref2VA Patch-In HF 1.02 Experimental ComfyUI single-file checkpoint derived from MiniMaxAI/MiniMax-H3 Ref2VA and the Comfy-Org/MiniMax-H3 INT8 ConvRot repack. [!CAUTION] This is an experimental weight modification, not an official MiniMax release and not a proven “de-oil”, “de-wax”, restoration, or quality-fix model. The original checkpoint remains the recommended default. [!IMPORTANT] MiniMax H3 is governed by the MiniMax H3 Community License Agreement, which defines excluded territories and mandatory redistribution conditions. Before publishing or redistributing this derivative, include the official LICENSE, keep the modification notice and NOTICE, and confirm that the intended distribution method and audience are authorized. A public Hugging Face repository may be reachable from excluded territories; a repository gate alone is not necessarily geographic access control. This model card is not legal advice. 中文说明 这是 MiniMax H3 Ref2VA 的实验性 ComfyUI 单文件衍生模型。它没有训练、微调或蒸馏,只对 视频 patch 输入投影中的 2×2 空间高频分量增加 2%。两组固定条件测试都出现了很弱的皮肤 高频代理正增益,但肉眼仍未确认能够消除 Ref2VA 的油感或蜡感,所以只能作为 EXP 对照模型。 发布前必须阅读上游许可证。该许可证对适用地域、公开分发、商业使用、安全措施、修改声明、 LICENSE 和 NOTICE 都有要求。 Model description What was changed MiniMax H3 video latents are patchified with a 1 × 2 × 2 patch. For every latent channel, the four spatial input columns of videopatchproj.weight were transformed in an orthonormal 2×2 Haar basis: - the DC/common component is preserved at gain 1.00; - the three non-DC H/V/D components are multiplied by 1.02; - the result is transformed back and written into the input projection; - no output head, audio tensor, shared Transformer block, VAE, text encoder, or ComfyUI node was changed. The equivalent 4×4 transform has diagonal 1.015 and off-diagonal -0.005. Its all-ones/DC eigenvector has gain 1.00; the three orthogonal spatial-detail eigenvectors have gain 1.02. This is not an output sharpening filter. It changes the video latent input projection used during each joint audio-video denoising forward pass. It cannot reconstruct real texture that the model does not generate. Checkpoint integrity Source The source size and SHA-256 match the LFS object published by Comfy-Org/MiniMax-H3. Modifie

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makert8star
TypeVideo models
Variants1
Runs withminimax-h3
Based onMiniMaxAI/MiniMax-H3, Comfy-Org/MiniMax-H3
Released2026-08-10
Popularity3k downloads / month
Likes2
LicenceCommercial licence needed

How it works

How video models work

Prompt / imagestart pointTemporal diffusionframes over timeVideoMP4 clipA video model generates a sequence of coherent frames from your prompt or a starting image.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
minimax_h3_ref2va_patchin_hf102BF16Weights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys minimax-h3-ref2va-patchin-hf102 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (minimax-h3-ref2va-patchin-hf102 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/videos/generations \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"minimax-h3-ref2va-patchin-hf102","prompt":"a drone shot over a forest"}'

Details

Languages, data & research

Languages

en zh

Tags

minimax-h3 ref2va comfyui diffusion-single-file int8-convrot synchronized-audio-video experimental image-text-to-video en zh

Licence

Commercial licence needed

The weights are open but its licence needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗

Sources

Weights & code

Want minimax_h3_ref2va_patchin_hf102 on EU-owned hardware?

Request a licence + hosting quote See what’s served now

Explore

More video models

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms