Alex1343543 commited on
Commit
e64afa3
·
verified ·
1 Parent(s): 11000d1

Document FC2 sidecar and modifications

Browse files

Document artifact hashes, exact-output performance measurements, automatic source validation, attribution, and the MiniMax-required modification notice.

Files changed (2) hide show
  1. NOTICE +10 -0
  2. README.md +33 -9
NOTICE ADDED
@@ -0,0 +1,10 @@
 
 
 
 
 
 
 
 
 
 
 
1
+ MiniMax H3 is licensed under the MiniMax H3 Community License Agreement, Copyright © 2026 MiniMax. All Rights Reserved.
2
+
3
+ Modification notice
4
+
5
+ PulpCut transferred the credited FL2VA Turbo adapter to the Ref2VA transformer
6
+ and requantized it to the documented INT8 ConvRot format. PulpCut also derived
7
+ the optional FC2 input-major sidecar by transposing the storage of the 50 INT8
8
+ mlp.fc2.weight tensors. The sidecar does not change tensor values, scales,
9
+ reference conditioning, or model behavior. See README.md for sources, hashes,
10
+ measurements, and reproduction instructions.
README.md CHANGED
@@ -16,25 +16,45 @@ base_model: MiniMaxAI/MiniMax-H3
16
  pretty_name: PulpCut MiniMax H3 Ref2VA Turbo INT8 ConvRot
17
  ---
18
 
19
- # MiniMax H3 Ref2VA Turbo · pruned INT8 ConvRot (single file)
20
 
21
  ## What this repository is
22
 
23
- A single-file MiniMax H3 **Ref2VA** diffusion transformer — the omni-reference
24
- variant, which conditions generation on ordered reference images — with the
25
- lightx2v **turbo step-distillation merged into the weights**, quantized in the
26
- same pruned **INT8 ConvRot** layout as the Comfy-Org release. It is a drop-in
27
- replacement for `minimax_h3_ref2va_pruned_int8_convrot.safetensors` in any
28
- runtime that reads the optimized INT8 single-file layout — including
 
29
  [H3ddle](https://github.com/AlexanderIstomin/h3ddle), the open-source native
30
  macOS app it was built for.
31
 
32
- This file is **not a standalone model**. It needs the rest of the optimized
33
  package (Qwen3-VL-32B INT8 text encoder, video/audio VAEs, tokenizer) from
34
  [Comfy-Org/MiniMax-H3](https://huggingface.co/Comfy-Org/MiniMax-H3), and the
35
  FL2VA transformer alongside it if you also want prompt-only and keyframe
36
  generation.
37
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
38
  ## Why this merge was made and republished
39
 
40
  Every published turbo LoRA for MiniMax H3 targets the **FL2VA** transformer.
@@ -100,7 +120,7 @@ Derivative of MiniMax H3 weights; the
100
  [MiniMax H3 Community License Agreement](https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/939557dc319dd91227e30195a763f272ba7f8765/LICENSE)
101
  applies. By downloading you agree to its terms.
102
 
103
- ## What this file is used for in H3ddle
104
 
105
  H3ddle installs it as a managed model alongside the reference-capable package:
106
  the app verifies the SHA-256 below, reuses the shared package files it already
@@ -119,6 +139,7 @@ of the MiniMax H3 Community License apply unchanged.
119
  | File | Bytes | SHA-256 |
120
  |---|---|---|
121
  | `minimax_h3_ref2va_pruned_turbo_int8_convrot.safetensors` | 20,970,379,854 | `e64cef63bc2785bcd72e6103c52aa78c6cd2c4f9870a7ce79675083fd65cf2e7` |
 
122
 
123
  ## Reproducibility references
124
 
@@ -127,6 +148,9 @@ The conversion is a single dependency-free Python script,
127
  in the H3ddle repository, including the strength-0 self-check used to validate
128
  the pipeline against the official file.
129
 
 
 
 
130
  ## Contact
131
 
132
  Open an issue in the [H3ddle repository](https://github.com/AlexanderIstomin/h3ddle/issues).
 
16
  pretty_name: PulpCut MiniMax H3 Ref2VA Turbo INT8 ConvRot
17
  ---
18
 
19
+ # MiniMax H3 Ref2VA Turbo · pruned INT8 ConvRot
20
 
21
  ## What this repository is
22
 
23
+ An optimized MiniMax H3 **Ref2VA** package centered on the omni-reference
24
+ diffusion transformer, which conditions generation on ordered reference
25
+ images. It has the lightx2v **turbo step-distillation merged into the weights**
26
+ and uses the same pruned **INT8 ConvRot** layout as the Comfy-Org release. The
27
+ primary transformer is a drop-in replacement for
28
+ `minimax_h3_ref2va_pruned_int8_convrot.safetensors` in any runtime that reads
29
+ the optimized INT8 layout — including
30
  [H3ddle](https://github.com/AlexanderIstomin/h3ddle), the open-source native
31
  macOS app it was built for.
32
 
33
+ The transformer is **not a standalone model**. It needs the rest of the optimized
34
  package (Qwen3-VL-32B INT8 text encoder, video/audio VAEs, tokenizer) from
35
  [Comfy-Org/MiniMax-H3](https://huggingface.co/Comfy-Org/MiniMax-H3), and the
36
  FL2VA transformer alongside it if you also want prompt-only and keyframe
37
  generation.
38
 
39
+ ## H3ddle FC2 performance sidecar
40
+
41
+ `minimax_h3_ref2va_pruned_turbo_int8_convrot_fc2_input_major.safetensors`
42
+ is an optional H3ddle performance sidecar derived from the transformer in
43
+ this repository. It contains only the 50 INT8 `mlp.fc2.weight` matrices, with
44
+ their storage transposed from `[output, input]` to `[input, output]`. Values,
45
+ quantization scales, reference conditioning, and model behavior are unchanged.
46
+
47
+ H3ddle installs it beside the transformer and selects it automatically after
48
+ validating its format version, source file size, exact source-header
49
+ fingerprint, and all 50 tensor schemas. It cannot silently be used with a
50
+ different checkpoint. The original transformer remains available as the
51
+ fallback and for runtimes that do not understand the sidecar.
52
+
53
+ A real 512-class, 50-block Ref2VA parity run produced identical baseline and
54
+ sidecar hashes: video `85a5ccfc5a4d8075`, audio `731e24ae9dc2e7ec`.
55
+ The matched cold pair measured 43.753 seconds without the sidecar and 40.736
56
+ seconds with it; the broader FL2VA A/B/B/A benchmark measured a 7.15% gain.
57
+
58
  ## Why this merge was made and republished
59
 
60
  Every published turbo LoRA for MiniMax H3 targets the **FL2VA** transformer.
 
120
  [MiniMax H3 Community License Agreement](https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/939557dc319dd91227e30195a763f272ba7f8765/LICENSE)
121
  applies. By downloading you agree to its terms.
122
 
123
+ ## What these files are used for in H3ddle
124
 
125
  H3ddle installs it as a managed model alongside the reference-capable package:
126
  the app verifies the SHA-256 below, reuses the shared package files it already
 
139
  | File | Bytes | SHA-256 |
140
  |---|---|---|
141
  | `minimax_h3_ref2va_pruned_turbo_int8_convrot.safetensors` | 20,970,379,854 | `e64cef63bc2785bcd72e6103c52aa78c6cd2c4f9870a7ce79675083fd65cf2e7` |
142
+ | `minimax_h3_ref2va_pruned_turbo_int8_convrot_fc2_input_major.safetensors` | 3,853,522,260 | `0ad6a5673abdf842c39d4d8de7c34c971a420b64bd5f79eb6f4331c5bfb5cd97` |
143
 
144
  ## Reproducibility references
145
 
 
148
  in the H3ddle repository, including the strength-0 self-check used to validate
149
  the pipeline against the official file.
150
 
151
+ The FC2 sidecar is reproducible with
152
+ [`Scripts/optimize-h3-fc2-sidecar.py`](https://github.com/AlexanderIstomin/h3ddle/blob/main/Scripts/optimize-h3-fc2-sidecar.py).
153
+
154
  ## Contact
155
 
156
  Open an issue in the [H3ddle repository](https://github.com/AlexanderIstomin/h3ddle/issues).