Rows are hardware profiles
Each row names the card class and the environment that was actually used. The owned RTX 3060 row is a record, not a promise about every RTX 3060.
Measured here · one protocol · raw values
Every row below is a run performed here. Cells we have not measured say not tested; we do not infer support from a neighbouring result, and we do not turn community numbers into site measurements.
How to read
Each row names the card class and the environment that was actually used. The owned RTX 3060 row is a record, not a promise about every RTX 3060.
Each column combines a model or precision track with a task. A measured cell links to its test_id; it is a projection of that run record.
not tested means this site has not run the combination. It does not mean “unsupported,” “impossible,” or “will fail.”
The rows do not share one machine, workflow, or memory budget. We place facts beside one another and do not calculate a speed ranking across them.
Hardware × task
Measured cells show the first formal run in the series and link to the full record below. Repeats remain separate records; the matrix is not a second set of measurements.
| GPU profile | REF2VA INT8 · R2V | FL2VA · T2V | FL2VA · I2V | Turbo LoRA |
|---|---|---|---|---|
| RTX 3060 12GB owned bench · GPU 0 |
2,581.8 s11,649 MiB · 94.8% 1344×768 · 124 frames |
2,175.5 s11,023 MiB · 89.7% 1344×768 · 124 frames |
2,373.7 s11,125 MiB · 90.5% 1344×768 · 124 frames |
528.6 smedian · 4-step Turbo 1344×768 · 124 frames |
| 8GB rental class not tested · no support verdict |
not testedRental-class run queued under Requirements §7.2 item 6. | not testedRental-class run queued under Requirements §7.2 item 6. | not testedRental-class run queued under Requirements §7.2 item 6. | not testedNo 8GB Turbo A/B; no support verdict. |
The 12GB row is site-measured on one owned bench. The 8GB row is a queued inventory item, not an inference from the 12GB measurements.
Full ledger · 26 records
Formal runs, smoke checks, shared baselines, cache-node tests and failures live together here so a result cannot disappear merely because it was inconvenient. Each anchor is the stable test_id.
These records use the site benchmark track and the official Ref2VA INT8 workflow family. The GPU 1 incident remains a first-class failure record; the two GPU 0 formal runs are independent repeats.
GPU 1 smoke check · 512×288 / 22 frames · comparison telemetry before the formal attempt.
site_benchmarkSITE-3060-R2V-B1-SMOKE-api.json; derived smoke workflow, hash retained in the pinned R2V baseline report.res_multistep / simple / 1.0submit.json; SageAttention absent; Turbo LoRA absent; cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9 · custom nodes N/Acontaminated=falseGPU 1 formal attempt · retained hardware incident · VM froze after roughly nine minutes.
site_benchmarkVM_FROZEN · no output · no valid wall timeres_multistep / simple / 1.0bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9 · custom nodes N/AGPU 0 smoke comparison · 512×288 / 22 frames · same workload family as the GPU 1 smoke.
site_benchmarkres_multistep / simple / 1.0bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=falseGPU 0 formal run 1 · 1344×768 / 124 frames · the first formal site R2V record.
site_benchmarkcc2a6f75…be6060fSITE-3060-R2V-B1-api.json · f52a1b919c2fb29acfbe9affb7fa82ceef1ad2a57f2d68427383d25bc8257790minimax_h3_ref2va_pruned_int8_convrot.safetensors · Ref2VA SHA-256 9255f52b6677845ad238f20dfaafa94727053694127ab7f255c048f0f9365779; Qwen3-VL encoder and VAEs from the pinned manifest.res_multistep / simple / 1.0submit.json; SageAttention absent; Turbo LoRA absent; cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9 · custom nodes N/Acontaminated=falseGPU 0 formal run 2 · independent repeat of SITE-3060-R2V-GPU0-B1.
site_benchmarkcc2a6f75…be6060fSITE-3060-R2V-B1-api.json · f52a1b919c2fb29acfbe9affb7fa82ceef1ad2a57f2d68427383d25bc8257790res_multistep / simple / 1.020260812; SageAttention absent; Turbo LoRA absent; cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. GPU 0 peak temperature 74 °C.repeat_of=SITE-3060-R2V-GPU0-B1. The two raw values are intentionally shown separately, never collapsed into a single figure.These six records use the pinned Comfy-Org template revision and the site benchmark track. Each formal task crossed the 30-minute threshold and therefore has two independent raw runs. The two smoke records verify the template, input and audio path; they are not formal baseline statistics.
T2V smoke · 512×288 / 22 frames · native audio on.
site_benchmarkres_multistep / simple / 1.0 · fixed seed 20260820bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=falseT2V formal run 1 · pinned official template · native audio on.
site_benchmark3d582cbcf88b3aa8d60e93cea4c65bba445ab475c8c3243591107c7dab4ee8021b083dfa1a4540b3805b998a66f5585f4337d5fbe14b441ca7ac0a2e1bf545a2Comfy-Org/workflow_templates@7837633a31c1aa40495ed7f9171e9c97358a1d93minimax_h3_fl2va_pruned_int8_convrot.safetensors SHA-256 e889202c41dafb67b10d67b97f0d8541508036a6090af23425a5c2615d03c47a; Qwen3-VL encoder and video/audio VAEs match the pinned manifest.res_multistep / simple / 1.0 · fixed seed 20260820bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · cache state mixedT2V formal run 2 · independent repeat of SITE-3060-FL2VA-T2V-GPU0-B1.
site_benchmark3d582cbcf88b3aa8d60e93cea4c65bba445ab475c8c3243591107c7dab4ee8021b083dfa1a4540b3805b998a66f5585f4337d5fbe14b441ca7ac0a2e1bf545a220260820, 1344×768, 124 frames, 20 steps, audio on.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · cache state mixedrepeat_of=SITE-3060-FL2VA-T2V-GPU0-B1. The 0.7-second difference is shown as raw values, not collapsed into a mean.I2V smoke · official input image · 512×288 / 22 frames · native audio on.
site_benchmarktransparent_rgb_gaming_mouse.png input; smoke output retained49696748d2fff0e8c9b63c7173c6d6282b70eac0195def5b39402f9564410e75.res_multistep / simple / 1.0 · fixed seed 20260820bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=falseI2V formal run 1 · official first-frame input · native audio on.
site_benchmark8681605254c00719f7af488c13adb0124136de177c0dde02e906b198cc5cb47f22648d0a8d42cae3207e9d5b0e9def8e55afd2ae4f7d6a2b809271141de30e40Comfy-Org/workflow_templates@7837633a31c1aa40495ed7f9171e9c97358a1d93 · transparent_rgb_gaming_mouse.png, SHA-256 49696748d2fff0e8c9b63c7173c6d6282b70eac0195def5b39402f9564410e75res_multistep / simple / 1.0; fixed seed 20260820.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · cache state mixedI2V formal run 2 · independent repeat of SITE-3060-FL2VA-I2V-GPU0-B1.
site_benchmark8681605254c00719f7af488c13adb0124136de177c0dde02e906b198cc5cb47f22648d0a8d42cae3207e9d5b0e9def8e55afd2ae4f7d6a2b809271141de30e4020260820.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · cache state mixedrepeat_of=SITE-3060-FL2VA-I2V-GPU0-B1. The 4.4-second difference is retained as a raw observation, not a mean.These five records freeze the FL2VA T2V baseline as the A side and add the third-party 4-step Turbo LoRA as the B side on the same RTX 3060 GPU 0. Only the LoRA, steps (20 → 4) and two shift schedules (video 6 / audio 3) changed; prompt, seed, canvas, frames and audio were held identical. B-side wall times were under 30 minutes, so the protocol's three-run branch applies and a median plus range is published. This is a site_benchmark pair on one 12GB card, not a cross-environment comparison and not a reproduction of any community claim.
B-side smoke · same LoRA, 4 steps and shift 6/3 as the formal runs · 512×288 / 22 frames · not part of formal statistics.
site_benchmarkd701cee5…7193 · 134,730 bytesfbd544c2…bfded1; canvas and frames only changed from the formal file.minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16.safetensors at strength 1.0 (site value; upstream gives no recommendation); 4 steps; shift video 6 / audio 3.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9comfy-aimdo 0.4.10 · cache node none · SageAttention absent · Torch compile off · GPU 1 113→113 MiB · contaminated=falseB-side formal run 1 of 3 · 4-step Turbo LoRA · same T2V workload as the A-side baseline.
site_benchmarke3f19c0d…60bc2 · 2,650,811 bytes735972ca…d10a; reused byte-identical for all three formal runs.minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16.safetensors at strength 1.0 (site value; upstream gives no recommendation); 4 steps; shift video 6 / audio 3; res_multistep / simple; fixed seed 20260820.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9comfy-aimdo 0.4.10 · cache node none · SageAttention absent · Torch compile off · GPU 1 113→113 MiB · contaminated=false · cache state mixedB-side formal run 2 of 3 · independent repeat of SITE-3060-FL2VA-T2V-TURBO-GPU0-B1.
site_benchmarke3f19c0d…60bc2 · 2,650,811 bytes735972ca…d10a.20260820 as run 1.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9comfy-aimdo 0.4.10 · cache node none · SageAttention absent · Torch compile off · GPU 1 113→113 MiB · contaminated=false · cache state mixedRetained orchestration failure · never entered inference · FileNotFoundError on a host-path workflow before prompt submission.
site_benchmark/opt/minimax-h3/data/user/bench/…, which does not exist inside the container; FileNotFoundError. Queue stayed empty; GPU0/GPU1 peaks 331/113 MiB; no inference started./data/user/bench/… as test_id B1-RUN3-RETRY1; workflow bytes and workload parameters unchanged.11e3bf6a…09a2749 retained; node_errors not applicable because the prompt was never submitted.B-side formal run 3 of 3 · retry of the retained RUN3 failure with the corrected container path.
site_benchmarke3f19c0d…60bc2 · 2,650,811 bytes735972ca…d10a; workload parameters identical to the failed RUN3.20260820 as runs 1 and 2.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9comfy-aimdo 0.4.10 · cache node none · SageAttention absent · Torch compile off · GPU 1 113→113 MiB · contaminated=false · cache state mixedThese records share the 864×480 / 124-frame baseline family. The baseline has no cache node; Phase B records test individual third-party nodes. A node's author claim is not converted into a site claim, and the failed C5 run remains in the ledger.
Shared cache-node baseline A · cache none · 864×480 / 124 frames.
site_benchmarkmdat SHA-256 62392754…85ddbBASE-864-A-api.json · 565dfdbf326bf226b701b6fbd27eb47fe0e17ce175d1537606252222d29f0676res_multistep / simple / 1.0 · seed 20260814 · cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · reference run for Phase B speedup calculations; no cross-device ranking.Shared cache-node baseline B · independent baseline observation.
site_benchmarkmdat SHA-256 62392754…85ddbBASE-864-A-api.json · 565dfdbf326bf226b701b6fbd27eb47fe0e17ce175d1537606252222d29f067620260814, sampler res_multistep, scheduler simple, denoise 1.0, cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. Kept as a separate raw baseline observation; baselines are not collapsed into a single figure.Post-compose baseline check · verifies the deployment change did not alter the output payload.
site_benchmarkmdat box SHA-256 62392754f4b6e0ae232a68e1b86b375f20d7e4dfa681c06e1a04ccad6fd85ddbBASE-864-A-api.json · 565dfdbf326bf226b701b6fbd27eb47fe0e17ce175d1537606252222d29f067620260814, res_multistep / simple / 1.0, cache none.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. Full mdat box matched the baseline; this record is not a new cache-node result.TE-Speed-MiniMaxH3-OSS · completed with the required core patch inactive.
site_benchmarkc1dacf47bc02cb9326f7b93c69280529b93d391bACTIVE_NO_GAIN · 0.999× against the 597.0 s baseline; no acceleration conclusionmdat changed from baseline · output SHA retained in the cache reportPHASEB-C1-api.json · 95cb19d724fa65fe333792603f0e3bdb1cd6ffc7ede1ca52e066b71f2e4c9c5f20260814 · official FL2VA track · audio on.patch_model.py write failed on the read-only ComfyUI core. The node log said it returned the model unpatched, so this is not evidence of a working speed patch.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. C1 is both a compatibility and execution record; the patch failure is deliberately visible.FirstBlockCache · active cache run.
site_benchmark725973c3bfd9de6dce249bc93dc5fe27f820df31ACTIVE · 1.644× against the fixed 597.0 s comparison valuePHASEB-C2-api.json · e268ada3c0b55e18c500c41533da4c43335af1f46a7ec0fc521edaf40cccdf0520260814 · official FL2VA track · audio on.ApplyMiniMaxH3FirstBlockCache; measured preset H3 Fast — 0.10 / max 2; console recorded cached 10/20 steps.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. This is a site result under this workflow, not a claim about the repository author's advertised speed.TeaCache · independent run 1.
site_benchmark4cbb50d69c73a19a5d6ec42c5aec1989d5a04b6fACTIVE · 1.937× against the fixed 597.0 s comparison valuePHASEB-C4-RUN1-api.json · 01a093b0858f085485d43158da9dd46e08a996e9dde93e2ea91f090e6b327bd020260814 · official FL2VA track · audio on.rel_l1_thresh=0.15 · start_step=2 · end_step=-2 · total_steps=20; API node MiniMaxH3TeaCache.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. The result is a local site measurement, not a confirmation of a third-party claim.TeaCache · independent run 2; raw repeat shown beside run 1.
site_benchmark4cbb50d69c73a19a5d6ec42c5aec1989d5a04b6fACTIVE · 1.935× against the fixed 597.0 s comparison valuePHASEB-C4-RUN1-api.json · 01a093b0858f085485d43158da9dd46e08a996e9dde93e2ea91f090e6b327bd020260814, 20 steps and audio path as PB-C4-RUN1.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false · repeat_of=PB-C4-RUN1. The page retains 1.937× and 1.935× as two separate observations.comfyui-speed-minimaxH3 · compatibility failure on ComfyUI 0.31.0.
site_benchmark2f507d687cd6767212ae003d272042bf886a3cd3FAILED · no successful ComfyUI wall time · no outputAttributeError: ... time_shift_slope runtime hook missingPHASEB-C5-api.json · b17a050fd4f6d4459f7d0c2b4be50d7feb53beb0d3ef40685fd4d85cc1fd766c20260814 · official FL2VA track · audio on.MiniMax H3 Speed Cache ... estimated 1.00x; then AttributeError: module 'comfy.ldm.minimax.model' has no attribute 'time_shift_slope'. No core source was patched and no retry altered the environment.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. Client telemetry is retained, but no speedup or output hash is computed from a failed execution.FBcache-shendumao · cache profile measured without SageAttention or Sol-Attn.
site_benchmarka19b53c4507192a9cb164263335b238a454c9690ACTIVE · 1.336× against the fixed 597.0 s comparison valuePHASEB-C6-api.json · 6afd00e8095eee95c8749f0c5718b159f0d22456db329c0255a8d3a9cc7bd89320260814 · official FL2VA track · audio on.H3CombinedAcceleration; Sage Disabled, cache profile Balanced, Sol-Attn Disabled; console recorded 14 full steps and 6 cache skips.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. Local result only; no author-number comparison is calculated.AdaptiveCache · balanced preset · no Turbo LoRA or Sol-Attn.
site_benchmarkdc35065676617055aa90c4acd8b926cb265873deACTIVE · 1.581× against the fixed 597.0 s comparison valuePHASEB-C7-api.json · 42a8a8495823d54f6f1ad8bd9ddd60ce881af0e3a97106560fe62621b4f4f3bf20260814 · official FL2VA track · audio on.MiniMaxH3AdaptiveCache; measured preset balanced; cache_device=auto; advanced node not used.bf4c9a08fc854df6d3b2bef1b92b509e2ef2d2c9contaminated=false. Local result only; no cross-GPU ranking.Cache-node compatibility · checked 2026-08-14
C1–C8 are compatibility-screening objects, not eight invented performance records. Where a Phase B run exists, its test_id links to the ledger. “Not measured” means no performance run was generated.
| Node | Commit | Install | Load / visible | Phase B record | Observed note |
|---|---|---|---|---|---|
| C1 · TE-Speed-MiniMaxH3-OSS | c1dacf47…d391b | Partial | Yes / yes | PB-C1 | Core patch failed on read-only filesystem; completed run returned to stock speed. |
| C2 · FirstBlockCache | 725973c3…0df31 | Yes | Yes / yes | PB-C2 | API node visible; active result recorded. |
| C3 · MiniMaxH3-Cache | 8a45e096…73819 | Yes | Yes / yes | Not measured | Dynamic in-memory patch observed; README/documentation boundary kept it out of Phase B. |
| C4 · TeaCache | 4cbb50d6…04b6f | Yes | Yes / yes | PB-C4-RUN1 + RUN2 | Two raw active results; 1.937× and 1.935× are both retained. |
| C5 · comfyui-speed-minimaxH3 | 2f507d68…a3cd3 | Yes | Yes / yes | PB-C5 failed | Runtime hook raised missing time_shift_slope on ComfyUI 0.31.0. |
| C6 · FBcache-shendumao | a19b53c4…c9690 | Yes | Yes / yes | PB-C6 | Balanced cache profile; Sage and Sol-Attn disabled. |
| C7 · AdaptiveCache | dc350656…873de | Yes | Yes / yes | PB-C7 | Balanced preset; advanced node not used. |
| C8 · TE-Speed-MiniMaxH3 | e8b6549b…61b135 | Yes | No / no | Not measured | ModuleNotFoundError for the repository's Windows binary on this Linux environment. |
Compatibility source: cache-node benchmark report. Compatibility labels describe this checked environment; they are not universal support claims.
What these runs establish
Each statement is tied to records above and carries its conditions. None is a cross-GPU ranking or a community-result verdict.
In the R2V GPU 0 smoke-to-formal pair, the frame-pixel workload is approximately 40×. Peak VRAM moved from 11,591 to 11,649 MiB (+0.5%), while peak RAM moved from 43,176 to 43,587 MiB (+1.0%). The smoke was cold and the formal run warm, so this is an observation about this pair, not a scaling law.
The seven Phase B execution records, including the failed C5 telemetry record, span 11,163–11,773 MiB of recorded peak VRAM. C5 has no successful output, and C3/C8 have no Phase B performance record; the range is not evidence that all eight nodes work.
The two formal GPU 0 R2V runs each produced the same output SHA-256 prefix recorded in the cards: cc2a6f75…be6060f. Under this fixed seed, workflow and hardware setup, the outputs were byte-identical. The wall times remain 2,581.8 s and 2,579.8 s as separate observations.
Protocol in one screen
site_benchmark, not a community reproduction.Full protocol: site-reproduction-protocol.md, version 0.2-draft.
Boundaries
Community reports remain community-reported. Their setup, timing and claims are not inserted into a site-measured cell.
Different cards, environments and workflows cannot be reduced to one leaderboard. The 8GB row is not a support or failure verdict.
The Turbo LoRA A/B ran once on this owned RTX 3060 12GB bench with the same prompt, seed, canvas, frames and workflow as the FL2VA T2V baseline, changing only the LoRA, steps and shift schedules. The same-environment pairing gives a 4.113×–4.132× wall-time range; it is not a cross-environment comparison and it is not placed beside any community "5×" claim. All three B-side runs peaked above 8,192 MiB, so this page makes no 8GB support or failure claim.
Repeat values are separate elements with separate IDs. If a future run changes, the raw ledger can show what changed.
FAQ
It means MiniMax H3 Tutorial has not run that GPU, precision and task combination. It is not a statement that the combination is unsupported or impossible.
The site protocol requires an independent repeat for a formal benchmark lasting at least 30 minutes. Both test IDs and both original values stay visible; this page does not replace them with an average.
No. These records use the site_benchmark track and their own pinned workflows, settings and hardware conditions. They are not presented as a reproduction, validation or confirmation of a community result.
No conclusion is made here. The 8GB rental-class row is explicitly not tested, and a result on the owned RTX 3060 12GB bench is not an extrapolation to another card, memory size, precision or task.
On one owned RTX 3060 12GB bench, the same FL2VA T2V workload with the 4-step Turbo LoRA took a median 528.6 seconds (range 526.7 to 528.9 seconds) against an A-side baseline of 2,175.5 and 2,176.2 seconds — a same-environment 4.113× to 4.132× wall-time range, not an average. All three Turbo runs peaked above 8,192 MiB VRAM, so the page makes no 8GB support or failure claim.
Next checks