image imagewidth (px) 512 1.02k |
|---|
- π§ Repository Map & Quick Navigation
- π¬ MiniMax-H3 Center (
h3-center/) - π¨ FLUX.2 & Klein 9B RefMods (
klein9/) - β‘ Z Image Base vs. Z Image Turbo Training Suite (
zimage-turbo-vs-base-training/) - π οΈ Production Training Scripts & Templates (
training-scripts/) - π₯ Video Samples & Benchmarks
- π Tutorials & Articles Collection
- π Connect with malcolmrey
malcolmrey's Various AI Model, Architecture & Research Repository
Welcome to the central research and asset repository of malcolmrey. This repository hosts cutting-edge tools, custom architectures, RefMod latent adapter systems, video synthesis engines, training configurations, benchmark suites, cinematic scripts, and comprehensive educational guides spanning MiniMax-H3, FLUX.2 / Klein 9B, WAN 2.1, LTX-Video, Z-Image, SDXL, and Stable Diffusion.
π§ Repository Map & Quick Navigation
| Section / Directory | Focus Area & Description | Key Resources & Direct Links |
|---|---|---|
π¬ h3-center/ |
MiniMax-H3 Video & Cinema Hub Complete prompt guides, RefMod stacking, 1,500+ character tests, and 9-episode Crossovers cinematic series. |
β’ Prompting Guide β’ RefMod Stacking Guide β’ Multistacking Benchmark Demos β’ Known Characters Usability Index β’ Crossovers Cinema Series |
π¨ klein9/ |
FLUX.2 & Klein 9B RefMods Instant reference latent adapter ecosystem, custom ComfyUI nodes, CLI batch extractor, and benchmark suites. |
β’ RefMod Architecture Guide β’ ComfyUI Node β’ Batch Extractor Tool β’ Workflows | Visual Benchmarks |
β‘ zimage-turbo-vs-base-training/ |
Z Image Base vs. Turbo Benchmark Head-to-head training comparison across OneTrainer, MalTrainer, and AI Toolkit with datasets and weights. |
β’ Benchmark Overview & Configs β’ Datasets, AdamW/Prodigy configs, & .safetensors weights |
π οΈ training-scripts/ |
Production Training Configs Ready-to-use recipes for OneTrainer, Musubi Tuner, AI Toolkit, and MalTrainer. |
β’ OneTrainer Templates (Ernie, FK9, Krea2) β’ Musubi Pipelines (Ernie, LTX23) β’ AI Toolkit & MalTrainer |
| π₯ Sample Media & Benchmarks | Model Generation Sample Libraries High-fidelity outputs and verification test suites. |
β’ ltx23-samples/ (20 LTX-Video 2.3 renders)β’ ernie-samples/ (ernie-samples.zip)β’ samples/fk9/ (Rose Byrne & Sydney Sweeney outputs) |
| π Guides & Articles | Comprehensive Knowledge Base In-depth technical guides for model training, optimization, and prompt engineering. |
β’ WAN 2.1 LoRA Training Tutorial β’ CivitAI Articles Collection |
π¬ MiniMax-H3 Center (h3-center/)
The MiniMax-H3 Center is an all-in-one research laboratory and production studio for the MiniMax Hailuo 01 / H3 text-to-video, image-to-video (I2V), reference-to-video (R2V), and clip-to-video (C2V) architectures.
π In-Depth Guides & Technical Documentation
- π MiniMax-H3 Video Prompt Writing Guide (v1.3.0): The authoritative standard for prompting MiniMax-H3. Covers pinpoint actor casting, character tag binding (
<Subject 1>), timestamped action choreography (through 00:08.000), diegetic soundscapes, non-diegetic audio rules, camera vectors, and anti-patterns. - 𧬠RefMod Stacking & Multi-Subject Architecture Guide: Guide on injecting multiple pre-encoded latent references (
.safetensors), single-persona Triple-RefMod fidelity stacking, conceptual body shape boosters, dual-persona shared frame conditioning, dual audio lip-syncing, and polyphonic counterpoint duets. - π¬ Multi-Stacking Benchmark Demos & Video Suite: Verified
.mp4benchmarks demonstrating single-persona Triple RefMods (Billie Eilish), dual conversations (Billie & Miley Cyrus), simultaneous two-shots, dual audio conditioning, and counterpoint singing duets with local RefMod latent files. - π C2V Continuous Chain Generation Guide: Pipeline for chaining 5-second to 15-second video clips losslessly into seamless continuous takes.
- π΅ Audio-Synchronized & C2V Music Videos Guide: Techniques for aligning video cadence and singing motion with pre-recorded vocal tracks.
- βοΈ RefMods Installation & Usage Guide & RefMod Creation Guide: Step-by-step instructions for extracting RefMod
.safetensorsfrom images and loading them into ComfyUI pipelines. - βοΈ RefMods vs. Reference Images Comparison: Architectural analysis of speed, latent fidelity, and token efficiency comparing pre-computed RefMods against raw pixel reference workflows.
π₯ Known Characters & Usability Database
- π Usability Index (INDEX.md) & Unusable Index (INDEX_BAD.md): Empirical test results evaluating over 1,500+ characters, actors, and public figures directly in MiniMax-H3:
- Good (540+ subjects): Characters with zero-shot likeness locks (e.g., House, Dexter, Michael Scott, Walter White, Gandalf, Wednesday Addams, Dean Winchester).
- On the Fence (90+ subjects): Subjects requiring targeted prompt tuning or RefMod assistance.
- Bad (920+ subjects): Subjects that fail without custom LoRAs or dedicated RefMods.
π₯ Crossovers Episodic Cinema Series
- πΏ Crossovers Series Master Catalog: 9 full-length narrative episodes created with MiniMax-H3, complete with rendered
.mp4video files, cast listings, and synopses:- Episode 1: The Contagion of Secrets (House, Monk, Dwight, Dexter, Malcolm Reynolds, Mulder, Penny)
- Episode 2: The Cosmic Extradition (Dean Winchester, Mal Reynolds, Sherlock, Pam Beesly, Lucifer, 10th Doctor, Saul Goodman, Steve Rogers, John Locke, Tony Stark)
- Episode 3: Lockdown at Sabre Tower (Michael Scott, Wednesday Addams, John McClane, Seeley Booth, James Bond, Steve Rogers, Elliot Alderson, Lucifer, Jack Bauer)
- Episode 4: The War for the Obsidian Portal (Maleficent, Daenerys, Gandalf, Dr. Strange, Hermione, Maximus, Jack Sparrow)
- Episode 5: The Flat Earth Experience (Joe Rogan, Kanye West, Neil deGrasse Tyson, Dave Chappelle, Sir Anthony Hopkins)
- Episode 6: Dave Chappelle: The Flat Earth Special (Stand-up comedy special)
- Episode 7: Dr. Robert Ford: On Meaning (Philosophical monologue from Westworld)
- Episode 8: Ellis Boyd 'Red' Redding: On Hope and Love (Shawshank monologue)
- Episode 9: Captain Malcolm Reynolds: Alone on Serenity (Firefly character study)
- π Show & Character Guides (
h3-center/crossovers/docs/shows/): Over 250+ show guides detailing character wardrobe, distinctive traits, and scene rules.
π¨ FLUX.2 & Klein 9B RefMods (klein9/)
RefMods (Reference Latent Adapters) provide zero-shot facial fidelity and concept locking for FLUX.2, Klein 4B, Klein 8B, and Klein 9B without the VRAM overhead or training time required by full fine-tunes or LoRAs.
- π Interactive Model Browser: https://huggingface.co/spaces/malcolmrey/browser (Explore 1,400+ pre-extracted RefMods)
- π€ Model Weights Hub: https://huggingface.co/malcolmrey/klein9 (Download
.safetensorsadapters) - π ComfyUI Custom Node:
klein9/comfyui/custom_nodes/ComfyUI-Flux2Klein9Mod/(Also available on GitHub) - β‘ CLI Generator & Mass Batch Extractor:
klein9/generate_flux2_klein9_refmod.py - π Architecture & Extraction Guide:
klein9/docs/FLUX2_KLEIN9_REFMODS_GUIDE.md - π Visual Benchmarks & Comparisons:
klein9/docs/VISUAL_BENCHMARKS_AND_COMPARISONS.md - π Production Workflows:
- Pure RefMod Text-to-Image:
klein9/workflows/workflow_klein9_refmod.json - Hybrid RefMod + LoRA Stack:
klein9/workflows/workflow_klein9_refmod_lora.json
- Pure RefMod Text-to-Image:
- πΌοΈ Sample Generations:
samples/fk9/(Rose Byrne and Sydney Sweeney benchmark runs)
β‘ Z Image Base vs. Z Image Turbo Training Suite (zimage-turbo-vs-base-training/)
An empirical training study benchmarking Z Image (Base) against Z Image (Turbo) using identical curated datasets (Felicia Day, Olivia).
- π Overview & Documentation
- π Dataset: Curated 23-image high-resolution dataset
- βοΈ OneTrainer Configs: AdamW and Prodigy optimization profiles (
.json) - π€ MalTrainer & AI Toolkit Recipes: Full training configurations (
.yaml) - π¦ Trained Weights: Verified
.safetensorsmodels for both Base and Turbo versions
π οΈ Production Training Scripts & Templates (training-scripts/)
Curated configurations for high-efficiency model training across popular trainers:
- π’ OneTrainer (
training-scripts/onetrainer/):fk9_template.json&fk9_prodigy_template.jsonβ Klein 9B / FLUX.2 LoRA templatesernie_template.jsonβ Ernie training configurationkrea2prodigy_template.jsonβ Krea 2 Prodigy training template
- π£ Musubi Tuner (
training-scripts/musubi/):musubi/ernie/β Multi-resolution Ernie training scriptsmusubi/ltx23/β LTX-Video 2.3 video model training configurations
- π΅ AI Toolkit (
training-scripts/aitoolkit/):- Config templates for WAN 2.1, Flux, and SDXL
- π MalTrainer (
training-scripts/maltrainer/):- Training configs for malcolmrey's proprietary trainer
π₯ Video Samples & Benchmarks
- ποΈ
ltx23-samples/: 20 full-motion test videos evaluating coherence, prompt fidelity, and motion dynamics with LTX-Video 2.3. - π¦
ernie-samples/: Comprehensive benchmark sample package (ernie-samples.zip). - π¬
h3-center/crossovers/: 8 rendered.mp4full narrative crossover episodes. - πΌ
h3-center/docs/multistacking/: 8 dual-audio and counterpoint singing benchmark videos.
π Tutorials & Articles Collection
π WAN 2.1 LoRA Training Tutorial
Complete step-by-step guide on training WAN 2.1 LoRA models using AI Toolkit:
- Dataset preparation (20 optimal images, 2500 steps)
- Resolution bucketing & learning rate schedules
- VRAM management & cloud GPU configurations
- ComfyUI inference workflow integration
π CivitAI Articles Collection
Curated directory of malcolmrey's published training and generation guides:
- DreamBooth / LyCORIS / LoRA Complete Guide
- SDXL LoRA Training on RunPod
- Textual Inversion / Embedding Training Guide
- Flux Guide, Part I - LoRA Training
- Improving Results with Multi-Model Blending (Turning it to 11!)
- Quality Deep Dive (Bringing it up to Twelve!)
- ADetailer, Inpainting & Facial Restoration Best Practices
π Connect with malcolmrey
- π¬ Discord Community: malcolmrey's place
- β Support / Priority Requests: Buy Me a Coffee
- π€ Hugging Face Hub: https://huggingface.co/malcolmrey
- π Interactive Model Browser: https://huggingface.co/spaces/malcolmrey/browser
- π¨ CivitAI Profile: https://civitai.com/user/malcolmrey
- π¬ Reddit Community: r/malcolmrey
- Downloads last month
- 62,169