Dataset Viewer
Auto-converted to Parquet Duplicate
Search is not available for this dataset
image
imagewidth (px)
512
1.02k

malcolmrey's Various AI Model, Architecture & Research Repository

Welcome to the central research and asset repository of malcolmrey. This repository hosts cutting-edge tools, custom architectures, RefMod latent adapter systems, video synthesis engines, training configurations, benchmark suites, cinematic scripts, and comprehensive educational guides spanning MiniMax-H3, FLUX.2 / Klein 9B, WAN 2.1, LTX-Video, Z-Image, SDXL, and Stable Diffusion.


🧭 Repository Map & Quick Navigation

Section / Directory Focus Area & Description Key Resources & Direct Links
🎬 h3-center/ MiniMax-H3 Video & Cinema Hub
Complete prompt guides, RefMod stacking, 1,500+ character tests, and 9-episode Crossovers cinematic series.
β€’ Prompting Guide
β€’ RefMod Stacking Guide
β€’ Multistacking Benchmark Demos
β€’ Known Characters Usability Index
β€’ Crossovers Cinema Series
🎨 klein9/ FLUX.2 & Klein 9B RefMods
Instant reference latent adapter ecosystem, custom ComfyUI nodes, CLI batch extractor, and benchmark suites.
β€’ RefMod Architecture Guide
β€’ ComfyUI Node
β€’ Batch Extractor Tool
β€’ Workflows | Visual Benchmarks
⚑ zimage-turbo-vs-base-training/ Z Image Base vs. Turbo Benchmark
Head-to-head training comparison across OneTrainer, MalTrainer, and AI Toolkit with datasets and weights.
β€’ Benchmark Overview & Configs
β€’ Datasets, AdamW/Prodigy configs, & .safetensors weights
πŸ› οΈ training-scripts/ Production Training Configs
Ready-to-use recipes for OneTrainer, Musubi Tuner, AI Toolkit, and MalTrainer.
β€’ OneTrainer Templates (Ernie, FK9, Krea2)
β€’ Musubi Pipelines (Ernie, LTX23)
β€’ AI Toolkit & MalTrainer
πŸŽ₯ Sample Media & Benchmarks Model Generation Sample Libraries
High-fidelity outputs and verification test suites.
β€’ ltx23-samples/ (20 LTX-Video 2.3 renders)
β€’ ernie-samples/ (ernie-samples.zip)
β€’ samples/fk9/ (Rose Byrne & Sydney Sweeney outputs)
πŸ“š Guides & Articles Comprehensive Knowledge Base
In-depth technical guides for model training, optimization, and prompt engineering.
β€’ WAN 2.1 LoRA Training Tutorial
β€’ CivitAI Articles Collection

🎬 MiniMax-H3 Center (h3-center/)

The MiniMax-H3 Center is an all-in-one research laboratory and production studio for the MiniMax Hailuo 01 / H3 text-to-video, image-to-video (I2V), reference-to-video (R2V), and clip-to-video (C2V) architectures.

πŸ“– In-Depth Guides & Technical Documentation

  • πŸ“˜ MiniMax-H3 Video Prompt Writing Guide (v1.3.0): The authoritative standard for prompting MiniMax-H3. Covers pinpoint actor casting, character tag binding (<Subject 1>), timestamped action choreography (through 00:08.000), diegetic soundscapes, non-diegetic audio rules, camera vectors, and anti-patterns.
  • 🧬 RefMod Stacking & Multi-Subject Architecture Guide: Guide on injecting multiple pre-encoded latent references (.safetensors), single-persona Triple-RefMod fidelity stacking, conceptual body shape boosters, dual-persona shared frame conditioning, dual audio lip-syncing, and polyphonic counterpoint duets.
  • 🎬 Multi-Stacking Benchmark Demos & Video Suite: Verified .mp4 benchmarks demonstrating single-persona Triple RefMods (Billie Eilish), dual conversations (Billie & Miley Cyrus), simultaneous two-shots, dual audio conditioning, and counterpoint singing duets with local RefMod latent files.
  • πŸ”— C2V Continuous Chain Generation Guide: Pipeline for chaining 5-second to 15-second video clips losslessly into seamless continuous takes.
  • 🎡 Audio-Synchronized & C2V Music Videos Guide: Techniques for aligning video cadence and singing motion with pre-recorded vocal tracks.
  • βš™οΈ RefMods Installation & Usage Guide & RefMod Creation Guide: Step-by-step instructions for extracting RefMod .safetensors from images and loading them into ComfyUI pipelines.
  • βš–οΈ RefMods vs. Reference Images Comparison: Architectural analysis of speed, latent fidelity, and token efficiency comparing pre-computed RefMods against raw pixel reference workflows.

πŸ‘₯ Known Characters & Usability Database

  • πŸ“Š Usability Index (INDEX.md) & Unusable Index (INDEX_BAD.md): Empirical test results evaluating over 1,500+ characters, actors, and public figures directly in MiniMax-H3:
    • Good (540+ subjects): Characters with zero-shot likeness locks (e.g., House, Dexter, Michael Scott, Walter White, Gandalf, Wednesday Addams, Dean Winchester).
    • On the Fence (90+ subjects): Subjects requiring targeted prompt tuning or RefMod assistance.
    • Bad (920+ subjects): Subjects that fail without custom LoRAs or dedicated RefMods.

πŸŽ₯ Crossovers Episodic Cinema Series

  • 🍿 Crossovers Series Master Catalog: 9 full-length narrative episodes created with MiniMax-H3, complete with rendered .mp4 video files, cast listings, and synopses:
    • Episode 1: The Contagion of Secrets (House, Monk, Dwight, Dexter, Malcolm Reynolds, Mulder, Penny)
    • Episode 2: The Cosmic Extradition (Dean Winchester, Mal Reynolds, Sherlock, Pam Beesly, Lucifer, 10th Doctor, Saul Goodman, Steve Rogers, John Locke, Tony Stark)
    • Episode 3: Lockdown at Sabre Tower (Michael Scott, Wednesday Addams, John McClane, Seeley Booth, James Bond, Steve Rogers, Elliot Alderson, Lucifer, Jack Bauer)
    • Episode 4: The War for the Obsidian Portal (Maleficent, Daenerys, Gandalf, Dr. Strange, Hermione, Maximus, Jack Sparrow)
    • Episode 5: The Flat Earth Experience (Joe Rogan, Kanye West, Neil deGrasse Tyson, Dave Chappelle, Sir Anthony Hopkins)
    • Episode 6: Dave Chappelle: The Flat Earth Special (Stand-up comedy special)
    • Episode 7: Dr. Robert Ford: On Meaning (Philosophical monologue from Westworld)
    • Episode 8: Ellis Boyd 'Red' Redding: On Hope and Love (Shawshank monologue)
    • Episode 9: Captain Malcolm Reynolds: Alone on Serenity (Firefly character study)
  • πŸ“š Show & Character Guides (h3-center/crossovers/docs/shows/): Over 250+ show guides detailing character wardrobe, distinctive traits, and scene rules.

🎨 FLUX.2 & Klein 9B RefMods (klein9/)

RefMods (Reference Latent Adapters) provide zero-shot facial fidelity and concept locking for FLUX.2, Klein 4B, Klein 8B, and Klein 9B without the VRAM overhead or training time required by full fine-tunes or LoRAs.


⚑ Z Image Base vs. Z Image Turbo Training Suite (zimage-turbo-vs-base-training/)

An empirical training study benchmarking Z Image (Base) against Z Image (Turbo) using identical curated datasets (Felicia Day, Olivia).

  • πŸ“‹ Overview & Documentation
  • πŸ“ Dataset: Curated 23-image high-resolution dataset
  • βš™οΈ OneTrainer Configs: AdamW and Prodigy optimization profiles (.json)
  • πŸ€– MalTrainer & AI Toolkit Recipes: Full training configurations (.yaml)
  • πŸ“¦ Trained Weights: Verified .safetensors models for both Base and Turbo versions

πŸ› οΈ Production Training Scripts & Templates (training-scripts/)

Curated configurations for high-efficiency model training across popular trainers:


πŸŽ₯ Video Samples & Benchmarks

  • 🎞️ ltx23-samples/: 20 full-motion test videos evaluating coherence, prompt fidelity, and motion dynamics with LTX-Video 2.3.
  • πŸ“¦ ernie-samples/: Comprehensive benchmark sample package (ernie-samples.zip).
  • 🎬 h3-center/crossovers/: 8 rendered .mp4 full narrative crossover episodes.
  • 🎼 h3-center/docs/multistacking/: 8 dual-audio and counterpoint singing benchmark videos.

πŸ“š Tutorials & Articles Collection

πŸ“˜ WAN 2.1 LoRA Training Tutorial

Complete step-by-step guide on training WAN 2.1 LoRA models using AI Toolkit:

  • Dataset preparation (20 optimal images, 2500 steps)
  • Resolution bucketing & learning rate schedules
  • VRAM management & cloud GPU configurations
  • ComfyUI inference workflow integration

πŸ“š CivitAI Articles Collection

Curated directory of malcolmrey's published training and generation guides:


🌐 Connect with malcolmrey

Downloads last month
62,169

Models trained or fine-tuned on malcolmrey/various