Skip to content

fix(minimax-h3): align Ref2VA audio and conditioning paths - #81

Merged
Komorebi623 merged 3 commits into
THU-MIG:mainfrom
Komorebi623:main
Aug 18, 2026
Merged

fix(minimax-h3): align Ref2VA audio and conditioning paths#81
Komorebi623 merged 3 commits into
THU-MIG:mainfrom
Komorebi623:main

Conversation

@Komorebi623

Copy link
Copy Markdown
Collaborator

No description provided.

- read prompt files before generation when no inline prompt is supplied
- encode MiniMax-H3 reference audio per stereo channel and resample standalone audio consistently
- build Ref2VA presentation as ordered reference segments for image/video/audio inputs
- share condition noise RNG across reference latents and add debug dumps for alignment checks
- enforce a minimum MiniMax-H3 generation canvas to avoid invalid tiny outputs
- document MiniMax-H3 output resolution constraints
Remove temporary tensor-load/dump and Qwen debug-target plumbing that was accidentally included with the Ref2VA fixes. Keep the functional audio, conditioning, prompt-file, and resolution changes intact.
@Komorebi623
Komorebi623 merged commit 1822d78 into THU-MIG:main Aug 18, 2026
2 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant