跳到主内容
@wquguru
精选80Chubby♨️模型发布/更新多源精选 ×3

MiniMax发布33B视频模型H3,可生成同步立体声

Holy, those insanae releases: MiniMax released a 33B video model that generates…

原文
发到 X

Holy, those insanae releases: MiniMax released a 33B video model that generates synchronized stereo audio and runs on one RTX 5090.

H3 combines text, images, video and audio references for generation and editing, with clips up to 15 seconds.

ComfyUI’s optimized stack is roughly 40GB, using dynamic RAM/SSD offloading to fit consumer GPUs. Early 5090 (!) tests produced five seconds at native 768p-class resolution in around 5.5 minutes.

MiniMax calls it open source, but several core pieces remain server-side: context orchestration, 2K regeneration and sparse attention.

Open weights. Closed quality stack. Restricted geography. However, H3 is a major step for local video, but also a reminder that “downloadable” does not mean open source in general.

Btw: It also cannot legally be used under its public license in Germany, the EU, US, UK or South Korea.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →