frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

https://blog.comfy.org/p/minimax-h3-day-0-support-in-comfyui
61•vblanco•1h ago

Comments

SV_BubbleTime•51m ago
I saw the samples people have posted. Immediately deleted LTX2 and WAN folders. Those are completely worthless now.

There is some debate on the license for those in the US, UK, EU, plus… no comment other than whew those samples though!

Maxious•26m ago
"Regions such as the EU, UK, South Korea, and the US are currently developing or enforcing AI-related regulations that may have specific implications for generative video models"

You just have to pinkie promise you won't make disney mad and they will send you a licence https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/Q...

SV_BubbleTime•19m ago
I do work animations for fun and internal use only (mostly jokes). I MAY reach out to them.
rvz•48m ago
Hollywood and the film industry on red alert. Too bad.

This is AGI.

trwhite•40m ago
The example video just looks like the highly produced art (TV, commercials, games) other people have created. I find it impossible to believe this wasn't trained on other people's work, and there is no protection for it. Terribly sad. A lack of original thinking is coming.
MSFT_Edging•39m ago
Can't wait for netflix to be topped on unimaginative nonsense meant to have in the background while you scroll your phone.
toasty228•26m ago
Unlimited slop "content" to fill decomposed brain shaped vessels, the future is bright!
kyriakos•15m ago
Unfortunately Hollywood been feeding us human generated slop for a long time already. Transition won't be hard.
Mashimo•37m ago
> The result gives a total memory footprint reduced by 66%, from 123.6 GB in full precision to 42.5 GB with the smallest models variants. Combining this with our dynamic VRAM offloading enables a next-generation 2K video model to run locally on a GPU like the RTX 3060.

Pretty cool.

But assuming you have a 16GB 3060, how long would it take to generate a 15 second clip?

sheesdev•36m ago
The mouse render is surprisingly good. Several of those clips stood out to be a pretty big leap in terms of current SOTA models.

The only one that looks "off" is the beverage ad video during the can opening clip, it still has that "AI smoothening" effect. Good thing this can be done pretty well using traditional rendering.

I feel like for a good while now we'll transition into a process that uses traditional "close-up" rendering/shots + AI generated wide-shots or quick cuts.

Exciting, but also troubling. This being open-weights is a massive win for the community though.

vblanco•35m ago
Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.
robbru•22m ago
1 minute for a second of footage, that is awesome! Thanks for sharing.
mwigdahl•20m ago
If you wouldn't mind sharing, what's your Comfy workflow for this? I have the same video card setup and would like to give it a shot.
vblanco•19m ago
just the default one in the link for image-to-video
Maxious•19m ago
FWIW on a 5080 16GB it takes 3 minutes for 10 seconds 480p video (the mouse video workflow with length changed from 5 seconds to 10 seconds)
ddevnyc•18m ago
I am particularly curious how multimodal models will work with types of knowledge that are inherently non-text. For example, SOTA LLMs really suck at electronics, especially analog electronics.

Is MiniMax H3 capable of logical / technical reasoning, or is it purely art oriented?

fodkodrasz•29m ago
On one hand: impressive. On the other hands aesthetically it all looks painfully bland and generic.
SV_BubbleTime•16m ago
This, today, is the absolute worse this model will ever be. Chill.
embedding-shape•19m ago
> We found that the model's modulation weights (~40% of the total parameters) could be pruned and replaced with a functionally equivalent lookup table, dramatically shrinking the memory footprint with no loss in output quality.

Is this a common approach to reducing weights with "no loss in output quality", assuming this is true? Seems almost too simple to work. If this is doable, would this be applicable to LLMs as well?

Neat with native frame-to-frame generation, but wonder how easy it is to "link" together clips at the intersection, typically the models kind of lose the "momentum" across these stiches, being able to merge things with frame-to-frame between clips might help with this it feels like.

_diyar•3m ago
Also begs the question whether this is applicable for high-throughput applications on FPGAs, which are to my novice mind basically LUTs, right?

I remember a paper which was posted on HN a few weeks ago where somebody implemented KAN networks in FPGAs, since those can readily be approximated as LUTs.

satvikpendem•17m ago
I've said it before and I'll say it again, human directors are still valuable, as they use AI video editing tools to generate the shots they want and put them together in a cohesive way. Previously they might've used film and actors but if they can just prompt the AI (or create workflows as seen with ComfyUI) then they arrange them together just like how an EDM producer doesn't actually play the instruments but instead the creativity is in the arrangement.

I suspect it'll be quite a while until AI gets a good enough aesthetic sense to do this, as even with static HTML websites humans can easily see that it's AI slop.

How Does the Qatar-Donated Air Force One Compare with Other Presidential Jets?

https://www.wsj.com/politics/national-security/how-does-the-qatar-donated-air-force-one-compare-w...
1•thm•58s ago•0 comments

Seven books I keep close because I love them

https://blog.plover.com/book/elbow-shelf.html
1•Brajeshwar•4m ago•0 comments

The Deleted Face: Why Facial Recognition Must Erase You First

https://smarterarticles.co.uk/the-deleted-face-why-facial-recognition-must-erase-you-first
1•dxs•4m ago•0 comments

Text and Test Anything Protocol

http://dtap.sparrowhub.io/doc/README
1•melezhik•5m ago•0 comments

The Mental Models I Use to Work with AI

https://metedata.substack.com/p/015-the-mental-models-i-use-to-work
1•young_mete•6m ago•0 comments

Fix Your Analog Life First

https://katherinemartinko.substack.com/p/fix-your-analog-life-first
1•speckx•7m ago•0 comments

Quantization hurts knowledge nonlinearly – Qwen3.6 27B case study

https://quesma.com/blog/quantization-hurts-knowledge/
1•stared•7m ago•0 comments

WAN 3.0 Deep Dive: Video Model Evolution from Open-Source Breakthrough

https://wan3video.art/blog/wan-3-0-deep-dive
2•wangneo276•8m ago•0 comments

How the heck do you catch a fly ball?

https://perthirtysix.com/how-the-heck-do-you-catch-a-fly-ball
1•robmoore•9m ago•0 comments

Europe opens bidding for seven AI 'gigafactories' in a €30B bid to catch up

https://thenextweb.com/news/eu-ai-gigafactories-call-30bn
1•doener•9m ago•0 comments

Show HN: Lunar, a "fast", memory-efficient Lua 5.1 VM written in Go

https://github.com/mmcdole/lunar
2•drakenot•11m ago•1 comments

My Personal Software Journey: Self-Hosting Agent-Built Apps

https://metedata.substack.com/p/016-my-personal-software-journey
2•young_mete•11m ago•0 comments

Why does Mail app contact iCloud when sending a non-iCloud email?

https://lapcatsoftware.com/articles/2026/8/2.html
2•cdrnsf•11m ago•0 comments

NASA's New 3D Model Shows the Earth Is a Lumpy Mess

https://www.wired.com/story/nasas-new-3d-model-shows-lumpy-earth/
1•haritha-j•11m ago•0 comments

Google plans to exempt sanctioned nations from Android developer verification

https://arstechnica.com/gadgets/2026/07/google-plans-to-exempt-sanctioned-nations-from-android-de...
2•nostrademons•13m ago•0 comments

IBM: Three Demonstrations Prove Quantum Advantage Has Been Reached

https://www.nextplatform.com/compute/2026/07/31/ibm-three-demonstrations-prove-quantum-advantage-...
1•rbanffy•14m ago•0 comments

A 21-hour dip could help explain Tabby's star's decade-long dimming mystery

https://phys.org/news/2026-07-hour-dip-tabby-star-decade.html
1•rbanffy•14m ago•0 comments

Restoring the Java Ring on a Japanese TV Show [video]

https://www.youtube.com/watch?v=LF5PrzP9YTs
1•kidmin•15m ago•0 comments

Show HN: jit - just-in-time secrets for your Mac, gated by Touch ID

https://github.com/jitpass/jit
1•bukershok•15m ago•1 comments

Sovereign Tech Fellowship for Rust Maintenance (June-July 2026 Report)

https://kobzol.github.io/rust/2026/08/03/stf-june-july-2026.html
1•Tomte•16m ago•0 comments

Dbctl – a terminal-native DB client with built-in SSH/SSM tunneling

https://github.com/stivio00/dbctl
1•stivio00•17m ago•0 comments

Transfer Files with QR Codes

https://decimen.app/
1•ramsj•19m ago•0 comments

Hank Green apologizes for relying too heavily on ChatGPT

https://mashable.com/life/hank-green-chatgpt-ai-apology
1•bjourne•19m ago•0 comments

Google Chrome may soon block New Tab hijacker extensions by default

https://www.bleepingcomputer.com/news/google/google-chrome-may-soon-block-new-tab-hijacker-extens...
2•Brajeshwar•19m ago•0 comments

3.2-Gigapixel LSST Camera Captures over Half a Million Galaxies

https://petapixel.com/2026/08/02/3-2-gigapixel-lsst-camera-captures-over-half-a-million-galaxies/
2•larodi•20m ago•1 comments

eek! it's a tiny rust LLM proxy in ~1k loc

https://github.com/Liana64/eek
1•saidarembrace•20m ago•1 comments

A quine in Piet – a GIF image that prints itself [video]

https://www.youtube.com/watch?v=GwMtzhjCzyc
1•surprisetalk•20m ago•0 comments

DARPA Lift Challenge – Aug 3 – Main Broadcast [video]

https://www.youtube.com/watch?v=KEGWvve7rls
1•stefan_•21m ago•0 comments

AUR malware wave forces Arch Linux to disable every package push

https://runtimewire.com/article/aur-malware-wave-forces-arch-linux-to-disable-every-package-push
3•ryanmerket•22m ago•0 comments

PostScript: Idle Frontier and on developing games with AI

https://ljvmiranda921.github.io/projects/2026/08/03/idle-frontier/
1•ljvmiranda•22m ago•1 comments