frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

https://blog.comfy.org/p/minimax-h3-day-0-support-in-comfyui
62•vblanco•1h ago

Comments

SV_BubbleTime•55m ago
I saw the samples people have posted. Immediately deleted LTX2 and WAN folders. Those are completely worthless now.

There is some debate on the license for those in the US, UK, EU, plus… no comment other than whew those samples though!

Maxious•29m ago
"Regions such as the EU, UK, South Korea, and the US are currently developing or enforcing AI-related regulations that may have specific implications for generative video models"

You just have to pinkie promise you won't make disney mad and they will send you a licence https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/Q...

SV_BubbleTime•22m ago
I do work animations for fun and internal use only (mostly jokes). I MAY reach out to them.
rvz•51m ago
Hollywood and the film industry on red alert. Too bad.

This is AGI.

trwhite•43m ago
The example video just looks like the highly produced art (TV, commercials, games) other people have created. I find it impossible to believe this wasn't trained on other people's work, and there is no protection for it. Terribly sad. A lack of original thinking is coming.
echelon•2m ago
> A lack of original thinking is coming.

As with the arts, 99.9% of people can't use these models to express vision, get attention, or achieve distribution.

The game is the same as it has always been. You still need hard work, taste, something important to say, the ability to articulate it, good timing, and luck.

Nothing has changed. We can just build faster.

MSFT_Edging•42m ago
Can't wait for netflix to be topped on unimaginative nonsense meant to have in the background while you scroll your phone.
toasty228•29m ago
Unlimited slop "content" to fill decomposed brain shaped vessels, the future is bright!
kyriakos•18m ago
Unfortunately Hollywood been feeding us human generated slop for a long time already. Transition won't be hard.
Mashimo•40m ago
> The result gives a total memory footprint reduced by 66%, from 123.6 GB in full precision to 42.5 GB with the smallest models variants. Combining this with our dynamic VRAM offloading enables a next-generation 2K video model to run locally on a GPU like the RTX 3060.

Pretty cool.

But assuming you have a 16GB 3060, how long would it take to generate a 15 second clip?

sheesdev•39m ago
The mouse render is surprisingly good. Several of those clips stood out to be a pretty big leap in terms of current SOTA models.

The only one that looks "off" is the beverage ad video during the can opening clip, it still has that "AI smoothening" effect. Good thing this can be done pretty well using traditional rendering.

I feel like for a good while now we'll transition into a process that uses traditional "close-up" rendering/shots + AI generated wide-shots or quick cuts.

Exciting, but also troubling. This being open-weights is a massive win for the community though.

vblanco•38m ago
Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.
robbru•25m ago
1 minute for a second of footage, that is awesome! Thanks for sharing.
mwigdahl•24m ago
If you wouldn't mind sharing, what's your Comfy workflow for this? I have the same video card setup and would like to give it a shot.
vblanco•22m ago
just the default one in the link for image-to-video
Maxious•22m ago
FWIW on a 5080 16GB it takes 3 minutes for 10 seconds 480p video (the mouse video workflow with length changed from 5 seconds to 10 seconds)
embedding-shape•2m ago
FWIW on a 6000 Pro it takes 68 seconds for 10 seconds 480p video, cold startup, same demo workflow as you used. Set "megapixels" to 2.0 (1920x1088) and same video seems to take 5+ minutes, not sure if everything is right/correct at the moment.

As far as I can tell, the current ComfyUI nodes don't even do compilation, and I haven't looked into what attention mechanism they're using, but I'm sure with time these durations will come down even more.

fodkodrasz•32m ago
On one hand: impressive. On the other hands aesthetically it all looks painfully bland and generic.
SV_BubbleTime•19m ago
This, today, is the absolute worse this model will ever be. Chill.
_diyar•1m ago
I agree. But is that the model or the prompt?

I remember reading a report where people running AI-model instagram account were using insanely long and detailed prompts about the setting, lighting, makeup, pose, disposition, clothing, etc. about their models. Presumably with some reference image of the face / body to remain consistent across images.

It‘s not clear to me whether a sufficiently detailed prompt can generate actually interesting video with a natural ”texture” (for lack of a better word).

embedding-shape•22m ago
> We found that the model's modulation weights (~40% of the total parameters) could be pruned and replaced with a functionally equivalent lookup table, dramatically shrinking the memory footprint with no loss in output quality.

Is this a common approach to reducing weights with "no loss in output quality", assuming this is true? Seems almost too simple to work. If this is doable, would this be applicable to LLMs as well?

Neat with native frame-to-frame generation, but wonder how easy it is to "link" together clips at the intersection, typically the models kind of lose the "momentum" across these stiches, being able to merge things with frame-to-frame between clips might help with this it feels like.

_diyar•6m ago
Also begs the question whether this is applicable for high-throughput applications on FPGAs, which are to my novice mind basically LUTs, right?

I remember a paper which was posted on HN a few weeks ago where somebody implemented KAN networks in FPGAs, since those can readily be approximated as LUTs.

satvikpendem•20m ago
I've said it before and I'll say it again, human directors are still valuable, as they use AI video editing tools to generate the shots they want and put them together in a cohesive way. Previously they might've used film and actors but if they can just prompt the AI (or create workflows as seen with ComfyUI) then they arrange them together just like how an EDM producer doesn't actually play the instruments but instead the creativity is in the arrangement.

I suspect it'll be quite a while until AI gets a good enough aesthetic sense to do this, as even with static HTML websites humans can easily see that it's AI slop.

hnlqpx99l9•4m ago
Learned something, upvoted
nfnmema•1m ago
Any tutorial for me to learn how to use
ddevnyc•21m ago
I am particularly curious how multimodal models will work with types of knowledge that are inherently non-text. For example, SOTA LLMs really suck at electronics, especially analog electronics.

Is MiniMax H3 capable of logical / technical reasoning, or is it purely art oriented?

Critical CVE issued for hallucinated SQLite vulnerability

https://research.jfrog.com/post/sqlite-critical-cves-or-llm-slops/
380•ymir_e•3h ago•123 comments

MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

https://blog.comfy.org/p/minimax-h3-day-0-support-in-comfyui
62•vblanco•1h ago•26 comments

Don't be a meat proxy

https://gruhn.me/blog/2026-08-03/
1152•ngruhn•8h ago•491 comments

Qwen3.8-Max: A New Bar for Coding and Cowork

https://qwen.ai/blog?id=qwen3.8
830•ai2027•12h ago•423 comments

AirLLM 70B inference with single 4GB GPU

https://github.com/lyogavin/airllm
62•Anon84•3h ago•29 comments

SPF Record Syntax: Mechanisms, Qualifiers, Modifiers, and Macros

https://dmarcguard.io/blog/spf-record-syntax/
13•meysamazad•1h ago•3 comments

The Abandoned Fish Sauce Terrorizing a Small Canadian Town

https://defector.com/abandoned-fish-sauce-canada-interview
81•ohjeez•2d ago•51 comments

Prevent cognitive debt by manually retyping LLM-generated code

https://ankursethi.com/blog/prevent-cognitive-debt-by-manually-retyping-llm-generated-code/
206•mpweiher•5h ago•153 comments

Bonsai: Janestreet's UI Library

https://github.com/janestreet/bonsai
154•KolmogorovComp•6h ago•55 comments

Rust project goals: Immobile types and guaranteed destructors

https://github.com/rust-lang/rust-project-goals/blob/main/src/2026/move-trait.md
150•paavohtl•8h ago•57 comments

The true power of regular expressions (2012)

https://www.npopov.com/2012/06/15/The-true-power-of-regular-expressions.html
49•uneven9434•5h ago•38 comments

What DMARC Protects You From, and What It Does Not

https://senderledger.com/articles/what-dmarc-actually-protects-you-from
64•adulion•5h ago•18 comments

Walk on Decomposed Subdomains

https://clementjambon.github.io/wods/index.html#blogpost
11•E-Reverance•6d ago•0 comments

Show HN: Nightcrawler – A local AI pentesting agent running on a smartphone

https://github.com/garagehq/nightcrawler/
52•NickySlicks•3h ago•19 comments

Finding zombies in our systems: A real-world story of CPU bottlenecks

https://medium.com/pinterest-engineering/finding-zombies-in-our-systems-a-real-world-story-of-cpu...
5•fagnerbrack•23h ago•0 comments

Octane – React's programming model, compiled

https://octanejs.dev
79•nnx•6h ago•25 comments

Games at the press of a button: The Rip-O-Bot (1989)

https://blog.gingerbeardman.com/2026/08/02/games-at-the-press-of-a-button-the-rip-o-bot/
7•msephton•18h ago•1 comments

Train Simulator Controller

https://z80.me/blog/tsc-2026-july/
54•austinallegro•3d ago•2 comments

Show HN: Isopolis – Isometric pixel map of SF

https://sf.isopolis.city/
284•nuwandavek•13h ago•62 comments

The Future, Made in China

https://www.newyorker.com/magazine/2026/08/10/the-future-made-in-china
12•littlexsparkee•41m ago•2 comments

Ask HN: What are the viable alternatives to DuckDuckGo?

21•ldng•1h ago•34 comments

Show HN: We Fixed UniFi's Slow PPPoE Performance with PPPoE Half-Bridge

https://arcbox.dev/blog/unifi-pppoe-half-bridge-acceleration
44•uneven9434•6h ago•16 comments

Characterizing Warp Divergence from Pascal to Blackwell

https://arxiv.org/abs/2607.23402
15•matt_d•3d ago•1 comments

Why we write our own C and C++ inference engines

https://localai.io/blog/why-we-write-our-own-engines/
93•eatonphil•2d ago•39 comments

9front "This Was Supposed to Be Fun" Released

https://9front.org/releases/2026/08/02/0/
61•birdculture•3h ago•20 comments

Show HN: ssh ssh.place

https://ssh.place
155•jeninh•14h ago•86 comments

Situational Awareness and the Impending Stock Market Volatility

https://www.emergingtrajectories.com/lh/situational-awareness-bigger-picture/
49•cl42•8h ago•37 comments

CP/M-386 – CP/M for 386 protected mode, derived from CP/M‑68K

https://github.com/johnsonjh/cpm386
87•TMWNN•14h ago•42 comments

SwiftUI After 7 Years

https://ykvm.com/2026/07/swiftui-a-story-of-mediocrity/
230•mpweiher•19h ago•232 comments

Show HN: Kakehashi – Experimental userspace to run macOS binaries on Linux ARM

https://github.com/wie-project/kakehashi
237•vlad_kalinkin•22h ago•60 comments