frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Qwen-Image-2.1: Compact, efficient, and unified image creation

https://qwen.ai/blog?id=qwen-image-2.1
62•jmillikin•1h ago

Comments

fishfasell•35m ago
The capabilities of local LLM text-to-image is honestly pretty damn impressive. IMO, I think local image generation is currently ahead of local code generation. I can get an image in seconds locally with the quality being way higher than what I'd expect from a local model. However with coding it's much slower and much less impressive. I'm sure there's a reason for this and I'm not an AI expert so I'll let the smarter folks tell me why, but that's just been my observation thus far.
victorbjorklund•33m ago
I mean I’m sure it’s the reverse for an artist. They would be less impressed with the image and more impressed with the code quality
gedy•5m ago
To generalize, LLMs are great at what you are not skilled at.
mft_•17m ago
I've played with diffusion models on and off since the first release of Stable Diffusion - just for amusement, without a particular goal.

Recently, I've been helping a friend's wife with some basic vector images for her sewing hobby (she has what is essentially a CNC sewing machine) and have been super-impressed with FLUX.1-Kontext, which I've been running on my Macbook Pro with mflux. Its ability to (for example) take a photo of a human or an animal and return a line drawing which is recognisably them (rather than just a generic similarish image as I've experienced with other models) is excellent.

It's an older model now, but (AIUI) has the text-handling features baked in, and in my various testing is very reliable at giving me the outputs that I want, without the randomness I've experienced previously. It's big and relatively slow (~3 mins per 512x512 image edit on my M1 Max Mac) but excellent to work with. It's also very straightforward to set up, without the harness complexity of e.g. comfyui.

hn45e7pbij•33m ago
Image gen you eyeball one frame and stop, code needs hundreds of tokens all correct in sequence, one bad line and the whole thing fails.
Hard_Space•27m ago
Interesting in the example of assembling the Cheers team how the otherwise great result genericizes Shelley Long.
TomGarden•27m ago
Very impressive, and kind of worrying a 7B model can have such capabilities. The implications are huge. And Qwen does no watermarking (yet) yeah?
d2kx•25m ago
God I love the Qwen team. Easily the most diverse set of models from all the Chinese labs. Only Gemini/DeepMind comes close.
mdp2021•21m ago
How do you use this model locally, similarly to using `llama-server -m <model>`?

(Of course I mean: outside direct use of Python, and in the most efficient way.)

embedding-shape•18m ago
Probably ComfyUI is one of the easiest way to get started with local image/video models. Or perhaps vLLM, if they have support for it already, would be something like `vllm serve <model> --omni --port 9080`
utopiah•13m ago
why not just as you suggested i.e. https://qwen.readthedocs.io/en/latest/run_locally/llama.cpp.... then get the result either via a UI or wget/curl it back?
mdp2021•9m ago
I am not sure that llama.cpp also supports image generation models.
utopiah•6m ago
it's multimodal, see https://github.com/ggml-org/llama.cpp/blob/master/docs/multi...
fp64•13m ago
spottedmarley•19m ago
Boy do I love waking up to find a new awesome toy from the Qwen team waiting for me to play with! Pulling it now
trains39472•16m ago
A 7B diffusion model can now render CJK text better than Microsoft Windows.
tomjen3•5m ago
Just think about how recently we got that feature in the official ChatGPT image gen. And now we have that running locally — assuming that is, I can figure out how to get this running on my Mac — blows my mind.
hgufj•13m ago
I am really grateful to the Chinese Labs for open sourcing their best models. If it was left to the Americans, we would be forced to pay obscene API fees to use them.
on the linked GitHub page they list support Diffusers, ComfyUI, vLLM-Omni, SGLang, and LightX2V with links to each

Let me Jev that for you

https://ljtfy.dev
1•rbaudibert•22s ago•0 comments

Conquering Entropy: Reducing Risk

https://kaeruct.github.io/posts/2026/09/19/conquering-entropy-reducing-risk/
2•kaeruct•2m ago•0 comments

Vibecoded Game with Love

https://martino.im/Dungeotto
2•martinopiaggi•3m ago•0 comments

The Patriots' kicker scoreboard trick debunked

https://perthirtysix.com/essay/patriots-scoreboard-kicker-trick
2•potench•3m ago•0 comments

The senior engineer death spiral

https://sunilpai.dev/posts/the-senior-engineer-death-spiral/
2•nsavage•3m ago•0 comments

Our mission: Teaching language to learners who happen to not be human

https://www.theycantalk.org/home
2•bookofjoe•5m ago•0 comments

Procedural islands from fractal rock and 150k raindrops

https://ilands.ai/content/359897434698027008
2•yoichi_rig•7m ago•0 comments

Hair test proves Thomas Jefferson fathered children with slave, book claims

https://www.telegraph.co.uk/us/news/2026/09/19/thomas-jefferson-father-children-slave-sally-hemin...
2•Tomte•10m ago•0 comments

What if my Git host were a static site generator?

https://char.lt/blog/2026/09/sorcery-repo-viewer/
2•homarp•11m ago•0 comments

Scientists revive 100M-year-old microbes from deep under seafloor (2020)

https://www.reuters.com/article/world/scientists-revive-100-million-year-old-microbes-from-deep-u...
2•thunderbong•11m ago•0 comments

What I Remember

https://stevenwaterman.uk/what-i-remember/
2•StevenWaterman•11m ago•0 comments

Crawl Cosplay Trunk Tournament

https://www.crawlcosplay.org/cctt
2•Bluestein•14m ago•0 comments

Any2nix: Serialization-Powered Nix Converter

https://github.com/gbr-ufs/any2nix
2•gs-101•16m ago•1 comments

System Design in Depth – 200 topics, 118 diagrams, interactive demos

https://system-design-in-depth.pages.dev
4•innovatorved•18m ago•0 comments

Show HN: A minimal Pareto-optimal OpenRouter model router for pi, based on Jev

https://github.com/philippdubach/pi-jev-router
2•7777777phil•19m ago•0 comments

IBM Buys Hughs Research Lab for Quantum Dev

https://newsroom.ibm.com/2026-07-23-ibm-to-acquire-hrl-laboratories-to-power-the-future-of-quantum
3•jimbobbam•28m ago•1 comments

From 10-97% SoC in 9 minutes: Sunwoda joins the flash charging race

https://electrek.co/2026/09/20/from-10-97-soc-in-9-minutes-sunwoda-joins-the-flash-charging-race/
3•cisc•32m ago•0 comments

Laya MLX

https://github.com/mizorewww/laya-mlx
2•blueaquilae•32m ago•1 comments

With Adapted Sailing, the Sea Is for Everyone

https://reasonstobecheerful.world/adapted-sailing-school/
3•bookofjoe•34m ago•0 comments

Sherrod Brown Tries a New Retort to Criticism on Transgender Issues

https://www.nytimes.com/2026/09/19/us/politics/sherrod-brown-transgender-ad.html
2•whack•34m ago•0 comments

No More Code Dumps

https://www.reddit.com/r/rust/comments/1wkmzun/no_more_code_dumps/
3•mellosouls•36m ago•0 comments

Lawsuit: Anthropic, OpenAI, SpaceXAI and Google made illegal slowdown agreement

https://www.pbs.org/newshour/nation/lawsuit-says-anthropic-openai-spacexai-and-google-made-illega...
2•rdmuser•36m ago•0 comments

Mathematical Billiards (2024)

https://structures.uni-heidelberg.de/blog/posts/2024_01_costa/index.php
2•vismit2000•38m ago•0 comments

Cognition and consciousness arise from analog computations, says new theory

https://www.technology.org/2026/09/19/cognition-and-consciousness-arise-from-analog-computations-...
2•hackandthink•39m ago•1 comments

A Brief History of Koinometry

https://benjaminschneider.ch/2026/09/19/koinometry.html
2•bschne•40m ago•0 comments

OpenPrinter Teaser Update [video]

https://www.youtube.com/watch?v=wSV5XT4sl2A
2•rhmn_mdov•41m ago•0 comments

Winamp will be reborn in 2027 to take on streaming

https://www.engadget.com/2259236/winamp-coming-back-2027-deezer-partnership-streaming-music/
3•NordStreamYacht•43m ago•2 comments

Yes They Want to Kill Us: The Worshippers of Machines versus Human Beings

https://www.meditationsinanemergency.com/the-worshippers-of-machines-versus-the-lovers-of-nature/
3•eustoria•44m ago•0 comments

Show HN: Easymacros – free calorie and macro tracker, no ads, no subscription

https://easymacros.fit
3•hussein987•44m ago•1 comments

Which Regular Shapes can you draw on any* Grid? [video]

https://www.youtube.com/watch?v=BJXvaEUZvW0
2•vismit2000•44m ago•0 comments