frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Muse Spark 1.3

https://developer.meta.com/ai/models/muse-spark/
149•bvaldivielso•1h ago•78 comments

Gemini 3.8 Flash and 3.8 Flash Cyber

https://blog.google/innovation-and-ai/models-and-research/gemini-models/3-8-flash-and-3-8-flash-c...
674•bratao•5h ago•397 comments

I wanna live an NPC life

https://signalundefied.bearblog.dev/i-wanna-live-an-npc-life/
39•conferza•55m ago•20 comments

Google avoids a breakup of its ad tech business

https://www.nytimes.com/2026/09/02/technology/google-ad-tech-remedies.html
111•donohoe•6h ago•50 comments

Wendell Berry has died

https://www.nytimes.com/2026/08/31/us/wendell-berry-dead.html
65•Curiositry•1d ago•34 comments

Can I opt out of my input or output data being used for training?

https://help.mistral.ai/en/articles/455207-can-i-opt-out-of-my-input-or-output-data-being-used-fo...
322•teekert•8h ago•138 comments

Fable 5.1 World Modeling

https://github.com/PhiloLabs/fable51-worlds
26•surreal_•59m ago•5 comments

Biggest dark matter detector spots a single weird particle

https://www.science.org/content/article/world-s-biggest-dark-matter-detector-spots-single-weird-p...
209•randycupertino•7h ago•59 comments

Embedded Rust RTOS vs. C RTOS

https://tweedegolf.nl/en/blog/65/async-rust-vs-rtos-showdown/
29•kooi•2h ago•10 comments

Qantas Airbus A380 catastrophic engine failure in 2010 (2023)

https://admiralcloudberg.medium.com/a-matter-of-millimeters-the-story-of-qantas-flight-32-bdaa62d...
25•gumby•2h ago•8 comments

Exit the Cave

https://turtlespace.blog/p/exit-the-cave
160•akkartik•6h ago•42 comments

A practical guide to running 8x RTX PRO 6000's

https://www.gpupartner.com/blog/a-practical-guide-to-running-8x-rtx-pro-6000s
12•Retell15•2d ago•9 comments

Aging Brains Blend Memories Together Instead of Just Forgetting Them

https://studyfinds.com/aging-brains-blend-memories-together-instead-of-forgetting-them-study-finds/
158•mdp2021•7h ago•70 comments

SteamdDB Joins Nexus Mods

https://www.nexusmods.com/news/15597
84•HelloUsername•5h ago•47 comments

Commodore 64 released September 1, 1982

https://dfarq.homeip.net/commodore-64-released-september-1-1982/
297•giuliomagnifico•12h ago•152 comments

We could save petabytes of cache storage with Zstandard and Pingora

https://blog.cloudflare.com/cache-transcoding/
35•torutofu•1d ago•11 comments

AI Agents and the Refactoring That Never Happens

https://www.rosenfeld.page/articles/programming/2026_09_02_ai_agents_and_the_refactoring_that_nev...
12•rosenfeld•57m ago•15 comments

A Selection of Los Alamos Rolodex Business Cards

https://clui.org/collections/los-alamos-business-cards/selection-cards
117•1970-01-01•2d ago•25 comments

Making the Internet Boring

https://cemrehancavdar.com/2026/08/30/making-the-internet-boring/
41•zdw•3d ago•20 comments

Poisson Disk Sampling

https://stripeacross.com/posts/poisson-disk-sampling/
102•vismit2000•7h ago•16 comments

Paint.net 5.2 alpha now runs on Linux

https://forums.paint.net/topic/134562-paintnet-52-alpha-build-9739/
142•judah•3h ago•116 comments

A note on subscription prices from LWN

https://lwn.net/Articles/1090585/
649•rwky•7h ago•128 comments

WebLLM: high-performance in-browser LLM inference engine

https://github.com/mlc-ai/web-llm
65•saikatsg•6h ago•15 comments

Three sites made 215,128 “best software” pages for AI. Perplexity cites them

https://trellner.com/reports/manufactured-sources-behind-ai-recommendations/
247•jakobgreenfeld•6h ago•116 comments

I Don't Have a Smartphone

https://ploum.net/2026-09-02-i_dont_have_a_smartphone.html
167•speckx•2h ago•163 comments

Altair Basic Interpreter Source Code (1975) [pdf]

https://images.gatesnotes.com/12514eb8-7b51-008e-41a9-512542cf683b/34d561c8-cf5c-4e69-af47-3782ea...
5•Eridanus2•40m ago•0 comments

Show HN: FrontierHarness Eval – 9 harness, same model, cost per pass varies 17x

https://frontierharness.org
58•shiqimei•4h ago•31 comments

Using Cloudflare Workers and reCAPTCHA v3 for a Static Site Contact Form

https://nooshu.com/blog/2026/03/09/using-cloudflare-workers-and-recaptcha-v3-for-a-static-site-co...
9•speckx•2h ago•2 comments

Discontinuation of third level domain registrations for the .name TLD [pdf]

https://itp.cdn.icann.org/en/files/consensus-policies/rsep-2026013-name-request-15-04-2026-en.pdf
43•greyface-•1d ago•34 comments

Telli (YC F24) is hiring engineers and designers [Berlin, on-site]

https://careers.telli.com/
1•sebselassie•12h ago
Open in hackernews

Muse Spark 1.3

https://developer.meta.com/ai/models/muse-spark/
149•bvaldivielso•1h ago

Comments

frozenseven•1h ago
This should probably be primary:

https://news.ycombinator.com/item?id=49541149

simonw•1h ago

  llm -m meta-ai/muse-spark-1.3 "Generate an SVG of a pelican riding a bicycle"
https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

4.2266 cents, 38 seconds.

For comparison here's Muse Spark 1.2, which animated it without me asking it to: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

The 1.3 one is definitely better - better bicycle frame, better wing, better pelican hat.

UPDATE: Here's another one with five pelicans for each of the five Muse Spark 1.3 reasoning levels: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

The most expensive was reasoning level xhigh - 7.5 cents, 1m34s.

And I ran five pelicans at all reasoning levels for 1.2 as well, here: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

jmkni•1h ago
lol

Definitely an upgrade over 1.2

drusepth•1h ago
Is there a reason these pelicans always have roughly the same composition (side-view, 2d, biking right, flat ground beneath, etc)? I don't see any of that detailed in the prompt, yet they all seem to generate roughly the same image of differing quality.
simonw•1h ago
It's really interesting, isn't it? They almost always cycle from left to right - but I have had a few which cycle in the other direction.

The 2D / flat ground feels reasonable for a SVG, which implies a vector illustration.

piker•34m ago
I was going to ask the exact same question earlier but deleted it after thinking “I’m sure Simon has done some sort of discussion on this.” Since it does seem novel to you, too, it would be really interesting to read more about this phenomenon.
m12k•13m ago
It's my impression that it's common in western culture, where text is read left to right, and timelines are visualized as going from left to right, to also animate things going from left to right, since westerners thus have an instinct that "right = forward", so it "feels right" (familiar). I wonder to which degree this is reflected in the training data? And if you'd be more likely to get left-facing pelicans if you prompted it in Hebrew, Arabic or another right-to-left language?
Gecko4072•1h ago
Used Muse Spark 1.2 and was not impressed at all. Fast and cheap but even GPT 5.6 Terra felt much more capable. Also not really looking to support a company that was just forced to pay $18B for mental health damages.
WASDx•54m ago
I'm party using 1.2 to reverse engineer and re-implement an old game binary and it has been quite good and fast. The contributor pricing is very attractive, excited to try 1.3 and see if I feel a difference. 1.2 can get stuck outputting similar sounding thought summaries with no apparent progress when asked to solve bugs. Then I've switched to GLM-5.3-Flash which for this use case has been clearly better at finding suspected causes and following tracks.
wxw•56m ago
“contributor” pricing at $0.10/$0.20 is crazy cheap if it’s measuring up to Sol.

Definitely shows how important a user data flywheel is for RL and model improvement.

finnjohnsen2•56m ago
So one model is "Not used to improve our products" and is 10-20 times more expensive to the "Used to improve our products"-model.

Given this is Meta, my immediate assumptions that one is cheap because it lets me "be the product". I know I'm rushing to conclusions but there is zero trust here. The brain will do its thing. And the wording here is giving the brains a lot of wiggle room.

whimsicalism•54m ago
the meaning is pretty obvious - they want to train on your chats & tasks and are willing to subsidize for the privilege of doing so.
thefreeman•53m ago
aren't they explicitly saying this with both their pricing and their wording? I'm not sure what you are alluding to?
bigyabai•45m ago
Given OpenAI and Anthropic's behavior, do you really expect them to be singled out for this practice? Zero trust has been in "LGTM" territory for years now. Meta's bet against people taking a principled stance arguably paid off great.
Jcampuzano2•38m ago
I'm confused what your surprise is here. It's plain and simple right to the point wording.

I don't see the wiggle room at all.

duplessitous•1m ago
majerep•56m ago
The previous version was, in my experience, the best free model available on OpenCode. It's been very good at simple/moderate tasks where I am precise in my ask and it doesn't need to make a ton of undefined assumptions. Hopefully this new version is also available on opencode for free.
geooff_•54m ago
Could this be best intelligence / $ if you're willing to let zuck digest your data?
meerita•53m ago
I declined the use of cookies and everything went black. No content at all. Dissapointed.
7734128•52m ago
Practically free for "contributors" at 0.2 usd/mtok. That's going to be hard to say no to for hobbyists.
superfrank•49m ago
I started using Spark 1.2 for development because if you're willing to let Meta train on your data it was dirt cheap and was actually really pleasantly surprised with it. It's not a frontier model by any means, but for work that didn't require a top of the line model, I really enjoyed using it.

I'm anthropomorphizing it a bit, but it felt like it knew its weaknesses and didn't try to impose it's opinions on me. What I mean by that is that it did what I told it and if there was something unexpected in the code that it put out it was often because I gave it ambiguous or conflicting instructions. It didn't try to go above and beyond and just acted like a tool, which is what I want from a coding agent 90%+ of the time. I also felt that it did a much better job of following established patterns in my code than many of the other current models do. I'm a huge fan of OpenAI's models and Spark 1.2 is what I expected 5.6 Luna to be.

I'm curious and a little excited to use 1.3, but honestly a little worried that as Meta pushes for better benchmarks that Spark will start to fall into the trap of trying to be "helpful" in ways I don't want it to be.

Tangential, but when I first started using Spark 1.2, it made me realize how much I miss 5.3 Codex. That model was the peak of coding models, IMO, in that it knew how to write good code, but didn't try to overstep or be "helpful" in unexpected ways. That got me thinking about how the major labs seem to be stepping away from coding focused models toward more general purpose ones and how I can't help but feel like that's a mistake.

tinyhouse•48m ago
I had no idea Meta has a coding agent harness. Does anyone have experience with it and can comment? The 1.3 contributor prices look very attractive. I'll probably start using their API if performance is good and the API is reliable with decent rate limits.
bertili•46m ago
DeepSWE scores 75.4 - that's the best score so far. And it's crazy cheap! Google held the top a few hours today with Gemini 3.8 Flash, but now second to Spark 1.3. All this competition will drive prices down!
cbg0•41m ago
But is the score really reflective of the quality or are both models benchmaxxing?
bermudi•4m ago
Muse 1.2 wrote a terrible "smart summaries" extension for my pi setup. It was sending every single steamed chunk for summarization instead of waiting for the full CMD.

This is an error I would expect from sonnet 4, not a model that was supposedly just a few points behind sol.

dominotw•35m ago
how much of it is from reallocation of staff to ai training and labeling
WASDx•28m ago
With the contributor pricing being more than 10x cheaper than the standard, that would make it best and cheapest on the DeepSWE leaderboard! It feels fast in my experience too. LLMs keep improving at an insane pace.
Lucasoato•44m ago
A model that (at least in benchmarks) is getting closer to SOTA. A clear separation between what’s used to improve their products and what’s not (at least this is what they claim).

Good job Meta! Seriously. This is almost making me forget about the 18B$ lawsuit for children social media addiction.

dbbk•2m ago
How is it not SOTA? It's beating 5.6 Sol.
ChrisArchitect•43m ago
Blog post: https://research.meta.ai/blog/introducing-muse-spark-1-3 (https://news.ycombinator.com/item?id=49541149)
sunaookami•15m ago
>Previously available reasoning modes are available today with max reasoning coming shortly after we finish additional safety testing

Lmao. And their benchmark table only shows max reasoning.

mromanuk•35m ago
I didn't like 1.2, It make some mistakes in a web app, so I quickly went back to Claude, Kimi K3 or Deepseek V4. Hope this one can clear agentic development, because Muse Spark models are fast and cheap.
tyre•26m ago
Meta is one of those companies where, if there is anything remotely comparable, I'm happy to pay more to not use them. They've had a profoundly negative impact on society and Zuckerberg is not who I want controlling the future at the top of AI.

I feel the same about Grok w/ Elon. I will pay extra to use someone else.

I'm not an Amodei stan, but of all of these people he seems to have the most ethical focus. Again, not everything done perfectly and I have my gripes, but of the leaders of frontier labs, I'll vote with my money.

And, yeah, I wouldn't trust sama to watch my bag while I went to the bathroom.

loeg•22m ago
"Avoid generic tangents" / "Please don't complain about tangential annoyances."
reaperducer•16m ago
"Avoid generic tangents" / "Please don't complain about tangential annoyances."

That's pretty much 90% of HN these days.

Apple releases a new iPhone? Here comes the flood of decade-old complaints about long-discontinued Mac butterfly keyboards and walled gardens.

Microsoft releases a new version of Windows? Here come the gripes about Azure.

Google changes something in GMail? Play Store!

It's like there's an army of bots out there determined to reduce the productivity of the Western tech bubble by diverting everyone into endless circular arguments about absolutely nothing of relevance to the topic at hand.

_diyar•15m ago
How is this a tangential annoyance or a generic tangent?

> Meta announces they have a new model, demonstrating its capabilities.

> Parent comment states „regardless of this model‘s specific capabilities, if I can avoid it I will.“

souvlakee•25m ago
Why they didn't use LLM to create html table instead of https://lookaside.fbsbx.com/elementpath/media/?media_id=1048...?
jumploops•23m ago
The "contributor" pricing is the standout here at a ~20x discount, if you allow training on your data.

The model seems on par with Sol and Opus 5 on paper (admittedly on some older/saturated benchmarks, but very competitive for $).

Stats:

1M context, $0.10 input/$0.002 cached, $0.20 output (Mtok)

2001zhaozhao•15m ago
I have a feeling that Meta is not gonna like what people actually use the contributor model for lol.

(It's probably going to be a bunch of repetitive batch jobs like web search that have no training value)

fibonacci112358•14m ago
Is everyone rushing to launch something before Astra tomorrow?
johntb86•9m ago
Someone studied this (among other thigns): https://dylancastillo.co/posts/pelicanmaxxing.html . Pelicans on bikes always face right in this test, but other animals on other transportation methods sometimes face left.
vunderba•58m ago
The more generic your prompt, the more generic the response. It's a regression to the "mean" of the training data aka GIGO for AI.

It's like when you ask your average person off the street to draw a house - it'll almost always be square with a triangle roof, one door, and two windows.

In the pelican/bike example, it's probably a bit of a self-perpetuating snowball too. If the earliest examples were bike left-to-right, flat ground, etc. then they are also being scraped up in future LLMs.

polyterative•30m ago
as a kid I did them like this. nobody told me to do that. are we all so similar?
srcreigh•8m ago
The adults brainwashed us

https://www.ikea.com/ca/en/p/barndroem-box-beige-70560615/

https://www.ikea.com/ca/en/p/vallaby-rug-green-10548216/

collabs•8m ago
I sincerely believe I've never had a single original thought™ in my whole life.

There is this scene in the HBO series Westworld where a "host" says some words in sequence which is shown on a display as she says it. Of course, even me thinking of this scene and connecting it to your comment was not original, someone else clearly had the same programming as me.

A medium blog post says

> Pair what with me?” — the moment Maeve (a humanoid android) uttered those words in Westworld (Season 1, Episode 6: “The Adversary”), something clicked. Not for the average viewer, but for me, a STEM educator and AI enthusiast who, just weeks earlier, had read Stephen Wolfram’s seminal essay, What Is ChatGPT Doing … and Why Does It Work?

postalcoder•55m ago
Search Google Images for "bicycle". Almost all bicycle product shots are staged the same way: side view, going left-to-right. It makes sense to me that given that skew in the training data, the model grounds itself in the bicycle.
daemonologist•49m ago
and furthermore, this is because the drivetrain is ~always on the right side of the bike - if you want to inspect or admire a bicycle you look at the right side, as you might look under the hood of a car.

(Why the drivetrain is on the right, I don't know. But most bike parts follow open standards so it's quite entrenched.)

georgemcbay•17m ago
> and furthermore, this is because the drivetrain is ~always on the right side of the bike

While I'm sure this factors into things for advertisements for bike components, there is also just a general preference that westerners have for left-to-right motion. Not just in bike ads, but all ads with (or suggesting) movement. And also not just ads, but movies where directors believe left-to-right motion is associated with progression and right-to-left motion is regressive.

kibae•12m ago
Since most languages read from left to right, rightward movement tends to read as forward progression. So when showing a bicycle in side profile, having it face right feels more naturally like it’s moving forward.
threetonesun•14m ago
Product shots yes, people riding them its more like 50/50. Also if you search for a specific bicycle race you'll find more going right to left.
ModernMech•38m ago
Yes, I do a thing where I ask the machine to generate responses in the form of a lizard talking to a cat. The lizard is always a green gecko and the cat is always orange, which I never specify.
optimalsolver•25m ago
Sun is missing a few rays and not wearing sunglasses.
reaperducer•6m ago
Is there a reason these pelicans always have roughly the same composition

Because they're computers. They don't have an imagination and the ability to create things from whole cloth the way humans do.

Much like a mother pelican, they regurgitate what they've been fed.

tomrod•53m ago
What does the mean pelican look like at this point?

Also 3X token use vs. 1.2

_puk•13m ago
Red eyes and a tattoo?
jonahx•51m ago
If you have a grading rubric, huge points off for adding arms instead of using the wings as arms!
Fergusonb•47m ago
I think it's hilarious that this detail is enough for me to dismiss looking into the model, but here we are, and it is.
jonplackett•26m ago
Has any ab tried to game this yet and just made the most amazing pelican by hand and always reply with that?
NoOneCares44•24m ago
No one cares.

These benchmarks you guys invent for yourselves prove nothing.

A small model like Mistral 7b is just as useful to the end task as any larger model, if not more so because it’s faster.

Your big model may be able to draw pelicans or solve some esoteric nonsense but it cannot do real work in the real world.

These phoney benchmarks and experiments mean nothing.

None of the LLMs can replace a software engineer nor even a barista or car mechanic etc. not even close.

Instead of inventing fake benchmarks do something tangible and tell me how it performs.

Before laying off half the country and going full retard on AI

aldanor•22m ago
Someone did care enough to create a throwaway account to vent here, it seems.
tonyhart7•5m ago
how tf you can have green names with negative karma ???
jttnr•14m ago
I wonder, given Simons reputation in AI benchmarking, whether model providers try to train or tweak their models to perform better at drawing bicycles and pelicans?
EugeneOZ•10m ago
Absolutely BRUTAL! :)

Thank you for doing this, I love your benchmark the most!

What is the confusion? They directly state that you are the product if you use their discounted offering. It isn't an assumption that should lead you to this, it is Meta's very direct communication that should lead you to this
loeg•12m ago
Grandparent comment has zero to do with the article. It's just GP generically bitching about Meta. (Your "quote" of the comment does not appear anywhere in the actual comment.)
optimalsolver•18m ago
If it was up to Dario we'd all be banned from using open-weight models, and we'd have to be investigated for PRC connections before sending our allotted five API queries a week.
redox99•14m ago
Google, Zuck, Sama, Elon, Amodei (in no particular order).

They all suck. Pick your poison.

kenjackson•12m ago
They don't all suck equally.

Here's the order, from best to worst.

Amodei

Google

SamA

Zuck

Elon

scottyah•8m ago
SamA better than Zuck? Zuck was at least a kid when he made a lot of his bad decisions, and he seems to be getting much better. Sam is on the reverse trajectory.
redox99•8m ago
You can ask 100 people and they'll all give you a different list. It's subjective.

I think a less personal ranking would be, as a business owner, which of those providers is more dependable? As in, you don't care about evil, just your stuff working. I think maybe OpenAI?

a2ff6eeb0•1m ago
Google. They have experience operating at scale, and AI is a big enough focus that they won't wind it down.
drob518•12m ago
Gotta be honest that I’m tired of the “I hate Zuck and Meta so much” comments every time Meta does anything. Ditto Elon/X. Fine, I get it. I don’t like Zuck either. But the post is about Muse Spark 1.3. What do you think about that? If you don’t like it because Meta made it, then maybe just don’t use it and stay silent.
idiotsecant•9m ago
Not liking something because the embodiment of corporate malfeasance made it is a rational way to decide what products to support.
duplessitous•7m ago
Technology doesn't just spring into being, there will always be comments on the organizations that developed it. If you don't like them or find them repetitive, it is far easier to collapse them and move on then bend a stranger to your will
canadaduane•5m ago
I get it, but the underlying problem is: we don't have a society-wide, effective solution to counterbalancing extractive systems. Lacking a reliable label, we have to constantly signal what's on the ingredients list.
jesse_dot_id•2m ago
Staying silent is unfortunately how fascism festers.
fouc•12m ago
How have they had a negative impact? How about google?
aftbit•3m ago
Okay but Muse Glimmer 30B is one of the best small open weight models today, and IMO the best from a US lab (only real comparison is Gemma4 dense right now).