frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Apple Pass Designer

https://developer.apple.com/pass-designer/
102•soheilpro•52m ago•46 comments

Around 2-6% of World Bank foreign aid got siphoned into crypto wallets

https://www.nber.org/papers/w35655
52•bko•1h ago•6 comments

Court agrees with EFF: Utah's VPN law demands a technical impossibility

https://www.eff.org/deeplinks/2026/10/court-agrees-eff-utahs-vpn-law-demands-technical-impossibility
302•hn_acker•21h ago•138 comments

With most information hidden, the game Stratego had stumped AI until now

https://arstechnica.com/science/2026/10/ai-finally-beat-the-best-stratego-player-in-history-and-d...
76•PaulHoule•5h ago•20 comments

A 12-year sequence of telescope images of a star and four planets orbiting

https://bsky.app/profile/theplanetaryguy.com/post/3mwucf5ert22f
49•mariuz•8h ago•8 comments

Greg Kroah-Hartman – Security in the LLM Age [video]

https://www.youtube.com/watch?v=NnV_cWeoo5Q
63•usernomdeguerre•17h ago•8 comments

Loss of cell identity drives human aging: Two new papers

https://erictopol.substack.com/p/loss-of-cell-identity-drives-human
73•bookofjoe•23h ago•11 comments

On social reality in China

https://www.lesswrong.com/posts/b5cSYh4emQb2qrGmK/on-social-reality-in-china
46•thicTurtlLverXX•8h ago•25 comments

Mike Tomlin spent 12 years building a Minecraft city

https://www.nytimes.com/athletic/7648198/2026/10/01/mike-tomlin-minecraft-nfl-coach/
66•CoryOndrejka•1d ago•14 comments

Sites in ChatGPT

https://chatgpt.com/features/sites/
122•polvi•21h ago•135 comments

FLUX 3 Image

https://bfl.ai/models/flux-3-image
197•minimaxir•1d ago•45 comments

Blogging with Gleam, Org-Mode and Pandoc

https://byzantine-systems.github.io/blogging-with-gleam-org-mode-and-pandoc/
18•schonfinkel•8h ago•2 comments

One month coding with GLM 5.3 Flash

https://wagtail.org/blog/one-month-on-glm-53-flash/
33•ThibWeb•4h ago•26 comments

From the creator of Redis; run LLM locally with ds4

https://dwarfstar.sh/
24•fibo•1h ago•1 comments

The Legend of von Neumann (1973) [pdf]

https://gwern.net/doc/math/1973-halmos.pdf
209•suopspaces•6h ago•122 comments

Our Project Suncatcher prototype satellite is in orbit

https://blog.google/innovation-and-ai/models-and-research/google-research/project-suncatcher-prot...
22•pantalaimon•8h ago•24 comments

The first packet sent via RFC1149 avian carrier is up for auction at Christie's

https://onlineonly.christies.com/s/fine-printed-books-manuscripts-science/carrier-pigeon-internet...
12•peter_hansteen•7h ago•1 comments

How accurately calibrated is Jev?

https://maximumeffort.substack.com/p/jev-is-poorly-calibrated
21•dblack12705•4h ago•7 comments

STS-51-F Abort-to-Orbit (1985)

https://en.wikipedia.org/wiki/STS-51-F
11•schoen•3h ago•3 comments

Show HN: Giving Opus 5.5 a simulated paint canvas

https://stillwet.art/
139•alstonite•19h ago•46 comments

"The only intuitive interface is the nipple" (2012)

https://www.greenend.org.uk/rjk/misc/nipple.html
13•ibobev•40m ago•8 comments

Venice’s failed war against Constantinople led to the first bond market

https://bigthink.com/books/a-fabulous-debt/
16•RickJWagner•6h ago•5 comments

100 years of student radio history in the DLARC college radio collections

https://blog.archive.org/2026/10/02/100-years-of-student-radio-history-in-the-dlarc-college-radio...
14•HieronymusBosch•7h ago•1 comments

What if we stopped using GPUs? [video]

https://www.youtube.com/watch?v=xc2FTBGRSJo
12•sandslash•1h ago•3 comments

GrapheneOS has fixed the Android 17 QPR1 kernel performance regression

https://discuss.grapheneos.org/d/42511-grapheneos-has-fixed-the-massive-android-17-qpr1-kernel-pe...
9•Cider9986•13m ago•0 comments

Show HN: Pyxel – A Python retro game engine with built-in art and sound editors

https://github.com/kitao/pyxel
24•kitao•20h ago•4 comments

Anatomy of a Lean proof for software engineers

https://agostbiro.net/posts/2026-10-anatomy-of-a-lean-proof/
24•abiro•1d ago•0 comments

Supabase is acquiring Turso

https://supabase.com/blog/supabase-is-acquiring-turso
171•cvburgess•4h ago•90 comments

Updates to Full Disk Access in macOS

https://developer.apple.com/news/?id=p6zjojqw
11•notfirstpost•22m ago•1 comments

F.02 Decommission

https://www.figure.ai/news/f-02-decommission
18•ad_hockey•9h ago•5 comments
Open in hackernews

One month coding with GLM 5.3 Flash

https://wagtail.org/blog/one-month-on-glm-53-flash/
33•ThibWeb•4h ago

Comments

ThibWeb•4h ago
It was a bit of a silly challenge, wasn’t sure how workable, learned a lot in the process about what actually drives usage / costs, and how to keep both under control
cvburgess•4h ago
Considering they were your top two models, how did the flash and non-flash versions compare? Did you use them for different tasks?
gunalx•37m ago
When text wuality or for pure but adwansed coding is concerned i always pick glm 5.3. The flash is awesome for everything that dosent really matter though.
ThibWeb•21m ago
I have a hard time justifying GLM 5.3 these days. It’s slightly better than Flash but rarely enough to justify the much steeper price. We chose to use usage-based billing only so are very sensitive to model price.
sampullman•1h ago
Worth a comparison with DeepSeek v4.1 flash, if you've got another month to spare!
lnenad•42m ago
They've got an image of their homegrown benchmark in the post that lists DS4.1.
samtheprogram•34m ago
DS v4.1 Flash is roughly equivalent. It's going to get some things right/better that GLM flash doesnt and vice versa.
ThibWeb•17m ago
Yep, I think next month will be on that. It feels slightly better from a few days of use, and in our WIP benchmarking it scores way higher
nowittyusername•5m ago
Ill tell you from my personal use glm 5.3 flash was better then deepseek 4.1 flash, also deepseek liked to yap in his reasoning traces soo fucking much, the yapping was fast but the task was so slow to complete...
throw930rmdkdk•1h ago
Flash is pretty decent coder, but it should be paired with good planner and reviewer. I would pick astra low for planning and sol 6.1 medium for reviews.
eikenberry•59m ago
What would you use if you wanted to stay (at least) open weight?
esafak•44m ago
Deepseek 4.1 Flash and Mimo 2.6 Flash.
verdverm•42m ago
qwen3.8, kimi3, kimi2.7, GLM-5.3 are all good families I use in my coding team

I'm mainly using flash varients, at least as the default, bump.up to stronger model as needed (less often these days)

Geof25•32m ago
qwen3.8 27B or 2.4T? they are completely different models with completely different pricing
verdverm•27m ago
I use all the qwen!

It's my favorite model family to interact with, it's prose is the best imo, it makes me laugh from time-to-time (like when it said it would "crib" some code from another project, lul)

I currently have qwen-flash working on an NES emulator harness so qwen-little can play my first RPG (ff1)

(tho I have used all the others I mentioned, happenstance I'm using qwen this iteration/task)

aktenlage•36m ago
> Unfortunately there are still consequences to it. I chose the 'wrong' model for the prototype, and we spent 450M tokens / $150 / 5kWh of energy use almost overnight. The MCP server itself works well and we now have a great demo of the capabilities, so it’s not for nothing:

> Nonetheless, it’s a good reminder to be careful with model selection and with agentic patterns. We could have achieved similar results for most likely 5x less cost with not that much more effort. Lessons learned! We need to budget for this, and be more careful. Could have seen it coming, but now we know.

I don't get it. Why was it wrong? Which one would have been better? What was the lesson and how could you have foreseen it?

ThibWeb•25m ago
Hmm I might need to rephrase. Initial challenge was to use GLM 5.3 Flash and I was on the non-Flash version for the whole vibe coded build. Just wasn’t paying attention and I didn’t realize that one session was a quarter of the month’s spend and 150% of the budget (big price difference between models)
gpugreg•32m ago
How did you measure energy usage?

Edit: I found a linked article that mentions the inference provider who does the measurements.

ThibWeb•24m ago
Yes, all from Neuralwatt, GPU energy use only. Makes models’ “efficiency” much more visible than tokens.
epistasis•21m ago
One thing about these numbers that's absolutely shocking to me is how low the energy use is:

> That model’s usage was well within our budget ($68, about 4kWh of energy use / 365 grams of carbon emissions).

The energy cost is literally 1% of the total cost. For context, 4kWh of energy would drive you about 15 miles in an EV, about half of the average person's driving miles. It's boiling 10 gallons of water.

With the talk of AI Data Center's impact on the world, you'd think this would be 10x to 100x the amount of energy in order to get the effects they're using here.

My takeaway: the AI data center buildout is an overbuild probably at least as large as the fiber buildout that left us with so much dark fiber. If not even bigger. The only thing that will save the economy is the inability of NVIDIA and chip fabs to produce enough chips to match the buildout planned.

ThibWeb•6m ago
I agree but it does worry me how fast my usage is increasing. Two months ago I was using 10x less tokens and probably not much more than 5kWh on inference. This month about 30kWh on inference. If it becomes more affordable, is there going to be another jump? Not quite sure
disiplus•5m ago
Was this post generated with LLM, did he properly mention anywhere why exactly did it fail with example or i have trouble reading.
monksy•2m ago
GLM5.3-flash has been fantastic for me to make minor fixes in ambigious ways. "Fix x feature, whats going wrong. " It does the job.
esafak•45m ago
Flash is plenty good for planning and reviewing, for my needs. In fact, I use it for that because it's too slow for execution, despite the name.

edit: I subscribe to z.ai, I don't host.

gunalx•36m ago
Will agree on this. From z.ai i have found the non flash to have way more consistent performance.
jminnl•32m ago
> I use it for that because it's too slow for execution

What kind of hardware and what particular quant?