frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Yale med school prof tells his postdocs: Work extra hours at my command

https://statmodeling.stat.columbia.edu/2026/09/02/yale-med-school-prof-tells-his-postdocs-work-ex...
1•Tomte•35s ago•0 comments

ChatGPT Subsidy Chart reveals $3k monthly spend nets $62k API-equivalent value

https://aicharts.io/gpt-subsidy
1•thoughtpeddler•1m ago•0 comments

Tradeviction – Long or short startup conviction

https://tradeviction.com/about
1•mantegna•1m ago•1 comments

A New Book Uncovers LA's Most Unique and Unexpected Museums (2024)

https://lamag.com/arts-and-entertainment/a-new-book-uncovers-l-a-s-most-unique-and-unexpected-mus...
1•spking•1m ago•0 comments

NATO's next nightmare: The French presidential election

https://www.politico.eu/article/nato-next-nightmare-the-french-presidential-election/
1•robtherobber•2m ago•0 comments

TruckSmarter shutting down: Dispatch app goes dark on Friday

https://www.freightwaves.com/news/trucksmarter-shutting-down
1•crescit_eundo•2m ago•0 comments

Category Name Needed for OpenClaw, Hermes, Grok Bot

https://claude.ai/share/35ec8519-06eb-4303-bf01-083bd2d4f8ac
1•VikRubenfeld•4m ago•1 comments

Claude Fable 5.1 watermark: It has a blind spot developers can't ignore

https://thenewstack.io/fable-5-1-watermark/
1•Brajeshwar•6m ago•0 comments

Software Factories in September 2026

https://igoro.com/archive/software-factories/
1•c0ff•7m ago•0 comments

Show HN: Blognami – a Markdown based blogging platform for node

https://blognami.com/docs
1•jodysalt•7m ago•0 comments

Show HN: Aura – a Rust agent that investigates and fixes production incidents

https://github.com/mezmo/aura
1•jvogt•8m ago•0 comments

You can be wrong, even on your birthday

https://www.rawsignal.ca/newsletter-archive/you-can-be-wrong-even-on-your-birthday/
1•mooreds•8m ago•0 comments

Linode's $48 Standard 4 Got Faster on Retest – But Storage Still Moved Around

https://webbynode.com/articles/linodes-48-standard-4-got-faster-on-retest
1•gsgreen•8m ago•0 comments

Show HN: Verity.md quality gates, memory and cost control for Claude Code

https://verity.md/
1•ARayOutOfBounds•8m ago•0 comments

AI Doomers Are Buying Potemkin Articles

https://twitter.com/brianchau57/status/2095128785955987698
2•MrBuddyCasino•9m ago•0 comments

Unplug America [Tod Maffin]

https://engageq.notion.site/unplug
1•JdeBP•9m ago•1 comments

Welcome to the Madhouse [audio]

https://softwaremadhouse.com/episodes/episode-1-welcome-to-the-madhouse/
1•mooreds•10m ago•0 comments

China's fiery incinerator boom [video]

https://www.youtube.com/watch?v=AuIEBSb2xiE
1•eamag•10m ago•0 comments

AI Killed My Blog Traffic

https://clintmcmahon.com/blog/ai-killed-my-blog-traffic
1•speckx•10m ago•0 comments

Uber is laying off 10% of its workforce

https://techcrunch.com/2026/09/02/uber-is-laying-off-10-of-staff-or-3300-people/
4•rglullis•12m ago•1 comments

Financial Accounting Concepts for Developers

https://docs.tigerbeetle.com/coding/financial-accounting/
1•bramadityaw•12m ago•0 comments

Microspeak: Funded / Unfunded

https://devblogs.microsoft.com/oldnewthing/20260901-00/?p=112662
1•ibobev•12m ago•0 comments

I made my package in Python called qyl

https://pypi.org/project/qyl/
1•iwannabefamous•12m ago•0 comments

CERN Transitioning to Debian After Being a Longtime RHEL Institution

https://www.phoronix.com/news/CERN-Goes-Debian-Leaving-RHEL
2•jbd•13m ago•1 comments

Climate crisis is creating hellish conditions for waste pickers at Nairobi dump

https://www.independent.co.uk/climate-change/kenya-aid-waste-nairobi-dump-b2838750.html
3•eamag•13m ago•0 comments

Gemini 3.8 Flash

https://twitter.com/googleai/status/2095175759606231439
1•zhiQ•14m ago•1 comments

AI Slop is saturating the internet and things are becoming dramatic quickly

https://sites.google.com/view/sources-aislop
4•embedding-shape•14m ago•2 comments

Share Foundation: Students and Opposition Politicians Targeted by Spyware

https://sharefoundation.info/en/share-foundation-students-and-opposition-politicians-targeted-by-...
1•Kostic•16m ago•0 comments

Genome Duplication Is an Evolutionary Gamble

https://www.quantamagazine.org/genome-duplication-is-a-radical-evolutionary-gamble-20260902/
1•ibobev•18m ago•0 comments

New Things You Should Know About HTML Here in Mid 2026

https://blog.master.dev/new-things-you-should-know-about-html-here-in-mid-2026/
1•ibobev•18m ago•0 comments
Open in hackernews

Gemini 3.8 Flash

https://deepmind.google/models/model-cards/gemini-3-8-flash/
132•bratao•50m ago

Comments

Mashimo•36m ago
It's 404 now.
freedomben•32m ago
Came and went in a flash
k8sToGo•22m ago
Because they are preparing Gemini 3.9 Flash
kingstnap•27m ago
The blog post is gone but I can currently use it in the gemini chat website.
OG_BME•32m ago
What did it say?
realist_not•32m ago
Anyone has a cached page / mirror ? 404
Namahanna•30m ago
Page - https://web.archive.org/web/20260902151410/https://deepmind.... PDF Card - https://web.archive.org/web/20260902150007/https://storage.g...
yipinwong•29m ago
"Page not found"...
tacomonstrous•29m ago
Looks like Google's given up on frontier models for external consumption?
ok123456•26m ago
Given up frontier models for selling compute.
iamdelirium•25m ago
How can you say that when a Flash model is benchmarking close to Opus and Sol?
heyjamesknight•25m ago
Gemini 4 pre training is underway: https://x.com/OfficialLoganK/status/2079594867161022817

My guess is we skip 3.5 and go straight to 4 Pro. With the monthly Flash releases, releasing 4.0 Flash and Pro in 6-8 weeks would be a nice buildup.

(I work at Google but don't know anything that isn't already public)

WarmWash•21m ago
Latest rumor is that 3.5 pro was struggling to be meaningfully better than flash, since iterations on flash were moving much faster than iterations on pro, likely due to model size (flash is estimated to be in the 200-400B range).
VirusNewbie•9m ago
I found 3.5 pro to be much better than 3.5 flash, but 3.7 flash with high reasoning is comparable and way way faster.
mattlondon•29m ago
Wow this comes after what - 3 or 4 weeks since 3.7 Flash, which was also 3 or 4 weeks after 3.6 Flash IIRC?

I eagerly wait more info but sounds like Deepmind without Demis calling the shots has been unleashed and are operating at full speed? Shocker!

At this point it is a meme of course, but where is 3.5 Pro :)

meetpateltech•10m ago
According to the WSJ, 3.5 Pro is reportedly being skipped entirely, making Gemini 4 the next flagship model after post-training.

https://x.com/AndrewCurran_/status/2094937419615502370

sva_•28m ago
model card https://news.ycombinator.com/item?id=49537354 (doesn't 404)
mattlondon•27m ago
Also https://web.archive.org/web/20260902150007/https://storage.g... in case it goes again
Barbing•15m ago

  [1] For tone and instruction following, a positive percentage increase represents an improvement in the tone of the model on sensitive topics and the model’s ability to follow instructions while remaining safe compared to Gemini 3 Flash. We mark improvements in green and regressions in red.
Gemini 3 Flash?! So is Gemini 3.8 Flash less safe than 3.7 Flash in all areas besides Text to Text Safety (and identical on Image to Text Safety)?

Why bother with a column “Gemini 3.8 Flash vs. Gemini 3.7 Flash” when you’re going to disregard the label for 20% of it? Also is the “Tone” label short for “Tone and Instruction Following”?

Chartcrime, the major AI lab tradition.

andai•24m ago
Wait, I didn't realize 3.7 Flash was already beating Sol on a bunch of the benchmarks. Isn't it a way smaller models?
ipsod•20m ago
IDK if it's smaller, but I know it's way faster. In one test I did, Flash 3.7 high was ~9.4x faster than Luna High.

But, also... Sol crushes Flash 3.7 at writing code in a codebase of any size beyond "tiny".

Flash is my go-to for prototyping, and basically anything that isn't writing production code.

ramon156•12m ago
The only company with a proper TPU set-up is bound to have the fast models, now add a market cap like Google to the mix.
ipsod•9m ago
They've been my bet to win the AI race for a while. I was starting to doubt, but this 3.6, 3.7, and 3.8 arc has anchored me.
esafak•6m ago
Luna is way slow. I don't remember an OpenAI model ever being this slow.
realist_not•20m ago
It's pretty good if you can actively steer it , its actually really really good , the antigravity free tier and pro tiers are generous as well . I'm shocked at how fast it generates tokens.
xnx•23m ago
Seem like a great, no-compromise, upgrade over 3.7 which is already a bargain, fast, and doesn't have the brain-damaged writing style of Claude.
fitsumbelay•20m ago
that's certainly what it's looking like so far. kind of mind boggling ...
fitsumbelay•21m ago
shows up in /models though and encourages you to use it over 3.7 Flash I prefer this over reading specs: the "just show me" way
mattlondon•20m ago
Currently top at https://deepswe.datacurve.ai - beating Opus 5!

https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5!

Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use.

Gecko4072•19m ago
Google - we're so back
WarmWash•15m ago
The benchmark also doesn't include speed. You almost think something has gone wrong when using it because it returns full responses so incredibly fast.
scrlk•9m ago
Not just speed, also reliability. IME, Gemini's speed or quality doesn't degrade badly during weekday working hours compared to OAI and especially Anthropic.
ttul•14m ago
Crushing it on DeepSWE is a very big deal. Excited to give this a try.
satvikpendem•13m ago
We'll see about that. I suspect benchmaxxing as all the labs do as I haven't found Gemini models to be nearly as good in agentic engineering compared to Claude or GPT models.
advenn•19m ago
But where is Gemini 3.5 pro?
GaggiX•18m ago
Gemini 3.5 pro is never going to be released, it was a failure.
simonsarris•9m ago
it is most likely that 4 pro will be released pretty soon instead, since pre-training for 4 began in late July.

https://x.com/OfficialLoganK/status/2079594867161022817

mythz•19m ago
It's just another mid-tier flash model, nothing exciting, but Antigravity has very generous quotas so it's a good workhorse model when your Claude/OpenAI subs run out.

And whilst it's a fast model, having to baby sit through and approve prompts every few seconds ends up making it slower than the Auto approve modes of Claude/ChatGPT - they definitely need an auto approve mode.

shuvrojit•18m ago
Gemini is getting less useful with each update. I could edit a pdf with the 3-pro model before but 3.1-pro couldn't edit the given pdf nor it could generate one for me.
ipsod•18m ago
3.5 pro doesn't exist yet?
shuvrojit•15m ago
Sorry my bad, I messed up the numbers, 3 and 3.1 pro. All of these model numbers have me confused
leumon•17m ago
You probably mean 3.5-flash? Pro is still good for a lot of use cases, but it seems it's still officially in the "preview" phase.
pwython•17m ago
Is there any reason to even use 3.1 Pro now?
bitexploder•5m ago
It is still going to be better at text work, skills, document review, deep reasoning, architecture review, etc. It is only 6 months old, it isn’t like its world knowledge and software knowledge is really out of date. Use it to churn on harder design problems.
kelvinjps10•15m ago
I see benchmarks beating sol terra and sonnet. But is actually better? Has someone used it? I don't see actually much people that use Gemini for coding.
ASinclair•14m ago
From personal experience it feels much more capable than 3.7 Flash.
leumon•14m ago
[delayed]
satvikpendem•13m ago
Is the Gemini CLI still terrible compared to Claude Code and Codex? The harness the main thing holding back Google models as they could've been the best given all the advantages in compute capacity and training data they initially had, where now even the Google CEO said they're falling behind in agentic tasks, which is sort of a vicious cycle because RLHF relies on human usage.
pshirshov•11m ago
There is no Gemini CLI anymore, nor you can use Gemini with your own harness unless you pay per-token.
zipy124•6m ago
It was superseded by the antigravity CLI.
rancar2•6m ago
That was sunset and replaced by Antigravity. FWIW until I abandoned it knowing the sunsetting, I was able to get good behavior out of Gemini CLI with overriding the system prompt. The default prompt crippled the harness with very poor instructions, but there was a hidden ENV to override it. Replacing it with Claude Code like prompts based on the model selected, it ran at a much higher intelligence level full stack with significantly less errors.
f311a•13m ago
Is the google infra stable enough right now? At the start of the year, the flash model was unusable for a whole month via gemini CLI. They could not fix it for a whole month and I was a paid customer.
ipsod•5m ago
I haven't had any issues lately.
meh2frdf•12m ago
The flash models, for coding are reckless in my experience. I have a Ultimate subscription, get good quota, but still use Opus 4.6 as it's much more reliable if you manage the context window carefully.
onlyrealcuzzo•8m ago
> The flash models, for coding are reckless in my experience.

My experience is that antigravity is awful and reckless - but that the model itself isn't.

upcoming-sesame•7m ago
If by reckless you mean commit, push, deploy without me asking it to, the I agree!
hmokiguess•11m ago
Model card https://storage.googleapis.com/deepmind-media/Model-Cards/Ge...
deanc•11m ago
And yet again another failed launch from Google. I pay for their AI plus Google one package to get more cloud storage (have no interest in their AI bundle but you have to pay). and all I see in the Gemini app is 3.6-flash
WarmWash•9m ago
Google has been doing staged roll outs on all their products since forever.
simonw•6m ago
Pelican (thinking effort high): https://tools.simonwillison.net/markdown-svg-renderer?url=ht... - 8.9742 cents

Here's a 3.7 thinking effort high pelican for comparison: https://tools.simonwillison.net/markdown-svg-renderer.html?u... - 8.4387 cents

thisisauserid•19m ago
They don't want to release a frontier model that requires data sharing with the government and right now it looks like they'd have to.
onlyrealcuzzo•8m ago
And the benchmarks agreed with you... until now.

So, yes, maybe it's still not - but this would be the only time it would be highly suspicious / obvious benchmaxxing / obviously bad benchmarks.

sunaookami•12m ago
>shows an intelligence score of 59, the same as Opus 5!

...on Medium reasoning. Claude Opus 5 (high) is the default in e.g. Claude Code and scores 61. Still very impressive.

onlyrealcuzzo•9m ago
The rumor is that 3.9 is an equal improvement in all directions, and that it should be another fast follow on like 3.7 and 3.8 were.
markasoftware•9m ago
On artificial analysis it's only equal to opus 5 medium effort. Opus 5 max scores 63.

Further, opus 5 medium outputs 4x fewer tokens to achieve the same result, negating a lot of the speed difference.