frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Fixing the Portobello Police Station Clock

https://pointinthecloud.com/2026-04-11-211700.html
165•avidly•2h ago•35 comments

28% of job postings on company career sites have been open over 90 days

https://unlisted.careers/ghost-jobs/report/2026-09
41•rubatrejo•47m ago•41 comments

Radicle: Disclosure of Vulnerability in the Network Protocol

https://radicle.dev/2026/09/23/disclosure-of-vulnerability-in-network-protocol
38•lostmsu•1h ago•13 comments

Gemini 3.8 text-to-speech

https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech/
84•swolpers•1h ago•50 comments

Stripe's Knowledge AI Platform

https://stripe.dev/blog/meet-stripes-knowledge-ai-platform
109•ltononro•3h ago•62 comments

Jev in 25 Lines of Python

https://www.nobodywho.ai/posts/jev-in-25-lines/
499•bashbjorn•9h ago•153 comments

Strands Harness

https://strandsagents.com/blog/introducing-strands-harness/
82•zuckerborg0101•2h ago•55 comments

Claude Code reads AGENTS.md only when telemetry is on [fixed]

https://blog.szypowi.cz/p/claude-code-reads-agents.md-only-when-telemetry-is-on/
365•pszypowicz•5h ago•201 comments

GPT-6 Sol and Luna

https://openai.com/index/introducing-gpt-6-sol-and-luna/
1680•OfficialTurkey•23h ago•805 comments

Z80 REPL (2018)

https://abagames.github.io/z80-repl/index.html
110•adunk•6h ago•14 comments

Claude Opus 5.5

https://www.anthropic.com/claude-opus-5-5
1704•km144•1d ago•1040 comments

I don't want the details

https://michaelheap.com/i-dont-want-the-details/
207•mooreds•4h ago•127 comments

Tokens Too Cheap to Meter

https://jyn.dev/tokens-too-cheap-to-meter/
132•teoruiz•8h ago•103 comments

GPT-6 Astra has gained the ability to drive a car

https://drivingbench.com/
167•plurby•2h ago•136 comments

Web-based IBM 1620 emulator and IPL-V from 1963

https://github.com/pkimpel/retro-1620
17•abrax3141•17h ago•4 comments

Seattle City Council votes to ban surveillance pricing in sale of groceries

https://advocacy.consumerreports.org/press_release/seattle-city-council-votes-to-ban-surveillance...
88•ortusdux•3h ago•30 comments

QuestDB (YC S20) Is Hiring a Sales Engineer

https://questdb.com/careers/pre-sales-engineer-north-america/
1•nhourcard•5h ago

OpenAI GPT–6 Astra breaks Enigma message that has resisted solution since 2005

https://www.cryptocellar.org/bgac/the-mvueh-break.html
709•sohkamyung•1d ago•427 comments

Transit rewards

https://waymo.com/blog/2026/09/transit-rewards/
220•raybb•14h ago•277 comments

The GitHub wiki is an anti-pattern (2022)

https://michaelheap.com/github-wiki-is-an-antipattern/
125•ibobev•4h ago•73 comments

What California is learning from solar panels built over irrigation canals

https://www.kqed.org/science/2002033/heres-what-california-is-learning-from-solar-panels-built-ov...
314•Jtsummers•1d ago•611 comments

How did AMD Ryzen get 50% faster in two years?

https://lemire.me/blog/2026/09/18/how-did-amd-ryzen-get-50-faster-in-two-years/
440•ibobev•4d ago•177 comments

Microsoft killed FoxPro in 2007. Anyway, here's FoxPro revived

https://foxscript.org/
427•boredjohnny•20h ago•238 comments

'We hacked the FBI:' Hackers say they have data on all FBI employees

https://www.404media.co/we-hacked-the-fbi-hackers-say-they-have-data-on-all-fbi-employees/
755•spenvo•23h ago•545 comments

ReBarUEFI: Resizable BAR for almost any UEFI system

https://github.com/xCuri0/ReBarUEFI
213•nateb2022•2d ago•66 comments

SAML: A fractal of bad design

https://blog.trailofbits.com/2026/09/21/saml-a-fractal-of-bad-design/
314•aray07•22h ago•163 comments

Data-only attacks are easier than you think (2024)

https://www.usenix.org/publications/loginonline/data-only-attacks-are-easier-you-think
87•segfaultbuserr•13h ago•33 comments

WordPress: Unauthenticated path traversal leading to conditional RCE

https://github.com/WordPress/wordpress-develop/security/advisories/GHSA-7hp8-65ch-5whp
224•vntok•1d ago•123 comments

Pentagon says overreliance on AI contributed to missile strike on Iran school

https://www.bloomberg.com/graphics/2026-iran-school-attack/
841•devonnull•22h ago•450 comments

Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)

https://artificialanalysis.ai/models/claude-opus-5-5
320•theanonymousone•1d ago•101 comments
Open in hackernews

Gemini 3.8 text-to-speech

https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-text-to-speech/
81•swolpers•1h ago

Comments

thangalin•1h ago
Here's a video of my Emotive Audiobook Creator, KeenLore, a locally hosted web app:

https://www.youtube.com/watch?v=WAeHgE94rVo

No cloud, no tokens to pay. Reads a book using a full cast of characters. Quotation attribution detection (for my novel) is at 97.2% accuracy (485/499 quotes identified and assigned correctly). The autofill of character voice descriptions uses the prose to determine how the character sounds.

Employs Gemma 4[1] for the prose analysis (voice fills, quotation detection) and Qwen3 TTS Voice Design[2] for creating voice samples. Runs on an 8GB NVIDIA T1000 GPU card, 96 GB RAM, and a AMD Ryzen 5 7600.

[1]: https://deepmind.google/models/gemma/gemma-4/

[2]: https://huggingface.co/spaces/Qwen/Qwen3-TTS-Voice-Design

Multicomp•1h ago
<grumble grumble people putting in links they expect you to follow to arbitrary goatse youtube videos for all I know>

The title of the video is 'KeenLore - Emotive Audiobook Creator Demo' and it appears to be a web UI and some local stack that reads text files.

loremm•57m ago
It's cool technology and I read a lot of audiobooks, even hundreds of hours of TTS. I feel like my brain can fill in the character voices from the text - on the page it's not like they're different fonts.

I understand audiobook narrators often do it, and that's fun. But it's not so critical in my opinion

talon8635•55m ago
Our imaginations and minds continue to rot under the weight of endless and effortless entertainment
simonw•1h ago
> Voice replication: Recreate consistent vocal profiles from just a 30-second audio sample of your voice or a voice you have the rights to use, backed by built-in consent verification, SynthID watermarking, and C2PA credentials to protect both developers and their vocal talent.

I guess voice cloning is widely enough available now from other providers that Google are no longer hesitant to ship it.

Multicomp•1h ago
They probably do something similar to GPT-Live where they expect a given voice profile to send them a sample saying 'This is the owner of this voice and I consent for synthetic samples to be made of it'

and/or local voice cloning is good enough as is so Google doesn't grant a uniquely liable ability?

gruez•8m ago
>and/or local voice cloning is good enough as is so Google doesn't grant a uniquely liable ability?

Probably the latter. Cat's already out of the bag to the extent that you can synthesize with a specific voice in one go and it sounds decent. Even if you need commercial models for better intonation or whatever, you can probably get the commercial models to first generate with a generic voice, then use a local model to transfer that to voice you're cloning. That'll probably get rid of any C2PA watermarks too.

miltonlost•46m ago
"Don't be evil... unless other companies are doing it first"
imjonse•34m ago
voice cloning is a tool, it is not necessarily evil, even though the scenarios it can be used for nefarious purposes outnumber the legitimate ones.
perrohunter•1h ago
Gemini 3.8 "Flash" says hello
112233•59m ago
"Super tinny monotone robotic voice" does not sound neither tinny nor monotone. Compared to what TTS from 90s sounded like. Or even how actors impersonated robots in movies. Has the model been eating too much hype DJs?
burkaman•28m ago
None of these examples are really what the prompt asked for. It's just like image models, once you get over how unbelievable it is that a computer produced this you realize the result isn't actually what you want.
Multicomp•58m ago
I direct my own extended daydream Star Trek fanfic (okay, I'm on season 2 episode 17) and recently I looked to see if I could have each scene file be read aloud a la an audiobook or radio drama.

Getting GPT-Live to have unique enough voices and to be expressive with how I imagine the voices going in my head is hard to direct, there's not enough control there.

So this Gemini 3.8 specific large voice library and ability to tightly control (if you are willing to write a script) is nice to find, and while I'm not sure which of the 5,286 Gemini products this is, nor how to onboard and get started feeding this my own text files, nor what training will happen to my data if I did somehow use it, I love that the state of the industry is such that Google can do this and release it publicly, because that means eventually an equivalent product can come from someone else and be used locally / confidently that the generated audio or inputs won't be retained and misused.

ghostbrainalpha•43m ago
Original series or Next Generation?
hungryhobbit•16m ago
Please don't say Nu Trek.
k12sosse•12m ago
Lower decks obviously
exhilaration•34m ago
You might find this interesting, it seems to be exactly what you need to make an audio drama: https://github.com/Finrandojin/alexandria-audiobook

Also the Qwen3-TTS demo is cool, you can describe the voice you want: https://huggingface.co/spaces/Qwen/Qwen3-TTS

I came across both on this subreddit, it's very active: https://www.reddit.com/r/TextToSpeech/

I'm personally using this locally: https://github.com/mateogon/pdf-narrator (it's a Python frontend for Kokoro) on my M1 Macbook Air (from 2020, with 8GB RAM) and it's incredible. I make my own audiobooks now - for free!

My favorite voice is am_michael and here's a sample: https://voicerankings.com/voice/kokoro-82M/male/am_michael/s...

talon8635•57m ago
Great. Now in additional to AI email responses I will get AIs impersonating my contacts on the phone too. Lovely.
sgc•53m ago
Sorry if this is in that article, but I am on my phone and can't see it. How much would this cost to batch generate an audiobook? Right now I just listen to things in the 11 labs app which is free, but I would rather just generate audio files.
thevinter•50m ago
Roughly 5-10$ for 10h, assuming you few-shot it.

Price per hour:

- 3.8 Flash TTS, standard: $0.81

- 3.8 Flash TTS, batch: $0.41

- 3.8 Flash‑Lite TTS, standard: $0.54

- 3.8 Flash‑Lite TTS, batch: $0.27

xnx•50m ago
Would be great if this would power the Google Books app feature. The voice system there is pretty out of date.
mamudo•43m ago
Yes, I am quite disappointed by seeing all this cool AI stuff and yet the same Play Books. Come on, it is the best place to apply AI, in my opinion.
laweijfmvo•43m ago
the ratio of new voice models i see on hackernews to the number actually deployed in any product i use is approximately infinity.
andrewstuart•47m ago
I’ve never found a TYS that does convincing British accents.

They all sound like Americans putting in their best fake British accent.

maelito•42m ago
Related, for embedding small models, this lib is incredible.

Having a voice under 1Mo is crazy, even if it sounds robotic.

https://tts.ampixa.com/sanoTTS/

drewbitt•38m ago
Great price at least until December 31 too.
burkaman•34m ago
Seems like voice actors are safe for now. This is technologically incredible, but the results are really not very good, and usually not particularly close to the prompt. In basically all of these examples some core part of the prompt is completely ignored.
avazhi•24m ago
> Seems like voice actors are safe for now. This is technologically incredible, but the results are really not very good,

Um, what?

burkaman•15m ago
Technologically incredible as in "I cannot believe it's possible for a computer to do this" and not very good as in "these examples are not what was prompted and I can't think of a use case where these would be acceptable".

Imagine someone showing you that they've trained their dog to hold a paintbrush and paint. There would be no contradiction between "this is incredible" and "these paintings suck".

nater5000•29m ago
It's giving me an error when I try to generate a voice with Voice Design in AI Studio. It also says voice replication isn't available in my region.

Also weird that there are no "neutral gender" voices in the English language. There's also limited "use cases," like the "Gaming" use case is empty?

And there's no pricing listed anywhere.

I don't know, I guess their roll out is a bit sloppy. It's a bit of a shame, though, since the voices which are available all sound like generic Gemini voices to me. Nothing stands out is being particularly interesting or impressive about this.

accountrequired•28m ago
"users must provide a verbal consent recording from the voice owner that matches the reference speaker before a voice can be created"

How long is this stored? What could go wrong? :P

kmoser•3m ago
Does it detect synthesized consent recordings?
m3kw9•25m ago
Still sounds AI, you can tell they exaggerate all the tone and trailing "high scoring expressive sounds" like your job depends on it.
seemaze•24m ago
My primary use case for TTS is converting written content (blogs, articles, etc.) in to clips I can listen to on the go.

Is there a good browser extension that does this with a flexible TTS backend? I know Qwen, Kokoro, and VibeVoice all have decent quality..

kyrra•23m ago
If you use Chrome on Android, there is a "Listen to this page" option in the menu.

https://support.google.com/chrome/answer/14768725?hl=en

xmorse•23m ago
Based on this video I think this model was trained on p*rn

https://storage.googleapis.com/gweb-uniblog-publish-prod/ori...

droidjj•21m ago
It’s just a woman’s voice…
hmokiguess•21m ago
That argument says as much about the training as it says about your sexuality, you do realize that right.
xmorse•19m ago
you can't use this model if it randomly moans between phrases
fullstackwife•23m ago
It would be nice to have sound effect generation (use case: games)
nitroedge•18m ago
Couldn't see this in the article, does it support API calls for one-shot conversation type responses like ElevenLabs offers?

$0.50 per hour pricing could last a long time with back and forth conversation use.

cainxinth•5m ago
They keep announcing new 3.8 variants. I wonder why they still haven't updated 3.1 Pro yet.
OutOfHere•4m ago
Pricing isn't noted.
simonw•3m ago
I vibe coded a playground UI for trying this out. The conversation mode is neat, and it's very expensive - most of my experiments have cost less than a cent.

https://tools.simonwillison.net/gemini-tts-playground#compos...

bakies•22m ago
is the legit ones just like... putting carrie fisher in star wars?
gegtik•12m ago
usually people bring up anyone who is losing physical control of their voice faculties, so they can have a synthetic voice that matches their natural voice
jolan•7m ago
As an example of this, Sarah Langs uses a synthesized voice due to the progression of her ALS symptoms.

https://en.wikipedia.org/wiki/Sarah_Langs

simonw•2m ago
The iPhone has a built in voice cloner hidden in the accessibility settings for exactly this use-case: creating a backup of your voice in case you need it in the future.
gruez•6m ago
I mean yeah? If google offers a cloud nmap tool, should everyone get in a tizzy about how google is "evil", even though it saves baddies maybe 5 minutes of work?
kmoser•5m ago
Google granted themselves the authority be evil back in 2018. https://en.wikipedia.org/wiki/Don%27t_be_evil