frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Clef: our open-source decision models

https://blog.cloudflare.com/clef-decision-models/
94•jasondavies•59m ago•27 comments

RIP, vector database

https://turbopuffer.com/blog/rip-vector-database
75•razin•1h ago•17 comments

StreetComplete on iOS is now in public beta

https://github.com/streetcomplete/StreetComplete/issues/5421
392•Snowly•6h ago•83 comments

Git 3.0's upcoming SHA-256 default will be a costly mistake

https://blog.gitbutler.com/git-3-sha-256
14•chmaynard•21m ago•1 comments

RacketCon Is Saturday

https://con.racket-lang.org/
57•spdegabrielle•2h ago•17 comments

How to speed up the Rust compiler in September 2026

https://nnethercote.github.io/2026/09/30/how-to-speed-up-the-rust-compiler-in-september-2026.html
167•trickypr•4h ago•81 comments

Identity Management for Agentic AI [pdf] (2025)

https://openid.net/wp-content/uploads/2025/10/Identity-Management-for-Agentic-AI.pdf
41•cgeier•2h ago•6 comments

Cloudflare K2: serverless event streams

https://blog.cloudflare.com/cloudflare-k2-streams/
69•elffjs•3h ago•20 comments

Gemini 4 Argon

https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/
1582•bradleyg223•21h ago•1055 comments

Polyedergarten: Garden of Paper Polyhedron Models

https://www.polyedergarten.de/e_index.htm
21•isaacimagine•2h ago•2 comments

GPT-Synopsys: Frontier Intelligence to Revolutionize Chip Design

https://news.synopsys.com/2026-09-30-OpenAI-and-Synopsys-Announce-GPT-Synopsys-Frontier-Intellige...
131•giuliomagnifico•6h ago•70 comments

Red Hat being phased out of existence?

https://techrights.org/n/2026/10/01/Red_Hat_Being_Phased_Out_of_Existence_Like_Many_Other_Compani...
74•amcclure•1h ago•33 comments

ParadeDB Search Performance Improvements

https://www.paradedb.com/blog/opening-a-closed-tin
5•craigkerstiens•13m ago•0 comments

Figma restricts MCP access to whitelisted clients, excluding Pi

https://twitter.com/GayaniFigma/status/2105295629941350454
97•thdr•2h ago•45 comments

Micron CEO Says Memory Supply Will Be Much Tighter in 2027 and 2028 Than in 2026

https://www.techpowerup.com/353296/micron-ceo-says-memory-supply-will-be-much-tighter-in-2027-and...
205•speckx•4h ago•245 comments

OpenDLSS: A Vulkan Reimplementation of Nvidia's DLSS 5 Neural Rendering Network

https://github.com/maanHimself/OpenDLSS-NR
210•sagacity•1d ago•100 comments

Cops Can Bypass iPhone's Automatic Reboot to Get into Locked Phones

https://www.404media.co/cops-can-bypass-iphone-automatic-inactivity-reboot-graykey/
122•speckx•2h ago•77 comments

Various Projects Find Hidden SDR Capabilities in ESP32 Microcontrollers

https://www.rtl-sdr.com/various-projects-independently-find-hidden-sdr-capabilities-in-esp32-micr...
10•nkw•2h ago•0 comments

FTC is investigating OpenAI, Anthropic and other AI companies over product risks

https://www.cnbc.com/2026/09/30/ftc-ai-probe-openai-anthropic.html
142•dgellow•4h ago•88 comments

Truemetrics (YC S23) Is Hiring a GTM Founder's Associate

https://www.ycombinator.com/companies/truemetrics/jobs/THLEzXI-gtm-founder-s-associate
1•truemetricsIngo•7h ago

Book of Shapes – Collection of minimal, generative and customizable SVG-patterns

https://bookofshapes.com/
200•eustoria•2d ago•14 comments

Returning from vacation? The government can search your phone without a warrant

https://arstechnica.com/tech-policy/2026/09/immigration-advocate-sues-border-agents-for-demanding...
264•rbanffy•6h ago•256 comments

Lightweight PDF parser with layout, tables, formulas and bounding boxes

https://github.com/beatrizalmeidaf/papero-pdf-text-extractor
3•beatrizalmeidaf•1h ago•0 comments

Canada fast-tracks Pacific oil pipeline to reduce US dependence

https://apnews.com/article/alberta-canada-carney-pipeline-68539133d6e0245fad3622263afd4aeb
40•geox•1h ago•25 comments

Why the Bronze Age Collapsed

https://www.worksinprogress.news/p/why-really-caused-the-bronze-age
360•AnodicElegy•2d ago•237 comments

The top secret URSALA, RAQUEL, and FARRAH satellites (2025)

https://www.thespacereview.com/article/4951/1
286•Bluestein•19h ago•147 comments

Ask HN: Who wants to be hired? (October 2026)

19•whoishiring•2h ago•97 comments

Adding Floating-Point Decimals for Fun and Profit

https://blog.vero.site/post/float
50•ibobev•2d ago•26 comments

Before pixels: Modular industrial dashboards

https://unsung.aresluna.org/before-pixels-modular-industrial-dashboards/
260•leephillips•22h ago•47 comments

Ask HN: Who is hiring? (October 2026)

21•whoishiring•2h ago•40 comments
Open in hackernews

RIP, vector database

https://turbopuffer.com/blog/rip-vector-database
75•razin•1h ago

Comments

OutOfHere•53m ago
It would be nice to have a page that actually loads. This one doesn't. RIP.
syndacks•49m ago
loads just fine on my $10k laptop with 10g internet here in NYC
uproarchat•41m ago
Also loads fine on my beater in the sticks :)
alexjplant•36m ago
Takes 11 seconds to load on Firefox on Linux with 3G-level throttling enabled in Dev Tools.
throwawy0352•41m ago
Loads really fast for me. (MacBook Air, average internet)

If you still have issues, try https://web.archive.org/web/20261001100105/https://turbopuff...

wilj•31m ago
It has a pagespeed insights score of 55 and noticeably sluggish on my m3 max.

And what's with the throwaway account for this one comment? Is this becoming reddit with throwaway shills now?

phoghed•25m ago
fucking shills, making helpful comments and promoting seemingly nothing, what's this place coming to?
throwawy0352•19m ago
Yes, I get big money from the Internet Archive to promote their services. It's the new scheme that shills like me go for.

The reason is that I have no account on HN and rarely comment. I create a new account a few times a year because I don't remember or care about my previous account.

I could have made an account named john2026 and you would not think twice. Instead, I let people know upfront what type of account this is. Quite the opposite of what a true shill would do.

I got a Lighthouse score of 99 in Chrome. Believe it or not, I won't spend more of our time on this. (relevant XKCD: https://xkcd.com/386/ )

First Contentful Paint 0.7 s

Largest Contentful Paint 0.9 s

Speed Index 0.7 s

It makes a lot of requests, and some are stopped by my ad blocker, but most of them don't seem to make an difference. It is almost instant from my point of view. I disabled the ad blocker and didn't notice any visual difference.

sreekanth850•52m ago
I find very little reason to use a pure vector database for enterprise retrieval. We built an enterprise retrieval engine on top of a SQL database with native vector support, and the flexibility is something we cannot ignore. Vector similarity is just one query primitive alongside full text search, filters, joins, ordering and normal relational predicates. Tenant/app/collection isolation becomes part of the query itself. ACLs, document versions, categories, metadata constraints and temporal filters are ordinary predicates rather than something you have to bolt onto a vector store. SQL is already going to be part of almost any enterprise system. Adding a separate vector database introduces another moving part and syncing two system whenever you update your data is the most difficult thing to get right.
blakeashleyjr•51m ago
This sounds like the Postgres vs. InnoDB argument 10 years later. Postings pointed at physical location (the ANN slot), so every SPFresh rebalance rewrote every index touching that doc. InnoDB solved this by pointing secondary indexes at the PK and eating an extra lookup on read. Curious what that extra lookup costs you when it's an S3 GET instead of a B-tree hop.

"Updating one vector can move hundreds of attributes and their indexes" is basically Uber's 2016 Postgres write amplification post, but for search. Same fix too: stop pointing indexes at where the row lives.

So ANN becomes a secondary index that points at a doc ID, and vector search now needs a hop to complete. Do clusters keep their own copy of the vectors so the search itself stays local, and only result fetch pays the indirection? Otherwise cold p99 seems like it gets worse.

gopalv•32m ago
> This write amplification is large enough that our efforts to tune indexing throughput have started to hit diminishing returns.

> don't key on the ANN address. That is precisely the change turbopuffer v3 makes. As you can imagine, it is not a trivial change.

This is a direct parallel to how Postgres and Mysql built indexes.

Your design choice went from a Postgres design pattern to a Mysql one. The difference is the reindexing cost vs the lookup cost - Postgres optimized for lookup and Mysql does for indexing on writes. Or more accurately, Postgres was better with good schema design using joins & mysql was optimized for a bad design with less normalization where many indexes exist for the same table.

Postgres always points an index to a row-id within postgres which is an arbitrary value which changes on each update.

Mysql, always assuming the storage engine is pluggable, points to the primary index entry and adds an extra indirection to the lookup.

This means that you point the mysql index to a stable id, so unless you go update the primary key for a row, you won't have to update the indexes for all the attribute lookups you might have made to data.

I don't do databases any more that much, but the design for NIMBLE file format has a lot of quirks which are relevant to this specific idea (wide tables).

But the old Uber post about switching from Postgres to Mysql to prevent index amplification[1] is a direct mirror to this post.

[1] - https://www.uber.com/us/en/blog/postgres-to-mysql-migration/

phoghed•26m ago
> mysql was optimized for a bad design

TIL I should have been using mysql the whole time

woadwarrior01•23m ago
Richard Gabriel's "Worse is better" vibes.
drewlanenga•29m ago
the multi-vector duplication thing makes sense, copying every attribute once per vector explodes quickly. what's the new primary index?
gk1•26m ago
Vector databases were always more about retrieval than either vectors or data storage. But the term stuck all too well and companies held on to it a tad too long. Sorry :)
anuptalwalkar•17m ago
Anyone using pure vector dbs at this day and age is shooting themselves in the foot, but there is more to it than just deprecating vectors.

I built a corrective memory layer for our agents which is using filtering, hybrid/ranked fusion search and strongly typed predicates to provide the LLMs context to correct themselves in case of errors.

Small plug, if anyone wants to try it out- https://polign.com/recall

I struggled quite a bit relying on pure vector DBs, so this is a welcome change. You still need vectors to reach close enough areas to fetch the context though.

ActorNightly•2m ago
Im not full read up on RAG pipelines, but has anyone ever tried to make the database a neural net itself?

I see something like taking an auto encoder, cutting it in half to get the discrete latent space, and then mapping documents across that latent space in terms of threshold values.

When storing documents, you just compute their latent space representation, and for each value in the latent space, you have a map of threshold -> document.

When doing a query, you simply map the query into the space, and then filter each value on the activation threshold.

Then when you do a query, that gets mapped to latent space. Then you sequentially check every v

Meanwile do

That way to index documents, you just have to compute their latent space threshold values.