frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Show HN: We Beat MLPerf: Modern Storage for KV Offload and LLM Training

https://www.theopenlake.com/blog/openlake-leads-mlperf-storage-v3-0
2•arnav__1•2m ago•1 comments

Structured Procrastination

https://www.structuredprocrastination.com/
1•Tomte•2m ago•0 comments

Data Center Backlash Floods Campaign Trail as Hill Action Stalls

https://news.bgov.com/bloomberg-government-news/data-center-backlash-floods-campaign-trail-as-hil...
2•01-_-•3m ago•0 comments

How Much of the Internet Is Written with AI?

https://www.pewresearch.org/data-labs/2026/08/20/how-much-of-the-internet-is-written-with-ai/
1•giuliomagnifico•4m ago•0 comments

Personal Development Can Improve Productivity

https://medium.com/@flaviopilotodasilva/how-personal-development-can-improve-productivity-d7a621c...
1•flaviopilotodas•4m ago•0 comments

Building Games with Astra

https://developers.openai.com/blog/how-to-build-games-with-astra
1•Garbage•5m ago•0 comments

Music courses are dead (tanks AI) [video]

https://www.youtube.com/watch?v=PCqkJ0HuHTM
1•dgellow•6m ago•0 comments

Why nearly all songs are in a minor key now [video]

https://www.youtube.com/watch?v=982k27De5xc
1•fallinditch•9m ago•0 comments

The Toothbrush Test

https://danromero.org/the-toothbrush-test.html
1•tosh•9m ago•0 comments

Ask HN: Outside of work, what are your social groups?

1•Gecko4072•12m ago•1 comments

It is important who teaches you

https://bastiangruber.ca/posts/it-is-important-who-teaches-you/
2•recvonline•14m ago•0 comments

Find a meat processing plant near you

https://meatprocessingplant.com/
2•g_langenderfer•14m ago•1 comments

Coding-agent skills that cut your worthless tests and write fewer, better ones

https://github.com/haukebri/mira-test-engineer
2•haukebri•17m ago•0 comments

Fable 5.1: 3277 elo UCI NNUE Engine in 24 hrs

https://talkchess.com/viewtopic.php?t=86707
3•EvgeniyZh•17m ago•0 comments

OpenAI boosts Astra's eval metrics, and continues to change others

https://fortune.com/2026/09/04/openai-quietly-boosts-some-of-astras-evaluation-metrics-amid-rare-...
1•enraged_camel•18m ago•0 comments

Praetor – Privacy-first open-source system prompt for CV self-assessment

https://github.com/simonesan-afk/CV-Praetorian-Guard
1•simonenespolosa•18m ago•0 comments

The most epic bank statement, 6 months later (2016)

https://medium.com/@yanismydj/the-worlds-most-epic-bank-statement-6-months-later-6e749c100435
3•Tomte•24m ago•0 comments

Subagents on Subagents: How Many Layers Deep Is Too Many?

https://twitter.com/JoshARosen/status/2087944178558791874
1•gmays•27m ago•0 comments

The Giant Explanation of Possession (1981)

https://filmcolossus.com/possession-explained-1981
2•Bluestein•29m ago•0 comments

When Traceroutes Lie

https://labs.ripe.net/author/shk/when-traceroutes-lie-how-ttl-jumps-can-give-us-a-false-impressio...
1•jruohonen•29m ago•1 comments

A proof of a rectangle partition conjecture using two-bar slabs

https://zenodo.org/records/22349886
1•shaikhmubin•30m ago•0 comments

Hallucinated Citations and Accuracy of Papers

https://veruscite.com/blog/hallucinated-citations-and-accuracy-of-papers
2•apwheele•30m ago•2 comments

Show HN: A browser-based public speaking simulator with WebXR support

https://www.generateppt.com/free-tools/public-speaking-presentation-vr-simulator
1•fer_momento•32m ago•0 comments

Deft: A gradual type system for Janet

https://codeberg.org/zzkt/deft
1•birdculture•32m ago•0 comments

Cortex – I wrecked my cofounder's codebase with agents, so we built this

https://workspace.socra.com/products/cortex
1•eduardadotai•33m ago•1 comments

Pile of Index Cards

https://www.flickr.com/photos/hawkexpress/albums/72157594200490122/
2•Tomte•33m ago•0 comments

Architecting memory and storage in the AI era

https://www.technologyreview.com/2026/09/04/1140872/architecting-memory-and-storage-in-the-ai-era/
1•joozio•33m ago•0 comments

Ask HN: How have interviews changed over the last year?

2•mikodin•33m ago•0 comments

Show HN: TossedTheTVKeptTheRemote – VIA-like configuration to reuse IR remotes

https://github.com/Brisk4t/TossedTheTVKeptTheRemote
1•Brisk4t•34m ago•0 comments

Ask HN: Source Code of Emerald Editor

1•fithisux•36m ago•0 comments
Open in hackernews

Show HN: xbrl-tag mapping for 10ks and 10qs, using EDGAR database

https://xbrlmetrics.com
1•Loris_jk•44m ago
Hello HN! I built a pipeline which fetches SEC EDGAR 10k and 10qs for more than 600 companies, maps xbrl-tags onto around 35 financial concepts and computes growth- , fundamental- and valuation-multiples (prices provided by yfinance). This project originated bc I wanted clean, non-third-party financial data in order to judge a trading thesis of mine. I could not find any app providing such data (for free), let alone being open source. Turns out there are no data providers because xbrl of 10ks and 10qs SUCKS(!!!!). It is messy, unregulated and devastated. Its the wild west of finance. A project like this was 5 years ago not possible, at least not with less than ~1000 people. Thats because of the nature of xbrl-tags. Since they are unregulated, a certain xbrl-tag could fit perfectly for a certain company, while providing nothing, only a partial amount, or complete garbage for another company. You would have to iterate through every financial statement ever relesed by a company, searching every tag possbile, comparing, once found, a tag and constantly checking, once more companies are added to the obserable universe, if the tag also produces meaningfull data for those new companies. As I said, unmanagable task. Except for LLMs. I used Anthropics Claude Fable 5 mostly, giving strict, non-regression focused tag research tasks, constantly letting it iterate through the entire ticker universe, searching and comparing tags. And it worked. It took me around 3 weeks just to produce wrong data, manually searching for a ticker universe of 3 (AAPL, MSFT, NVDA). Now this pipeline produces correct, cleaned data (more on that later) for more than 600 companies.

That was just the main problem tho. Notable problems taht also had to be fixed: -companies tagging everything as fy ull year; -cash flow being filed cumulative; -q4 is never filed; -a quarter isnt always the same length; tag changing by a company; -STOCK SPLITS (god i hate them); -companies reporting wrong (a few 0s more for example);

Still got a lot of work to do, UI and UX are...rudimentary. Appreciate any feedback/critics. Thanks for taking your time reading this.