frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

TCP is failing AI, but Stanford's Homa is here to help

https://www.theregister.com/networks/2026/10/01/tcp-is-failing-ai-but-stanfords-homa-is-here-to-h...
1•joebuckwilliams•18s ago•0 comments

Folder structure has two jobs, but only does one

https://blog.sterlingfreeman.net/posts/folder-structure-two-jobs/
1•alwaysDrafting•52s ago•0 comments

Cold War spy satellite named after Farrah Fawcett just exploded in space

https://www.popsci.com/science/farrah-fawcett-satellite-explosion-explained/
1•CGMthrowaway•1m ago•0 comments

SvelteKit 3 Is Here

https://svelte.dev/blog/sveltekit-3-is-here
1•sampsn•5m ago•0 comments

ArXiv's Updated Rate Limit Policy

https://blog.arxiv.org/2026/10/01/updated-rate-limit-policy/
2•50kIters•7m ago•0 comments

The failure modes of Claude Code in a guided 60h project

https://hmijail.substack.com/p/building-a-semantic-fuzzer-for-obsidian-sync-in-spite-of-claude-2
1•hmijail•9m ago•1 comments

Antarctica's winter sea ice peaks at 3rd lowest level on record

https://www.nationalheraldindia.com/environment/antarcticas-winter-sea-ice-peaks-at-3rd-lowest-le...
1•wslh•9m ago•0 comments

Show HN: HeadWater B2B Data and AI Kit – A Production Architecture Shortcut

https://github.com/PunkiePal/HeadWater-AI-Digital-Products/blob/main/headwater-b2b-data-purificat...
1•PunkiePal•10m ago•0 comments

cua-speedrun: Standardized Benchmarking of the Speed of Computer-Use Agents

https://cuaspeedrun.com/
1•matt_d•10m ago•0 comments

Gboard Conveyor Belt Version

https://www.youtube.com/watch?v=DAn34l_YrUM
1•adrianhon•12m ago•0 comments

How the Bayh-Dole Act Transformed Tech Transfer

https://research.umn.edu/units/techcomm/news/changing-rules-discovery-how-bayh-dole-act-transform...
1•CGMthrowaway•13m ago•0 comments

Loss of Cell Identity Drives Human Aging

https://erictopol.substack.com/p/loss-of-cell-identity-drives-human
1•bookofjoe•16m ago•0 comments

The Game Theory of AI Pacing

https://www.paradigm.xyz/writing/the-game-theory-of-ai-pacing
1•ronfriedhaber•17m ago•0 comments

Roman Dodecahedron

https://en.wikipedia.org/wiki/Roman_dodecahedron
1•softwaredoug•20m ago•0 comments

Random TV Time

https://play.google.com/store/apps/details?id=com.randomtvtime.app&hl=en_US
1•eoleol•22m ago•0 comments

Show HN: MoA as Universal IL –> 30yrs from PhD to 5-device validated compiler

https://github.com/womenflyplanes/moa-compiler
1•lmullin•22m ago•0 comments

Free hosted API for Laya, the open-weight decision model

https://vercel.com/ai-gateway/models/laya
1•boundlesshq•22m ago•1 comments

Kcc – a C17 compiler built solo with an LLM on $100/month, no agents

https://github.com/LiterateDrivenDevelopment/kcc
3•temo_pacheco•26m ago•1 comments

Show HN: Paper(Local CLI for Agent Debugging)

1•varman11•28m ago•0 comments

Getting started with Claude Code mods

https://claude.dev/blog/getting-started-with-claude-code-mods/
2•mfiguiere•31m ago•0 comments

SvelteKit 3 Released

https://github.com/sveltejs/kit/releases/tag/%40sveltejs%2Fkit%403.0.0
4•thunderbong•37m ago•0 comments

Show HN: What 482 hospitals charge vs. what insurers pay, from their own files

https://www.billmender.com/hospital-prices
1•curatedmcp•37m ago•0 comments

AI fabrication suspected of writing articles in Canadian publications

https://www.theglobeandmail.com/world/article-ai-suspected-of-writing-dozens-of-articles-in-canad...
3•rdmuser•39m ago•0 comments

We Used AI to Find Out Just How Much Americans Hate AI

https://www.wsj.com/politics/elections/we-used-ai-to-find-out-just-how-much-americans-hate-ai-2ca...
3•dzohrob•39m ago•0 comments

Show HN: Cascade, Mac OS window cascade utility

https://github.com/hasansayani/MacOSApps-Cascade
1•hhsayani•41m ago•0 comments

Open Source PWA of Doom rebuild from zero

https://github.com/spipu/doom-js
1•Spipu45•41m ago•1 comments

War Bros Didn't Always Rule Silicon Valley

https://www.wired.com/story/war-bros-didnt-always-rule-silicon-valley/
2•gmays•43m ago•0 comments

Etymon – Get rid of all your Claude.md, AGENTS.md, cursor/rules, mcp.json, etc.

https://github.com/landoncrabtree/etymon
2•landoncrabtree2•44m ago•2 comments

Layercake converts OpenStreetMap data into tables you can SQL query via a URL

https://layercake.openstreetmap.us/
3•raybb•45m ago•0 comments

I'm Calling for a Pause in the Development of the Universal Bone Dissolver

https://www.theatlantic.com/newsletters/2026/10/ai-slowdown-bone-dissolver/688854/
4•Jtsummers•45m ago•3 comments
Open in hackernews

Apex, a coding model specialized in mobile and web with React

https://www.callstack.com/blog/apex-agetic-coding-for-mobile-and-web-spcialied-in-react
5•thymikee•1h ago

Comments

thymikee•1h ago
Hi HN, I’m Michal. Together with Piotr, Artur, Lech, and Mike, for the last 8 months we've been building Apex, a coding model for mobile and web development. A model which we trained on React Native expertise and turned out to be good at React web too.

We're a small R&D team at Callstack, where we spend a lot of time working on React Native apps and create open-source libraries. We wanted to see how far that experience could take an open-weight model through domain-specific training. Could we achieve frontier-level performance for specific tasks at a much lower cost? After a few base model iterations, we sticked to Qwen as a foundation, trained on curated framework documentation, source code, APIs, and repository-grounded conversations, evaluated by our engineers.

On our React Native Evals (rn-evals.vercel.app), Apex scores 87.8%, compared with 88.6% for Claude Opus 5.5. The reported evaluation cost is $7.72 versus $28.75. These are our benchmarks, so we’d encourage you to try tasks from your own repository too.

Over the weekend we released a stealth version of Apex, Pixel Canary on Vercel AI Gateway. In Vercel’s Next.js evaluation, that preview passed 28 of 31 tasks, or 30 with documentation in context, matching GPT-6 Astra. Opening free access to the model on AI Gateway smashed our infrastructure and in these first days Pixel Canary was slow. We're not DevOps by train, but over these 5 days we had a quick crash course on scaling the model performance on B200s with custom prefill and decode setups to optimize load accordingly to the traffic we made, allowing us to achieve ~100 TPS at 4s p50 latency on the last day, with up to 80B tokens / day throughput.

The API is now available for business users (we need few more weeks to sort out EU consumer rights). It works with OpenAI- and Anthropic-compatible coding tools. We charge $0.50/M input, $0.20/M cache, and $3/M for output.

We’re interested in where specialization pays off and where it falls short. We're planning more specialized models for native iOS and Android development, and more.

I'm using this model for other languages and it's just as good as running Sonnet or Sol on medium. It's still Qwen, but tuning seems to make it better than barebones version on coding tasks.

If you try it, I’d love to hear which tasks it handles well and where you still reach for another model.