frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

About Z AI "Benchmaxxing"

https://twitter.com/natolambert/status/2088272938361606397
1•tosh•1m ago•0 comments

Ask HN: Should hard-tech founders join a problem or technology-first PhD lab?

1•misterballer•2m ago•0 comments

Show HN: Omegle with (full-duplex) Voice Bots

https://www.aimegle.xyz/
1•MichelBartels•2m ago•0 comments

Show HN: A painting engine that lets you sculpt drawings (tech demo)

https://softedge-techdemo.jign.workers.dev/
1•jignb•4m ago•0 comments

When Genius Fails: The Intellectual Arrogance of the AI Labs

https://weightythoughts.com/p/when-genius-failsthe-intellectual
2•gmays•5m ago•0 comments

AI Watermark removers can't verify themselves I built the tool that can

https://github.com/Neeeophytee/ai-watermarks-reality-check
1•sromana14•7m ago•0 comments

Senv: Sandboxed Python environments with the uv workflow

https://medium.com/@Koukyosyumei/senv-sandboxed-python-environments-with-the-uv-workflow-44b6b0c4...
1•syumei•8m ago•0 comments

Until Further Notice

https://docs.eventsourcingdb.io/blog/2026/08/03/until-further-notice/
1•goloroden•8m ago•0 comments

Ask HN: Why not expand accessibility laws to include the right to use AI?

1•amichail•9m ago•0 comments

Dots3-Note Preview

https://studio.dots.ai/dots/dots3-en.html
2•tosh•9m ago•0 comments

2004 RuneScape fit a multiplayer RPG into 56k dial-up

https://jkm.dev/posts/how-2004-runescape-fit-a-multiplayer-rpg-into-56k-dialup/
1•birdculture•9m ago•0 comments

Every Fucking Website

https://lxe.github.io/everywebsite/
3•doubletwoyou•9m ago•1 comments

The Geometry of Transformer Weights – A Canonical Basis for LLMs

https://zenodo.org/records/21935673
1•todot_ge•9m ago•0 comments

LLM Batch APIs: The Half-Price Lane Nobody Budgets

https://www.digitalapplied.com/blog/llm-batch-api-pricing-landscape-2026
1•peter_d_sherman•10m ago•0 comments

I closed pull requests and issues on my projects, and I'm happier

https://en.andros.dev/blog/e83e9103/i-closed-pull-requests-and-issues-on-my-projects-and-im-happier/
2•andros•13m ago•0 comments

Show HN: Version 2 of the Savitar macOS MUD client

https://github.com/jkoutavas/Savitar2
1•shazeubaa•13m ago•0 comments

Fast, on Device Agentic AI with Muse Glimmer on ExecuTorch

https://pytorch.org/blog/fast-ondevice-agentic-ai-with-executorch/
1•andrewstetsenko•13m ago•0 comments

Canva slashes valuation by $10B as AI reality bites

https://www.msn.com/en-au/money/news/sobering-markdown-canva-slashes-valuation-by-10b-as-ai-reali...
2•ilamont•14m ago•1 comments

Show HN: Opensourcing APH Engine and Servers in Rust and N Lang

https://github.com/squillo/aph
2•scott-b•15m ago•0 comments

Zero Platform Cut, Flat Fee: How Z-Text Channels Work

https://z-text.org/zero-platform-cut-flat-fee-z-text-channels/
1•ztextzksnarks•15m ago•0 comments

Ask HN: What's the AI Revolution "Model T Ford"?

1•ComputerPerson•15m ago•0 comments

Show HN: WASI-Env – Single-file HTML with 90 Suckless C binaries and SQLite3

https://wasi-env.netlify.app/
1•badfox•16m ago•1 comments

2027 Is Genuinely Looking Apocalyptic [video]

https://www.youtube.com/watch?v=gFFn6XwN2RY
2•xbmcuser•17m ago•0 comments

Show HN: Safe TensorBoard Trace Reducer(90%+ Footprint Cut) with Fail-Safe Guide

https://github.com/PastToFuture-Whisperer/xprof-cubism-reducer/discussions/3
1•PTF-W•17m ago•0 comments

AI Model Atlas – visualizing populations of ML models as interconnected 3D graph

https://run.cosmograph.app/public/ca9fd1ad-fe83-4238-8b69-b707c633aef0
2•bj-rn•17m ago•0 comments

Cursor and SpaceXAI: the fastest iterating team wins

https://www.a16z.news/p/cursor-spacexai-fastest-iterating-team
2•jbredeche•18m ago•0 comments

I turned my RSS feeds into an e-ink newspaper to stop reading on my phone

https://heyjonny.dev/posts/rss-to-eink-newspaper/
1•speckx•19m ago•0 comments

Show HN: We mapped AI patterns before watermark removal became a mainstream

https://founderstoday.org/claude-watermark-remover
1•RichardOdds•21m ago•0 comments

'Masturbation Consultants' Were Hired to Pleasure Themselves with AI

https://www.wired.com/story/these-masturbation-consultants-were-hired-to-pleasure-themselves-usin...
1•pavel_lishin•22m ago•1 comments

Your Contacts All Know Each Other. That's the Problem

https://julienreszka.com/blog/your-contacts-all-know-each-other-that-s-the-problem/
1•julienreszka•23m ago•1 comments
Open in hackernews

Show HN: Is AI Dumber Today? An index of AI model experience from user's opinion

https://isaidumber.today/
2•schafberg•49m ago

Comments

schafberg•49m ago
Hi HN,

I started to build Is AI Dumber Today because I kept seeing people complaining like "GPT/Claude/Gemini get dumber, is it just me?", usually after an update, or before new models launch.

We saw benchmarks on lots of platforms, but they don't keep tracking models qualities, and the model channels, harness, configurations make such tracking more complex.

So I built this tool to track how people describe their real experiences hourly, from the platforms like Reddit, Hacker News, Zhihu, Rednote etc., which are the major platforms in English and Chinese communities. It's even more interesting to see how people feel and react on the same models in different regions.

You can use it to: - See which models perform better or worse than usual and the reasons from user opinions, with hourly latency (Because I am a data engineer) - View model performance from user opinions across the lifecycle of model - Compare models based on user’s feedback and recommend top models in categories like coding, reasoning, speed, etc. - Report your own experience with one tap

For the limitation, I would like to say it is an observational experience index, but not a real benchmark. Although I don't think benchmarks can help us to make decisions to choose models, I also don't think Is AI Dumber Today is more accurate. It just provides us with information from another perspective. I believe as the AI developing, and the iteration of this product, we can extract more useful opinions, like which models do users prefer exactly, and how they are using them. We can learn from others easily.

I would appreciate feedback on the methodology, and what would make the tool more useful or trustworthy.

I also write weekly newsletters based on the data I collected and my observations here: https://isaidumbertoday.substack.com/

You can also follow me on X: https://x.com/isaidumber