frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Millions of fish tested Canada's largest nuclear plant

https://www.cbc.ca/news/canada/ontario-bruce-power-climate-change-gizzard-shad-fish-event-lake-hu...
1•geox•17s ago•0 comments

Show HN: HSL-Noise – A noise generator mapped to the HSL color space

https://hsl-noise.mitpit.com/
1•MitPitt•30s ago•0 comments

Friday – a personal AI agent that lives in your text messages

https://askfriday.io
1•claire_befreed•34s ago•0 comments

English: A vs. An

https://www.redblobgames.com/blog/2026-09-16-english-a-vs-an/
2•ingve•5m ago•0 comments

Brevo supply-chain attack injected ClickFix scripts on customer sites

https://www.bleepingcomputer.com/news/security/brevo-supply-chain-attack-injected-clickfix-script...
1•sbulaev•8m ago•0 comments

After the Write Hole (Btrfs)

https://claude.ai/code/artifact/90f24687-21ff-4709-a3e0-5bf50ba547c2
1•airhangerf15•9m ago•0 comments

Plugin4Shell – Zero Click RCE Vulnerability found in top four coding agents

https://www.air.security/blog-posts/plugin4shell
2•fishthethis•9m ago•0 comments

The open houses at the edge of disaster

https://heated.world/p/the-open-houses-at-the-edge-of-disaster
1•sebg•10m ago•0 comments

Being Human After AI: Daft Punk, Pope Leo and 'Magnifica Humanitas'

https://www.americamagazine.org/music/2026/08/14/daft-punk-pope-leo-magnifica-humanitas/
1•pseudolus•11m ago•0 comments

The Pi0-Fast Libero-10 Baseline Is 25 Points Low

https://topicqueue.substack.com/p/the-pi0-fast-libero-10-baseline-is
1•DISCURSIVE•11m ago•0 comments

Show HN: What if you don't feel overwhelmed reading big articles

https://chromewebstore.google.com/detail/kora-–-focus-reader-websi/kbkcobapaoaoilpfgbaaeaedlinc...
1•AkshayS96•12m ago•0 comments

Aberrant excitatory neuronal ERBB4 promotes Alzheimer's disease pathology

https://www.nature.com/articles/s41586-026-10964-z
1•wslh•13m ago•0 comments

Fastembed – Rust library for generating vector embeddings, reranking locally

https://github.com/Anush008/fastembed-rs
1•modinfo•13m ago•0 comments

The FAA's $875M Plan to Use AI to Ease Air-Traffic Woes

https://www.wsj.com/business/airlines/the-faas-875-million-plan-to-use-ai-to-ease-air-traffic-woe...
1•JumpCrisscross•14m ago•0 comments

Rdap-Java – A ~5KB Java 11 RDAP client to check domain availability

https://github.com/slxca/rdap-java
1•slxca•15m ago•0 comments

Show HN: Hishel – HTTP Caching for Python

https://hishel.com
1•karpetrosyan•17m ago•0 comments

CC is an AI agent for families and groups

https://blog.google/innovation-and-ai/models-and-research/google-labs/cc-expanding-to-groups/
2•xnx•18m ago•0 comments

Deals: The Most Active Investors in European Defence Tech in 2026

https://foxandlion.pub/analysis/68-deals-the-most-active-investors-in-european-defence-tech-in-2026
1•fascinnn•19m ago•0 comments

Why do Microsoft job levels start in the high 50s?

https://devblogs.microsoft.com/oldnewthing/20260915-00/?p=112700
1•ibobev•21m ago•0 comments

Wanix – WASM-native Unix sandboxing for the web

https://wanix.dev/
1•orangea•21m ago•0 comments

In Montana, a revolt against corporate money could reshape political spending

https://www.reuters.com/legal/government/deep-trump-country-revolt-against-corporate-money-could-...
1•JumpCrisscross•23m ago•1 comments

Meta Launches Subscription Bundles

https://www.theverge.com/tech/995453/meta-one-subscriptions-ai
1•busymom0•24m ago•0 comments

Show HN: A map of 68 programs and jobs in AI safety research

https://cleverhack.com/ai-research-ai-safety-ai-talent
1•jlark77777•25m ago•0 comments

ClickHouse's New TimeSeries Engine: Drop-In Prometheus Replacement

https://clickhouse.com/blog/introducing-promql
1•elmazout•25m ago•0 comments

Show HN: Multiplayer Mode for AI Agents

https://gotincan.com/
1•dko•27m ago•2 comments

What makes entrepreneurs entrepreneurial? (2001) [pdf]

https://effectuation.org/hubfs/Journal%20Articles/2016/06/what-makes-entrepreneurs-entrepreneuria...
1•wslh•28m ago•0 comments

Run QWEN3.8 27B on 16gb Nvidia GPUs

https://github.com/MiaAI-Lab/Qwen3.8-27B-16gb-NVIDIA-GPUs-one-click-install
2•Pragmata•28m ago•1 comments

A14 Technology

https://www.tsmc.com/english/dedicatedFoundry/technology/logic/l_A14
1•JumpCrisscross•28m ago•0 comments

Std: Call_once vs. Std:Async

https://devblogs.microsoft.com/oldnewthing/20260917-00/?p=112706
1•ibobev•31m ago•0 comments

Everybody's Lost Their Minds

https://www.netmeister.org/blog/everybodys-lost-their-minds.html
2•ibobev•31m ago•0 comments
Open in hackernews

Show HN: MCPJam - the first testing & evaluations platform for MCP servers

https://www.mcpjam.com
7•prathmeshmcp•51m ago
Prathmesh, CEO of MCPJam here.

Users now start in ChatGPT, Claude, Cursor, and other AI clients. They reach your product through your MCP server.

That means your users often aren’t in your product anymore. You can’t see what they prompted for, how the agent interpreted it, or whether your server helped them get the result they wanted.

I saw this firsthand leading MCP technical strategy at Asana, including our ChatGPT and Claude launches. We were building high-stakes enterprise integrations, but we had no reliable way to test them the way we test normal software- or to know whether they worked once they reached real users.

I started using MCPJam for those problems after re-connecting with my former coworker who created the project. brought it to more of our developers, and worked it into our CI/CD pipeline. I joined the team because I kept hearing the same issue from other companies building for agents.

So, what does “good” look like for MCP? For us, it means users reliably get the outcome they came for, across the AI clients they use.

That’s what we’ve been building toward. MCPJam now helps you test the full workflow, from the first prompt to the expected result:

* Swarms: Simulate users with different goals and prompts to find where workflows break across AI clients. * User Testing: Watch how real users interact with your MCP product, where they get stuck, and how they feel about the results. * Evals: Turn those workflows into repeatable tests that check whether users get the expected outcome. * CI/CD: Run those evals across AI clients before each release to catch regressions.

Over 106,000 developers and +300 enterprises use our open-source solution to see how their servers behave locally across major AI clients. MCPJam has grown from a debugging tool into a continuous testing and evaluation workflow for MCP servers.

If you’re building an MCP server or agent-facing product, give MCPJam a try. What is the hardest thing for you to test? We love hearing about your MCP server builds!

Comments

vigjam•41m ago
love the direction, but the problem for me has been about creating stronger evals and knowing what I should be checking for. does this help me understand that?