frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Toolbase – Build reliable AI teammates by example, not instruction

2•David1238•1y ago
Hey HN, we’re David and Ethan, co-founders of Toolbase.

Toolbase is an AI agent and workflow builder that helps you quickly create production-grade AI automations — batteries included.

Try it without registering here: https://gettoolbase.com/

------- The dev cycle in Toolbase -------

1. Define a rough goal.

2. Connect any API or MCP server (we have thousands, or bring your own).

3. Teach the AI what valid input and output examples look like (these later become your unit tests!).

4. Let Toolbase generate the perfect prompt, code, workflow, or agent.

5. Deploy your project as an API, MCP server, or chat interface!

Coding optional. Sharing encouraged.

Note: If you’re familiar with Cursor/Windsurf and the core concepts behind frameworks like Mastra (which Toolbase runs on), you already know how to use Toolbase. You retain the flexibility of coding but avoid boilerplate and plumbing tasks (integration, validation, context mapping, testing, etc.) unless you explicitly choose to do them.

------- Demo -------

https://www.loom.com/share/540c61b2c5634996b088ebbb16989cf0?...

This simple agent validates company billing addresses in our CRM (Pipedrive) by researching them through Tavily search. If there’s an address mismatch, it asks a human to pick the right one via email (watch on 2x speed):

Output email for the example shown in the demo: https://gettoolbase.com/assets/demo-screenshot.png

Producing code and more deterministic workflows follows a similar process.

------- Why another agent/workflow builder? -------

We started building Toolbase out of frustration with existing frameworks, especially for production use, and the lack of IDE support for MLOps artifacts, such as prompts, golden data, workflows, and evaluations. Since much of the code for agentic systems can be dynamically generated and validated using these artifacts, they often become even more important than the code itself.

After speaking with other builders, we also realized that manually coding workflows, experimenting with prompts through trial-and-error, and setting up infrastructure/integrations all took far more time than they should. Tools like Cursor and Windsurf help, but extracting meaning from AI-generated code is slow. Chatbots whipping up arcane code potions in tinted chat windows, which is the other end of the spectrum, demos really well but isn’t maintainable at all (sorry, vibe-coding). So we went with something in the middle: an AI-assiststed visual builder with full code fallback.

------- What do you think? -------

We’re excited for any feedback, thoughts, or questions from the HN community.

Let us know what you think in the comments!

- David & Ethan

Comments

David1238•1y ago
If you want to jump straight to trying it, you can preview it here:

https://gettoolbase.com

Try clicking the play button at the bottom right of workflow pane that appears next to the big blue text.

Clicking each node after you submit your query shows you its results for the run.

Full-screening the experience (icon on the top right) and expanding a node (same icon within a step of the workflow) lets you train that prompt in the same way we did in the demo.

The first node (“User Intent Extractor”) is purposefully vague so you can train it yourself!

Show HN: TypeSeer – on-device autocomplete for every text field on macOS

https://typeseer.com/
1•animateus•49s ago•0 comments

Show HN: Rediagram – put the diagrams you have on your brand

https://rediagram.app/ai-diagram-generator
1•marrouchi•1m ago•0 comments

We Must Create the Shit Machine

https://www.mcsweeneys.net/articles/we-must-create-the-shit-machine
1•beepbooptheory•2m ago•0 comments

Intel Appears to End Its Bug Bounty Program

https://www.phoronix.com/news/Intel-Bug-Bounty-Program-Ends
1•nh2•6m ago•0 comments

DeepSeek v4.1 Flash avg 102 tps on 4x RTX6000 pro max-q, 2.1x up from v4-flash

https://forum.level1techs.com/t/llm-inference-workstation-4x-rtx6000-blackwell-pro-max-q-384gb-vr...
2•ambientlight•8m ago•1 comments

We were right (about passkeys) all along

https://mailpace.com/blog/musings/we-were-right
4•albertgoeswoof•9m ago•0 comments

Stack Overflow relaunched Developer Story (who certifies that a human wrote it?)

https://stackoverflow.blog/2026/09/10/re-introducing-developer-story/
1•rkovashikawa•10m ago•0 comments

Friday Facts #446 – An ARM and a Frame

https://factorio.com/blog/post/fff-446
1•robotnikman•10m ago•0 comments

The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It

https://arxiv.org/abs/2609.16247
1•1317•17m ago•0 comments

I Have Been a DelGuard

1•bh0k4l•20m ago•0 comments

The Mic Is On. So Is Live Auto-Tune.

https://www.nytimes.com/2026/09/11/arts/music/live-auto-tune-jojo-flo-singing.html
1•bookofjoe•22m ago•1 comments

Institutional Parasitism in Open Technology Communities

https://garden.s3.wasabisys.com/__garden/bacee012-5d7d-4f1a-bcf8-67864e663877
2•theyoungsir•22m ago•0 comments

UK could force phone companies to add 'anti-theft protections'

https://www.bbc.com/news/articles/ckqxvy98x578o
3•electrum•25m ago•0 comments

Using jev to improve product experiences is pretty crazy

https://www.elvex.com/blog/early-experimentation-using-jev-to-rethink-harness-ux
4•sak84•29m ago•3 comments

What I learned from using FreeBSD as a main OS for a summer

https://divanv.com/post/adventures-in-bsd/
2•divanvisagie•29m ago•0 comments

A Lack of Honesty Is the Ultimate Killer

https://phillipspobrien.substack.com/p/a-lack-of-honesty-is-the-ultimate
2•JumpCrisscross•31m ago•0 comments

Ask HN: What do you think of Noul, a new decision primitive

1•hbarka•32m ago•0 comments

Wealth Taxes Can Make Capital Markets More Efficient

https://www.promarket.org/2026/09/14/wealth-taxes-can-make-capital-markets-more-efficient/
4•paimapi•33m ago•0 comments

German Hospitals Prepare for War, Drones and Mass Casualties

https://www.bloomberg.com/news/features/2026-09-16/germany-s-hospitals-prepare-for-war-and-mass-c...
4•vrganj•33m ago•1 comments

Meta Muse Hits #1 in Apple App Store

https://www.businessinsider.com/meta-muse-personal-ai-agent-top-app-store-charts-2026-9
3•pizzathyme•34m ago•0 comments

You need more than just vanilla RAG

https://medium.com/@NikoZero11/your-rag-pipeline-doesnt-need-more-retrieval-it-needs-better-decis...
3•Savanjj•35m ago•0 comments

Claude Code now reads AGENTS.md if there is no Claude.md

https://code.claude.com/docs/en/changelog
76•datadrivenangel•35m ago•40 comments

Your 401(k) Is Propping Up the AI Bubble

https://www.promarket.org/2026/05/05/your-401k-is-propping-up-the-ai-bubble/
2•paimapi•35m ago•0 comments

PHP, on a Whim #2: Stop Calling Everything an Array

https://carthage.software/en/blog/article/PHP-on-a-Whim-2-Stop-Calling-Everything-an-Array
1•spacebuffer•36m ago•0 comments

Fat Bear Week 2026

https://explore.org/fat-bear-week
1•slater•38m ago•0 comments

Automattic names interim CFO after exec departures

https://techcrunch.com/2026/09/18/automattic-names-interim-cfo-after-exec-departures/
2•cdrnsf•39m ago•0 comments

What is a System One model and why we need it?

https://stackness.dev/blog/what-is-a-system-one-model-and-where-does-it-go-in-your-stack
1•gosen•40m ago•0 comments

Show HN: ReacherX – Open-source platform to find and reach the right people

https://github.com/VecterAI/reacher-x
1•noobships•40m ago•0 comments

Sam Altman to brief UN Security Council next week

https://www.reuters.com/business/openais-sam-altman-to-brief-un-security-council-next-week-during...
2•vertigoruntime•40m ago•1 comments

Silex: Laser Enrichment Between Promise and Proliferation Risk

https://nuclearnetwork.csis.org/silex-laser-enrichment-between-promise-and-proliferation-risk/
2•looofooo0•44m ago•0 comments