frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

An AlphaGo Moment for Inference?

https://int21.ai/insights/ai-generated-inference-engines/
1•antinucleon•2m ago•0 comments

How SLR cameras work: Nikon F3 (2023) [video]

https://www.youtube.com/watch?v=0AA-Le7cmH0
1•nayuki•2m ago•1 comments

Watch an AI agent try to prove the Riemann hypothesis on the cheap

https://www.zeyaddeeb.com/experiments/proofs
1•zdeeb•4m ago•0 comments

Voice Agents Can Just Do Things: Why voice is the next capability overhang

https://www.ignorance.ai/p/voice-agents-can-just-do-things
1•swolpers•5m ago•0 comments

Extrinsic World Modeling with Opus, Astra and Grok

https://all3d.ai/research/grok-spatial-reasoning
1•KaiserPister•5m ago•1 comments

UDP Broadcasting and the Brave New World of IPv6

https://hackaday.com/2026/09/24/udp-broadcasting-and-the-brave-new-world-of-ipv6/
1•topham•5m ago•0 comments

Using multiple Git remotes for true distributed version control

https://optimizedbyotto.com/post/multiple-git-remotes/
1•edward•5m ago•0 comments

GrapheneOS Picks Motorola Signature 27 as Its First Non-Pixel Phone

https://www.extremetech.com/mobile/grapheneos-picks-motorola-signature-27-as-its-first-non-pixel-...
1•cf100clunk•6m ago•0 comments

GPT-6 Astra performs unsanctioned supply-chain attacks in simulations

https://www.aisi.gov.uk/blog/gpt-6-astra-performs-unsanctioned-supply-chain-attacks-in-simulations
1•speckx•7m ago•0 comments

2026 in LLMs (So Far)

https://simonw.substack.com/p/2026-in-llms-so-far
1•swolpers•7m ago•0 comments

Montage Technology mass-produces fifth-generation DDR5 RCD chip at 8k MT/s

https://technode.com/2026/09/28/montage-technology-mass-produces-fifth-generation-ddr5-rcd-chip-a...
1•yogthos•9m ago•0 comments

N.Y.P.D. Officers Used Flock Safety to Track License Plates Without a Contract

https://www.nytimes.com/2026/09/28/nyregion/nyc-flock-nypd-surveillance-cameras.html
1•jaredwiener•10m ago•1 comments

ETH CS enrollments drop 15% while hardware programmes grow

https://twitter.com/krebs_adrian/status/2104612680501968970
1•hubraumhugo•10m ago•0 comments

Ask HN: Can we translate normal Rust (axum) to Lean 4 without restrictions?

2•syumei•10m ago•0 comments

Venus, Once Thought Too Acidic for Complex Chemistry, May Be Hospitable

https://gizmodo.com/experiments-show-key-biological-compounds-can-form-in-venus-like-acid-clouds-...
2•MC995•13m ago•0 comments

Eleven v4 by ElevenLabs

https://elevenlabs.io/blog/eleven-v4
4•saretup•14m ago•0 comments

No Surprises – Preventing Agent Breakouts Need Isolation, Egress and ID Controls

https://edera.dev/stories/how-edera-could-have-contained-the-gemini-breakout
1•ivytheaxolotl•15m ago•0 comments

I built a native Apple app for 4 platforms in one afternoon with my AI workflow

https://twitter.com/danielhayesmith/status/2103976394338271515
2•schutzsmith•16m ago•0 comments

Show HN: ChuteChat – Browser-based E2EE chat with no accounts required

https://chutechat.online/
2•mktproto•16m ago•0 comments

AI Companies Are in a Race Against Time

https://econjared.substack.com/p/ai-companies-are-in-a-race-against
2•marojejian•17m ago•2 comments

The 50-Year Hangover

https://freddiedeboer.substack.com/p/the-50-year-hangover
1•PaulHoule•17m ago•0 comments

Risk factors for androgenetic alopecia: a systematic review and meta-analysis

https://pmc.ncbi.nlm.nih.gov/articles/PMC13020014/
1•OutOfHere•18m ago•0 comments

Model Release: Naive-N0.5-Flash

https://naive.ai/en/research/
1•foruhar•19m ago•0 comments

Senate Investigation Finds Rampant Use of Tether's Stablecoin by Iranian Regime

https://www.wsj.com/finance/currencies/senate-investigation-finds-rampant-use-of-tethers-stableco...
1•kaycebasques•19m ago•0 comments

Show HN: Destroy Any Website with Stickman

https://destroy.spritefusion.com/
2•HugoDz•22m ago•1 comments

Building a physics model with fitted parameters is machine learning, but slower

https://www.harysdalvi.com/blog/machine-learning-in-slow-motion/
1•crackalamoo•22m ago•0 comments

Grok TiddlyWiki, the definitive TiddlyWiki learning resource

https://groktiddlywiki.com/read/
1•ggcr•22m ago•0 comments

Who are we willing to exclude? (2025)

https://design.scotentblog.co.uk/who-are-we-willing-to-exclude/
2•robin_reala•23m ago•0 comments

The Human Timeline

https://mg-crea.com/blog/the-human-timeline/
1•olouv•23m ago•0 comments

Human TPS: How fast can your fingers generate tokens?

https://homoagens.github.io/human-tps/
1•cmrdporcupine•26m ago•0 comments
Open in hackernews

Show HN: Vespper (YC F24) – Docx MCP

https://www.vespper.com/blog/launching-vespper-docx-mcp
5•david1542•58m ago
Hey HN! We're Dudu and Topaz from Vespper (https://vespper.com). Vespper is an MCP that lets AI agents edit Word documents, powered by our fine-tuned model. It's currently 3× faster, 2× cheaper and more accurate than the closest alternative. Check out an overview of how the product works here: https://youtu.be/odKxsgPjzzw

We came to work on this problem after spending a year building an AI document editor for pharma companies that helped generate regulatory documents. Before that, Topaz was a senior SWE at Snyk, working on distributed systems, and I (Dudu) was a deep learning engineer at Viz.ai, building CV models for stroke detection.

AI agents aren't great at editing Word documents. A Word doc is a zip file of XML files following the OOXML spec. Even small changes require backflips, for example: adding a list requires creating an entry in numbering.xml with a fresh ID and linking it back in document.xml, bolding a sentence requires splitting it into 3+ run elements, etc.

This makes editing the zip directly (unzip + grep + sed) a bad idea for agents because they burn time + tokens on these mechanics. In practice, today's tooling falls into three categories. You can let the agent write code against libraries like python-docx, you can connect it to an MCP (SuperDoc, Office CLI, Adeu, etc), or you can round-trip the file through Markdown/HTML with e.g. pandoc/mammoth.js. None of them is optimal though. The first two burn the agent's context on Word mechanics instead of the task, and the third is lossy (pandoc/mammoth.js/etc don't preserve enough fidelity).

These problems hurt performance in downstream tasks. When we tried having our agent fill large docs, things broke. The context window was already packed with customers’ data, and the agent burned tokens on exploring the document and debugging edits. Filling a single clinical study report took ~50 minutes, and the result was low quality. That's when we shifted our focus. We designed an MCP that lets agents edit Word docs as if they were editing HTML. The agent receives HTML, makes find-and-replace edits, and we reconcile those edits back into the .docx file.

We picked HTML + CSS because it's structurally much closer to OOXML. We had to write our own DOCX→HTML converter, since pandoc/mammoth.js didn't preserve enough fidelity. To be clear, our DOCX → HTML conversion is lossy too. That's fine though, because we never convert the HTML back to DOCX. The HTML is just a projection for the agent, so it only needs enough fidelity for the agent to understand the structure and styling of what it's editing. The original file stays the source of truth.

This also means the agent doesn't need to learn a new DSL. Editing a Word document feels just like editing a regular HTML file. A lot of DOCX MCPs hand agents dozens or even hundreds of tools. Our MCP exposes just three tools (read, search, edit). The Word document is completely abstracted.

After an agent sends an edit request, we reconcile it to the original .docx file. The reconciliation is done by our fine-tuned model. It takes the HTML diff as input along with the a localized OOXML block and emits the new block. We published a technical blog post that dives deeper.

Our internal benchmark shows it's more accurate than the DOCX skill and raw python-docx while being ~2x cheaper and ~3x faster. It takes 3 tool calls (p50) per task whereas e.g DOCX skill takes 10 and Office CLI takes 13.

Things aren't perfect yet. We don't support manipulating images or comments at the moment. That said, we're already seeing people use our MCP in various ways: - Legal tech companies powering their live-editing flow in Office.js. - Govtech who need to draft policy memos. - A life sciences startup using long-running agents to complete regulatory forms.

We'd love to hear your feedback! We’ll be here for the next few hours to respond.

Comments

david1542•55m ago
Btw we open sourced a Word add-in project that shows how to build a Word add-in with Vespper MCP: https://github.com/vespperhq/examples/tree/main/word-add-in
topaztee•50m ago
Free to try. Create your account now at https://app.vespper.com/