frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

From OpenAPI spec to MCP: How we built Xata's MCP server

https://xata.io/blog/built-xata-mcp-server
45•tudorg•1y ago

Comments

_pdp_•1y ago
I mean there are 2 other posts related to data exfiltration attacks against MCP severs on the main page of HN at the time of this comment - at this point I think you want to involve a security person to make sure it is not vulnerable to stupid things.
Atotalnoob•1y ago
The MCP attacks are really just due to bad token scoping.

If you allow Y to do X, if an attacker takes control of Y, of course they can do X.

wild_egg•1y ago
Can you elaborate on "bad token scoping"?

I don't think your XY phrasing fully describes the GitHub MCP exploit and curious if you think that's somehow a "token scoping" issue.

fkyoureadthedoc•1y ago
I'm unaware of the GitHub MCP "exploit", but given the overall state of LLM/MCP security FUD, there's probably some self promotion blog post from a security company about an LLM doing something stupid with GitHub data that the owner of the LLM using system didn't intend.

For example, let's say I create an application that lets you chat with my open source repo. I set up my LLM with a GitHub tool. I don't want to think about oauth and getting a token from the end user, so I give it a PAT that I generated from my account. I'm even more lazy so I just used a PAT I already had laying around, and it unfortunately had read/write access to SSH keys. The user can add their ssh key to my account and do malicious things.

Oh no, MCP is super vulnerable, please buy my LLM security product.

If you give the LLM a tool, and you give the LLM input from a user, the user has access to that tool. That shrimple.

wild_egg•1y ago
https://news.ycombinator.com/item?id=44097390

Also currently on the front page. It's mainly that this tool hits the trifecta of having privileged access, untrusted inputs, and ability to exfiltrate. Most tools only do 1-2 of those so attacks need to be more sophisticated to coordinate that.

rexer•1y ago
I think this downplays the security issue. It's true that scoping the token correctly would prevent this exploit, but it's not a reasonable solution under the assumptions that are taken by the designers of MCP. LLM+MCP is intended to be ultra flexible, and requiring a new (differently scoped) token for each input is not flexible.

Perhaps you could have an allow/deny popup whenever the LLM wanted to interact with a service. But I think the end state there is presenting the user a bunch of metadata about the operation, which the user then needs to reason about. I don't know that's much better; those OAuth prompts are generally click throughs for users.

truemotive•1y ago
GitLab Duo got hit with an oopsie, "AI agent runs with same privilege to site content as the authenticated user" kinda oopsie where you could just exfiltrate private repo information via a pixel gif.

I knew it would get bad, but this bad already? I yearn for rigor haha

alooPotato•1y ago
i really dont get why we cant just feed the openapi spec to the LLM instead of having this intermediate MCP representation. Don't really buy the whole 'the api docs will overwhelm an LLM" - that hasn't been my experience.
wild_egg•1y ago
I haven't looked at MCP payloads properly to compare but often the raw OpenAPI spec is overly verbose and eats context space pretty quick.

Really trivial to have the LLM first filter it down to the sections it cares about and then condense those sections though.

Wrap that process in a small tool and give that to the LLM along with a `fetch` tool that handles credentials based on URLs and agent capabilities explode pretty rapidly.

crystal_revenge•1y ago
I see this question frequently related to MCP, but I'm guessing these questions come from people who haven't built a lot of products using LLMs?

Even if you're LLM could learn the openai spec, you still have to figure out how to concretely receive a response back. This is necessary for virtually any application build using an LLM and requires support for far, far more use cases than just calling an API.

Consider the following use case: - You need to include some relevant contextual data from a local RAG system. - There are local functions that you want the model to be able to call - The API example you describe - You need to access data from a database

In all of these cases, if you have experience working with LLMs, you've implemented some ad hoc template solution to pass the context into the model. You might have writing something like "Here is the info relevant to this task {{info}}" or "These are the tools you can use {{tools}}", but in each case you've had to craft a prompting solution specific to one problem.

MCP solves this by making a generic interface to sending a wide range of information to the model to make use of. While the hype can be a bit much, it's a pretty good (minus the lack of foresight around security) and obvious solution to this current problem in AI Engineering.

lmeyerov•1y ago
Slightly different experience here

We have been adding MCP remote server to louie.ai, think a semantic layer over DBs for automating investigations, analytics, and viz over operational systems. MCP is nice so people can now use from Slack, VS Code, CLI, etc, without us building every single integration when they want to use it outside of our AI notebooks. And same starting point of openAPI spec, and even better, fastapi standard web framework for the REST layer.

Using frameworks has been good. However, for chat ergonomics, we find we are defining custom tools, as talking directly to REST APIs is better than nothing, but that doesn't mean it's good. The tool layer isn't that fancy, but getting the ergonomics right matters, at least in our experience. Most of our time has been on security and ergonomics. (And for fun, we had an experiment of vibe coding this while hitting enterprise-level quality goals.)

ENGNR•1y ago
Agreed, I’ve only implemented one endpoint, but even on that the amount of data coming back was too high, and the json shape ate up context

I think MCP responses will be high level, aggregated, sorted, etc. Also strongly considering YAML over JSON

matt-attack•1y ago
Why? Does the a sense of quotes and commas really make a difference in context size?
jedisct1•1y ago
If you got an OpenAPI spec and want to expose it as MCP, https://jedisct1.github.io/openapi-mcp/ is an easy way to do it.
otabdeveloper4•1y ago
Just ask the model to respond with JSON. Give it a template example response.

You don't need a spec.

For sending prompts to the LLM you will absolutely need to hand-craft custom prompts anyways, as each model responds slightly different.

wild_egg•1y ago
> you still have to figure out how to concretely receive a response back

Isn't that handled by whatever Tool API you're using? There's usually a `function_call_output` or `tool_result` message type. I haven't had a need for a separate protocol just to send responses.

truemotive•1y ago
If you're working from OpenAPI, ideally you want to be able to process any, potentially full of shit formatting spec file. I find that half the integrations I run into have some old weird version of Swagger, and the rest work like hell to stay up to date with the 3.x spec track.

I agree, I wish, it will be a solved problem eventually. Just feeding a complex data model like that to the paper shredder that is the LLM, for making decisions about whether DELETE or POST is used is just asking for trouble.

Whistle: Speech to Text in 16.9 MB

https://cactuscompute.com/blog/whistle
249•gmays•3h ago•68 comments

The value of not getting to the point (2015)

https://ken.arneson.name/2015/11/the-value-of-not-getting-to-the-point/
26•NaOH•54m ago•4 comments

Show HN: K10s – A Clickable Kubernetes TUI (Go, Bubble Tea)

https://github.com/p10node/k10s
43•pierreneter•1h ago•23 comments

Vitalik Buterin backs crypto ‘bunker mode’ amid rapid AI math advances

https://cointelegraph.com/news/justin-drake-urges-crypto-bunker-mode-as-ai-could-break-wallet-sec...
14•firstcomm•41m ago•1 comments

Why isn't the industry freaking out about DeepSeek 4.1 Flash?

https://www.dgt.is/blog/2026-10-07-deepseek-freek-out/
37•jonotime•19h ago•20 comments

Trump administration is suspending Microsoft from a green card program

https://apnews.com/article/h1b-visa-program-vance-microsoft-e7b3a407f822702b269ee277d21343ea
672•alephnerd•4h ago•1133 comments

Step 5 Preview, a 1M-context MoE from StepFun, shows up on OpenRouter

https://openrouter.ai/stepfun/step-5-preview
53•AnneWodell•3h ago•18 comments

Making a flexible "neon" t-shirt with LED filaments

http://scottbezek.blogspot.com/2026/10/making-flexible-neon-t-shirt-with-leds.html
12•scottbez1•3h ago•4 comments

Beauty in DVD Menus

https://vale.rocks/posts/dvd-menus
204•speckx•6h ago•128 comments

OpenAI annualised revenues $20B less than previously signalled

https://www.cnbc.com/2026/10/08/open-ai-revenue-nvidia-oracle-coreweave.html
225•mfiguiere•3h ago•133 comments

“Math 2.0” will need to value mathematical progress more holistically

https://mathstodon.xyz/@tao/117395269325940185
574•ent101•14h ago•596 comments

US man given prison sentence for bot-farming music streams

https://thequietus.com/news/us-man-given-prison-sentence-for-bot-farming-music-streams/
11•cdrnsf•18h ago•6 comments

Archaeologists Are Reconstructing the 'Invisible' Technologies of the Stone Age

https://www.smithsonianmag.com/science-nature/archaeologists-are-reconstructing-the-invisible-tec...
73•Hooke•22h ago•36 comments

OLED burn-in test: 30-month update

https://www.techspot.com/article/3178-oled-burn-in-test/
39•baal80spam•9h ago•18 comments

DuckDB Ducklake

https://github.com/duckdb/ducklake
55•saikatsg•1d ago•3 comments

Man discovers his parents' coffee machine used 1TB of data in 10 days

https://www.dexerto.com/entertainment/man-discovers-his-parents-coffee-machine-used-1tb-of-data-i...
104•ck2•1d ago•51 comments

I hired an illustrator to draw my house. Now it's my Home Assistant dashboard

https://antonfrolov.substack.com/p/i-hired-an-illustrator-to-draw-my
94•soheilpro•1d ago•4 comments

A 5.3M-year-old deep-sea whale necropolis in the Diamantina Zone

https://www.nature.com/articles/s41586-026-10546-z
22•bryanrasmussen•1d ago•0 comments

I gave Opus 5.5 one prompt and six hours to visualize Invisible Cities

https://quesma.com/blog/invisible-cities-one-shot/
299•stared•7h ago•156 comments

Theranos.World

https://www.theranos.world/
7•kbyatnal•2h ago•0 comments

License update: AI derivation prohibited on all my art, lore, stories, comics

https://www.davidrevoy.com/article1178/license-update-ai-derivation-prohibited-on-all-my-art-lore...
3•frizlab•15m ago•0 comments

Show HN: I Put an AI Agent on a Nokia 110

https://github.com/anupray95/AI-Agent-on-a-NOKIA
15•anupray•5h ago•3 comments

New CRAM method offers giant boost to compressed memory reads

https://www.tomshardware.com/software/linux/new-linux-tech-compresses-memory-in-ram-as-ram-for-45...
32•danny00•6h ago•15 comments

The Deeply Impersonal Personalized Recruiter Mail

https://blog.pentlander.com/the-deeply-impersonal-personalized-recruiter-mail/
27•speckx•2h ago•24 comments

Ask HN: What do you run on a $5 VPS that's worth keeping online 24/7?

36•mariocesar•1d ago•40 comments

4-hour battery storage is cheaper to install than gas turbines all across globe

https://www.solarpowerworldonline.com/2026/10/4-hour-battery-storage-is-cheaper-to-install-than-g...
191•01-_-•3h ago•99 comments

Vanillin provides a sweet solution for chronic wound healing

https://news.flinders.edu.au/blog/2026/10/06/vanillin-provides-a-sweet-solution-for-chronic-wound...
12•geox•8h ago•1 comments

Orkut.com

https://orkut.com/
94•andreynering•5h ago•69 comments

What ArtCraft's Vibe-Coded Apps Say About Adobe

https://tedium.co/2026/10/08/artcraft-vibe-coding-creative-cloud-remake/
23•speckx•1h ago•29 comments

The Slow Formation of Durable Software

https://newsletter.dancohen.org/archive/the-slow-formation-of-durable-software/
220•benbreen•2d ago•94 comments