frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Show HN: I built an MCP that gives DeepSeek V4 Flash eyes (bad ones)

https://github.com/theJian/squint-mcp
1•theJian•1m ago•0 comments

The White House Is Going to Expand It's AI Policy

https://www.wired.com/story/the-white-house-is-going-to-expand-its-ai-policy/
1•jazzypants•2m ago•0 comments

FeatureScript MCP Server: The Fastest Path to AI-Driven Design

https://www.onshape.com/en/blog/featurescript-mcp-server-enables-text-code-cad
1•hapax_legomena•7m ago•0 comments

Lord of the Rings fan fighting for a refund gets it from Google, sort of

https://www.polygon.com/lotr-lord-of-the-rings-digital-refund-google-denied/
2•HelloUsername•8m ago•0 comments

Say Less (agents are keeping score)

https://theterminalbet.com/p/say-less
1•drupeek•9m ago•0 comments

Claude users are mad that Anthropic's new watermarks will catch them using it

https://techcrunch.com/2026/08/12/some-claude-users-are-mad-that-anthropics-new-watermarks-will-c...
8•ashurandi•14m ago•3 comments

Agento: The missing dashboard for Claude Code

https://github.com/shaharia-lab/agento
1•shahariaa•15m ago•0 comments

The greatest story of promotion incentive hacking at FAANG

https://twitter.com/deedydas/status/2085953995525476597
2•tosh•18m ago•0 comments

A content integrity framework based on ML encoding

https://lyfe.ninja/news/introducing-blkseal-a-practical-trust-layer-for-digital-content/
1•lyfeninja•19m ago•0 comments

Google's AI Doctor Experiment Separates Conversation from Clinical Reasoning

https://ai-updates.net/google-video-doctor-separates-conversation-clinical-reasoning/
1•ashurandi•19m ago•0 comments

Rapid warming may tip AMOC at 2°C, slower warming may avert collapse

https://phys.org/news/2026-08-rapid-atlantic-circulation-2c-slower.html
2•maxboone•20m ago•0 comments

My MacBook can't divide by 20: an Apple Metal compiler bug

https://twitter.com/demircancelebi/status/2087844651847888898
1•demircancelebi•23m ago•0 comments

What we're learning about AI memory (and why most of it doesn't work yet)

https://www.reddit.com/r/staynimble/s/XKEYgCpwln
1•theaniketmaurya•25m ago•0 comments

Sherlock – AI Face Search

https://play.google.com/store/apps/details?id=com.popyakter.sherlock&hl=en
1•pranto12345•25m ago•1 comments

The Odyssey running on a Psion 5mx – as Nolan intended

https://www.reddit.com/r/OldHandhelds/s/AmzMAZTzg7
1•dterm•25m ago•0 comments

Mojo in Mid-2026: From Python-Inspired Experiment to Production Systems Language

https://www.reddit.com/r/AgentContext_dev/comments/1vn5xef/mojo_in_mid2026_from_pythoninspired_ex...
1•javaeeeee•30m ago•0 comments

Operational Guidelines for DNS Transport in Mixed IPv4/IPv6 Environments

https://datatracker.ietf.org/doc/rfc10001/
1•m_montazeri•33m ago•0 comments

FoodGuessr

https://www.foodguessr.com
1•choult•36m ago•1 comments

Potatoes 'boil' in the ground as record heatwave sweeps across swathes of Asia

https://www.theguardian.com/world/2026/aug/13/asia-heatwave-south-korea-potatoes-boil-ground-reco...
3•akbarnama•38m ago•0 comments

Cyberattack on Taiwan Exposes the Execution-Finality Gap

https://zenodo.org/records/21915825
1•sangamdas1982•39m ago•0 comments

NightmareEclipse Publishes New Windows Defender Zero Day

https://cyberplace.social/@GossiTheDog/117082623896479140
2•_tk_•42m ago•0 comments

Romania to shut down nuclear plant due to low Danube level

https://www.lemonde.fr/en/international/article/2026/08/13/romania-to-shut-down-nuclear-plant-due...
2•geox•45m ago•0 comments

Lawsuit fights Ellison-Trump corruption's harm to news outlets

https://freedom.press/issues/lawsuit-fights-ellison-trump-corruptions-harm-to-news-outlets/
3•abdelhousni•47m ago•0 comments

AI 2027 Tracker: Tracking predictions from the AI 2027 scenario against reality

https://ai2027-tracker.com/
2•merksittich•51m ago•0 comments

Gnome Shell Design Dreams

https://blogs.gnome.org/shell-dev/2026/08/11/gnome-shell-design-dreams/
1•birdculture•51m ago•0 comments

Terminate-and-Stay-Resident Program

https://en.wikipedia.org/wiki/Terminate-and-stay-resident_program
2•Bluestein•55m ago•0 comments

Rate of climate change affects stability of the AMOC

https://www.eurekalert.org/news-releases/1139863
2•TechTechTech•57m ago•0 comments

List of Unsolved Problems in Mathematics

https://en.wikipedia.org/wiki/List_of_unsolved_problems_in_mathematics
2•frozenseven•59m ago•0 comments

Affordable AI+human insight helped us close 40yr sporadic group Galois problem

https://www.shaowuzhang.com/files/how-we-found-m23.pdf
3•oliculipolicula•1h ago•0 comments

Show HN: Wala Vibes – scene-based radio for nostalgic Indian music

https://walavibes.wtf
1•yunweiguo•1h ago•0 comments
Open in hackernews

If I own Claude's outputs why can't I train my own model on them?

https://support.claude.com/en/articles/12326764-can-i-use-my-outputs-to-train-an-ai-model
57•DarenWatson•1h ago

Comments

nicbou•29m ago
I was not asked for permission when they trained their model on my output.
stingraycharles•21m ago
Fairly certain you actually did give away quite a few of your permissions when you create an account and/or post things online.

I don’t think Google needs permission to index these comments either.

theshackleford•15m ago
A few, not all, and in fact I gave many of them away under very specific licenses, of which they have paid exactly zero mind too.

If I allow my friends into my home for a visit, does that mean I should shrug when I return home from holiday and find out they’ve helped themselves to its usage without my knowledge or permission to do so?

spiderfarmer•27m ago
I added © 2024 - No Rights for Hypocrites to the footer of my website. Claude scraped all pages at least once every week since it came online.

Now what.

ferrouswheel•27m ago
You can, they are just confused about what effective altruism means.
lelanthran•26m ago
Corollary: If you can't use the output to train, then you don't own them.
j45•12m ago
Yup, either you own it, or you don’t own it and ownership is being redefined as a quasi license.
shikck200•26m ago
They stole all the data, and then dont want you to steal it back. Its basically Robin Hood all over again.
philipallstar•15m ago
Piracy is not theft.
eoverride•11m ago
Piracy is theft.

Copyright infringement isn't piracy. Nor theft.

broccoluvr•25m ago
yikes
hackmack10•25m ago
Honestly, the entire dev community should save question and answers, uploaded them to a shared repo anonymously, then just use that data to distill further models and provide them to the public for free.

Theoretically, this should be legal and ethical, when comparing to Anthropic's own behavior. That said, the reason you can't is Anthropic states in their terms that they don't want you to do this.

Anthropic's entire business model is skating on thin ice.

ratmilk•24m ago
This seems like the standard AI company hypocrisy. Hopefully Claude users disregard this nonsense.
stavros•22m ago
Yeah, no. I was OK with them scraping everything if it means we get AI, but, conversely, they don't get to control what happens to their outputs.

Hell, arguably they should release their weights (or at least the weights of their older models), since they trained them on the concentrated knowledge of humankind.

qsera•18m ago
And they are probably also training them using the current human interaction as well...
f6v•12m ago
> I was OK with them scraping everything if it means we get AI

Except there’s no “we” since their models aren’t open.

maaaaattttt•22m ago
If more and more of the web's content is AI generated, AI companies are bound to train on each other's data.

Or, what if I generate content with Claude/ChatGPT/Gemini, warp it in HTML using an open model, put this on my website conveniently dedicated to "Best practices in prompt and AI answers" for example, then train my own model that only scraps my website?

theyliesoeasily•22m ago
You downloaded millions of books from shadow libraries and trained on them.

I don't care about your policies, go fuck yourself.

ForgotMyUUID•22m ago
> When customers use Claude to generate Outputs that then train competing models, they're essentially using our infrastructure and investment to build direct competitors to our service

We did so, please do not repeat it at home.

krona•21m ago
From my interpretation, in order to get the data you 'own' (it's not theirs to give away since they can't claim the copyright on it), you need to use their services. The agreement the user has with Antropic is for the service, not a restriction on how the data is used.
moomin•21m ago
So can you write GPL content using Claude? MIT? Because if this condition applies to the output, I don't see how it's compatible with FOSS.
stingraycharles•18m ago
Yes, obviously, you own the output. Training on those outputs is a violation of the service agreement, and that’s separate from who owns the outputs.
Hamuko•16m ago
I hope I see the day I start seeing AI companies suing other AI companies for training their models on GitHub repositories that were written using the AI companies' models. If Anthropic says that I can't train using Claude's output, then surely OpenAI can't train on a repository that's 100% vibecoded with Claude.
riffraff•12m ago
I think they care about the reasoning traces and such being used for training, not the effectively final output.
sourdecor•14m ago
But LLMs are being trained on GitHub libraries that leveraged Claude?
j45•13m ago
Transferring actual and complete ownership of the output doesn’t allow any further interpretations of the use of it.

That sentence seems to be in a blurry line between outright “ownership” and licensing.

This type of explanation leans towards the reality being you don’t own the outputs from Claude.

This kind of an explanation is like trying to be half pregnant.

abuanwar072•21m ago
Now that is standard AI company hypocrisy
NicuCalcea•20m ago
Imagine that you bought an axe, but the manufacturer banned you from making more axes with it.
cybice•19m ago
Требую исключить весь мой код из обучения всех ваших моделей
woadwarrior01•19m ago
All this posturing is all for naught, because people can do proper logit distillation into smaller models with open-weight models.
gostsamo•18m ago
There must be a clear difference between terms of the Anthropic service and the legal standing of the ai output. The output is mine and I'll do with it whatever I want. The service is Anthropic's and they can do business with whoever they want. Everything else is hallucination.
TekMol•18m ago
Wouldn't such language and reasoning from Anthropic be an argument that they needed written permission to train their model on data from websites?

Has any individual somewhere around the world tested this in court by now? Sued Anthropic for copyright infringement because Claude can reproduce information that is only available on their website?

It shouldn't be that expensive, right? If you sue them for - say - $10000 then what would the costs of such a court case be?

Personally, I think "learning" is not a copyright violation. But if they themselves make it one, then they should also face the consequences, no?

Arnt•17m ago
But you can give them all to me and I can train my model on them. The restriction on you is not related to your ownership of that days, it's related to the contract you "signed".

Or you can publish them on the web and countless others will do it.

lifthrasiir•16m ago
Apart from the usual hypocrisy,

> Safety controls may be lost – models trained on Claude's Outputs won't have our safety measures, potentially leading to harmful or dangerous AI systems. We also have no visibility into deployment, meaning we cannot monitor how these distilled models are used or prevent misuse.

So was that why your models caused three real-world security incidents?

wolvoleo•15m ago
I don't think they care about the little guy doing a bit of fine-tuning.

What they don't want is DeepSeek training their models with Claude output at scale. That's why they forbid it. Gives them a legal basis to cut off accounts doing that.

Not that it's effective because it's being done anyway.

And yes it's super hypocritical but that's another issue IMO.

esjeon•15m ago
One obvious loophole: upload your full outputs publicly. Anyone can feed those to their models.

Wait, we're already doing it. /s

npn•15m ago
sorry LLM output is considered public domain in my country.
Traster•15m ago
The shear hypocrisy of this is quite staggering. Ok, so we own the outputs, but you get to decide how we use them, and you don't trust us to use them. But the people whose books you stole to build your psychopath machine, they didn't trust you, did they. They didn't really want you to build a machine that can generate thousands of books to compete with their work, but you did do it anyway. Are you going to commit to auditing every input that was used to train your model and attain positive consent for their use for training?

It's also pretty wild to call this standard practice. It's not. I can grab any of the open weights models and train to my hearts content. So it's not standard is it. You'd like it to be standard because you don't want to compete.

And you don't trust us, but it is your company that's been going around telling us how excited you are that your model goes out onto the internet and hacking people.

This continual authoritarian bent from the least trustworthy people in the world is deeply problematic and the only saving grace is their absolute total and complete failure to enforce the restrictions they wish to place on us.

Your ability to build the God machine doesn't magically endow you with the moral authority or judgement to decide how it's used, and the fact that these people believe it does is a great indicator that they aren't to be trusted.

fbrncci•15m ago
Cute :)
____mr____•14m ago
> Why we restrict model training > When Outputs are used to train new models without our oversight, additional risks emerge. Safety controls may be lost - ...

> What you can do with Outputs > You can use Claude's Outputs to train models that don't compete with Anthropic's own models.

Why even include that bullshit at top? It is brazenly obvious to anyone with a brain that Anthropic and other frontier AI Labs have no real way to monetize or recoup their investment unless they strongly guard the usage of their models.

j45•11m ago
This means you can train smaller models than Claude’s but not competing.

I’m not sure how an AI company feels that’s safer for their business. Specialization will always best generalization trying to do the same.

riffraff•10m ago
I had the same thought.

"We need to control the model to protect you from evil robots, except it's fine if the evil robots are not competing with our business" is hilariously hypocritical.

isusmelj•13m ago
Any AI safety experts here? I'm wondering if this claim here really holds: > Anthropic invests significantly in making Claude safe, helpful, and harmless. We conduct rigorous pre-release testing, implement multiple safety layers, and continuously monitor our models' behavior. When Outputs are used to train new models without our oversight, additional risks emerge. Safety controls may be lost – models trained on Claude's Outputs won't have our safety measures, potentially leading to harmful or dangerous AI systems.

From my understanding distillation pretty much copies behaviour. If someone intends to distill from a Model, it can't extract unsafe behaviour, but would learn the same safety mechanisms, no?

bnj•11m ago
I guess this is another angle on why they are interested in watermarking the generated content
Laurel1234•10m ago
> When customers use Claude to generate Outputs that then train competing models, they're essentially using our infrastructure and investment to build direct competitors to our service. Like other software and service providers, we expect that our services won't be used to undermine our product offerings.

I hope Claude wrote this because if a human being typed this we need to kill not just them but everyone that so much as shook their hands.

hackmack10•9m ago
Hmmm... the mods already took this off the front page. What a bunch of bullshit.
xyzal•9m ago
Nah, just pass the outputs to a person not related to Anthropic in any way.

Or -- maybe better -- just fsck the TOS.

neuroticnews25•8m ago
Can I just sell the outputs to someone with no contractual relation with Anthropic, therefore not bound by their TOS?
cladopa•8m ago
I have released some of my projects as Open Source, I also have a company with privative software.

Claude and AI partners have taken all what they could from the Open Source projects without giving credit or respecting the licenses. They have increased the traffic on websites in an absolute disrespectful way increasing the hosting cost in inefficient and ridiculous ways.

They have taken all the important books and not asked permission from the authors.

Fair enough. Fair use.

Of course I would create a competitor software to Claude or any others if I could. Using Claude(and others) of course.

I am not paying you 200 dollars/month for you to tell me that I could not create code that competes with you. If you try to go to court in Europe with this you will lose.

It is just the same fair use you proclaim for taking the data from others.

smallerize•8m ago
The only penalty here is being banned for a ToS violation. Maybe they could sue for fraud or something? That doesn't change who owns the data.