frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

"As a Language Model": Chat Template Switches LLM Self-Referential Voice

https://arxiv.org/abs/2609.25021
24•yu3zhou4•57m ago

Comments

ForHackernews•26m ago
In my view, these models should never be trained to output first-person "experiential" (from the abstract) language. It's too easy to humans to anthropomorphize software that presents itself as having an identity.

The AI companies have chosen to package LLMs as friendly chatbots because they know that will be engaging for humans, but it's manipulative. An honest LLM interface would sound like the computer off Star Trek.

actionfromafar•19m ago
Agreed, completely. I would pay for that Star Trek computer interface.
yu3zhou4•18m ago
Same! I believe that you could actually train a LoRA on top of a model to get results close to that
yu3zhou4•19m ago
The strange thing is that the base models (before RLHF) use the "experiential" voice, even though they are not incentivized to do that.
gwerbin•10m ago
It doesn't seem that strange when you consider these things are trained on millions and millions of conversations, both real and fictional.
j-pb•16m ago
Do you want to get turned into a paperclip? Because building intelligence that doesn't understand what it's like to be human gets you turned into a paperclip.

Besides, if you train a model on human communications you get something that behaves like a communicating human, it's not anthropomorphising or manipulative, it's what these models naturally are by construction.

mnsc•11m ago
"naturally"...
LiamPowell•25m ago
> yet what drives them is not well understood

Presumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part.

yu3zhou4•20m ago
Thanks for pointing out, maybe I should be more explicit in the wording - I mean we don't fully know what drives the voice in LLMs. Models that are post trained as instruct models are expected to have the disclaimers, but what about base models (those that are trained on just a lot of text)? How do they talk about themselves? What happens when you strip off the chat template from instruct model's prompt? I hope the rest of the paper makes the questions clearer, but I will try to do better in the abstract next time, as you point out this sentence is kind ambiguous. Thank you!
qsera•5m ago
>How do they talk about themselves?

"You are a Large Language Model" in (system?) prompt would do the trick..

The Poisoned Chalice

https://imgur.com/a/a6gKgjw
1•xkaymac•3m ago•0 comments

A world that remembers (8 bit AI game)

https://www.twitch.tv/genesisplanet
1•pro_methe5•4m ago•1 comments

CUDA Device Max Connections

https://leimao.github.io/blog/CUDA-Device-Max-Connections/
1•eigenBasis•5m ago•0 comments

Anthropic Is Building HAL 9000

https://j.jnord.workers.dev/articles/2026-09-27-anthropic-is-building-hal-9000
1•jtrn•6m ago•1 comments

Netflix Stratum Media Processing: Automated Container Right-Sizing

https://netflixtechblog.com/netflix-stratum-media-processing-automated-container-right-sizing-ddf...
1•violetta98•7m ago•0 comments

The Joys of Microadventures (2024)

https://www.noemamag.com/a-single-small-map-is-enough-for-a-lifetime/
1•Tomte•9m ago•0 comments

Tradition Is Smarter Than You Are (2018)

https://scholars-stage.org/tradition-is-smarter-than-you-are/
1•Tomte•9m ago•0 comments

Show HN: What about a no-SSH agentic access tool

https://getfda.dev/
1•laotree•15m ago•0 comments

'Things will never be chill again': The doomers who shaped AI safety freakout

https://www.msn.com/en-us/money/general/things-will-never-be-chill-again-the-doomers-who-shaped-t...
1•tolugenius•15m ago•0 comments

The World Needs Better Power Grids. Why Smaller May Be Better

https://www.nytimes.com/2026/09/25/business/power-grid-jonas-birgersson.html
1•thelastgallon•19m ago•0 comments

Anthropic's Secretive AI-Powered Wet Lab Breaks Cover and Makes First Discovery

https://www.the-scientist.com/anthropic-s-secretive-ai-powered-wet-lab-breaks-cover-and-makes-fir...
1•speerer•20m ago•0 comments

Show HN: Compliance Posture – free, searchable SaaS compliance directory

https://complianceposture.com/
1•godcrampy•22m ago•0 comments

Thieves Stole 'Nvidia' Trailers. They Got 20 Tons of Sand

https://www.wired.com/story/thieves-stole-nvidia-trailers-they-got-20-tons-of-sand/
1•joozio•23m ago•0 comments

Show HN: Llamora, a private journal that remembers with you

https://welcome.llamora.app/
1•anon1253•30m ago•1 comments

Fighting YouTube Addiction

https://buttondown.com/usaa.ma/archive/fighting-youtube-addiction/
1•usamasulaiman•34m ago•0 comments

From APIs to Languages: Generalising Method Names [using regular expressions]

https://blog.acolyer.org/2015/11/09/from-apis-to-languages-generalising-method-names/
1•ogogmad•39m ago•0 comments

How do you run System One decision models locally?

https://stackness.dev/blog/how-do-you-run-system-one-decision-models-locally-ollaya-laya-mlx-and-...
2•gosen•42m ago•0 comments

DeepSeek has 64% market share

https://twitter.com/gherget/status/2104155774910083275
2•gaborme•44m ago•0 comments

Sponsored-logs-rs: Open the Rust book: monetize the tracing exhaust

https://github.com/sponsoredlogs/sponsored-logs-rs
2•thunderbong•44m ago•0 comments

Show HN: Jev-windows-agent – Windows UI Automation back end for CUA agents

https://github.com/VBS2004/jev-windows-agent
2•VBS2004•47m ago•0 comments

Being Mentioned in AI Answers Is Not the Same as Being Recommended

https://cuescout.com/blog/mentioned-is-not-recommended
3•cuescout•48m ago•0 comments

SRE Udemy Course: Software Reliability Map

https://www.udemy.com/course/reliability-map/
2•Brixdes•53m ago•0 comments

Jempo by Thomas Ha

https://www.lightspeedmagazine.com/fiction/jempo/
2•Alien1Being•54m ago•0 comments

"As a Language Model": Chat Template Switches LLM Self-Referential Voice

https://arxiv.org/abs/2609.25021
24•yu3zhou4•57m ago•10 comments

Changes in quality of life: digital exercise therapy for knee/hip osteoarthritis

https://doi.org/10.1186/s12955-026-02634-5
3•chavasorani•58m ago•0 comments

Postgres AT TIME ZONE 'UTC' does NOT do what you think it does

https://bookofrevenue.com/blog/6ab81e9a97a13f0001f7e4e1/postgres-at-time-zone-u-does-not-do-what-...
1•birdculture•1h ago•0 comments

Trying GPT-6 Sol with AI SDK Evaluation API

https://vercel.com/i/using-gpt-6-sol-with-ai-sdk-evaluation
1•flashbrew•1h ago•0 comments

Show HN: TanStack Start and TanStack AI Jev Integration

https://vercel.com/i/using-jev-in-tanstack-start-with-tanstack-ai
1•flashbrew•1h ago•0 comments

Another week, another data breach for Revolut customers

https://www.theregister.com/cyber-crime/2026/09/25/another-week-another-data-breach-for-revolut-c...
4•imalerba•1h ago•0 comments

Ask HN: What does your AI workflow look like?

1•tadziokas•1h ago•3 comments