frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Prompting Claude Opus 5.5

https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/prompting-claude-opus-5-5
55•Michelangelo11•1h ago

Comments

prodigycorp•53m ago
Opus 5.5 is a good model, but I've tried to understand the extreme hype about it on social media about Opus' ability to do 2d work, as we got with Astra doing 3d work. In both releases, the models required extensive access to third party apis to generate assets for it, and a lot of the models work was essentially coordinating everything.

There's so many "x generated this in one shot, this is agi" stuff that gives you the impression that you can vibe operate modern models the same way you operated last year's models. There's so much more to it than that. It requires you to put a faith in the leap in the capability of models, one that would've surely been a waste of time in previous models.

Not sure where i'm going with this other than I think most can relate that it's exhausting keeping up with. I cant imagine what it'd be like parenting a kid that went from toddler to puberty in the span of a year and planning for them to go to college the next year. This industry is moving so fast that it's becoming fact that it's the user that's "holding it wrong" every six months.

XenophileJKO•42m ago
Opus 5.5 does NOT need anything other than some javascript/typescript libraries to make very detailed 2d and 3d visualizations. I've spent a week worth of tokens just feeling out what it can do.

The step function change on Opus 5.5 for visual work shocked me.. and I haven't been surprised like this in a long time with LLMs.

EDIT: When I first saw the "P(DOOM)" video and some of the other animations I was VERY skeptical that Opus 5.5 without a lot of tools could make something like that.. until I tried it for myself. It can.. 100%.

prodigycorp•40m ago
It's very good, yes, but I expected it to produce midjourney type results out of the box. That did not happen. The models are definitely granular stuff now though. They must be training off a ton of digital artist stroke data now.

But it other cases, like the music videos, much of the magic is done by access to elevenlabs and suno apis.

Edit: just saw your edit about the pdoom video. Can you share how you prompted it? Would be helpful to know.

collabs•35m ago
I feel like we have different expectations from these frontier models. I don't use Claude code or any agent that has acts to my local machine. I roll up the code and give it the text file that contains all the code. I asked Claude Opus 5.5 max to make me a 2D terminal based racing game with no assets drawings or audio and it exceeded my expectations. Only one failed unit test and that one too it said the test was faulty rather than the code.

I'm still more worried about the malice and any malicious acts by the people at these frontier labs than the models at the frontier labs.

alansaber•28m ago
"a lot of the models work was essentially coordinating everything." - I don't see anything wrong with that personally. It's still extremely challenging to build a model harness, and having a model-mediated everything is clearly wishful thinking. It's exhausting to keep up with, but also somewhat exciting, all depends on your perspective of course.
prodigycorp•18m ago
ya, was not speaking negatively of the models. Reading what I just said I think my comment would be properly classified as rambling.
mathisfun123•51m ago
This "how to prompt" shit changes like every 3 months. Remember when earlier this year it was critical to tell Claude to keep going because it would just give up. It's amazing this is really considered a product - imagine having to relearn how to drive your car every 3 months.
ulfw•47m ago
All this overhyped crap will explode and most of us will be poor but hey whatever. Some billionaires will be better off. That should make us all content
ddosmax556•44m ago
Not sure how you missed that but models are fundamentally changing in features, scope, intelligence, pricing, communication style. Of course it changes every 3 months. Of course there is no product that lasts more than 3 months. We're in a race right now. It won't stop changing for a while.
mathisfun123•41m ago
> Not sure how you missed that but models are fundamentally changing in features, scope, intelligence, pricing, communication style

Not sure how you missed it but that's exactly what I'm calling out as asinine.

> We're in a race right now

Again: consider the analogy about cars... which are literally used for racing (occasionally).

UrineSqueegee•34m ago
>Not sure how you missed it but that's exactly what I'm calling out as asinine.

i dont understand how you're framing this. how is this a bad thing exactly? How is it asinine?

Bishonen88•49m ago
> Frontend design defaults

> Asked for frontend work without design direction, Claude Opus 5.5 falls back on a few default styles, and a general instruction such as "avoid a generic AI look" mostly swaps one default for another. It responds well to instructions that name specific patterns to avoid, as in the following example. Work iteratively: check which styles the first result used instead, and extend the list if needed.

I hardly ever read tips for prompting etc. because things change too quickly, the writeups are kindof big. Glad I read this one, because I often did exactly what they assume users would do. I write "don't make it look like generic ai slop" and that seemed to work nicely. Now I know why there was still a chance of seeing similar styles across apps. I reckon doing some manual work in terms of scouting dribbble/behance for nice layouts will yield better results.

emsixteen•29m ago
This is relatively useful to know, but I can't help but wonder how people are expected to be able to describe something that they probably have difficulty putting in to words. Maybe it's mostly useful for those who have a design eye, background, or experience.
skeledrew•26m ago
I really find it strange though. How does an AI know what "AI slop" is? Is it reasonable to tell a child not to do "wrong" if you haven't told them what things are wrong?
bronlund•48m ago
First thing it did when I tried it, was roaming through files in directories way outside of the project. I tried to get it to explain why it did it multiple times, but I never got anything resembling an explanation.
prodigycorp•45m ago
What we know after the openai incidents is that the RSI process involves models having access to user rollouts via tool calls. What's considered crappy training one year is another year's kompromat!
dhanushnehru•48m ago
The most useful feature for coding AI is the unattended run.

Just saying "continue" when it gets stuck usually makes it repeat the same error. A better way is to save its last action and result, then make it try a new approach. If it tries the exact same thing twice, it should stop and ask the user for help instead of wasting money on a loop

skeledrew•37m ago
All that keeps jumping out at me is how they've set it to refuse giving users thinking tokens and prompts for full reasoning in output. Just drives me further away; I may not stop using Claude completely for now, but I'll be moving even more of my primary workload to Chinese providers. That's where openness and freedom is now at.
shaan7•34m ago
Yeah its annoying. I need to pay for thinking, but I can't see it :/
OtomotO•31m ago
Ironic, especially given 100 hundred years of Hollywood proaganda telling the west that the US are the center of freedom.

Which was and is true to some extent.

And don't get me wrong, China is a dictatorship, and a tyranny for some.

But then again, the west is a tyranny for some.

muzani•28m ago
That's how the cycles happen. China realizes they could use a little more freedom and US realizes that they could do with a little less. The emerging/shrinking middle class of both countries also moves the sweet spot.
pllbnk•12m ago
Crazy how tables have turned. Life seems surreal since 2020.
rajeevk•5m ago
What Chinese models/providers are you using for this? I'm hitting Claude's weekly limits much sooner than I used to with roughly the same workload, so I'm interested in trying alternatives, especially ones with strong coding/agentic performance.
bob1029•34m ago
> Fourth, if long tool-calling turns still go quiet for longer than you want, have your harness ask for an update

I'm not sure I understand this complexity. In all harnesses I've ever used, tool calls themselves are surfaced to the user as an indication of progress. When the UI/UX around this is engineered well, the user should be able to infer roughly what is going on. Different tools have different ideal presentations. You can't reduce everything to plaintext blobs.

If I absolutely needed intra-turn progress updates, I'd accumulate a separate per-turn transcript and feed it into a cheaper model at deterministic intervals.

pookieinc•32m ago
Idea: Someone should just build a prompt generator that takes whatever the latest "Prompting" techniques are for each model and re-configure it to be as optimal as possible, adding in whatever is needed to get the highest quality result.

I say the above because I'm seeing entire worlds and games being one-shotted built on X and I just have no idea how they do it. I tried building a large prompt for Fable when it was first released and it didn't have anything close to resembling some of the stuff I'm seeing today.

derencius•28m ago
i've built a subagent that does that. i have a hook to force any promot writing to go through this subagent that has reference to all those docs from anthropic, openai and gemini.
user43928•31m ago
What bothers me most with Opus 5.5 is its verbosity.

Claude Code has an output style setting that I set to "Concise", with no apparent effect.

I am told this is merely something in the system prompt that the model tends not to pay attention to with large contexts.

Opus 5.5 writes whole essays at the end of the turn, with the important actionable steps somewhere at the bottom.

When prompted to give a concise summary, it usually overshoots into a super short summary and then you have to dig into the details again anyway.

In general I find Opus 5.5's writing to still have more "ticks" or "Claudisms" than the OpenAI models.

Its explanations often appear overcomplicated for simple concepts.

Sure, it's leagues above the ridiculous writing of Opus 5, but Anthropic still has a long way to go here.

derencius•29m ago
ask claude code to start new claude sessions while refining the output style. until it looks right. my claude code now writes well.
user43928•26m ago
Interesting.

I guess I could take some lengthy example explanation, and have it try various instructions and test what results in output that I find preferable.

Maybe I'll give that a try, thanks!

MrsPeaches•9m ago
This comment reminds me that one of the features of technological revolutions is just the sheer number of people who are on the bleeding edge of use.
wren6991•19m ago
> I am told this is merely something in the system prompt that the model tends not to pay attention to with large contexts.

IIRC it's a system reminder injected after every single turn.

It must be pretty ingrained to be so resilient against prompting. I think RL on relatively short-horizon programming tasks has given the model a tendency to write down absolutely everything, so it survives compaction. Longer-term (project-scale) tasks where this crap starts to pile up and cause problems are in the evolutionary shadow, so to speak.

kimseungyong•23m ago
I can feel that the token consumption has slowed down so that we’re able to cover more in a five-hour session than before. I'm using Korean, but sometimes the words or sentences are hard to read
preommr•20m ago
I am getting increasingly worried that coding is not solved, and that AI won't lead to some kind of coding singularity where we never have to read the code any time soon.

In which case we've royally fucked ourselves that the level of engineering we've reached is... prompts. Because there is a deadline where we have to show productivity to justify all the investment spending.

People need to build with tools in a reliable, constructive way. Not vodoo magic based off vibes. We need better structured output, better transparency on what these models can do, better controls overla, maybe new ideas on loops graphs, and ways to use the models. Like, at least people were trying new things with jev.

KellyCriterion•18m ago
I found out that my initial/system/"base" prompt is now only partially applied, it seems. While it was perfect for Opus 4.6, now the answers are much longer than before - does Anthropic this to sell me more tokens?

I used Opus 5.5 for some simpler tests and was quite angry when I saw that each of my question was above 10USd

simonw•17m ago
That "mark pasted text" thing is interesting: https://platform.claude.com/docs/en/build-with-claude/prompt...

  Summarize the main complaints in this thread.
  
  <pasted_content id="ab12">
  ...text the user pasted...
  </pasted_content id="ab12">
Where those IDs are randomly generated and unknown to the user, and the model is told to use that markup to help avoid it suffering prompt injection attacks.

In the past I've been very skeptical of this kind of protection. Anthropic have clearly trained their models for this though, so maybe Opus 5.5 is smart enough for this to work?

Will be interesting to see if minds more devious than mine can break it.

nialse•3m ago
Got to love the pseudo markup slop! An id attribute on an XML closing tag?!? Complete nonsense. Working nonsens, of course, but still nonsense.
mathisfun123•28m ago
For the third and final time: consider the analogy about cars.
bronlund•43m ago
Yeah. It is not so much like a "coding assistant", but more like a temp agency sending you different autists every other month.

Edit: Someone commented that this is insulting to autists, and I guess it kind of is - sorry. What I ment was an intellectual; one that can be an absolute retard, but have read an aweful lot.

anotha_one•41m ago
This is an insult to autistic people. AI isn't autistic, it's retarded.
kalleboo•26m ago
> imagine having to relearn how to drive your car every 3 months

Cars were just like that during their early years, with tillers and knobs. See the video where Top Gear finds the first car with controls we recognize https://www.youtube.com/watch?v=fkwGJzU5B-I

jmiskovic•9m ago
Did you know you can get certification from Anthropic? And it actually costs real money.
spiclk•2m ago
Remember how we were going to be "left behind" if we didn't "keep up"? I'm so glad I haven't wasted any time or effort learning how to kick each month's flavour of idiot assistant.
hannesv•2m ago
What Chinese provider would you use that is on par with Claude code?
aytigra•1m ago
With accumulated "writing style" memories after 5.0 the new 5.5 seem to be quite great, it is concise enough. But I am bothered by another thing, 5.5 seem to be over-eager and agreeable, when I ask stuff like "why is that like this?" it just goes and applies tons of edits instead of clarifying what I mean or what I want. And similarly it changes stuff and then asks if that is how I wanted to be, ignoring three memories that tell it to ask first.

Prompting Claude Opus 5.5

https://platform.claude.com/docs/en/build-with-claude/prompt-engineering/prompting-claude-opus-5-5
55•Michelangelo11•1h ago•46 comments

Owed a billion dollars in Nvidia stock

https://colo.to/nvidia-stock-narrative.html
661•Eric_Gullichsen•7h ago•276 comments

Thinking fast and slow in AI: The role of metacognition (2021)

https://arxiv.org/abs/2110.01834
91•teleforce•6h ago•26 comments

Ember-1

https://fireworks.ai/blog/ember-1
455•gmays•15h ago•211 comments

When did Google get so weird?

https://sancho.bearblog.dev/google-weird/
1259•sancho-panza•13h ago•708 comments

Malleable software: Restoring user agency in a world of locked-down apps (2025)

https://www.inkandswitch.com/essay/malleable-software/
83•evakhoury•14h ago•35 comments

Nissan's third generation e-POWER powertrain

https://www.nissan-global.com/EN/INNOVATION/TECHNOLOGY/ARCHIVE/E_POWER_GEN3/
52•mroche•6h ago•57 comments

Made by Mechanical Means

https://felixrieseberg.com/made-by-mechanical-means/
13•cocoflunchy•2h ago•2 comments

Footguns with Postgres "at time zone 'UTC'"

https://bookofrevenue.com/blog/6ab81e9a97a13f0001f7e4e1/postgres-at-time-zone-u-does-not-do-what-...
16•birdculture•23h ago•5 comments

Alan Kay's answer to “Did the ENIAC have a BIOS”?

https://www.quora.com/Did-the-ENIAC-have-a-BIOS/answer/Alan-Kay-11
128•midnightfish•13h ago•40 comments

Functional Mechanical Sympathy [video]

https://www.youtube.com/watch?v=jSHmDnUZCm8
23•surprisetalk•2d ago•1 comments

Maybe don't let Muse run your Facebook Marketplace account

https://www.threads.com/@matt.j.robb/post/DdxwAJnDhNy
26•embedding-shape•1h ago•17 comments

Self-Hosting on the Dark Web

https://david.alvarezrosa.com/posts/self-hosting-on-the-dark-web/
186•mooreds•13h ago•69 comments

Lunar Terminator Paradox

https://notes.secretsauce.net/notes/2026/09/27_lunar-terminator-paradox.html
68•dima55•12h ago•48 comments

Was silent reading unusual during Augustine's time?

https://www.historyofinformation.com/detail.php?entryid=4341
59•ryanscio•8h ago•30 comments

Guitar amp and effects pedal built on the Waveshare ESP32-S3-Touch-AMOLED-2.06

https://github.com/dashersw/coyopedal
74•arbayi•2d ago•32 comments

The state of SIMD in Rust in 2026

https://shnatsel.github.io/state-of-simd-rust-2026/
148•verdagon•2d ago•34 comments

Don't couple your Go code to GitHub

https://iain.rocks/blog/dont-couple-your-go-code-to-github
238•birdculture•16h ago•109 comments

Deterministic Concurrency [video]

https://www.youtube.com/watch?v=25x0UuSCKuU
34•surprisetalk•2d ago•3 comments

There is more to code review than (automatable) detection

https://www.adaptivecapacitylabs.com/2026/08/24/there-is-more-to-code-review-than-automatable-det...
137•utiiiD•1d ago•83 comments

Show HN: Lofi Cities – Pixel-art city nights with browser-generated lofi

https://loficities.com/
241•safaelmali•14h ago•112 comments

In an $80 motel room, a discovery to shed light on the origins of life

https://www.nytimes.com/2026/09/26/science/motel-science-discovery.html
244•danso•18h ago•93 comments

What I did at Recurse Center

https://thill.me/2026/09/11/what-i-did-at-rc.html
107•bingden•14h ago•32 comments

Imp is a full port of DSPy to the BEAM

https://github.com/deepfates/imp
76•mpweiher•13h ago•7 comments

Replacing the old battery on rechargeable bike lights

https://jvns.ca/blog/2026/09/27/replacing-the-old-battery-on-rechargeable-bike-lights/
172•surprisetalk•19h ago•92 comments

Reading’s Bayeux Tapestry

https://diamondgeezer.blogspot.com/2026/09/readings-bayeux-tapestry.html
46•zeristor•1d ago•17 comments

Previously unheard recordings of John Coltrane, captured by Frank Tiberi

https://www.jazzwise.com/content/news/john-coltrane-centenary-celebrations-see-impulse-records-re...
89•gregsadetsky•2d ago•33 comments

Oral history of John Chowning, inventor of FM synthesis [video]

https://www.youtube.com/watch?v=e1Xn3030IvM
77•Rochus•15h ago•21 comments

Writing Efficient C++ Code (2013)

https://asawicki.info/articles/writing_efficient_cpp_code.php
154•ibobev•2d ago•118 comments

Packing Binary Is Fun

https://hereticpleb.vercel.app/blog/packing-binary-is-fun-actually
18•BurnerBurner•7h ago•4 comments