frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Launch HN: EdotEnv (YC S26) – Quant Trading RL Envs to Teach LLMs Research

https://edotenv.com/
11•Mzzzzz•1h ago
We are Rui and Michael and we’re building EdotEnv (https://edotenv.com): self-improving RL environments from Quant Trading workflows.

With all the benchmaxxing around, evals saturate and become meaningless for model comparison. Useful benchmarks should increase in difficulty as models advance. Back in our Quant jobs, Michael and I saw that the market has exactly this property: markets became more efficient as people profited from trading inefficiencies, making new profitable strategies harder to find and old ones decay over time.

This makes markets an ideal, continuously evolving benchmark for LLM training. The hard part is to turn professional quant workflows into reliable training envs, as this is a very niche expertise.

In our environments, we give LLMs a quant trading workflow and evaluate their performance on out-of-sample data: build predictive features/ models, design a portfolio, backtest strategies, adapt continuously to market regimes. Each step is a task with different self-built tools. For example, a predictive feature building task gives the agent cleaned market data of time period [0,T] to research ideas, a backtesting tool to test created features at time t on [0, t], an execution tool to trade strategies with the new features on [t+1, T] and a final evaluation. Our reward isolates the agent's feature building skills and yet benefits from market properties.

From running SOTA models in our environments, we see that i) they seem to struggle with iterating deeply on research ideas, preferring broad shallow searches; ii) higher reasoning does not seem to increase performance and iii) agents do not understand trading, e.g. when losing money they stop trading instead of trading smarter. Check out our blogs for more details! https://edotenv.com/?tab=blog

Quant workflows are essentially applied ML research, long-horizon planning and continual learning. Through our envs, we teach these transferable research skills, rather than task specific answers. Our environments are closer to a realistic research workflow: we use real-world data instead of synthetic ones; our envs naturally contain noise and real trade-offs; our rewards are verifiable and immediate, with no need for an additional LLM judge or human expert.

We open sourced a sample task repository: https://github.com/MMcollab-dotcom/feature-engineering. We plan to sell continuously improving envs to AI labs/researchers/enterprises training their own agents, who are interested in ML modelling capabilities, continual learning, long horizon planning or Quant Research in general.

We'd love feedback from anyone trying out their own agents in our envs, for either eval or post training. And of course, we are always happy to discuss the future of trading with LLMs (and no, it should not be asking the LLM to read tea leaves and give you the stock to buy tomorrow). Looking forward to your comments!

Comments

Mzzzzz•44m ago
Here is an example rollout trace with gpt 5.6 luna. Checkout if you are interested what kind alphas the agent found XD https://hub.harborframework.com/jobs/af0299f9-a3bb-44ea-8ced...
ak_111•39m ago
if the data is not synthetic, how do you ensure that the LLM hasn't learnt about this data for example from training on the Financial Times.
Mzzzzz•24m ago
We do a 2 step anonymisation: 1. Mask all symbols, timestamps etc. So the agents cannot infer the assets/time periods. 2. Mathematically transform numerical values and returns. E.g. the market return targets are not the raw market returns, but neutralised and manipulated. So even the agents have certain bullish/bearish biases, it cannot make use of it, as we use the transformed values.

In addition, we did not observe such behaviour in our traces. An example: https://hub.harborframework.com/jobs/af0299f9-a3bb-44ea-8ced...

ak_111•2m ago
ah i thought so, interesting. I think the challenge is to do 2 while still keeping it realistic, which actually gets very close to synthetic data generation.

Show HN: Simple algorithm and color space to generate diverse skin tones

https://toneyalexander.github.io/inclusive-color-space/
310•automatoney•4h ago•69 comments

Waymo – Dallas Open to All

https://waymo.com/blog/shorts/dallas-open-to-all/
37•xnx•1h ago•23 comments

Investors in Situational Awareness deserved to lose their shirts

https://www.economist.com/finance-and-economics/2026/08/04/investors-in-situational-awareness-des...
25•Anon84•27m ago•18 comments

Why some people mow a lawn better than others

https://pudding.cool/2026/06/mow/
77•carlos-menezes•1h ago•63 comments

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

https://mistral.ai/news/shieldstral/
61•riadsila•3h ago•16 comments

Hop.earth – OpenStreetMap based car racing game

https://hop.earth/?server=lkhr7&route=fQ5nuu9R
36•faebi•1h ago•22 comments

The Warp Agent CLI

https://www.warp.dev/blog/introducing-the-warp-agent-cli-coding-agent
54•emschwartz•2h ago•27 comments

DeepSeek V4 Flash on a Single AMD MI300X

https://github.com/ryanzhou/deepseek-v4-flash-mi300x
320•zhoutong•9h ago•76 comments

Launch HN: EdotEnv (YC S26) – Quant Trading RL Envs to Teach LLMs Research

https://edotenv.com/
11•Mzzzzz•1h ago•4 comments

U.S. used 'virtually all' of its long-range precision missiles during Iran war

https://www.cnbc.com/2026/08/04/us-has-used-virtually-all-of-its-long-range-precision-missiles-re...
208•tcp_handshaker•8h ago•299 comments

Truemetrics (YC S23) Is Hiring in Berlin – GTM Lead

https://www.ycombinator.com/companies/truemetrics/jobs/bIQQ7tP-founding-gtm-lead
1•truemetricsIngo•2h ago

When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation

https://arxiv.org/abs/2602.16763
42•doppp•3h ago•53 comments

Stephen Wolfram's Wife Has Died

https://writings.stephenwolfram.com/2026/08/in-memory-of-my-wife-elise-cawley-1961-2026-with-than...
69•jdcampolargo•54m ago•3 comments

Vlt 1.0 and Hosted Package Registries

https://www.vlt.io/blog/1-0
36•patrikcsak•2h ago•7 comments

Web security is too hard

https://textslashplain.com/2026/08/04/security-is-hard-yall/
102•kevincox•1h ago•36 comments

Keyv and friends compromised in active Shai-Hulud supply chain attack

https://www.aikido.dev/blog/keyv-and-friends-compromised-in-npm-supply-chain-attack
195•cimi_•8h ago•98 comments

Apple says more ex-employees may have taken confidential data to OpenAI

https://techcrunch.com/2026/08/04/apple-says-more-ex-employees-may-have-taken-confidential-data-t...
223•thewebguyd•4h ago•167 comments

Xbox goes down. You can't play games you own on disc

https://birchtree.me/blog/xbox-goes-down-you-cant-play-games-you-own-on-disc/
466•surprisetalk•7h ago•521 comments

The Red Strings Club

https://www.vegard.net/the-red-strings-club-review/
5•meysamazad•1d ago•0 comments

Perspec 1.0

https://adriansieber.com/announcing-perspec-1-0/
29•surprisetalk•4h ago•8 comments

Blackmail Fail (2013)

https://gwern.net/blackmail
10•simonebrunozzi•1h ago•0 comments

Dates That Don't Exist (2015)

https://blog.yossarian.net/2015/06/09/Dates-That-Dont-Exist
89•EndXA•3d ago•60 comments

Online ad giant Adform was hacked, proving once again why ad blockers are needed

https://this.weekinsecurity.com/online-advertising-giant-adform-was-hacked-proving-once-again-why...
168•speckx•4h ago•59 comments

Harness engineering for self-improvement

https://lilianweng.github.io/posts/2026-07-04-harness/
261•tosh•13h ago•56 comments

Everything I Know (1975)

https://www.bfi.org/about-fuller/everything-i-know/
107•simonebrunozzi•8h ago•31 comments

There Will Come Soft Rains (1950) [pdf]

https://users.wpi.edu/~zrbutzke/Docs/BradburyStories(1).pdf
301•pmg101•20h ago•335 comments

What's Behind the Sharp Drop in Labor Force Participation?

https://www.stlouisfed.org/on-the-economy/2026/aug/what-is-behind-sharp-drop-labor-force-particip...
14•toomuchtodo•2h ago•15 comments

Where .env Went Wrong

https://secretspec.dev/blog/where-env-went-wrong/
79•domenkozar•4d ago•58 comments

Nobel Disease

https://en.wikipedia.org/wiki/Nobel_disease
72•num42•8h ago•71 comments

The Pedagogy Behind the Studio

https://gail.wharton.upenn.edu/gen-ai-studio/the-generative-ai-studio-pedagogy/
7•ray__•5d ago•1 comments