Viewing profile — RuiWang0811
RuiWang0811
HN member- Joined
- Sun, Mar 02, 2025, 8:13 PM UTC
- HN karma
- 2
- Public activity
- 14 items
- HN profile
- View on Hacker News ↗
About RuiWang0811
Recent public activity
-
comment
Comment #49185557
there are loads of env businesses for coding tasks, enterprise tasks, computer use etc etc. we offer different envs from a niche industry, which just so happens to be a very hard d…
-
comment
Comment #49184371
We have experiments showing that agents at least can learn from the environment by overfitting on train. But we do not yet have full post train runs, mainly due to time. But follow…
-
comment
Comment #49184341
We use real historical market data for the environments. There is no parametric modelling involved. The decay property refers to alpha that we give the agent for trade in the env -…
-
comment
Comment #49184308
No we sell our own research to AI labs as RL envs. Realistic RL envs grows in demand as labs seek better data train better models. It’s a complimentary business. Simply put: we sel…
-
comment
Comment #49177101
languagelearner, I think you need to spend more time learning languages
-
comment
Comment #49175527
this seems to be a common misconception, our envs use market data, but the goal is not (only) trading. Market data just happens to be a good source of hard data science tasks. Re t…
-
comment
Comment #49175457
we do affine transformations of the data, so all return/ pnl measures are still the same as with untransformed data. The transformation doesn’t change the conditional distribution …
-
comment
Comment #49175425
not sure about your background, the trace shows the feature engineering the LLMs did
-
comment
Comment #49175418
cofounder here - LLMs can do some model training, they train on ML competition data after all. But they do struggle with low signal to noise ratio of market data. But that’s exactl…
-
comment
Comment #49100367
Does this mean you have to retrain routing rules every time a new model gets released? I imagine since the price/token (or rather the amount of work that can be done per token) doe…
-
comment
Comment #49098732
With all the benchmaxxing happening, current evals are oversaturating and become meaningless for model comparison. I think there is only one test that can't be gamed: let them trad…
- story
-
comment
Comment #49088287
[dead]
- story