Viewing profile — david_shi
david_shi
HN member- Joined
- Sun, Jun 09, 2019, 10:13 PM UTC
- HN karma
- 284
- Public activity
- 202 items
- HN profile
- View on Hacker News ↗
About david_shi
you can probably guess my email
Recent public activity
- story
- story
- story
- story
- story
-
comment
Comment #48821645
Is this meant to be read in order?
-
comment
Comment #48782486
Is this the story of Johnny Rotten?
- story
-
comment
Comment #48669114
Interesting that there's not a single mention of cannabis, perhaps it's more of a musician's choice.
-
comment
Comment #48665283
> I believe eval startups can work when they're targeting safety benchmarks specifically. Are there any examples of successful startups doing this?
-
comment
Comment #48638802
Yeah. I'm realizing that the models are strong at drafting the overall shape of the writing but the specific phrases are grating once you've seen it hundreds of time in slop.
-
comment
Comment #48638790
Even with the examples, I've found that explicitly pointing out what not to do is moderately helpful if the model is given some time to self-evaluate. I wish this was something tha…
-
comment
Comment #48638781
Ah this is very helpful. I've been pointing out things that the model does, labeling it, and then adding them into a skill. The models (Opus, etc) are very good at labeling the pat…
-
story
Ask HN: How do you make AI writing usable?
Every time I ask the latest models to write something, it defaults to LLMisms like contrastive negation and long unnecessary lists. Is there a way to desloppify AI writing, or is t…
-
comment
Comment #48627573
This is a charitable read, but I think that being able to pick from a panoply of models will actually yield much better results in the long run. The same model that has been post-t…
-
comment
Comment #48627231
> GLM-5.2 cost a fraction as much. Opus finished in half the time and shipped a cleaner game. Off topic, but does anyone else instantly pick up on LLMisms like this? It seems like …
-
comment
Comment #48625592
It's similar to this: https://openrouter.ai/blog/announcements/fusion-beats-fronti... Basically, if you combine a bunch of near-frontier models (like GPT 5.5, etc) you can get perf…
-
comment
Comment #48625576
Their research around building a domain specific model is pretty cool, it's kind of like Karpathy's autoresearch but pointed at deciding the optimal model to use at each step of th…
-
comment
Comment #48625297
Whoa, say more about Fable sabotaging your codebase?
-
comment
Comment #48624345
These models don't seem very competitive, who's their target audience?
-
comment
Comment #48621641
Super helpful, thanks for sending.
- story
-
comment
Comment #48615165
Speaking from personal experience, I already know exactly what I’m getting with containers. Same with Postgres.
-
comment
Comment #48613802
The price can move after the IPO too
-
comment
Comment #48606355
The economics of working at a pre-IPO company that will likely have a successful IPO and a 20+ year post-IPO company are also very different.