Viewing profile — tshadley
tshadley
HN member- Joined
- Mon, Nov 28, 2016, 9:41 PM UTC
- HN karma
- 204
- Public activity
- 98 items
- HN profile
- View on Hacker News ↗
About tshadley
No profile information was provided.
Recent public activity
-
comment
Comment #46891608
https://en.wikipedia.org/wiki/Liquid_droplet_radiator
-
comment
Comment #46890954
> LLMs cannot offer that promise by design, so it remains your job to find and fix any deviations from the abstraction you intended. LLMs are clumsy interns now, very leaky. But we…
-
comment
Comment #44638719
As IMO medalists they would be expected to I'm sure. But this can be verified because the results are public: https://github.com/aw31/openai-imo-2025-proofs/
-
comment
Comment #44638533
Yes, OpenAI: https://x.com/alexwei_/status/1946477754372985146 > 6/N In our evaluation, the model solved 5 of the 6 problems on the 2025 IMO. For each problem, three former IMO med…
-
comment
Comment #44388191
Thank you, amazing, fresh.
-
comment
Comment #44174312
The goal here is not to replace transformers but combine them with RNN so you get both good short-term memory (self-attention) and much improved long-term memory (ATLAS recurrent m…
-
comment
Comment #44097913
100% agreed with your experience, AI provides little value to one's area of expertise (10+ years or more). It's the context length -- AI needs comparable training or inference-time…
-
comment
Comment #42911555
> To understand the capabilities of LLMs, we evaluate GPT3 (text-davinci-003) [11], ChatGPT (GPT-3.5-turbo) [57] and GPT4 (gpt-4) Oh dear, this is embarrassing. Anil Anathaswamy, a…
-
comment
Comment #42647410
Well the Franks study probably destroyed any chance for natural sleep conditions. Nedergaard is scathing: https://www.thetransmitter.org/glymphatic-system/new-method-... > The new …
-
comment
Comment #42489034
Seems to me o3 prices would be what the consumer pays, not what OpenAI pays. That would mean o3 could be more efficient in-house than paying subject-matter experts.
-
comment
Comment #42480973
I always get the feeling he's subconsciously inserting a "magical" step here with reference to "synthesis"-- invoking a kind of subtle dualism where human intelligence is just diff…
-
comment
Comment #42332704
> One random example to illustrate the distinction: training gaps can easily decrease uncertainty. You have lots of mammals in your training data, and none of them lay eggs. You as…
-
comment
Comment #42328394
> ...proving that this one particular piece of the hallucination problem may be conceptually simple. Everything mentioned in the article boils down to that one particular piece-- n…
-
comment
Comment #42323375
The article referenced the Oxford semantic entropy study but failed to clarify that the issue greatly simplifies LLM hallucination (making most of the article outdated). When we ar…
-
comment
Comment #41489333
"Why PCIe Risers suck and the importance of using SAS Device Adapters, Redrivers, and Retimers for error-free PCIe connections." I'm a believer! Can't wait to hear more about this.…
-
comment
Comment #40662749
Sure looks like a typo. Contact author? https://x.com/fchollet https://x.com/arcprize https://x.com/mikeknoop
-
comment
Comment #38069413
https://mathshistory.st-andrews.ac.uk/HistTopics/Bakhshali_m... has some examples. |One person possesses seven asava horses, another nine haya horses, and another ten camels. Each …
-
comment
Comment #37726303
So this is old news?
-
comment
Comment #36334747
All cynicism aside, there's vastly more in the collective writings of humans on empathy than medicine.
-
comment
Comment #35816603
From the article: "April 3, 2023 - Real Humans Can’t Tell the Difference Between a 13B Open Model and ChatGPT Berkeley launches Koala, a dialogue model trained entirely using freel…
-
comment
Comment #35808649
Ah, that's it; polite fictions are scored higher than uncomfortable facts.
-
comment
Comment #35808573
That's weird. Having the community study this would certainly help them. They're afraid this is giving too much insight into their proprietary training/modeling methods?
-
comment
Comment #35808455
That should be okay though, 10 good answers will still report the score of the best one chosen. I think the GPTs are using beam search which is projecting out a "beam" (looks more …
-
comment
Comment #35797298
> A probable guess will lower loss much better than "I don't know" or whatever equivalent. Guessing only reduces loss as much as the dataset allows -- a bad guess will give a highe…
-
comment
Comment #35763965
Earth's crust: not quite the same as Cu/Zn but way more than I expected: https://periodictable.com/Properties/A/CrustAbundance.an.htm... Lithium: 0.0017% Copper: 0.0068% Zinc: 0.00…