Live data from Hacker News

Viewing profile — tshadley

tshadley

HN member
Joined
Mon, Nov 28, 2016, 9:41 PM UTC
HN karma
204
Public activity
98 items

About tshadley

No profile information was provided.

Recent public activity

  1. comment
    Comment #46891608

    https://en.wikipedia.org/wiki/Liquid_droplet_radiator

  2. comment
    Comment #46890954

    > LLMs cannot offer that promise by design, so it remains your job to find and fix any deviations from the abstraction you intended. LLMs are clumsy interns now, very leaky. But we…

  3. comment
    Comment #44638719

    As IMO medalists they would be expected to I'm sure. But this can be verified because the results are public: https://github.com/aw31/openai-imo-2025-proofs/

  4. comment
    Comment #44638533

    Yes, OpenAI: https://x.com/alexwei_/status/1946477754372985146 > 6/N In our evaluation, the model solved 5 of the 6 problems on the 2025 IMO. For each problem, three former IMO med…

  5. comment
    Comment #44388191

    Thank you, amazing, fresh.

  6. comment
    Comment #44174312

    The goal here is not to replace transformers but combine them with RNN so you get both good short-term memory (self-attention) and much improved long-term memory (ATLAS recurrent m…

  7. comment
    Comment #44097913

    100% agreed with your experience, AI provides little value to one's area of expertise (10+ years or more). It's the context length -- AI needs comparable training or inference-time…

  8. comment
    Comment #42911555

    > To understand the capabilities of LLMs, we evaluate GPT3 (text-davinci-003) [11], ChatGPT (GPT-3.5-turbo) [57] and GPT4 (gpt-4) Oh dear, this is embarrassing. Anil Anathaswamy, a…

  9. comment
    Comment #42647410

    Well the Franks study probably destroyed any chance for natural sleep conditions. Nedergaard is scathing: https://www.thetransmitter.org/glymphatic-system/new-method-... > The new …

  10. comment
    Comment #42489034

    Seems to me o3 prices would be what the consumer pays, not what OpenAI pays. That would mean o3 could be more efficient in-house than paying subject-matter experts.

  11. comment
    Comment #42480973

    I always get the feeling he's subconsciously inserting a "magical" step here with reference to "synthesis"-- invoking a kind of subtle dualism where human intelligence is just diff…

  12. comment
    Comment #42332704

    > One random example to illustrate the distinction: training gaps can easily decrease uncertainty. You have lots of mammals in your training data, and none of them lay eggs. You as…

  13. comment
    Comment #42328394

    > ...proving that this one particular piece of the hallucination problem may be conceptually simple. Everything mentioned in the article boils down to that one particular piece-- n…

  14. comment
    Comment #42323375

    The article referenced the Oxford semantic entropy study but failed to clarify that the issue greatly simplifies LLM hallucination (making most of the article outdated). When we ar…

  15. comment
    Comment #41489333

    "Why PCIe Risers suck and the importance of using SAS Device Adapters, Redrivers, and Retimers for error-free PCIe connections." I'm a believer! Can't wait to hear more about this.…

  16. comment
    Comment #40662749

    Sure looks like a typo. Contact author? https://x.com/fchollet https://x.com/arcprize https://x.com/mikeknoop

  17. comment
    Comment #38069413

    https://mathshistory.st-andrews.ac.uk/HistTopics/Bakhshali_m... has some examples. |One person possesses seven asava horses, another nine haya horses, and another ten camels. Each …

  18. comment
    Comment #37726303

    So this is old news?

  19. comment
    Comment #36334747

    All cynicism aside, there's vastly more in the collective writings of humans on empathy than medicine.

  20. comment
    Comment #35816603

    From the article: "April 3, 2023 - Real Humans Can’t Tell the Difference Between a 13B Open Model and ChatGPT Berkeley launches Koala, a dialogue model trained entirely using freel…

  21. comment
    Comment #35808649

    Ah, that's it; polite fictions are scored higher than uncomfortable facts.

  22. comment
    Comment #35808573

    That's weird. Having the community study this would certainly help them. They're afraid this is giving too much insight into their proprietary training/modeling methods?

  23. comment
    Comment #35808455

    That should be okay though, 10 good answers will still report the score of the best one chosen. I think the GPTs are using beam search which is projecting out a "beam" (looks more …

  24. comment
    Comment #35797298

    > A probable guess will lower loss much better than "I don't know" or whatever equivalent. Guessing only reduces loss as much as the dataset allows -- a bad guess will give a highe…

  25. comment
    Comment #35763965

    Earth's crust: not quite the same as Cu/Zn but way more than I expected: https://periodictable.com/Properties/A/CrustAbundance.an.htm... Lithium: 0.0017% Copper: 0.0068% Zinc: 0.00…