Live data from Hacker News

Viewing profile — sshumaker

sshumaker

HN member
Joined
Fri, Oct 16, 2009, 8:25 PM UTC
HN karma
1,535
Public activity
352 items

About sshumaker

email ===== work: scott at persona-ai dot ai personal: sshumaker at gmail

Product-focused tech executive with 20+ years experience building consumer and enterprise products. Currently co-founder of a stealth startup building something outrageously ambitious.

Previously VP at Microsoft and at Credit Karma. Ex-Googler and former gamedev.

Recent public activity

  1. comment
    Comment #48441658

    There’s both medication (which takes a while to get dosage right but can be extremely helpful for some) and also techniques you can use to help compensate.

  2. comment
    Comment #48170828

    Are you serious? It’s multiple nodes to run a frontier model (a node is 8x GPUs), and they aren’t running on H100s. You are looking at 32+ GPUs.

  3. comment
    Comment #40700126

    It’s how LLMs work - they are effectively recursive at inference time, after each token is sampled, you feed it back in. You will end up with the same model state (not including no…

  4. comment
    Comment #40383544

    It depends on how large the input prompt (previous context) is. Also, if you can keep cache on GPU with a LRU mechanism, for certain workloads it's very efficient. You can also des…

  5. comment
    Comment #40383500

    They are almost certainly doing this internally for their own chat products. The simple version of this just involves saving off the KV cache in the attention layers, and restore i…

  6. comment
    Comment #40383452

    This is a pretty standard technique if you're running the models yourself. e.g. ChatGPT almost certainly does this. There's even work that is more sophisticated in this domain that…

  7. comment
    Comment #40225256

    Stealth startup | ML Engineer (edge inferencing) | Bay Area or LA (hybrid) | Full-time We're a stealth startup building something unbelievably ambitious in the AI space that blends…

  8. comment
    Comment #40053810

    I had the pleasure of working closely with Jeff Slutter early in my career. Was the first really fantastic engineer I worked with - he’s at Santa Monica Studios (God of War) these …

  9. comment
    Comment #39993200

    Monarchies usually result in long-term thinking? I’m sure we haven’t read the same history books. Most monarchs throughout history have been very self-interested and their decision…

  10. comment
    Comment #39846923

    Hamas gained a ton of support for their cause by the tragedy inflicted in the counterattack. I believe they celebrate the deaths of Palestinian innocents as much as they do the Isr…

  11. comment
    Comment #39655416

    It seems like rather than adding a random amount to each sample (which lets them compute a mean by recreating the same audio and extracting out the differences), Safari could inste…

  12. comment
    Comment #39566229

    Stealth startup | Lead mobile engineer, AI Engineer | Bay Area or LA (hybrid) | Full-time We're a stealth startup building something unbelievably ambitious in the AI space that ble…

  13. comment
    Comment #39502826

    You’re being downvoted because this is a hot take that isn’t supported by evidence. I just tried exactly that with dalle-3 and it worked well. More to the point, it’s pretty clear …

  14. comment
    Comment #38976856

    You can also look at Bertopic which has this functionality as an open source library: https://maartengr.github.io/BERTopic/index.html

  15. comment
    Comment #38976677

    Sometimes you can use a heuristic to estimate K, or use a variant that terminates at some distance threshold. That said, something like hdbscan doesn’t suffer from this problem.

  16. comment
    Comment #38220043

    I have an eightsleep, which cools the bed down dramatically (circulates cold water). Makes a huge difference in my ability up sleep.

  17. comment
  18. comment
    Comment #37626595

    I talked to Pavel about taking a role in my team earlier this year (I ended up leaving Microsoft myself since). He was passionate about making the developer experience excellent - …

  19. comment
    Comment #37357934

    I don’t know about [1]. I asked an example from the paper above to GPT-4: “[If you had to guess] how many thumbs did Lincoln’s maternal grandmother have?” Response: There is no wid…

  20. comment
    Comment #36191697

    If you ask it properly it gets it right. From a pure measurement standpoint, could Jupiter fit in the space between the earth and moon? The average distance from the Earth to the M…

  21. comment
    Comment #35962403

    I’m fairly confident this is untrue. At Microsoft at least, it’s a big deal when there is a privacy issue, even a small localized one on a single product - and creates a small fire…

  22. comment
    Comment #35852116

    I had the same issue, but support kept disconnecting me and I had to start all over again going through the entire “try all of the brain dead stuff like resetting” 3 times, before …

  23. comment
    Comment #35843248

    > They will instead pay a lot for MSFTs cloud service offering, which of course comes with the crucial promise that their data is safe and secured and handled in a way that is comp…

  24. comment
    Comment #35448652

    This is really a lot less necessary with GPT-4. What required careful prompting in 3.5 often you can give it something slapdash in 4 and it can do a great job figuring out intent.

  25. comment
    Comment #35162079

    Just use the Azure hosted solution, which has all of Azure's stronger guarantees around compliance. I'm sure it will update with GPT-4 pricing shortly. https://azure.microsoft.com/…