Viewing profile — sshumaker
sshumaker
HN member- Joined
- Fri, Oct 16, 2009, 8:25 PM UTC
- HN karma
- 1,535
- Public activity
- 352 items
- HN profile
- View on Hacker News ↗
About sshumaker
Product-focused tech executive with 20+ years experience building consumer and enterprise products. Currently co-founder of a stealth startup building something outrageously ambitious.
Previously VP at Microsoft and at Credit Karma. Ex-Googler and former gamedev.
Recent public activity
-
comment
Comment #48441658
There’s both medication (which takes a while to get dosage right but can be extremely helpful for some) and also techniques you can use to help compensate.
-
comment
Comment #48170828
Are you serious? It’s multiple nodes to run a frontier model (a node is 8x GPUs), and they aren’t running on H100s. You are looking at 32+ GPUs.
-
comment
Comment #40700126
It’s how LLMs work - they are effectively recursive at inference time, after each token is sampled, you feed it back in. You will end up with the same model state (not including no…
-
comment
Comment #40383544
It depends on how large the input prompt (previous context) is. Also, if you can keep cache on GPU with a LRU mechanism, for certain workloads it's very efficient. You can also des…
-
comment
Comment #40383500
They are almost certainly doing this internally for their own chat products. The simple version of this just involves saving off the KV cache in the attention layers, and restore i…
-
comment
Comment #40383452
This is a pretty standard technique if you're running the models yourself. e.g. ChatGPT almost certainly does this. There's even work that is more sophisticated in this domain that…
-
comment
Comment #40225256
Stealth startup | ML Engineer (edge inferencing) | Bay Area or LA (hybrid) | Full-time We're a stealth startup building something unbelievably ambitious in the AI space that blends…
-
comment
Comment #40053810
I had the pleasure of working closely with Jeff Slutter early in my career. Was the first really fantastic engineer I worked with - he’s at Santa Monica Studios (God of War) these …
-
comment
Comment #39993200
Monarchies usually result in long-term thinking? I’m sure we haven’t read the same history books. Most monarchs throughout history have been very self-interested and their decision…
-
comment
Comment #39846923
Hamas gained a ton of support for their cause by the tragedy inflicted in the counterattack. I believe they celebrate the deaths of Palestinian innocents as much as they do the Isr…
-
comment
Comment #39655416
It seems like rather than adding a random amount to each sample (which lets them compute a mean by recreating the same audio and extracting out the differences), Safari could inste…
-
comment
Comment #39566229
Stealth startup | Lead mobile engineer, AI Engineer | Bay Area or LA (hybrid) | Full-time We're a stealth startup building something unbelievably ambitious in the AI space that ble…
-
comment
Comment #39502826
You’re being downvoted because this is a hot take that isn’t supported by evidence. I just tried exactly that with dalle-3 and it worked well. More to the point, it’s pretty clear …
-
comment
Comment #38976856
You can also look at Bertopic which has this functionality as an open source library: https://maartengr.github.io/BERTopic/index.html
-
comment
Comment #38976677
Sometimes you can use a heuristic to estimate K, or use a variant that terminates at some distance threshold. That said, something like hdbscan doesn’t suffer from this problem.
-
comment
Comment #38220043
I have an eightsleep, which cools the bed down dramatically (circulates cold water). Makes a huge difference in my ability up sleep.
- comment
-
comment
Comment #37626595
I talked to Pavel about taking a role in my team earlier this year (I ended up leaving Microsoft myself since). He was passionate about making the developer experience excellent - …
-
comment
Comment #37357934
I don’t know about [1]. I asked an example from the paper above to GPT-4: “[If you had to guess] how many thumbs did Lincoln’s maternal grandmother have?” Response: There is no wid…
-
comment
Comment #36191697
If you ask it properly it gets it right. From a pure measurement standpoint, could Jupiter fit in the space between the earth and moon? The average distance from the Earth to the M…
-
comment
Comment #35962403
I’m fairly confident this is untrue. At Microsoft at least, it’s a big deal when there is a privacy issue, even a small localized one on a single product - and creates a small fire…
-
comment
Comment #35852116
I had the same issue, but support kept disconnecting me and I had to start all over again going through the entire “try all of the brain dead stuff like resetting” 3 times, before …
-
comment
Comment #35843248
> They will instead pay a lot for MSFTs cloud service offering, which of course comes with the crucial promise that their data is safe and secured and handled in a way that is comp…
-
comment
Comment #35448652
This is really a lot less necessary with GPT-4. What required careful prompting in 3.5 often you can give it something slapdash in 4 and it can do a great job figuring out intent.
-
comment
Comment #35162079
Just use the Azure hosted solution, which has all of Azure's stronger guarantees around compliance. I'm sure it will update with GPT-4 pricing shortly. https://azure.microsoft.com/…