Viewing profile — trsohmers
trsohmers
HN member- Joined
- Mon, Mar 14, 2011, 8:58 PM UTC
- HN karma
- 1,536
- Public activity
- 367 items
- HN profile
- View on Hacker News ↗
About trsohmers
2013 Thiel Fellow
2023-Present Positron AI (https://positron.ai)
2021-2023 Groq (https://groq.com)
2019-2020 Lambda Labs (http://lambdalabs.com); Managing hardware development and cloud/datacenter deployment
2013-2018 REX Computing(http://rexcomputing.com); Developed new processor architecture from scratch
@trsohmers
Follow me on twitter... @trsohmers
Recent public activity
-
comment
Comment #47840475
Software people, in my very direct experience, are terrible at hardware... While in jest, I do think most software engineer's understanding of hardware abstractions is pretty poor …
-
comment
Comment #47524066
It actually stands for "lizard brain"... it is (or at least was) an Infineon Aurix control and monitoring microcontroller, they may have changed to a newer one.
-
comment
Comment #46885982
Feel free to ask me any questions! Website: https://positron.ai
- story
- story
-
comment
Comment #44024530
Only with the oscillation overthruster flag enabled.
-
comment
Comment #43542871
I was put on it in 2015 after an acquaintance of mine that was previously on the list recommended me… I only heard from Forbes a few days before the list came out, they asked me fo…
-
comment
Comment #42180527
Based on their S1 filing and public statements, the average cost per WSE system for their (~90% of their total revenue) largest customer is ~$1.36M, and I’ve heard “retail” pricing…
-
comment
Comment #41852476
Do you think that the 16k GPUs get used once and then are thrown away? Llama 405B was trained over 56 days on the 16k GPUs; if I round that up to 60 days and assume the current mai…
-
comment
Comment #41397084
+1 this commenter. I just visited the UK for the first time at the beginning of this month and had a fantastic ~3 hours at Bletchley Park, but felt I had to cram TNMOC and the amaz…
- story
-
comment
Comment #40978304
They meant that there is no support for Codestral Mamba for llama.cpp yet.
-
comment
Comment #40519531
We had a basic LLVM backend that supported a slightly modified clang frontend and a basic ABI. We tried to make it drastically easier for both the programmer and compiler to handle…
-
comment
Comment #40519031
Founder of REX Computing here; I highly recommend checking out my interview on the Microarch Club podcast linked elsewhere on the thread; will also answer questions on this thread …
-
comment
Comment #40313677
This is a lesson that like all good Hitchhikers, you should always carry a towel.
-
comment
Comment #39682984
Significantly more than that; MFN pricing for NVIDIA DGX H100 (which has been getting priority supply allocation, so many have been suckered into buying them in order to get fast d…
-
comment
Comment #39673698
The quote from the linked press release is that they do training on TPUv4, while inference is running on GPUs. I have also heard this separately from people associated with Midjour…
-
comment
Comment #39553790
I’m right on the millenial/gen Z divide and an inner selfish purpose for me working on AI/ML is just to enable a creation of Jodorowsky’s 10 hour version of Dune with soundtrack by…
-
comment
Comment #39438579
Long story, but technically REX is still around but has not been able to continue to develop due to lack of funding and my cofounder and I needing to pay bills. We produced initial…
-
comment
Comment #39433379
I thought that was clear through my profile, but yes, Positron AI is focused on providing the best performance per dollar while providing the best quality of service and capabiliti…
-
comment
Comment #39432384
Groq states in this article [0] that they used 576 chips to achieve these results, and continuing with your analysis, you also need to factor in that for each additional user you w…
- story
-
comment
Comment #39318796
Yes; Mamba was a very easy match, with Hyena also being a good match, but could be greatly optimized with some minimal changes to the model architecture or hardware design.
-
comment
Comment #39311704
"The current round" of AI accelerators you are referring to are things that were designed 2015-2022; There are a number of startups (including my own) that are actually designing f…
-
comment
Comment #38740440
This article from less than a month ago says that it is on 576 chips https://www.nextplatform.com/2023/11/27/groq-says-it-can-dep...