Live data from Hacker News

Viewing profile — trsohmers

trsohmers

HN member
Joined
Mon, Mar 14, 2011, 8:58 PM UTC
HN karma
1,536
Public activity
367 items

About trsohmers

My website is trsohmers.com

2013 Thiel Fellow

2023-Present Positron AI (https://positron.ai)

2021-2023 Groq (https://groq.com)

2019-2020 Lambda Labs (http://lambdalabs.com); Managing hardware development and cloud/datacenter deployment

2013-2018 REX Computing(http://rexcomputing.com); Developed new processor architecture from scratch

@trsohmers

Follow me on twitter... @trsohmers

Recent public activity

  1. comment
    Comment #47840475

    Software people, in my very direct experience, are terrible at hardware... While in jest, I do think most software engineer's understanding of hardware abstractions is pretty poor …

  2. comment
    Comment #47524066

    It actually stands for "lizard brain"... it is (or at least was) an Infineon Aurix control and monitoring microcontroller, they may have changed to a newer one.

  3. comment
    Comment #46885982

    Feel free to ask me any questions! Website: https://positron.ai

  4. story
  5. story
  6. comment
    Comment #44024530

    Only with the oscillation overthruster flag enabled.

  7. comment
    Comment #43542871

    I was put on it in 2015 after an acquaintance of mine that was previously on the list recommended me… I only heard from Forbes a few days before the list came out, they asked me fo…

  8. comment
    Comment #42180527

    Based on their S1 filing and public statements, the average cost per WSE system for their (~90% of their total revenue) largest customer is ~$1.36M, and I’ve heard “retail” pricing…

  9. comment
    Comment #41852476

    Do you think that the 16k GPUs get used once and then are thrown away? Llama 405B was trained over 56 days on the 16k GPUs; if I round that up to 60 days and assume the current mai…

  10. comment
    Comment #41397084

    +1 this commenter. I just visited the UK for the first time at the beginning of this month and had a fantastic ~3 hours at Bletchley Park, but felt I had to cram TNMOC and the amaz…

  11. story
  12. comment
    Comment #40978304

    They meant that there is no support for Codestral Mamba for llama.cpp yet.

  13. comment
    Comment #40519531

    We had a basic LLVM backend that supported a slightly modified clang frontend and a basic ABI. We tried to make it drastically easier for both the programmer and compiler to handle…

  14. comment
    Comment #40519031

    Founder of REX Computing here; I highly recommend checking out my interview on the Microarch Club podcast linked elsewhere on the thread; will also answer questions on this thread …

  15. comment
    Comment #40313677

    This is a lesson that like all good Hitchhikers, you should always carry a towel.

  16. comment
    Comment #39682984

    Significantly more than that; MFN pricing for NVIDIA DGX H100 (which has been getting priority supply allocation, so many have been suckered into buying them in order to get fast d…

  17. comment
    Comment #39673698

    The quote from the linked press release is that they do training on TPUv4, while inference is running on GPUs. I have also heard this separately from people associated with Midjour…

  18. comment
    Comment #39553790

    I’m right on the millenial/gen Z divide and an inner selfish purpose for me working on AI/ML is just to enable a creation of Jodorowsky’s 10 hour version of Dune with soundtrack by…

  19. comment
    Comment #39438579

    Long story, but technically REX is still around but has not been able to continue to develop due to lack of funding and my cofounder and I needing to pay bills. We produced initial…

  20. comment
    Comment #39433379

    I thought that was clear through my profile, but yes, Positron AI is focused on providing the best performance per dollar while providing the best quality of service and capabiliti…

  21. comment
    Comment #39432384

    Groq states in this article [0] that they used 576 chips to achieve these results, and continuing with your analysis, you also need to factor in that for each additional user you w…

  22. story
  23. comment
    Comment #39318796

    Yes; Mamba was a very easy match, with Hyena also being a good match, but could be greatly optimized with some minimal changes to the model architecture or hardware design.

  24. comment
    Comment #39311704

    "The current round" of AI accelerators you are referring to are things that were designed 2015-2022; There are a number of startups (including my own) that are actually designing f…

  25. comment
    Comment #38740440

    This article from less than a month ago says that it is on 576 chips https://www.nextplatform.com/2023/11/27/groq-says-it-can-dep...