Viewing profile — timmyd
timmyd
HN member- Joined
- Thu, Feb 03, 2011, 3:55 PM UTC
- HN karma
- 626
- Public activity
- 164 items
- HN profile
- View on Hacker News ↗
About timmyd
Recent public activity
-
story
Qualcomm to Acquire Modular
https://investor.qualcomm.com/news-events/press-releases/new... https://www.modular.com/blog/qualcomm-to-acquire-modular https://x.com/clattner_llvm/status/2069769232477192354 , ht…
-
comment
Comment #48545825
[flagged]
- story
- story
-
comment
Comment #46619518
David Robertson took a quantization challenge designed for CUDA experts, and solved it in Mojo with AI assistance, and ended up 1.07x to 1.84x faster than the state-of-the-art C++/…
- story
-
comment
Comment #45329175
Co-founder here. There isn't any signup - that was 2+ years ago and we've been iterating a lot with the community and listening to feedback - which has been wonderful. Go freely an…
-
comment
Comment #45144168
You might also enjoy this series: https://www.modular.com/democratizing-ai-compute which goes into a lot of the details.
- story
- story
- story
- story
-
comment
Comment #44206364
thanks jsnell - i did they and they appreciated the comment above, and unflagged it. i appreciate it!
-
comment
Comment #44206156
FWIW I didnt take the blog as a dunk on CUDA, just as an impressive outcome from the blog writer in Mojo. It's awesome to see this on Hopper - if it makes it go faster thats awesom…
-
comment
Comment #44206045
[op here] To be clear: Yes, there are 3 kernels - you can see those in the linked github at the end of the article if you clicked that. These are: transpose_naive - Basic implement…
- comment
- comment
-
comment
Comment #44204855
Updated the title to the original. I did base the numbers on "This kernel archives 1437.55 GB/s compared to the 1251.76 GB/s we get in CUDA" (14.8%) which is still impressive
- story
- story
- story
- story
- story
- story
- story