Viewing profile — magic_at_nodai
magic_at_nodai
HN member- Joined
- Mon, Jul 01, 2013, 2:47 PM UTC
- HN karma
- 57
- Public activity
- 57 items
- HN profile
- View on Hacker News ↗
About magic_at_nodai
No profile information was provided.
Recent public activity
-
comment
Comment #47056503
yes lmk how i can help. at the minimum i can get you hw and help with PRs etc. firstname at amd.com to reach me.
-
comment
Comment #43209512
Im running ROCm ok on my 9070XT. You can build it from source today if you have a card. rocminfo: **** Agent 2 **** Name: gfx1201 Uuid: GPU-cea119534ea1127a Marketing Name: AMD Rad…
-
comment
Comment #42783787
ROCm on Radeon should work too and the poll above was to seek feedback on what to cards to support next.
-
comment
Comment #42783713
I will provide this feedback to the docs team to clean up. I found it hard when i was making that Poll :D but I looked harder instead of trying to fix the docs. So thank you for th…
-
comment
Comment #42783700
Is this the repo you are referring to https://github.com/amd/go_amd_smi ? Would having a prebuilt version there help you ?
-
comment
Comment #42783676
yes. We are behind on software support for all consumer cards and would love to support all cards. But are looking for guidance / feedback so we can prioritize.
-
comment
Comment #42783654
I have quad w7900s under my desk that work well for workloads on my desktop that translate well to MI300x. There are some perf gaps with FAv2, and FP8 but otherwise I get a seamles…
-
comment
Comment #42783594
We do care about software and acknowledge the gaps and will work hard to make it better. Please let me know any specific issues that are an issue for you and Im happy to push for i…
-
comment
Comment #42783560
PTX does provide a low level machine abstraction. However you still target some version of hardware ( https://arnon.dk/matching-sm-architectures-arch-and-gencode-... ). However a l…
-
comment
Comment #42783490
hey thats me. Happy to help answer anything here and look forward to your constructive feedback to make AMD software better. We got work to do and look forward to it.
-
comment
Comment #39568272
AMD Artificial Intelligence Group (AIG) | Remote / Global AMD Artificial Intelligence Group (AIG) leads AMD AI strategy and drives AI roadmap across client, edge, and cloud. We bui…
-
comment
Comment #35115288
We have it running as part of SHARK (which is built on IREE). https://github.com/nod-ai/SHARK/tree/main/shark/examples/sha...
-
comment
Comment #34084016
Can you give SHARK a try and let us know on our discord? We can try to help. People have been using it on older AMD GPUs back to Polaris arch.
-
comment
Comment #34083989
Here are a list of potential issues https://github.com/AUTOMATIC1111/stable-diffusion-webui/disc... That said we (Nod.ai team) will add support for xformers soon so you can opt in …
-
comment
Comment #33733286
Try SHARK on your AMD GPUs for SD. Follow the setup here: https://github.com/nod-ai/SHARK/tree/main/shark/examples/sha... . It works with Pytorch -> torch-mlir -> MLIR / IREE -> vu…
-
comment
Comment #30448056
unlikely since the interface from ANE is not public and it may change between hardware versions.
-
comment
Comment #30437768
I updated the blog with the reference. Basically it crashes to compile the model with https://github.com/NodLabs/shark-samples/blob/main/examples/... . The coremltools converter is…
-
comment
Comment #30435865
hear your pain and we really want to make it easy (after we make it work). //part of nod.ai / SHARK team.
-
comment
Comment #30435856
So with Tensorcores you use TF32 which is more like FP19-ish and the marketing makes you think you get 8x the performance. But if you want actual FP32 precision you will need somet…
-
comment
Comment #30435704
Yeah the ANE and AMX on cpu are wrapped behind Accelerate Framework and CoreML. So you will have to use CoreML (which wasn't able to compile the latest TF BERT). ANE is also infere…
-
comment
Comment #30435648
Thanks to: LLVM/MLIR --> For the awesome compiler infrastructure IREE --> For the awesome backend to MLIR SHARK/nod.ai --> For adapting IREE for use on various hardware and fine tu…
-
comment
Comment #30435616
This is not part of regular pytorch install. If you can build torch-mlir and SHARK from src you can use it. So hopefully soon we can make pip installable packages but for now the i…
-
comment
Comment #30435591
Here are the matmul sizes for the MiniLM model used for inference: https://github.com/mmperf/mmperf/blob/main/benchmark_sizes/b... These are the matmul sizes for the BERT training …
-
comment
Comment #29068683
nod.ai | Wherever you want in the US | https://nod.ai Come work on A.I Compilers, Runtimes and ML Systems in an _all_ engineer team. You will be working on the forefront on ML Fram…
-
comment
Comment #28723233
nod.ai | Wherever you want in the US | https://nod.ai Come work on A.I Compilers, Runtimes and ML Systems in an _all_ engineer team. No PHBs. We are looking for A.I Compiler Engine…