Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
atomicapple
on June 6, 2025
|
parent
|
context
|
favorite
| on:
Highly efficient matrix transpose in Mojo
I think the OP based the title off of "This kernel archives 1437.55 GB/s compared to the 1251.76 GB/s we get in CUDA" (14.8%) and not the final kernels for whatever reason
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: