Hacker Newsnew | past | comments | ask | show | jobs | submitlogin
Fleet: Hierarchical Task-Based Abstraction for Megakernels on Multi-Die GPUs (arxiv.org)
10 points by matt_d 30 days ago | hide | past | favorite | 1 comment


Nice read! Cool to see more GPU programming models that expose chiptet topology instead of treating it as a flat execution. Reminds me a little of Cerebras' CSL although far from as extreme as that.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: