兜忆轩
我是 Cedric,做 compiler / GPU / backend systems。
这里记录我拆 AI 编译器、Triton/TVM/MLIR、CUDA kernel 和工程工具的过程。
文章分为 lab notes、concept notes 和 essays;lab notes 优先给出可复现代码。
从 CUDA 软件栈测试全貌出发,给出一套面向 Triton/GPU compiler 的 fuzz 测试选型、架构、CI 分层和落地方案。
从 CUDA 软件栈测试全貌出发,给出一套面向 Triton/GPU Compiler 的 fuzz 选型与可落地方案:Hypothesis 生成合法程序,Reference/Differential 作为 oracle,Compute Sanitizer 检查内存与并发,mlir-reduce 自动缩减失败用例。
待补充:本文摘要
TVM TIR 中 ramp 向量表达与向量化地址生成笔记
待补充:本文摘要
待补充:本文摘要
待补充:本文摘要
待补充:本文摘要
待补充:本文摘要
待补充:本文摘要