-
Microsoft
- Bay Area, USA
- https://aayush-ankit.github.io/
Stars
CUDA Templates and Python DSLs for High-Performance Linear Algebra
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
PyTorch library to facilitate development and standardized evaluation of neural network pruning methods.
GPGPU-Sim provides a detailed simulation model of contemporary NVIDIA GPUs running CUDA and/or OpenCL workloads. It includes support for features such as TensorCores and CUDA Dynamic Parallelism as…
Github Pages template based upon HTML and Markdown for personal, portfolio-based websites.
This is originally a collection of papers on neural network accelerators. Now it's more like my selection of research on deep learning and computer architecture.
Models and examples built with TensorFlow