| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Forked from makslevental/openhls
PyTorch model to RTL flow for low latency inference
SystemVerilog 6
Forked from Liu-xiandong/How_to_optimize_in_GPU
This is a series of GPU optimization topics. Here we will introduce how to optimize the program on the GPU in detail. I will introduce several basic kernel optimizations, including: elementwise, re…
Cuda 1
Forked from microsoft/antares
Antares: an automatic engine for multi-platform kernel generation and optimization. Supporting CPU, CUDA, ROCm, DirectX12, GraphCore, SYCL for CPU/GPU, OpenCL for AMD/NVIDIA, Android CPU/GPU backends.
Python 1
Forked from Kyrie-Zhao/awesome-real-time-AI
This is a list of awesome edgeAI inference related papers.
Forked from eejlny/gemm_spmm
Hardware accelerator for pruned nertworks
C++ 1
A minimal tensor processing unit (TPU), inspired by Google's TPU V2 and V1
Notebooks and code for Neuromorphic Hardware Workshop at ISFPGA 2024.
Hardware and software implementation of Sparsely-active SNNs
[TCAD'23] AccelTran: A Sparsity-Aware Accelerator for Transformers
Hashed Lookup Table based Matrix Multiplication (halutmatmul) - Stella Nera accelerator
This organization has no public members. You must be a member to see who’s a part of this organization.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |