are-we-gfx1100-yet
Popular repositories Loading
-
-
-
composable_kernel
composable_kernel Public archiveForked from ROCm/composable_kernel
Composable Kernel: Performance Portable Programming Model for Machine Learning Tensor Operators
C++ 2
-
AutoGPTQ-rocm
AutoGPTQ-rocm Public archiveForked from AutoGPTQ/AutoGPTQ
An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.
Python 1
-
automatic
automatic Public archiveForked from vladmandic/sdnext
An opinionated SD.Next for Navi 3x
Python 1
-
tensorflow-rocm
tensorflow-rocm Public archiveForked from ROCm/tensorflow-upstream
TensorFlow ROCm port
C++
Repositories
- are-we-gfx1100-yet.github.io Public archive
- composable_kernel Public archive Forked from ROCm/composable_kernel
Composable Kernel: Performance Portable Programming Model for Machine Learning Tensor Operators
- pytorch Public archive Forked from pytorch/pytorch
Tensors and Dynamic neural networks in Python with strong GPU acceleration
- flash-attention Public archive Forked from ROCm/flash-attention
Fast and memory-efficient exact attention
- hipBLASLt Public archive Forked from ROCm/hipBLASLt
hipBLASLt is a library that provides general matrix-matrix operations with a flexible API and extends functionalities beyond a traditional BLAS library
- triton Public archive Forked from ROCm/triton
Development repository for the Triton language and compiler
- AITemplate Public archive Forked from facebookincubator/AITemplate
AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (NVIDIA GPU) and MatrixCore (AMD GPU) inference.
- AutoGPTQ-rocm Public archive Forked from AutoGPTQ/AutoGPTQ
An easy-to-use LLMs quantization package with user-friendly apis, based on GPTQ algorithm.
Top languages
Loading…
Most used topics
Loading…