-
-
triton Public
Forked from triton-lang/tritonDevelopment repository for the Triton language and compiler
MLIR MIT License UpdatedSep 30, 2026 -
FlagTree Public
Forked from flagos-ai/FlagTreeFlagTree is a unified compiler supporting multiple AI chip backends for custom Deep Learning operations, which is forked from triton-lang/triton.
Python MIT License UpdatedSep 29, 2026 -
LightX2V Public
Forked from ModelTC/LightX2VLightweight Image Video Action Generation Inference Framework
Python Apache License 2.0 UpdatedSep 29, 2026 -
tilelang Public
Forked from tile-ai/tilelangDomain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernels
Python Other UpdatedSep 25, 2026 -
FlagGems Public
Forked from flagos-ai/FlagGemsFlagGems is an operator library for large language models implemented in the Triton Language.
Python Apache License 2.0 UpdatedSep 23, 2026 -
-
-
FlagGems-sglang Public
Forked from flagos-ai/FlagGems-sglangFlagGems-sglang is part of FlagOS. It is a high-performance operator library designed for multiple hardware backends.
Python Apache License 2.0 UpdatedSep 21, 2026 -
FlagGems-vllm Public
Forked from flagos-ai/FlagGems-vllmPython Apache License 2.0 UpdatedSep 21, 2026 -
KernelGen Public
Forked from flagos-ai/KernelGenNext-Generation AI-Assisted Kernel Engineering for Multi-Chip Systems
Python Apache License 2.0 UpdatedSep 20, 2026 -
FlagAttention Public
Forked from flagos-ai/FlagAttentionA collection of memory efficient attention operators implemented in the Triton language.
Python Other UpdatedSep 20, 2026 -
SageAttention-AMD Public
Forked from zihaomu/SageAttention-AMDNative C++/HIP attention kernels, shape specializations, and validated performance practices across AMD GPUs.
C++ Apache License 2.0 UpdatedSep 18, 2026 -
KernelGenBench Public
Forked from flagos-ai/KernelGenBenchPython Apache License 2.0 UpdatedSep 17, 2026 -
ai_infra_notes Public
this repo is used to collect the attention schemes on rocm GPU
Python UpdatedAug 27, 2026 -
rocm-libraries Public
Forked from ROCm/rocm-librariessuper repo for rocm libraries
Assembly UpdatedAug 18, 2026 -
FlyDSL Public
Forked from ROCm/FlyDSLFlyDSL is the Python front‑end of the project: a Flexible Layout Python DSL for expressing tiling, partitioning, data movement, and kernel structure at a high level.
Python Other UpdatedAug 13, 2026 -
penguin-harness Public
Forked from Prism-Shadow/penguin-harnessAn Efficient Self-Improving Agent Harness That Learns from Experience
TypeScript Apache License 2.0 UpdatedAug 4, 2026 -
GEAK Public
Forked from AMD-AGI/GEAKGenerating Efficient AI-Centric Kernels
Python Apache License 2.0 UpdatedAug 4, 2026 -
Hyperloom Public
Forked from AMD-AGI/HyperloomAn agentic system that auto-optimizes LLM workloads on AMD GPUs.
Python Other UpdatedAug 4, 2026 -
-
-
opencv_contrib Public
Forked from opencv/opencv_contribRepository for OpenCV's extra modules
C++ Apache License 2.0 UpdatedJul 23, 2026 -
llama.cpp Public
Forked from ggml-org/llama.cppLLM inference in C/C++
C++ MIT License UpdatedJul 23, 2026 -
opencv Public
Forked from opencv/opencvOpen Source Computer Vision Library
C++ Apache License 2.0 UpdatedJul 22, 2026 -
TorchEasyRec Public
Forked from alibaba/TorchEasyRecAn easy-to-use framework for large scale recommendation algorithms.
Python Apache License 2.0 UpdatedJul 15, 2026 -
rocm-recsys-examples Public
Forked from NVIDIA/recsys-examplesExamples for Recommenders - easy to train and deploy on accelerated infrastructure.
Python Other UpdatedJul 14, 2026 -
openvino_contrib Public
Forked from openvinotoolkit/openvino_contribRepository for OpenVINO's extra modules
C++ Apache License 2.0 UpdatedJul 1, 2026 -
vla.cpp Public
Forked from VinRobotics/vla.cppA unified inference runtime for VLA models.
C++ Apache License 2.0 UpdatedJun 29, 2026 -


