Skip to content
View antiagainst's full-sized avatar

Organizations

@llvm @ROCm @iree-org @triton-lang @lightseekorg

Block or report antiagainst

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

TokenSpeed is a speed-of-light LLM inference engine.

Python 2,191 310 Updated Oct 6, 2026

Distributed Compiler and Optimized Parallel Kernels

Python 1,557 177 Updated Sep 18, 2026

Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.

Python 6,259 582 Updated Aug 22, 2025
Shell 1,595 14 Updated Aug 15, 2024

Small library for D3D12. Provides assert-like macro for HLSL that crashes the GPU.

C++ 56 5 Updated Aug 21, 2023

A curated list of awesome open source libraries to deploy, monitor, version and scale your machine learning

20,968 2,608 Updated Oct 3, 2026

📋 A list of open LLMs available for commercial use.

12,888 988 Updated Feb 13, 2025

StableLM: Stability AI Language Models

Jupyter Notebook 15,672 1,000 Updated Apr 8, 2024

mperf是一个面向移动/嵌入式平台的算子性能调优工具箱

C++ 199 32 Updated Aug 17, 2023

cuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.

Python 959 300 Updated Oct 6, 2026

An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.

Python 39,551 4,773 Updated May 1, 2026

GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.

C++ 77,384 8,274 Updated May 27, 2025

Locally run an Instruction-Tuned Chat-Style LLM

C 10,109 836 Updated Apr 19, 2023

LLM inference in C/C++

C++ 130,422 24,124 Updated Oct 6, 2026

A collection of modern/faster/saner alternatives to common unix commands.

33,020 827 Updated Sep 10, 2024

High-Performance Rendering Framework on Stream Architectures

C++ 1,051 108 Updated Oct 4, 2026

Making large AI models cheaper, faster and more accessible

Python 41,438 4,492 Updated Oct 5, 2026

MLPerf™ Mobile models

26 10 Updated Apr 30, 2026

A profiler to disclose and quantify hardware features on GPUs.

C++ 175 26 Updated May 15, 2022

📊 An infographics generator with 30+ plugins and 300+ options to display stats about your GitHub account and render them as SVG, Markdown, PDF or JSON!

JavaScript 17,267 2,293 Updated May 29, 2026

一个集高级重启、应用安装自动点击、CPU调频等多项功能于一体的工具箱。

Kotlin 1,534 163 Updated Dec 25, 2023

Customized matrix multiplication kernels

Jupyter Notebook 57 6 Updated Mar 5, 2022

Productive, portable, and performant GPU programming in Python.

C++ 28,403 2,390 Updated Oct 5, 2026

Eureka is a feature-rich and highly customizable Hugo theme.

HTML 955 186 Updated Oct 6, 2026

Sources for Arm Streamline's gator daemon, part of Arm Mobile Studio suite of performance analysis tools

C++ 154 69 Updated Apr 13, 2026

A debian-based shell environment designed for Android and adb

Shell 337 107 Updated Feb 4, 2023

Memory consumption and FLOP count estimates for convnets

MATLAB 929 112 Updated Jan 17, 2019

Library to load a DLL from memory.

C 3,159 825 Updated Jan 3, 2024

TNN: developed by Tencent Youtu Lab and Guangying Lab, a uniform deep learning inference framework for mobile、desktop and server. TNN is distinguished by several outstanding features, including its…

C++ 4,657 771 Updated May 9, 2025

Benchmarks for popular CNN models

Python 2,531 400 Updated Sep 25, 2017
Next