Skip to content
#

multi-gpu

Here are 121 public repositories matching this topic...

Efficient and Scalable Physics-Informed Deep Learning and Scientific Machine Learning on top of Tensorflow for multi-worker distributed computing

  • Updated Mar 1, 2022
  • Python

Qwen3.8-Flash-Next on 2× RTX 3090: with 128 GB RAM up to 4,191 tok/s prefill · 111.5 tok/s decode (131K prompt), full 256K window at 2,865 tok/s; with 64 GB RAM 3,410 tok/s prefill · 84 tok/s decode. New: opt-in uncensored mode (runtime abliteration, no new weights).

  • Updated Oct 2, 2026
  • Python
lilbee

The whole local AI stack in one executable: it runs and manages local AI models across every GPU, and it's a search engine you can talk to, with cited answers from your files, code, and the web. MCP server for coding agents, web crawler, TUI, CLI, REST API, Python library. No Ollama or LM Studio needed, works with both.

  • Updated Oct 1, 2026
  • Python

Add this topic to your repo

To associate your repository with the multi-gpu topic, visit your repo's landing page and select "manage topics."

Learn more