Gpu

Verified open-source repositories and technical documentation formatted for LLM context windows and autonomous AI coding agents (121 repositories indexed).

Code at the speed of thought โ€“ Zed is a high-performance, multiplayer code editor from the creators of Atom and Tree-sitter.

88,524 GitHub

๐Ÿ‘ป Ghostty is a fast, feature-rich, and cross-platform terminal emulator that uses platform-native UI and GPU acceleration.

59,567 GitHub

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

48,415 GitHub

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

42,918 GitHub

jax

Python Doc

Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more

36,154 GitHub

If you live in the terminal, kitty is made for you! Cross-platform, fast, feature-rich, GPU based.

34,382 GitHub

A GPU-accelerated cross-platform terminal emulator and multiplexer written by @wez and implemented in Rust

28,324 GitHub

Babylon.js is a powerful, beautiful, simple, and open game and rendering engine packed into a friendly JavaScript framework.

25,918 GitHub

Run frontier MoE models on hardware you already own โ€” pure C, zero deps, experts streamed from disk. Tiny engine, immense model. ๐Ÿฆ

24,345 GitHub

tfjs

TypeScript Doc

A WebGL accelerated JavaScript library for training and deploying ML models.

19,135 GitHub

ไธญๆ–‡LLaMA&Alpacaๅคง่ฏญ่จ€ๆจกๅž‹+ๆœฌๅœฐCPU/GPU่ฎญ็ปƒ้ƒจ็ฝฒ (Chinese LLaMA & Alpaca LLMs)

18,950 GitHub

A high-performance, zero-overhead, extensible Python compiler with built-in NumPy support

16,824 GitHub

Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability.

15,661 GitHub

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

14,346 GitHub

Suite of tools for deploying and training deep learning models using the JVM. Highlights include model import for keras, tensorflow, and onnx/pytorch, a modular and tiny c++ library for running math code and a java based math library on top of the core c++ library. Also includes samediff: a pytorch/tensorflow like library for running deep learn...

14,240 GitHub

ARIS โš”๏ธ (Auto-Research-In-Sleep) โ€” Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in โ€” works with Claude Code, Codex, OpenClaw, or any LLM agent.

14,148 GitHub

Lightweight Armoury Crate alternative for Asus laptops with nearly the same functionality. Works with ROG Zephyrus, Flow, TUF, Strix, Scar, ProArt, Vivobook, Zenbook, Expertbook, ROG Ally, and many more.

13,709 GitHub

Scalene: a high-performance, high-precision CPU, GPU, and memory profiler for Python with AI-powered optimization proposals

13,456 GitHub

John the Ripper jumbo - advanced offline password cracker, which supports hundreds of hash and cipher types, and runs on many operating systems, CPUs, GPUs, and even some FPGAs

13,403 GitHub

NVIDIAยฎ TensorRTโ„ข is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.

12,971 GitHub

The most powerful local music generation model that outperforms almost all commercial alternatives, supporting Mac, AMD, Intel, and CUDA devices.

11,317 GitHub