🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Pytorch
Verified open-source repositories and technical documentation formatted for LLM context windows and autonomous AI coding agents (157 repositories indexed).
Stable Diffusion web UI
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Tensors and Dynamic neural networks in Python with strong GPU acceleration
A high-throughput and memory-efficient inference and serving engine for LLMs
Deep Learning for humans
Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
Clone a voice in 5 seconds to generate arbitrary speech in real-time
Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
We write your reusable computer vision tools. 💜
Learn how to develop, deploy and iterate on production-grade ML applications.
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (ViT), MobileNetV4, MobileNet-V3 & V2, RegNet, DPN, CSPNet, Swin Transformer, MaxViT, CoAtNet, ConvNeXt, and more
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
You like pytorch? You like micrograd? You love tinygrad! ❤️
Visualizer for neural network, deep learning and machine learning models
Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
CLI tool for configuring and monitoring Claude Code
State-of-the-art 2D and 3D Face Analysis Project
The fastai deep learning library
Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch
Graph Neural Network Library for PyTorch
🏆 A ranked list of awesome machine learning Python libraries. Updated weekly.
ncnn is a high-performance neural network inference framework optimized for the mobile platform
✅(已完结)超级全面的 深度学习 笔记【土堆 Pytorch】【李沐 动手学深度学习】【吴恩达 深度学习】【大飞 大模型Agent】
🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
ONNX Runtime: cross-platform, high performance ML inferencing and training accelerator
Open standard for machine learning interoperability
FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Machine Learning Engineering Open Book
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama model family and using them on various provider services
Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.
Datasets, Transforms and Models specific to Computer Vision
A tiny scalar-valued autograd engine and a neural net library on top of it with PyTorch-like API
Computer Vision Annotation Tool (CVAT) is a leading platform for building high-quality visual datasets for vision AI. It offers open-source, cloud, and enterprise products, as well as labeling services, for image, video, and 3D annotation with AI-assisted labeling, quality assurance, team collaboration, analytics, and developer APIs.
Burn is a next generation tensor library and Deep Learning Framework that doesn't compromise on flexibility, efficiency and portability.
RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RNN and transformer - great performance, linear time, constant space (no kv-cache), fast training, infinite ctx_len, and free sentence embedding.
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
Suite of tools for deploying and training deep learning models using the JVM. Highlights include model import for keras, tensorflow, and onnx/pytorch, a modular and tiny c++ library for running math code and a java based math library on top of the core c++ library. Also includes samediff: a pytorch/tensorflow like library for running deep learn...
An open source implementation of CLIP.
Open3D: A Modern Library for 3D Data Processing
Nano vLLM