The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface. The fastest local inference engine in the world.
Updated 2026-10-11 21:24:44 +00:00
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Updated 2026-10-11 21:24:40 +00:00
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Updated 2026-10-11 21:18:21 +00:00
🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
Updated 2026-10-11 21:18:11 +00:00
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Updated 2026-10-11 21:18:10 +00:00
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
Updated 2026-10-11 21:11:56 +00:00
A high-throughput and memory-efficient inference and serving engine for LLMs
Updated 2026-10-11 21:09:45 +00:00
SD.Next: All-in-one WebUI for AI generative image and video creation, captioning and processing
Updated 2026-10-11 20:04:30 +00:00
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Updated 2026-10-10 13:11:35 +00:00
Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
Updated 2026-10-07 18:57:22 +00:00
🚀 Accelerate inference and training of 🤗 Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization tools
Updated 2026-10-06 09:24:49 +00:00
Large Language Model Text Generation Inference
Updated 2026-03-21 11:34:22 +00:00