🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Updated 2026-10-11 21:24:40 +00:00
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Updated 2026-10-11 21:20:22 +00:00
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Updated 2026-10-11 21:18:21 +00:00
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Updated 2026-10-11 21:18:10 +00:00
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
Updated 2026-10-11 21:13:49 +00:00
Faster Whisper transcription with CTranslate2
Updated 2026-10-10 03:39:41 +00:00
Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
Updated 2026-10-07 18:57:22 +00:00
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
Updated 2026-10-07 16:59:36 +00:00
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Updated 2026-10-05 17:17:20 +00:00
Large Language Model Text Generation Inference
Updated 2026-03-21 11:34:22 +00:00