A community-supported supercharged document management system: scan, index and archive all your documents
ai
angular
archiving
django
dms
document-management
document-management-system
llm
machine-learning
ocr
optical-character-recognition
pdf
Updated 2026-10-11 21:50:59 +00:00
💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
agents
ai
ai-agents
embeddings
information-retrieval
language-model
large-language-models
llm
nlp
python
rag
retrieval-augmented-generation
search
search-engine
semantic-search
sentence-embeddings
transformers
txtai
vector-database
vector-search
Updated 2026-10-11 21:29:39 +00:00
⚡ Zero-dependency Homebrew tap for high-performance LLMs (CachyLLama, ROCmFPX, llama-ai). Optimized for AMD ROCm 7 (Strix Halo, RDNA3/4), Apple Silicon Metal, & Vulkan RADV.
ai-inference
amd-gpu
apple-silicon
cachy-llama
homebrew
homebrew-tap
inference
llama-cpp
llm
local-ai
metal
quantization
rdna3
rdna4
rocm
rocm-gpu
rocmfpx
strix-halo
strix-point
vulkan-radv
Updated 2026-10-11 20:01:07 +00:00
Efficient Triton Kernels for LLM Training
Updated 2026-10-11 17:26:02 +00:00
How Python does AI. Agents, realtime voice, image generation, embeddings. Every model, every interface, typed end to end.
Updated 2026-10-11 05:08:37 +00:00
Go ahead and axolotl questions
Updated 2026-10-10 20:11:08 +00:00
Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.
agent-framework
agentic-ai
agentic-rag
agents
ai
ai-agents
context-engineering
framework
genai
generative-ai
information-retrieval
large-language-models
llm
mcp
multi-agent
orchestration
python
rag
retrieval-augmented-generation
semantic-search
Updated 2026-10-10 17:28:09 +00:00
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
asr
audio
chinese
emotion-recognition
funasr
mcp-server
multilingual-asr
openai-compatible-api
paraformer
punctuation
pytorch
real-time-asr
speaker-diarization
speech-recognition
speech-to-text
streaming-asr
transcription
vllm
voice-activity-detection
whisper-alternative
Updated 2026-10-10 13:11:35 +00:00
The LLM Evaluation Framework
evaluation-framework
evaluation-metrics
llm-evaluation
llm-evaluation-framework
llm-evaluation-metrics
python
Updated 2026-10-10 00:15:09 +00:00
20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.
Updated 2026-10-07 16:59:36 +00:00
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
ai-inference
deep-learning
generative-ai
inference-platform
llm
llm-inference
llm-serving
llmops
machine-learning
ml-engineering
mlops
model-inference-service
model-serving
multimodal
python
Updated 2026-10-05 17:17:20 +00:00
🚀 Automated nightly builds & multi-backend releases (ROCm 7 TheRock, Vulkan RADV, CUDA, Metal) for CachyLLama & llama-ai on AMD Strix Halo (8060S), Strix Point (890M), Phoenix (780M), and Steam Deck with Homebrew support.
amd-apu
cachyllama
deepseek-v4
hip
homebrew
kv-cache
llama-ai
llama-cpp
llm-inference
nightly-builds
portable-binaries
radeon-8060s
radeon-890m
radv
rocm
rocm-7
steam-deck
strix-halo
strix-point
vulkan
Updated 2026-09-05 19:31:50 +00:00
AetherTable — AI-native virtual tabletop & autonomous Game Master engine. LLM narrative generation decoupled from authoritative deterministic game state.
Updated 2026-08-31 12:15:06 +00:00
⚡ Automated nightly builds & portable ROCm 7 releases of Ember for DeepSeek-V4-Flash on AMD Strix Halo (gfx1151 / Radeon 8060S). Standalone C inference server with bundled ROCm runtime, continuous batching, and Homebrew support.
ai-inference
amd-apu
c-inference
continuous-batching
deepseek-v4
dflash
ember
gfx1151
gguf
hip
homebrew
llm-inference
nightly-builds
portable-binaries
radeon-8060s
rdna3-5
rocm
rocm-7
strix-halo
the-rock
Updated 2026-08-31 03:00:27 +00:00
⚡ Automated nightly builds & portable releases for lucebox (DFlash) inference server on NVIDIA CUDA (sm_75-120) and AMD ROCm 7 (Strix Halo gfx1151, RX 7900 XTX gfx1100, R9700 gfx1201) with zero-dependency $ORIGIN RPATH bundling.
ai-inference
cuda
deepseek-v3
deepseek-v4
dflash
gfx1100
gfx1151
gfx1201
hip
llm-inference
lucebox
nightly-builds
portable-binaries
rdna3
rdna4
rocm
rocm-7
rtx5090
rx7900xtx
strix-halo
Updated 2026-08-31 01:02:00 +00:00
Production-grade evaluation framework & turnkey GitHub Action for benchmarking AI coding harnesses (Claude Code, Gemini CLI, Antigravity, OpenCode, DeepSeek), Model Context Protocol (MCP) servers, LSP diagnostics, and agent plugins across CoderEval & Terminal-Bench.
ablation-study
ai-agents
antigravity
benchmark
claude-code
coder-eval
coding-assistant
deepseek
developer-tools
gemini-cli
github-action
language-server-protocol
llm-evaluation
lsp
mcp
model-context-protocol
opencode
terminal-bench
Updated 2026-08-22 20:33:17 +00:00
Unified LLM infrastructure, build registries, Docker compose stacks, and VRAM estimation tools
Updated 2026-08-13 16:20:21 +00:00
AI-powered LinkedIn and Indeed job application bot with Playwright automation, LLM integration, data anonymization, and Telegram reporting. Applies to ALL job types (not just Easy Apply), generates custom resumes, and provides skill statistics.
Updated 2026-07-01 15:32:17 +00:00