Complete isolated test system with 150/150 tests passing.
Production-ready after successful beta testing cycle.
See CHANGELOG.md for comprehensive details including:
- All critical issues from 1.1.0-beta3 resolved
- Enhanced test infrastructure with real model validation
- Multi-Python compatibility (3.9-3.13)
Issues Resolved:
• Issue #15: Token limits vs natural stop tokens race condition - FIXED
• Issue #16: Interactive vs server token limit policies - FIXED
Major Improvements:
• Automatic optimal token limits - no configuration needed
• Manual --max-tokens control still available when desired
• Eliminates old hardcoded 500/2000 token restrictions
• Performance gains: Up to 524x improvement for large context models
• Enhanced web client with model capabilities display and better UX
Additional Enhancements:
• Enhanced /v1/models API with context_length field
• Comprehensive test expansion: 114 → 131 tests (131/131 passing)
• Python 3.9-3.13 compatibility verified
Known Issues (Beta Status):
• Server deadlock possible under extreme concurrent model loading stress
• Workaround: Avoid simultaneous heavy model operations
Major milestone: First stable release with official PyPI distribution.
New Features:
- PyPI publication: Now installable via \`pip install mlx-knife\`
- Official CLI-only designation with clear API policy
- Absolute GitHub URLs for PyPI package display (logo + demo)
Documentation Updates:
- All docs updated to v1.0.0 (README, CHANGELOG, TESTING, SECURITY, CLAUDE.md)
- Added PyPI installation instructions to README
- Updated supported versions tables
- Clarified CLI-only usage policy
Release Highlights:
- Transition from 1.0-rc3 to stable 1.0.0
- Production-ready with 104/104 tests passing
- Global accessibility via PyPI distribution
- Comprehensive documentation overhaul
Ready for community adoption and production use.