Files
aha/docs/changelog.md
T
XiaoYang c582c4cd0a docs: update and optimize README and project documentation system
- Update README.md with improved formatting, logo, badges,
  and comprehensive documentation
- Add README.en.md with English translation of the documentation
- Include detailed quick start guide, CLI reference,
  and supported models table
- Add changelog information highlighting recent features
- Add script directory and adjust script file locations
2026-02-06 16:13:50 +08:00

1.8 KiB

Changelog

All notable changes to aha will be documented in this file.

The format is based on Keep a Changelog, and this project adheres to Semantic Versioning.

[0.2.0] - 2026-02-05

Added

  • Qwen3-ASR speech recognition model

[0.1.9] - 2026-01-31

Added

  • CLI list subcommand to show supported models
  • CLI subcommand structure support (cli, serv, download, run)
  • Direct model inference via new run subcommand

Fixed

  • Qwen3VL thinking startswith bug
  • aha run multiple inputs bug

[0.1.8] - 2026-01-17

Added

  • Qwen3 text model support
  • Fun-ASR-Nano-2512 speech recognition model

Fixed

  • ModelScope Fun-ASR-Nano model load error

Changed

  • Updated audio resampling with rubato

[0.1.7] - 2026-01-07

Added

  • GLM-ASR-Nano-2512 speech recognition model
  • Metal (GPU) support for Apple Silicon
  • Dynamic home directory and model download script

[0.1.6] - 2025-12-23

Added

  • RMBG-2.0 background removal model
  • Image and audio API endpoints

Changed

  • Performance optimizations for RMBG2.0 image processing

[0.1.5] - 2025-12-11

Added

  • VoxCPM1.5 voice generation model
  • PaddleOCR-VL text recognition model

[0.1.4] - 2025-12-09

Added

  • PaddleOCR-VL model support
  • FFmpeg feature for multimedia processing

[0.1.3] - 2025-12-03

Added

  • Hunyuan-OCR model support

[0.1.2] - 2025-11-23

Added

  • DeepSeek-OCR model support

[0.1.1] - 2025-11-12

Added

  • Qwen3-VL models (2B, 4B, 8B, 32B)

Fixed

  • Added serde default for tie_word_embeddings in Qwen3VL

[0.1.0] - 2025-10-10

Added

  • Initial release
  • Qwen2.5-VL model support
  • MiniCPM4 model support
  • VoxCPM voice generation model
  • OpenAI-compatible REST API
  • CLI interface for all model types