FunASR Releases
144 releases of modelscope/FunASR
- v0.5
Ultravox release achieving 60% improvement in transcription accuracy, expanding language support to 42 languages, and introducing Big Bench Audio evaluations.
Feb 11, 2025
- v2.0.0
A major update that completely removes PixiJS and Pixi Filters, replacing them with a custom 2D context rendering implementation to improve performance and stability.
Feb 1, 2025
- v1.5.1
Bug fix release for gpt-crawler addressing cookie handling issues to ensure correct cookie setting.
Jan 23, 2025
- 0.4.0
This release introduces a beta evaluation tool for the Chat Engine, adds support for new LLM and embedding providers like Gitee AI and Amazon Bedrock, and includes improvements for file handling and user interface.
Jan 3, 2025
- v11.15.2024
Adds new features for job application automation including regex-based blacklisting, Perplexity integration, and application tracking, along with documentation updates.
Nov 15, 2024
- v0.4.1
Ultravox release upgrading the Whisper encoder to Whisper-large-v3-turbo and adding support for 6 new languages (Chinese, Dutch, Hindi, Swedish, Turkish, Ukrainian).
Nov 12, 2024
- v1.1.1
Enhanced track management by allowing clips to be inserted at specific indices within stacked tracks without disrupting the order of existing clips.
Oct 29, 2024
- v1.1.0
Implemented functionality to efficiently retrieve and display waveforms for clips on the timeline to provide better visual feedback.
Oct 27, 2024
- v1.0.1
Fixed a logic issue where currently visible clips were incorrectly removed from the scene graph when their parent track was disabled.
Oct 23, 2024
- v1.0.0
Initial release introducing core video rendering capabilities with support for handling large videos and many cuts through on-demand demuxing.
Oct 20, 2024
- v0.4
Ultravox v0.4 upgrades the Whisper encoder from small to medium, expands multilingual support to nine languages, and significantly improves BLEU scores for speech translation.
Aug 27, 2024
- v0.3
Ultravox v0.3 introduces the model built on a frozen Llama 3.1 8B core, utilizing 2.5k hours of speech data, synthetic data augmentation, and Knowledge Distillation loss calculation.
Aug 23, 2024
- 0.2.0
This release provides links to its release notes, deployment guides, and upgrade instructions, though specific feature details are not included in the provided text.
Aug 22, 2024
- 0.1.0
This is the initial release, providing links to the release notes and deployment documentation.
Aug 12, 2024
- v1.5.0
The v1.5.0 release adds functionality to limit the git clone depth when running the crawler within a Docker container.
Jul 5, 2024
- v1.4.0
This update focuses on developer experience improvements by adding documentation for the server API and fixing linting issues.
Jan 15, 2024
- v1.3.0
v1.3.0 introduces a configuration option to exclude specific links from crawling based on defined patterns.
Jan 6, 2024
- v1.2.1
Jan 4, 2024
- ext2_image
May 16, 2023
- v0.3.00.3.0
Major release adding GPU runtime (nv-triton), CPU quantization, C++ gRPC services, and streaming inference capabilities. Multiple new domain-specific models including financial, audio-visual, and speaker diarization models were added.
Mar 16, 2023