FunClip
An open-source video and speech recognition tool with AI editing powered by large models.
- Region
- Domestic
- Pricing
- Free
- Open source
- Yes
- GitHub Stars
- ★ 5.8k
- Source
- GitHub
- Added
- 2026-06-06
- Last verified
- 2026-06-06

Overview
FunClip is an open-source video and speech recognition tool that integrates AI editing features based on large models. It addresses the time-consuming and labor-intensive nature of manual video editing by automatically recognizing speech content in videos, allowing users to freely select text segments or speakers for editing based on the recognition results. Ideal for users who need to quickly and accurately extract specific content from videos.
Key features
- ▪Integrated high-performance Chinese ASR model
- ▪Supports custom keywords to improve recognition accuracy
- ▪Features speaker identification capability
- ▪Offers multi-segment free editing
- ▪Supports processing of English audio files
Use cases
Pros
- +Simple and easy installation
- +High-accuracy speech recognition
- +Flexible editing options
- +Supports multiple languages
Limitations / notes
- -Requires local deployment to run
- -Relies on network API calls
Who it's for
This overview was compiled by AI from public sources and may contain inaccuracies — please refer to the official site.
FAQ
How do I get started with FunClip?
After downloading the code, follow the instructions in the README to install and run the launch.py script.
What languages does FunClip support?
Currently supports Chinese and English; more languages will be added in the future.
Something wrong? Let us know on the About page and we'll fix it.