提交工具
首页/工具库/AI音频工具

AI音频工具

语音合成、声音克隆、音乐生成、音频处理
80 个工具
按子类筛选:

ElevenLabs

拟真度极高的语音合成

免费增值海外

魔音工坊

AI 配音配乐工具,多情绪音色与字幕时间轴

免费增值国产

Speech

A scalable generative AI framework built for resear…

免费海外★ 1.8w

Orpheus TTS

Towards Human-Sounding Speech

免费海外★ 6.3k

Abogen

Generate audiobooks from EPUBs, PDFs and text with …

免费海外★ 5.8k

Fish Audio

中文声音克隆与 TTS

免费增值国产

Suno

文本生成完整歌曲

免费增值海外

Udio

AI 音乐生成与混音

免费增值海外

Ai Collection

The Generative AI Landscape - A Collection of Aweso…

免费海外★ 9.1k

Transformers

🤗 Transformers: the model-definition framework for…

免费海外★ 16.5w

Whisper

开源多语种语音识别

免费海外

Meetily

Privacy first, AI meeting assistant with 4x faster …

免费海外★ 3.0w

Inference

Swap GPT for any LLM by changing a single line of c…

免费海外★ 9.5k

Cactus

Quantization, kernels, runtime and inference engine…

免费海外★ 6.0k

Fastrtc

The python library for real-time communication

免费海外★ 4.6k

LLPlayer

The media player for language learning, with dual s…

免费海外★ 4.1k

MOSS TTS

MOSS‑TTS Family is an open‑source speech and sound …

免费海外★ 4.1k

Openless

Hold a key, speak, release — AI-polished text appea…

免费海外★ 3.4k

Musiclm Pytorch

Implementation of MusicLM, Google's new SOTA model …

免费海外★ 3.3k

TTS WebUI

A single Gradio + React WebUI with extensions for A…

免费海外★ 3.2k

Elevenlabs Python

The official Python SDK for the ElevenLabs API.

免费海外★ 3.1k

Awesome Whisper

🔊 Awesome list for Whisper — an open-source AI-pow…

免费海外★ 2.4k

Chat With Gpt

An open-source ChatGPT app with a voice

免费海外★ 2.4k

ReadAny

AI-powered cross-platform e-book reader with semant…

免费海外★ 2.3k

Kiro Gateway

👻 Proxy API gateway for Kiro IDE & CLI (Amazon Q D…

免费海外★ 2.3k

Kalosm

Instant, controllable, local pre-trained AI models …

免费海外★ 2.2k

WhisperJAV

ASR/STT subtitle generator. Uses Qwen3-ASR, local L…

免费海外★ 2.2k

Claude Real Video

Let Claude (or any LLM) actually watch a video — sc…

免费海外★ 2.1k

Openai Edge Tts

Free, high-quality text-to-speech API endpoint to r…

免费海外★ 2.1k

FireRedASR

Open-source industrial-grade ASR models supporting …

免费海外★ 2.0k

Dot

Text-To-Speech, RAG, and LLMs. All local!

免费海外★ 1.9k

Audio Ai Timeline

A timeline of the latest AI models for audio genera…

免费海外★ 1.9k

Openai Kotlin

OpenAI API client for Kotlin with multiplatform and…

免费海外★ 1.8k

Bailing

百聆 是一个类似GPT-4o的语音对话机器人,通过ASR+LLM+TTS实现,集成DeepSeek R…

免费海外★ 1.8k

Uzu

A high-performance inference engine for AI models

免费海外★ 1.7k

Soundstorm Pytorch

Implementation of SoundStorm, Efficient Parallel Au…

免费海外★ 1.5k

RCLI

Talk to your Mac, query your docs, no cloud require…

免费海外★ 1.5k

IOS_ML

List of Machine Learning, AI, NLP solutions for iOS…

免费海外★ 1.4k

Dragonfire

the open-source virtual assistant for Ubuntu based …

免费海外★ 1.4k

Audio Webui

A webui for different audio related Neural Networks

免费海外★ 1.2k

SincNet

SincNet is a neural architecture for efficiently pr…

免费海外★ 1.2k

Ai Audio Datasets

AI Audio Datasets (AI-ADS) 🎵, including Speech, Mu…

免费海外★ 962

Epub2tts

Turn an epub or text file into an audiobook

免费海外★ 956

BS RoFormer

Implementation of Band Split Roformer, SOTA Attenti…

免费海外★ 925

Genmusic_demo_list

a list of demo websites for automatic music generat…

免费海外★ 797

Orpheus FastAPI

High-performance Text-to-Speech server with OpenAI-…

免费海外★ 718

Voicebox Pytorch

Implementation of Voicebox, new SOTA Text-to-speech…

免费海外★ 703

Swift

Fast voice assistant powered by Groq, Cartesia, and…

免费海外★ 592

Open Musiclm

Implementation of MusicLM, a text to music model pu…

免费海外★ 559

E2 Tts Pytorch

Implementation of E2-TTS, "Embarrassingly Easy Full…

免费海外★ 517

Opentypeless

Open-source AI voice typing for macOS, Windows, and…

免费海外★ 487

Elevenlabs Js

The official JavaScript (Node) library for the Elev…

免费海外★ 440

GPTPortal

A feature-rich portal to chat with GPT-4, Claude, G…

免费海外★ 397

ViralCutter

Free tool to create viral videos from YouTube, gene…

免费海外★ 385

Sonobarr

Music discovery tool that integrates with Lidarr an…

免费海外★ 383

MultiMed

[LREC-COLING 2024 (Oral), Interspeech 2024 (Oral), …

免费海外★ 379

VoxNovel

VoxNovel: generate audiobooks giving each character…

免费海外★ 374

UTMOSv2

UTokyo-SaruLab MOS Prediction System

免费海外★ 365

AIUI

AIUI is a platform enabling seamless two-way verbal…

免费海外★ 355

Infinifi

infinifi plays gentle lofi music in the background …

免费海外★ 345

Vocalis

Speech-to-speech AI assistant with natural conversa…

免费海外★ 315

ComfyUI Qwen3 TTS

A ComfyUI custom node suite for Qwen3-TTS, supporti…

免费海外★ 299

Yapsnap

Snap any video URL or audio file into plaintext. No…

免费海外★ 296

Pyht

PlayHT Python SDK - AI Text-to-Speech Streaming & V…

免费海外★ 218