Home
Assets
My Works
Models
Voice Library
Developer
Blog
English
Pricing
Sign In

Models

Every generation model ListenHub runs. Pick one to open its tool, or take the model ID and call it from the API.

API docs
Best for videoSeedance 2.0 ProThe only one with both reference video and native audioBest for imagesNano Banana ProSteady text rendering and multi-reference consistencyBest for voiceLFlowSpeechOne engine behind podcasts, voiceover and dialogue

Capability: reference video · first/last frame · native audio · multi-reference · precise edit · multi-speaker · emotion · streaming
Access: has a tool page · API only

LListenHub: FlowSpeechPopular

The speech engine every voice tool shares. Standard voices run on our own engine, Pro voices on ElevenLabs; voices are a layer under it.

Input: textBilling: 按音频时长,约 4 积分 / 分钟Tools: text-to-speech · podcast · multi-speaker-tts
ByteDance: Seed Audio

A voice-acting model billed per second. Its tool page is called AI Voice.

Input: textBilling: 0.5 积分 / 秒,最低 1 积分Tools: ai-voice
LListenHub: Voice Clone

Clone your own voice from one recording; the result shows up under My voices in FlowSpeech.

Input: text · audioBilling: 见定价说明Tools: voice-cloning