Voicebox is a generative AI model for speech that can generalize to tasks it was not specifically trained for with state-of-the-art performance. Unlike existing speech synthesizers, it can be trained on diverse, unstructured data without requiring carefully labeled inputs. Voicebox uses a new approach called Flow Matching, which is a Meta's latest advancement on non-autoregressive generative models that can learn highly non-deterministic mapping between text and speech.
Starting price: Free. Features, limits, credits, seats, billing periods, taxes, regional availability, and promotions can change, so confirm the current official checkout or sales quote.
Research-model demonstrations are not necessarily public production services; verify availability license safety limitations and approved use.
You must be logged in to submit a review.
No reviews yet. Be the first to review!