Fugatto

Fugatto

NVIDIA
StepAudio 3

StepAudio 3

StepFun
+
+

Related Products

  • Adobe Firefly
    25,030 Ratings
    Visit Website
  • LALAL.AI
    5,355 Ratings
    Visit Website
  • Muzaic
    2 Ratings
    Visit Website
  • Google AI Studio
    41 Ratings
    Visit Website
  • LTX
    182 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • Screencapt
    140 Ratings
    Visit Website
  • Evertune
    1 Rating
    Visit Website
  • Gemini Enterprise Agent Platform
    999 Ratings
    Visit Website
  • Forethought
    166 Ratings
    Visit Website

About

Using text and audio as inputs, a new generative AI model from NVIDIA can create any combination of music, voices, and sounds. A team of generative AI researchers created a Swiss Army knife for sound, one that allows users to control the audio output simply using text. While some AI models can compose a song or modify a voice, none have the dexterity of the new offering. Called Fugatto, it generates or transforms any mix of music, voices, and sounds described with prompts using any combination of text and audio files. For example, it can create a music snippet based on a text prompt, remove or add instruments from an existing song, change the accent or emotion in a voice, and even let people produce sounds never heard before. Supporting numerous audio generation and transformation tasks, Fugatto is the first foundational generative AI model that showcases emergent properties.

About

StepAudio 3 is StepFun’s next-generation audio model family, built to understand, generate, and interact through voice, sound, and music. The lineup includes StepAudio 3 Realtime for natural full-duplex conversation, StepAudio 3 ASR for speech recognition, StepAudio 3 TTS for speech synthesis, StepAudio 3 Gen for general-purpose audio generation, and StepAudio 3 Music for long-form music creation. Realtime is designed around a continuous listen-converse-think-act loop, understanding not only words but also hesitation, laughter, emotion, pauses, backchannels, and interruptions. It can think while speaking, reason through harder questions without breaking conversational flow, and use tools to complete tasks once it understands the user’s intent. StepAudio 3 Gen unifies zero-shot TTS, voice design, vocal generation, sound effects, music, and mixed audio generation within one framework, while StepAudio 3 Music supports text-controlled songs, instrumentals, vocal arrangement, and more.

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Audience

Individuals and any user in search of a solution to generate music, voices, and sounds with AI

Audience

Developers, AI teams, and creators needing to build real-time voice agents, speech applications, transcription systems, and generative audio or music experiences

Support

Phone Support Supported
24/7 Live Support Not Supported
Online Supported

Support

Phone Support Not Supported
24/7 Live Support Not Supported
Online Supported

API

Offers API Not Supported

API

Offers API Not Supported

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version Not Supported
Free Trial Not Supported

Pricing

No information available.
Free Version Not Supported
Free Trial Not Supported

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation Supported
Webinars Supported
Live Online Not Supported
In Person Supported

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Not Supported

Company Information

NVIDIA
Founded: 1993
United States
blogs.nvidia.com/blog/fugatto-gen-ai-sound-model/

Company Information

StepFun
United States
static.stepfun.com/blog/stepaudio3/

Alternatives

Alternatives

StepAudio 3

StepAudio 3

StepFun
Seed-Music

Seed-Music

ByteDance
Seed-Music

Seed-Music

ByteDance
SongR

SongR

Riffit
Fugatto

Fugatto

NVIDIA

Categories

Generative AI Supported

Categories

AI Models Supported

Integrations

NVIDIA DGX Cloud Supported

Integrations

NVIDIA DGX Cloud Not Supported
Claim Fugatto and update features and information
Claim Fugatto and update features and information
Claim StepAudio 3 and update features and information
Claim StepAudio 3 and update features and information