Speech

A scalable generative AI framework for researchers and developers working on Large Language Models, Multimodal, and Speech AI.

Voice & SpeechUnknownOpen source
Visit tool

docs.nvidia.com

About Speech

Built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech).

Description summarised by AI from the sources listed below.

Key features

  • Automatic Speech Recognition (ASR)
  • Text-to-Speech (TTS)
  • Speaker Diarization
  • Speaker Recognition
  • Speech Language Models
  • Audio Processing

Pricing

Pricing model: Unknown — we have not been able to confirm pricing from the official website, so nothing is stated here.

Pricing summarised by AI from the sources listed below.