Skip to main content

Overview

60db’s Text-to-Speech (TTS) API converts written text into natural-sounding speech using advanced AI models. Our TTS engine supports multiple voices, languages, and customization options.

Features

Multiple Voices

Choose from 50+ pre-built voices or create custom voices

Voice Customization

Adjust speed, pitch, and other parameters

High Quality

Crystal-clear audio with natural intonation

Multiple Formats

Support for MP3, WAV, OGG, and FLAC output formats

Basic Usage

Voice Parameters

Speed

Control the speaking rate of the generated audio:
  • 0.5: Half speed (slow)
  • 1.0: Normal speed (default)
  • 2.0: Double speed (fast)

Pitch

Adjust the pitch of the voice:

Enhancement

Enable audio enhancement for better quality:

Output Formats

Supported audio formats:

Best Practices

  • Use proper punctuation for natural pauses
  • Break long texts into paragraphs
  • Use SSML tags for advanced control (coming soon)
  • Test multiple voices for your use case
  • Consider accent and gender for your audience
  • Use custom voices for brand consistency
  • Cache frequently used audio
  • Batch requests when possible
  • Use appropriate audio format for your use case
  • Enable enhancement for production use
  • Use WAV format for highest quality
  • Test with different speed settings

Use Cases

Voice Assistants

Content Narration

Accessibility

API Reference

For detailed API documentation, see:

Text to Speech

Standard TTS endpoint