Core Features


Feature Comparison


Getting Started

  1. Start the server:
  2. Open the web UI:
  3. Download required models:

Model Requirements

Different features require different models:

Next Steps

Choose a feature to learn more:

Voice Mode

Use Izwi for real-time voice conversations with speech recognition, chat, and text-to-speech.

Chat

Run local text and multimodal chat conversations through the Izwi CLI, web UI, and API.

Transcription

Convert audio to text with Izwi through the CLI, web UI, and local API.

Speaker Attributed ASR

Generate Granite Speech speaker-turn transcripts through the Transcription workspace and speech-text jobs API.

Voice Studio

Create, clone, design, preview, manage, and reuse voices from the unified Izwi Voices workspace.

Text-to-Speech

Generate natural speech from text with Izwi models, voices, streaming, and audio output formats.

Studio

Manage long-form text-to-speech projects, chapter workflows, and exports in Izwi Studio.

Settings and Onboarding

Configure appearance, updates, analytics, desktop system behavior, and first-run model setup in Izwi.

Diarization

Identify multiple speakers in audio and generate speaker-attributed transcripts with Izwi.

Voice Cloning

Create custom voices from reference audio and use them for local text-to-speech generation.

Voice Design

Design synthetic voices from text descriptions and use them in Izwi speech workflows.