> ## Documentation Index
> Fetch the complete documentation index at: https://mintlify.com/MateoRiosdev/Free-TTS-VozCraft/llms.txt
> Use this file to discover all available pages before exploring further.

# Using VozCraft - Complete Guide

> Comprehensive workflow guide for creating professional text-to-speech audio with VozCraft

# Using VozCraft - Complete Guide

This comprehensive guide covers complete workflows for using VozCraft effectively, from basic generation to advanced techniques for professional audio production. Whether you're creating content for education, business, or entertainment, this guide will help you maximize VozCraft's capabilities.

## Application Workflow Overview

VozCraft follows a simple, intuitive workflow:

```mermaid theme={null}
graph LR
    A[Access VozCraft] --> B[Configure Settings]
    B --> C[Enter Text]
    C --> D[Generate Audio]
    D --> E[Listen & Review]
    E --> F{Satisfied?}
    F -->|No| B
    F -->|Yes| G[Export Audio]
    G --> H[Manage History]
```

<Steps>
  <Step title="Access VozCraft">
    Open the application in your browser—no login or installation required.
  </Step>

  <Step title="Configure Settings">
    Choose language, voice type, speed, and mood for your audio.
  </Step>

  <Step title="Enter Text">
    Type or paste your content (up to 5,000 characters).
  </Step>

  <Step title="Generate Audio">
    Click the generate button to create audio instantly.
  </Step>

  <Step title="Listen & Review">
    Play back the audio to verify it meets your needs.
  </Step>

  <Step title="Export Audio">
    Download as MP3, WAV, or save transcript as TXT.
  </Step>

  <Step title="Manage History">
    Rename, organize, or export your audio library.
  </Step>
</Steps>

## Basic Workflows

### Workflow 1: Quick Audio Generation

**Use Case**: Generate a single audio file quickly

**Time Required**: 2-3 minutes

<Steps>
  <Step title="Open VozCraft">
    Navigate to VozCraft in your web browser.
  </Step>

  <Step title="Select Language">
    Click the **🌍 Voice / Accent / Region** dropdown and select your target language.

    Example: Choose "English (US)" for American English.
  </Step>

  <Step title="Keep Default Settings">
    For quick generation, use defaults:

    * Voice Type: Normal
    * Speed: Normal
    * Mood: Neutral
  </Step>

  <Step title="Enter Your Text">
    Type or paste your content into the text area. For a quick test:

    ```
    Welcome to VozCraft. This is a test of the text-to-speech system.
    ```
  </Step>

  <Step title="Generate">
    Click **"🎧 Generate Audio"** and listen to the result.
  </Step>

  <Step title="Export if Satisfied">
    Click **"MP3"** to download your audio file.
  </Step>
</Steps>

<Tip>
  **Quick Tip**: For most general content, the default settings (Normal voice, Normal speed, Neutral mood) provide excellent results.
</Tip>

### Workflow 2: Creating Multiple Audio Files

**Use Case**: Generate a series of related audio files (e.g., course lessons, podcast episodes)

**Time Required**: 5-10 minutes per file

<Steps>
  <Step title="Configure Base Settings">
    Set up your preferred voice configuration once:

    * Language: Your content language
    * Voice Type: Choose based on content style
    * Speed: Consider your audience
    * Mood: Match your content tone
  </Step>

  <Step title="Generate First Audio">
    1. Paste first text
    2. Click Generate Audio
    3. Listen to verify settings
    4. Adjust if needed
  </Step>

  <Step title="Rename in History">
    Click the ✏️ icon and name it descriptively:

    ```
    Lesson 1: Introduction
    ```
  </Step>

  <Step title="Generate Subsequent Files">
    For each additional file:

    1. Clear text area
    2. Paste next content
    3. Generate (settings persist)
    4. Rename in history

    Names like:

    ```
    Lesson 2: Basic Concepts
    Lesson 3: Advanced Techniques
    ```
  </Step>

  <Step title="Bulk Export">
    Once all files are generated:

    1. Export each as MP3/WAV
    2. Export history as JSON (backup)
    3. Save transcript files if needed
  </Step>
</Steps>

<Note>
  **Settings Persistence**: VozCraft remembers your last-used settings, making it easy to generate multiple files with consistent voice characteristics.
</Note>

### Workflow 3: Testing Different Voice Options

**Use Case**: Experiment with different voices to find the best fit

**Time Required**: 10-15 minutes

<Steps>
  <Step title="Prepare Test Text">
    Write a representative sample (100-200 characters) that includes:

    * Typical vocabulary from your content
    * Natural sentence structure
    * Any special terms or names

    Example:

    ```
    Welcome to our product demo. Today we'll explore the key features
    that make our solution unique. Let's get started with the basics.
    ```
  </Step>

  <Step title="Test Language Variants">
    Generate audio with different accents:

    1. English (US)
    2. English (UK)
    3. English (AU)

    Listen to each and note:

    * Clarity
    * Accent appropriateness
    * Pronunciation accuracy
  </Step>

  <Step title="Test Voice Types">
    Using your preferred language:

    1. Generate with Normal Voice
    2. Generate with High-pitched Voice

    Compare:

    * Which sounds more professional?
    * Which matches your brand?
    * Which is more engaging?
  </Step>

  <Step title="Test Moods">
    Try 3-4 relevant moods:

    * Neutral (baseline)
    * Happy (if content is positive)
    * Serious (if content is formal)
    * Energetic (if content is dynamic)
  </Step>

  <Step title="Select Winner">
    Choose the combination that:

    * Sounds most natural
    * Matches your content tone
    * Appeals to your audience
    * Meets your quality standards
  </Step>

  <Step title="Document Settings">
    Note your final settings:

    * Language: \_\_\_\_\_\_\_
    * Voice Type: \_\_\_\_\_\_\_
    * Speed: \_\_\_\_\_\_\_
    * Mood: \_\_\_\_\_\_\_

    Use these for all future content.
  </Step>
</Steps>

## Advanced Workflows

### Workflow 4: Long-Form Content Production

**Use Case**: Creating audiobook chapters, long articles, or extensive documentation

**Challenges**: 5,000 character limit, maintaining consistency

**Solution**: Split content intelligently and merge later

<Steps>
  <Step title="Prepare Your Content">
    1. Write complete content in text editor
    2. Check total length (character count)
    3. Plan splits at natural breakpoints:
       * End of paragraphs
       * Section breaks
       * Logical thought boundaries

    **Formula**: `numberOfSegments = totalCharacters / 4500`

    (Use 4,500 to leave buffer for safety)
  </Step>

  <Step title="Split Text Strategically">
    Create segments that:

    * End at sentence boundaries
    * Don't exceed 5,000 characters
    * Have slight overlap for smooth merging

    Example split points:

    ```
    Segment 1: [0-4500] chars - ends at "...and that concludes section one."
    Segment 2: [4480-8980] chars - starts 20 chars before end of Seg 1
    ```
  </Step>

  <Step title="Generate All Segments">
    For each segment:

    1. Paste text into VozCraft
    2. Verify character count is under 5,000
    3. Generate audio
    4. Rename clearly: "Chapter\_3\_Segment\_1", "Chapter\_3\_Segment\_2", etc.
    5. Export as WAV (for later editing)
  </Step>

  <Step title="Export and Organize">
    1. Export all segments as WAV
    2. Export history JSON as backup
    3. Organize files in project folder:

    ```
    Project/
    ├── Chapter_3_Segment_1.wav
    ├── Chapter_3_Segment_2.wav
    ├── Chapter_3_Segment_3.wav
    └── vozcraft-history.json
    ```
  </Step>

  <Step title="Merge in Audio Editor (Optional)">
    Using Audacity or similar:

    1. Import all segments
    2. Arrange on timeline
    3. Add slight crossfade between segments (1-2 seconds)
    4. Normalize volume across all segments
    5. Export final combined file
  </Step>
</Steps>

<Tip>
  **Consistency is Key**: Use identical settings (voice, speed, mood) for all segments to ensure seamless transitions.
</Tip>

### Workflow 5: Multilingual Content Creation

**Use Case**: Creating the same content in multiple languages

**Time Required**: 5-10 minutes per language

<Steps>
  <Step title="Prepare Translations">
    Translate your content into target languages:

    * Use professional translation service
    * Or Google Translate for basic needs
    * Verify translations with native speakers if possible
  </Step>

  <Step title="Create Reference Matrix">
    Document your plan:

    | Language         | Voice Type | Speed  | Mood    | Notes              |
    | ---------------- | ---------- | ------ | ------- | ------------------ |
    | English (US)     | Normal     | Normal | Neutral | Default            |
    | Español (México) | Normal     | Normal | Neutral | Match English      |
    | Français         | Normal     | Slow   | Neutral | Slower for clarity |
  </Step>

  <Step title="Generate Each Language">
    For each language:

    1. Select language in VozCraft
    2. Paste translated text
    3. Adjust settings per your matrix
    4. Generate and listen
    5. Rename: "ProductDemo\_EN", "ProductDemo\_ES", etc.
  </Step>

  <Step title="Export Organized Files">
    Structure your exports:

    ```
    ProductDemo/
    ├── EN/
    │   ├── ProductDemo_EN.mp3
    │   └── ProductDemo_EN.txt
    ├── ES/
    │   ├── ProductDemo_ES.mp3
    │   └── ProductDemo_ES.txt
    └── FR/
        ├── ProductDemo_FR.mp3
        └── ProductDemo_FR.txt
    ```
  </Step>

  <Step title="Quality Check">
    Have native speakers review:

    * Pronunciation accuracy
    * Naturalness
    * Appropriate voice characteristics
    * Cultural appropriateness
  </Step>
</Steps>

### Workflow 6: Voice Characterization

**Use Case**: Creating distinct character voices for storytelling, role-play, or educational scenarios

**Example**: Creating a story with 3 characters

<Steps>
  <Step title="Define Character Profiles">
    Plan voice characteristics:

    **Character 1: Narrator**

    * Voice: Normal
    * Speed: Normal
    * Mood: Neutral
    * Purpose: Authoritative, clear

    **Character 2: Young Character**

    * Voice: High-pitched
    * Speed: Fast
    * Mood: Happy
    * Purpose: Energetic, youthful

    **Character 3: Wise Elder**

    * Voice: Normal
    * Speed: Slow
    * Mood: Serious
    * Purpose: Thoughtful, experienced
  </Step>

  <Step title="Separate Dialogue">
    Split your script by speaker:

    ```
    Narrator_01.txt: "Once upon a time, in a faraway land..."
    Character2_01.txt: "Wow! Look at that amazing castle!"
    Character3_01.txt: "Patience, young one. All in good time."
    Narrator_02.txt: "And so the adventure began..."
    ```
  </Step>

  <Step title="Generate Each Voice">
    For each text segment:

    1. Configure settings for that character
    2. Generate audio
    3. Rename with character name and sequence number
    4. Export as WAV for editing
  </Step>

  <Step title="Assemble in Audio Editor">
    Using Audacity:

    1. Import all character audio files
    2. Arrange in story sequence on timeline
    3. Add pauses between speakers (0.5-1 second)
    4. Add background music or sound effects
    5. Export final story
  </Step>
</Steps>

<Info>
  **Voice Contrast**: Maximize differences between characters by using contrasting settings (High-pitched vs Normal, Fast vs Slow) for clear distinction.
</Info>

## Specialized Use Cases

### Use Case: Language Learning Content

**Goal**: Create audio for language learners with optimal clarity

**Recommended Settings**:

* **Speed**: Slow or Very Slow (0.75x or 0.50x)
* **Voice Type**: Normal (clearer pronunciation)
* **Mood**: Neutral (no emotional distraction)
* **Approach**: Pronunciation practice first, comprehension practice later

**Workflow**:

<Accordion title="Pronunciation Practice">
  **Settings**: Very Slow + Normal + Neutral

  **Content Format**:

  ```
  Word. Pause. Sentence.

  Apple. (pause) I have an apple.
  Book. (pause) She is reading a book.
  ```

  **Process**:

  1. Generate with maximum clarity (Very Slow)
  2. Export as MP3
  3. Use in flashcard apps or learning materials

  **Benefit**: Learners hear every sound clearly
</Accordion>

<Accordion title="Comprehension Practice">
  **Settings**: Normal + Normal + Neutral

  **Content Format**:

  ```
  Natural sentences and paragraphs at conversational speed.

  The weather today is sunny and warm. It's a perfect day
  for a walk in the park.
  ```

  **Process**:

  1. Generate at normal conversational speed
  2. Test learner comprehension
  3. Generate slow version for review if needed

  **Benefit**: Realistic listening practice
</Accordion>

<Accordion title="Progressive Difficulty">
  **Create 3 Versions**:

  1. **Version A**: Very Slow (0.50x) - Beginners
  2. **Version B**: Slow (0.75x) - Intermediate
  3. **Version C**: Normal (1.00x) - Advanced

  **Same Content, Different Speeds**:

  * Learners progress through versions
  * Builds confidence gradually
  * Accommodates different skill levels
</Accordion>

### Use Case: Podcast Intro/Outro

**Goal**: Create professional podcast intro and outro segments

**Recommended Settings**:

* **Intro**: High-pitched + Fast + Enthusiastic (energetic welcome)
* **Outro**: Normal + Normal + Neutral (professional closing)

**Workflow**:

<Steps>
  <Step title="Write Scripts">
    **Intro Script** (30-45 seconds):

    ```
    Welcome to [Podcast Name]! The show where we explore [topic].
    I'm your host [Name], and today we're diving into [episode topic].
    Let's get started!
    ```

    **Outro Script** (20-30 seconds):

    ```
    Thanks for listening to [Podcast Name]. If you enjoyed this episode,
    please subscribe and leave a review. Until next time!
    ```
  </Step>

  <Step title="Generate with Appropriate Mood">
    * **Intro**: Use Enthusiastic or Happy for energy
    * **Outro**: Use Neutral for professional close
  </Step>

  <Step title="Export as WAV">
    Export high-quality WAV files for podcast production
  </Step>

  <Step title="Enhance in Audio Editor">
    1. Add intro/outro music
    2. Apply compression for consistency
    3. Add reverb for polish (subtle)
    4. Normalize volume to -16 LUFS (podcast standard)
    5. Export final versions
  </Step>

  <Step title="Reuse for Every Episode">
    Use the same intro/outro files across episodes for brand consistency
  </Step>
</Steps>

### Use Case: IVR / Phone System

**Goal**: Create phone system prompts and menus

**Recommended Settings**:

* **Voice Type**: Normal (professional)
* **Speed**: Slow (clarity over phone)
* **Mood**: Neutral or Serious (professional tone)

**Workflow**:

<Accordion title="Main Menu">
  ```
  Welcome to [Company Name]. 

  For sales, press one.
  For support, press two.
  For billing, press three.
  To repeat this menu, press star.
  ```

  **Settings**: Normal + Slow + Neutral

  **Why Slow**: Phone audio quality is lower; slow speech ensures clarity
</Accordion>

<Accordion title="Hold Messages">
  ```
  Thank you for holding. Your call is important to us.
  A representative will be with you shortly.
  ```

  **Settings**: Normal + Normal + Neutral

  **Tip**: Keep hold messages calm and professional
</Accordion>

<Accordion title="Confirmation Messages">
  ```
  Your request has been submitted successfully.
  You will receive a confirmation email within 24 hours.
  Thank you for your business.
  ```

  **Settings**: Normal + Normal + Neutral

  **Export**: WAV format for best phone system compatibility
</Accordion>

## Best Practices

### Text Formatting for Best Results

<CardGroup cols={2}>
  <Card title="Use Proper Punctuation" icon="period">
    **Good**:

    ```
    Hello! Welcome to our store. How can I help you today?
    ```

    **Bad**:

    ```
    hello welcome to our store how can i help you today
    ```

    **Impact**: Punctuation creates natural pauses and intonation
  </Card>

  <Card title="Avoid Special Characters" icon="ban">
    **Good**:

    ```
    The price is twenty dollars.
    Call us at 5-5-5, 1-2-3-4.
    ```

    **Bad**:

    ```
    The price is $20.
    Call us at 555-1234.
    ```

    **Impact**: Write numbers and symbols as words for better pronunciation
  </Card>

  <Card title="Break Long Sentences" icon="scissors">
    **Good**:

    ```
    We offer many services. These include consulting, training, and support.
    Each service is customized to your needs.
    ```

    **Bad**:

    ```
    We offer many services including consulting training and support and
    each service is customized to your specific needs and requirements.
    ```

    **Impact**: Shorter sentences improve clarity and flow
  </Card>

  <Card title="Spell Out Abbreviations" icon="spell-check">
    **Good**:

    ```
    The United States of America.
    Doctor Smith will see you.
    ```

    **Bad**:

    ```
    The USA.
    Dr. Smith will see you.
    ```

    **Impact**: Spelling out ensures correct pronunciation
  </Card>
</CardGroup>

### Quality Control Checklist

Before exporting final audio:

<Steps>
  <Step title="Pronunciation Check">
    * [ ] All names pronounced correctly?
    * [ ] Technical terms handled well?
    * [ ] Numbers read naturally?
    * [ ] Acronyms spelled or spoken correctly?
  </Step>

  <Step title="Pacing Check">
    * [ ] Speed appropriate for content?
    * [ ] Natural pauses between sentences?
    * [ ] Comfortable listening pace?
    * [ ] Suitable for target audience?
  </Step>

  <Step title="Tone Check">
    * [ ] Mood matches content?
    * [ ] Emotional tone appropriate?
    * [ ] Professional/casual balance correct?
    * [ ] Voice type fits brand?
  </Step>

  <Step title="Technical Check">
    * [ ] Audio plays smoothly?
    * [ ] No cuts or glitches?
    * [ ] Volume consistent?
    * [ ] Exported in correct format?
  </Step>
</Steps>

### Naming Conventions

Develop a consistent naming system:

**Format**: `[Project]_[Type]_[Number]_[Language]_[Version]`

**Examples**:

```
Podcast_Intro_01_EN_v1
Course_Lesson_05_ES_v2  
Product_Demo_Main_FR_final
IVR_Menu_Main_EN_v3
```

**Benefits**:

* Easy to sort and find files
* Clear version tracking
* Language identification
* Professional organization

## Optimization Tips

### Speed Up Your Workflow

<Accordion title="Use History as Templates">
  Instead of configuring settings repeatedly:

  1. Generate audio with perfect settings
  2. Name it "TEMPLATE - \[Description]"
  3. For new content: Find template in history
  4. Check settings (they're displayed)
  5. Configure VozCraft to match
  6. Generate new audio

  **Saves**: 1-2 minutes per generation
</Accordion>

<Accordion title="Batch Text Preparation">
  Prepare all content before starting:

  1. Write all text in one document
  2. Run spell check
  3. Format numbers and abbreviations
  4. Split into segments if needed
  5. Copy/paste efficiently through VozCraft

  **Saves**: Reduces context switching, improves focus
</Accordion>

<Accordion title="Keyboard Shortcuts">
  Use browser shortcuts:

  * **Ctrl+A / Cmd+A**: Select all text
  * **Ctrl+C / Cmd+C**: Copy
  * **Ctrl+V / Cmd+V**: Paste
  * **Enter**: Confirm rename modal
  * **Esc**: Cancel rename modal

  **Saves**: Seconds per action, minutes over session
</Accordion>

### Maintain Consistency Across Projects

<Steps>
  <Step title="Document Your Standards">
    Create a voice style guide:

    ```
    Company Voice Standards:
    - Language: English (US)
    - Voice Type: Normal
    - Speed: Normal
    - Mood: Neutral (default) or Serious (important announcements)
    - Exceptions: Marketing uses Happy mood
    ```
  </Step>

  <Step title="Export Reference Audio">
    Create a reference file:

    1. Generate standard voice with sample text
    2. Name: "COMPANY\_VOICE\_STANDARD\_REFERENCE"
    3. Export MP3 and transcript
    4. Share with team
    5. Use for comparison
  </Step>

  <Step title="Regular Quality Reviews">
    Schedule periodic reviews:

    * Compare new audio to reference
    * Ensure settings haven't drifted
    * Update standards if needed
    * Train team members on standards
  </Step>
</Steps>

## Troubleshooting Common Workflows

<Warning>
  **Problem**: Generated audio sounds unnatural

  **Solutions**:

  1. **Check punctuation**: Add periods, commas, questions marks
  2. **Simplify text**: Break long sentences into shorter ones
  3. **Try different mood**: Switch to Neutral if using extreme moods
  4. **Adjust speed**: Normal speed often sounds most natural
  5. **Test different language variant**: Try another accent
</Warning>

<Warning>
  **Problem**: Can't generate consistent results

  **Solutions**:

  1. **Document settings**: Write down exact settings used
  2. **Export history JSON**: Backup contains all settings
  3. **Use history**: Reference previous successful generations
  4. **Check browser**: Same browser produces consistent results
  5. **Name systematically**: Include settings in filename
</Warning>

<Warning>
  **Problem**: Taking too long to produce content

  **Solutions**:

  1. **Batch preparation**: Prepare all text before starting
  2. **Use templates**: Reference previous good settings
  3. **Skip perfection**: "Good enough" may be sufficient
  4. **Learn shortcuts**: Use keyboard navigation
  5. **Optimize workflow**: Follow workflows in this guide
</Warning>

## Next Steps

<CardGroup cols={3}>
  <Card title="Voice Settings Guide" icon="microphone" href="/guides/voice-settings">
    Detailed guide for optimal voice configuration
  </Card>

  <Card title="Exporting Audio" icon="download" href="/guides/exporting-audio">
    Step-by-step guide for all export scenarios
  </Card>

  <Card title="Troubleshooting" icon="wrench" href="/guides/troubleshooting">
    Solutions to common problems
  </Card>
</CardGroup>
