User Manual
Complete guide to using the Vocametrix voice analysis platform. Learn how to create accounts, use analysis features, and more.
How to Create an Account
1. Go to the Registration Page
Visit vocametrix.com/registration or click Login then Create Account on the homepage.
2. Fill in the Registration Form
- First Name
- Last Name
- Email Address (used as your login username)
- Phone Number (optional)
- Password (minimum 8 characters)
- Confirm Password
- Account Type: Speech Therapist, Patient (adult), Parent, or Visitor
- If you select Patient or Parent, you will be asked if you already have a speech therapist or are looking for one. If you already have a therapist, you can optionally enter their email.
- Select language of your daily practice
- Select language for future communication
3. Agreements & Consent
- You must accept the Terms of Service and Privacy Policy to register.
- You can optionally consent to receive marketing communications.
4. Submit the Form
Click Create Account to submit your registration. If successful, you will see a confirmation and can log in.
Common Issues & Tips
- Passwords don't match: Make sure both password fields are identical.
- Email format error: Use a valid email address (e.g., name@domain.com).
- Missing required fields: Complete all required fields before submitting.
- Consent required: You must accept all required agreements to register.
How to Change the Interface Language
You can change the interface language of the Vocametrix platform through your profile settings. Currently, the platform supports English, French, Spanish, German, and Russian.
How to Access
- Log into your Vocametrix account
- Click on your profile picture or name in the top navigation
- Select Profile from the dropdown menu
- Navigate to the Profile Details tab (first tab)
Step-by-Step Instructions
1. Locate Language Settings
In the Profile Details section, scroll down to find the language settings. You will see two language options:
- Daily Practice Language: Used for the interface language and practice sessions
- Newsletter Language: Used for email communications and newsletters
2. Change Daily Practice Language
Click on the Daily Practice Language dropdown menu. Select your preferred language from the available options.
3. Save Your Changes
Click the Save Changes button at the bottom of the form to save your language preference.
4. Reconnect to Apply Changes
To see the interface in your selected language, you need to log out and log back in to the platform. Simply refreshing the page will not apply the language changes.
Supported Languages
The Vocametrix platform currently supports the following interface languages:
Important Notes
- Changes Take Effect After Logout/Login: The interface language will only change after you log out and log back in to the platform
- Profile Settings Only: Currently, language changes must be made through the Profile section - there is no quick language switcher in the main interface
- Account Required: You must be logged in to change your language preference
- Use Daily Practice Language: Change the "Daily Practice Language" setting, not the "Newsletter Language" setting, to modify the interface language
How to Get Statistics From a Text
The Text Statistics tool provides comprehensive analysis of written content without using any credits. This powerful feature helps analyze text complexity, readability, and linguistic patterns.
How to Access
- Log into your Vocametrix account
- Click on Clinical Tools in the main menu
- Select Text Analysis from the dropdown
- Or visit /ai/get-text-statistics directly
Getting Started
1. Enter or Paste Text
Type or paste the text you want to analyze into the provided input area. There is no character limit for this tool.
2. Run Analysis
Click the Analyze Text button. The system will instantly process your text and display a detailed breakdown of statistics.
Analysis Features
Basic Statistics
Quickly see the most important metrics for your text:
- Word Count: Total number of words
- Character Count: Total number of characters (with and without spaces)
- Sentence Count: Number of sentences detected
- Paragraph Count: Number of paragraphs
Readability & Complexity
Assess how easy or difficult your text is to read:
- Flesch Reading Ease: Standard readability score
- Flesch-Kincaid Grade Level: U.S. school grade level estimate
- Average Sentence Length: Words per sentence
- Average Word Length: Characters per word
Lexical & Linguistic Patterns
Explore vocabulary and language use:
- Unique Words: Count of distinct words
- Type-Token Ratio: Vocabulary diversity measure
- Most Frequent Words: Top words and their counts
- Parts of Speech: Distribution of nouns, verbs, adjectives, etc.
Sentence & Paragraph Structure
Analyze how your text is organized:
- Sentence Length Distribution: Histogram of sentence lengths
- Paragraph Length: Words per paragraph
- Longest/Shortest Sentence: Identify extremes in your text
Best Practices
- Use for Screening: Quickly check text complexity before clinical or research analysis
- Compare Samples: Analyze multiple texts to track changes over time
- Prepare for Analysis: Use before running syntax or pronunciation tools
- Educational Use: Teach students about text statistics and readability
Common Use Cases
- Speech Therapy: Assess written samples for therapy planning
- Clinical Documentation: Analyze patient narratives and reports
- Research: Quantify linguistic features in study data
- Education: Teach text analysis concepts to students
- Content Creation: Optimize readability for articles and materials
How to Analyze the Syntax of a Text
The Syntax Analyzer is an AI-powered tool that examines speech transcriptions to identify language issues in syntax, grammar, vocabulary, and discourse patterns. This professional tool helps speech-language pathologists analyze patient speech samples with detailed categorized feedback.
How to Access
- Log into your Vocametrix account
- Click on Clinical Tools in the main menu
- Select Syntax Analyzer from the Text Analysis section
- Or visit /ai/check-syntax directly
Getting Started
1. Provide Transcribed Text
Paste or type the speech transcription you want to analyze. The tool accepts up to 5,000 characters of text and works best with natural speech samples from therapy sessions or assessments.
2. Review Input
Ensure the text is accurate and representative of the speech sample. Remove any extraneous notes or non-speech content for best results.
3. Run Analysis
Click the "Analyze Syntax" button to process your text. The AI will examine the transcription and provide detailed feedback within seconds.
Analysis Categories
Syntax Issues
Identifies problems with sentence structure and word order:
- Word Order: Incorrect sequence of words in sentences
- Sentence Structure: Incomplete or malformed sentences
- Clause Formation: Issues with dependent and independent clauses
- Question Formation: Problems with interrogative structures
Grammar Issues
Examines grammatical accuracy and rule application:
- Verb Tense: Incorrect past, present, or future tense usage
- Subject-Verb Agreement: Mismatched singular/plural forms
- Articles: Missing or incorrect use of "a", "an", "the"
- Pronouns: Incorrect pronoun selection or reference
- Pluralization: Irregular or missing plural forms
Vocabulary Issues
Analyzes word choice and semantic appropriateness:
- Word Choice: Inappropriate or imprecise word selection
- Semantic Errors: Words used in wrong context
- Idioms: Incorrect or incomplete idiomatic expressions
- Colloquialisms: Register mismatches or informal language use
Discourse Issues
Evaluates overall communication effectiveness and coherence:
- Topic Maintenance: Difficulty staying on topic
- Cohesion: Poor connection between ideas and sentences
- Referencing: Unclear or ambiguous references
- Narrative Structure: Problems with story organization
Understanding Results
Issue Highlighting
Each identified issue includes the specific problematic text in quotes, making it easy to locate and understand the context of language difficulties.
Severity Ratings
Issues are categorized by severity level to help prioritize intervention areas:
- Minor: Small errors that don't significantly impact communication
- Moderate: Noticeable issues that may affect understanding
- Significant: Major problems that clearly impair communication
Brief Explanations
Each issue includes a concise explanation (maximum 15 words) describing why the highlighted text is problematic, providing immediate clinical insight.
Clinical Overview
Results begin with a 1-2 sentence overview of the speech sample, providing context for the detailed analysis that follows.
Important Considerations
Credit Usage
Each analysis consumes one credit from your account. Make sure you have sufficient credits before running the analysis. Your remaining credit balance is displayed after each analysis.
AI-Powered Analysis
This tool uses advanced AI (GPT-4) to analyze language patterns. While highly accurate, results should be reviewed by qualified professionals and used as part of comprehensive assessment protocols.
Text Requirements
The tool works best with natural speech transcriptions from therapy sessions, assessments, or spontaneous speech samples. Maximum input length is 5,000 characters.
Clinical Focus
The analysis focuses on linguistic patterns only and does not make diagnostic statements or suggest specific therapy approaches. Results maintain clinical objectivity and professional standards.
Common Use Cases
- Language Assessment: Systematic analysis of speech samples during evaluations
- Progress Monitoring: Track improvement in specific language areas over time
- Treatment Planning: Identify priority areas for therapeutic intervention
- Documentation: Support clinical reports with objective language analysis
- Research: Analyze language patterns for clinical research purposes
- Training: Educational tool for students learning language analysis
How to Convert Speech to Text
The Speech to Text tool provides accurate transcription of speech recordings with detailed word-level confidence scores and timestamps. This powerful tool uses Microsoft's Speech SDK (Cognitive Services) for high-accuracy asynchronous speech recognition, supporting batch processing of large audio files across 30+ language locales.
How to Access
- Log into your Vocametrix account
- Click on Clinical Tools in the main menu
- Select Speech to Text Transcription from the Sound Analysis section
- Or visit /sound-analysis/speech-to-text directly
Getting Started
1. Select Your Language
Choose the appropriate language/locale from the dropdown menu. The tool supports 30+ locales across multiple languages with region-specific dialect recognition for optimal transcription accuracy.
2. Provide Audio to Transcribe
You have two options for providing audio:
- Record Audio: Record yourself speaking directly using your microphone
- Upload File: Upload an existing audio or video file (audio will be extracted from video files)
3. Process Transcription
Click "Transcribe Audio" to process your recording. The system will upload your audio and perform asynchronous speech recognition, typically completing within seconds to minutes depending on file length.
Analysis Features
Text Recognition Fields
Multiple forms of transcribed text for different use cases:
- Raw Recognition: Unprocessed recognized words with original confidence scores
- Inverse Text Normalized: Transformations for numbers, abbreviations, and standardized forms
- Display Text: Cleaned output with proper capitalization and punctuation for reading
- Lexical Form: Basic word-level transcription without formatting
Analysis Capabilities
Detailed analytical features for professional assessment:
- Word-Level Confidence: Individual confidence scores (0-100) for each recognized word
- Precise Timing: Millisecond-accurate word start and end times
- Accent Recognition: High accuracy across multiple regional accents
- Noise Resistance: Robust recognition algorithms that handle background noise
Interactive Results
Comprehensive visualization and interaction options:
- Editable Transcription: Edit the transcribed text with automatic statistics recalculation
- Word-Level Breakdown: View individual word confidence and timing information
- Alternative Recognition: Access multiple recognition alternatives when available
- Visual Charts: Statistical visualizations of confidence levels and recognition patterns
- Audio Waveform: Visual representation of the audio with synchronized playback
File Management
Flexible options for saving and reusing transcriptions:
- Download Results: Save complete transcription data as JSON file
- Copy Text: Copy transcribed text to clipboard for use in other applications
- Upload Previous Results: Re-upload saved transcriptions to view without using additional time
- Export Integration: Direct link to Syntax Analyzer for language analysis
Understanding Results
Confidence Scoring
Each word receives a confidence score from 0-100 indicating the system's certainty in the recognition. Higher scores indicate more reliable transcription. Color-coded visual indicators help quickly identify areas that may need review.
Timing Information
Precise word-level timestamps enable detailed analysis of speech patterns, speaking rate, and pause analysis. This information is valuable for speech therapy assessment and progress tracking.
Multiple Recognition Options
When the system generates multiple possible transcriptions, you can review alternative interpretations to ensure the most accurate result for your specific use case.
Statistical Analysis
Comprehensive statistics include total words, average confidence scores, duration analysis, and speaking rate calculations that provide insights into speech characteristics and quality.
Important Considerations
⏱️ Usage Time Consumption
This tool consumes usage time from your account based on the duration of your audio recording. Check your remaining usage time before processing longer recordings. Usage is calculated per second of audio transcribed.
🎤 Audio Quality Requirements
For optimal transcription accuracy, ensure clear audio recording with minimal background noise. Use a quality microphone and record in a quiet environment. The system is designed to be noise-resistant but performs best with clean audio.
🤖 Microsoft Speech SDK Technology
This tool uses Microsoft's Speech SDK (Cognitive Services) for high-accuracy asynchronous speech recognition. The system provides detailed word-level confidence scores, precise timing information, and supports batch processing of large audio files.
📁 File Format Support
The tool accepts both audio and video files in standard formats. For video files, audio is automatically extracted for transcription. Supported formats include WAV, MP3, MP4, and other common media file types.
Best Practices
- Clear Speech: Speak clearly and at a natural pace for optimal transcription
- Quiet Environment: Minimize background noise and use quality recording equipment
- Appropriate Language: Select the correct language/locale that matches your audio
- File Management: Save transcription results for future reference without re-processing
- Review Results: Check confidence scores and review low-confidence words
- Edit When Needed: Use the editable transcription feature to correct any errors
- Integration: Use "Check Syntax" feature for additional language analysis
Technical Specifications
Language Support
Comprehensive language coverage with region-specific recognition:
- 30+ Locales: Support for major world languages and regional variants
- Dialect Recognition: Region-specific accent and dialect handling
- Mixed Content: Support for mixed-language content in single recordings
Processing Method
Asynchronous processing for reliable handling of all file sizes:
- Batch Processing: Optimized for large audio files
- Secure Upload: Files uploaded to secure Azure Blob Storage
- Asynchronous Results: Background processing with result retrieval
- Error Handling: Robust error detection and reporting
Output Features
Comprehensive transcription data with multiple formats:
- Multiple Text Forms: Raw, normalized, and display-ready text
- Confidence Metrics: Word-level and overall confidence scoring
- Timing Data: Precise start/end timestamps for each word
- Alternative Results: Multiple recognition candidates when available
Common Use Cases
- Speech Therapy: Transcribe therapy sessions for detailed analysis and documentation
- Clinical Assessment: Convert speech samples to text for language evaluation
- Progress Monitoring: Track speech improvement over time with accurate transcriptions
- Research Documentation: Create searchable text records of research sessions
- Patient Records: Generate written records of spoken interactions and assessments
- Educational Materials: Convert audio lessons and presentations to text format
- Accessibility: Create text versions of audio content for hearing-impaired users
- Content Analysis: Prepare speech samples for syntax and language analysis tools
How to Convert Speech to Text in Real Time
The Real-Time Speech to Text tool provides instant transcription of spoken words as you speak, using advanced streaming technology. This feature is ideal for live captioning, accessibility, and interactive applications where immediate feedback is required.
How to Access
- Log into your Vocametrix account
- Click on Clinical Tools in the main menu
- Select Real-Time Speech to Text from the Sound Analysis section
- Or visit /sound-analysis/transcription directly
Getting Started
1. Select Your Language
Choose the appropriate language/locale from the dropdown menu. This ensures the transcription is tailored to your spoken language.
2. Start Real-Time Transcription
Click the Start Transcription button and begin speaking into your microphone. The tool will display your words on the screen in real time.
3. Stop and Review
Click Stop when finished. You can review, copy, or download the transcribed text for further use.
Key Features
Live Transcription
See your words appear instantly as you speak, with minimal delay.
Multi-Language Support
Supports a wide range of languages and dialects for global accessibility.
Export & Integration
Copy, download, or integrate the live transcript with other tools and workflows.
Best Practices
- Speak Clearly: For best results, speak at a natural pace and enunciate words.
- Use a Quality Microphone: Minimize background noise for higher accuracy.
- Check Language Setting: Ensure the selected language matches your speech.
Common Use Cases
- Live Captioning: Provide instant subtitles for meetings, lectures, or events.
- Accessibility: Assist users with hearing impairments by displaying spoken content as text.
- Interactive Apps: Enable voice-driven features in real-time applications.
- Clinical Documentation: Record spoken notes or patient interactions instantly.
How to Assess Pronunciation
The Pronunciation Assessment tool provides detailed analysis of speech pronunciation accuracy combined with pitch analysis. This advanced tool uses Microsoft's Speech SDK (Cognitive Services) for pronunciation assessment combined with acoustic analysis to evaluate pronunciation quality, fluency, and prosodic features for speech therapy and language learning applications.
How to Access
- Log into your Vocametrix account
- Click on Clinical Tools in the main menu
- Select Pronunciation Assessment from the Sound Analysis section
- Or visit /sound-analysis/pronunciation-assessment directly
Getting Started
1. Select Your Language
Choose the appropriate language/locale from the dropdown menu. This ensures the pronunciation assessment is calibrated to the phonetic rules and patterns of the target language.
2. Enter Reference Text
Type or paste the text that will be read aloud. This reference text will be used to compare against the recorded speech for accuracy assessment. Use clear, well-structured sentences appropriate for the speaker's level.
3. Record Audio
Click the record button and read the reference text aloud clearly. Ensure you're in a quiet environment with a good quality microphone for optimal results. The tool will capture your speech for analysis.
4. Process Assessment
Click "Analyze Pronunciation" to process your recording. The AI will compare your speech against the reference text and provide comprehensive feedback within seconds.
Analysis Features
Pronunciation Accuracy
Detailed assessment of how accurately each word and phoneme is pronounced:
- Word-Level Scores: Individual accuracy ratings for each word
- Phoneme Analysis: Breakdown of individual sound accuracy
- Overall Score: Comprehensive pronunciation rating
- Error Detection: Identification of mispronounced segments
- Confidence Levels: Assessment reliability indicators
Fluency Assessment
Evaluation of speech flow, timing, and natural rhythm:
- Speaking Rate: Words per minute analysis
- Pause Detection: Identification of hesitations and breaks
- Rhythm Patterns: Natural speech flow assessment
- Continuity Score: Overall fluency rating
- Disfluency Markers: Detection of repetitions and false starts
Pitch & Prosody Analysis
Advanced acoustic analysis of pitch patterns and prosodic features:
- Fundamental Frequency: Detailed pitch contour analysis
- Intonation Patterns: Rising and falling pitch patterns
- Stress Analysis: Word and syllable stress patterns
- Pitch Range: Voice frequency span assessment
- Prosodic Features: Natural melody and rhythm evaluation
Completeness & Timing
Assessment of how completely and appropriately the text was read:
- Coverage Score: Percentage of reference text spoken
- Word Alignment: Timing synchronization with reference
- Omission Detection: Identification of skipped words
- Insertion Analysis: Detection of extra words or sounds
- Duration Analysis: Speaking time assessment
Understanding Results
Scoring System
Microsoft's Speech SDK provides scores on a 0-100 scale using the "HundredMark" grading system:
- Accuracy Score: How correctly phonemes and words are pronounced
- Fluency Score: How closely speech flow matches natural patterns
- Completeness Score: How much of the reference text was spoken
- Prosody Score: Natural rhythm, stress, and intonation assessment
- Overall Pronunciation Score: Weighted combination of all metrics
Visual Feedback
Results include color-coded visual indicators for quick assessment: green for accurate pronunciation, yellow for moderate issues, and red for significant pronunciation errors. Pitch contours are displayed as interactive graphs showing frequency changes over time.
Detailed Word Analysis
Each word in the reference text receives individual scoring and feedback, allowing for targeted practice on specific pronunciation challenges. Phoneme-level breakdowns help identify exact sound production issues.
Confidence Indicators
The system provides confidence levels for its assessments, helping clinicians understand the reliability of each measurement and make informed decisions about areas requiring further evaluation.
Important Considerations
⏱️ Usage Time Consumption
This tool consumes usage time from your account based on the duration of your audio recording. Check your remaining usage time before processing longer recordings. Usage is calculated per second of audio analyzed.
🎤 Audio Quality Requirements
For optimal results, ensure clear audio recording with minimal background noise. Use a quality microphone and record in a quiet environment. Poor audio quality may affect assessment accuracy and reliability.
🤖 Microsoft Speech SDK Technology
This tool uses Microsoft's Speech SDK (Cognitive Services) for pronunciation assessment. The system provides comprehensive analysis including accuracy scores, fluency assessment, completeness evaluation, and prosody analysis with phoneme-level granularity.
🌍 Multi-Language Support
The tool supports multiple languages and dialects for global accessibility and accurate results.
Common Use Cases
- Speech Therapy: Detailed pronunciation analysis for therapy sessions
- Language Learning: Practice and feedback for language learners
- Clinical Assessment: Objective scoring for clinical documentation
- Progress Monitoring: Track improvement in pronunciation and fluency
- Research: Analyze prosodic and phonetic patterns in speech studies
- Education: Support for speech-language pathology training
How to Get Voice Metrics
The Praat Calculator provides comprehensive voice analysis using advanced acoustic algorithms. This powerful tool offers multiple voice metric calculations for clinical assessment, research, and voice quality evaluation, with both recording and file upload capabilities.
How to Access
- Log into your Vocametrix account
- Navigate to the Sound Analysis section from the main dashboard
- Select Praat Calculator from the available tools
- Or visit /sound-analysis/praat-calculator directly
Available Voice Metrics
Available to All Users
Acoustic Voice Quality Index for comprehensive voice quality assessment
Dysphonia Severity Index for voice disorder evaluation
Cepstral Peak Prominence for voice quality analysis
Fundamental frequency and amplitude perturbation analysis
Harmonics-to-Noise Ratio across frequency bands
Advanced frequency domain voice characteristics
Vowel formant frequency analysis and stability
Maximum phonation time comparison for vocal efficiency
Glottal-to-Noise Excitation ratio measurement
First and second harmonic amplitude difference
Comprehensive voice stability and variation analysis
How to Use the Praat Calculator
1. Choose Your Analysis Type
Each metric is presented in an expandable card format. Click on any calculator card to expand and access its features:
- AVQI: Requires both Connected Speech (CS) and Sustained Vowel (SV) recordings
- DSI: Uses manual input parameters plus audio analysis
- Single Audio Metrics: Most other calculators require one audio sample
- S/Z Ratio: Requires separate 'S' and 'Z' sound recordings
2. Provide Audio Input
For each analysis, you can provide audio in two ways:
- Click the recording button
- Allow microphone access
- Record your voice sample
- Recording stops automatically or manually
- Click the upload button
- Select your audio file
- Supported formats: WAV, MP3, M4A
- File uploads automatically
3. Configure Parameters (if needed)
Some calculators require additional parameters:
- AVQI: Select version (v02.03 or v03.01)
- DSI: Enter patient age, gender, maximum F0, MPT, and minimum intensity
- Recording Duration: Set maximum recording time for direct recording
4. Run the Analysis
Click the Calculate button for your chosen metric. The system will process your audio and display detailed results including:
- Primary metric values with interpretations
- Additional acoustic parameters
- Clinical interpretation guides (where applicable)
- Comparative normative data
Recording Guidelines
Best Practices for Quality Results
- Environment: Record in a quiet space with minimal background noise
- Distance: Maintain consistent distance from microphone (6-8 inches recommended)
- Sustained Vowels: Record steady 'ah' sound for 3-5 seconds minimum
- Connected Speech: Use standardized reading passages or spontaneous speech
- S/Z Tasks: Sustain 'sss' and 'zzz' sounds as long as possible
- Volume: Speak at comfortable, consistent volume
Understanding Results
Common Metrics Explained
Lower values indicate better voice quality (0-10 scale)
Higher values indicate better vocal function
Lower percentages indicate more stable voice
Higher dB values indicate clearer voice quality
Important Notes
- Credit Usage: Each calculation uses analysis credits - check your usage indicator
- File Formats: Ensure audio files are in supported formats (WAV, MP3, M4A)
- Analysis Time: Complex metrics may take 30-60 seconds to process
- Clinical Use: Results should be interpreted by qualified speech-language professionals
- Data Privacy: Audio files are processed securely and not stored permanently
How to Analyze Speech with GeMAPS Features
What is GeMAPS Analysis?
GeMAPS (Geneva Minimalistic Acoustic Parameter Set) is a standardized toolkit that extracts 62 research-grade acoustic features from speech recordings. These features are widely used in speech research for analyzing voice quality, emotional expression, and speech characteristics.
Step-by-Step Guide
Access the GeMAPS Analyzer
Navigate to Sound Analysis → GeMAPS Feature Analysis from the main dashboard.
Upload Your Audio File
Click "Upload Audio File" and select an audio file from your device. Supported formats include MP3, WAV, M4A, and OGG.
Start Analysis
Click "Analyze Audio" to begin the feature extraction process. The system will automatically process your audio using the GeMAPS toolkit.
Review Your Results
Once analysis is complete, you'll see:
- Summary Statistics: Total features extracted, analysis duration, and processing method
- Feature Categories: Organized by F0 (pitch), Energy, Spectral, Formant, and Voice Quality
- Statistical Summary: Key features with mean, min, max, and standard deviation values
Export Your Data
Click "Export to CSV" to download your results for further analysis in spreadsheet applications or research tools.
Understanding Your Results
Feature Categories
- F0 (Pitch): Fundamental frequency measurements (mean, variation, range)
- Energy: Loudness and intensity characteristics
- Spectral: Frequency distribution and spectral properties
- Formant: Vowel formant frequencies (F1, F2, F3)
- Voice Quality: Jitter, shimmer, and harmonic-to-noise ratio
Processing Methods
- Single Segment: Short files analyzed as one unit
- Chunked Analysis: Longer files split into 8-second segments
- Statistical Aggregation: Chunked results combined using mean, min, max, and std dev
- Feature Count: Shows total extractions (62 × number of chunks)
Tips for Best Results
- Audio Quality: Use clear recordings with minimal background noise
- Duration: At least 3 seconds required; 10+ seconds recommended for reliable features
- Content: Speech content works best; pure music or silence may produce limited results
- Research Use: Results are compatible with research literature using GeMAPS v01b standard
Learn More
For detailed technical information about GeMAPS features, visit the official GeMAPS research paper by Eyben et al. (2016).
How to Visualize Pitch
The Pitch Detector provides real-time visualization and analysis of fundamental frequency (F0) patterns. This powerful tool helps monitor vocal pitch in real-time or analyze pitch contours from audio files, making it ideal for voice training, speech therapy, and vocal assessment.
How to Access
- Log into your Vocametrix account
- Navigate to the Visualization section from the main dashboard
- Select Pitch Detector from the available tools
- Or visit /visualization/pitch directly
Analysis Modes
Live pitch detection and visualization from your microphone input
- Instant pitch feedback
- Continuous frequency monitoring
- Perfect for voice training
- Real-time pitch contour display
Upload and analyze pitch patterns from pre-recorded audio files
- Supports WAV, MP3, M4A formats
- Complete pitch contour analysis
- Audio playback with sync
- Detailed frequency timeline
How to Use Real-time Mode
1. Select Real-time Microphone Mode
Click the Real-time Microphone button to activate live pitch detection.
2. Start Pitch Detection
Click Start Detection to begin:
- Grant microphone access when prompted
- The system will start detecting pitch in real-time
- Current frequency displays in Hz
- Pitch contour appears on the graph
3. Monitor Pitch Display
The interface shows two key indicators:
Live Hz reading updates in real-time
Shows 🎤 Recording or ⏸️ Stopped
4. Stop Detection
Click Stop Detection to end the session and review the recorded pitch contour.
How to Use File Analysis Mode
1. Select Audio File Mode
Click the Audio File button to switch to file analysis mode.
2. Upload Your Audio File
In the upload section:
- Click to select or drag and drop your audio file
- Supported formats: WAV, MP3, M4A
- The system will automatically analyze the pitch
- Analysis may take a few moments for longer files
3. View Analysis Results
Once uploaded, you'll see:
- Complete pitch contour visualization
- Audio player with synchronized playback
- Time-frequency graph with detailed tooltips
- File status showing 📁 File Loaded
4. Interactive Playback
Use the audio player to navigate through your file while watching the pitch cursor move along the visualization in real-time.
Understanding the Visualization
Chart Components
Shows elapsed time in seconds
Displays pitch in Hz (10 Hz to maximum detected)
Represents the pitch contour over time
Help read precise frequency and time values
Interactive Features
Tooltip Display
Enable the Show Tooltip checkbox to:
- See exact frequency values when hovering over the chart
- View precise time stamps for any point
- Get detailed readings during analysis
Chart Navigation
- Hover over any point to see frequency and time details
- Chart automatically scales to show the full frequency range
- Time axis adjusts based on recording or file duration
Pitch Analysis Applications
Voice Training
- Monitor pitch accuracy during singing
- Practice sustained note production
- Work on pitch range expansion
- Develop pitch stability
Speech Therapy
- Assess pitch variation patterns
- Monitor intonation therapy progress
- Analyze prosodic features
- Track voice disorder recovery
Research & Analysis
- Study pitch contours in speech
- Analyze emotional expression
- Compare before/after recordings
- Document vocal changes
Music & Performance
- Tune musical instruments
- Practice pitch matching
- Analyze vocal vibrato
- Study performance techniques
Technical Specifications
10 Hz - 1500 Hz (fundamental frequency)
Real-time with ~50ms refresh intervals
WAV, MP3, M4A (uncompressed preferred)
Modern browsers with WebAudio API support
How to Visualize Sound Intensity
The Sound Level Meter provides real-time measurement and analysis of audio intensity and volume patterns. This tool monitors sound levels in dBFS and normalized values, making it essential for vocal intensity assessment, volume control training, and analyzing dynamic speech patterns.
How to Access
- Log into your Vocametrix account
- Navigate to the Visualization section from the main dashboard
- Select Sound Level Meter from the available tools
- Or visit /visualization/intensity directly
Analysis Modes
Live sound level detection and measurement from your microphone input
- Instant intensity feedback
- Continuous volume monitoring
- Perfect for vocal training
- Real-time intensity charts
Upload and analyze intensity patterns from pre-recorded audio files
- Supports WAV, MP3, M4A formats
- Complete intensity timeline analysis
- Audio playback with sync
- Detailed volume measurements
How to Use Real-time Mode
1. Select Real-time Microphone Mode
When you access the tool, real-time mode is typically the default setting for immediate sound level monitoring.
2. Start Sound Level Detection
Click Start Detection to begin:
- Grant microphone access when prompted
- The system will start measuring sound levels in real-time
- Current intensity displays in multiple formats
- Live chart shows intensity patterns over time
3. Monitor Sound Level Displays
The interface shows three key measurement types:
0-1 scale, real-time intensity measurement
Decibel full scale, professional audio standard
Shows 🎤 Recording or ⏸️ Stopped
4. Use Debug Mode (Optional)
Enable Debug Mode to view:
- Raw RMS (Root Mean Square) values
- Additional technical measurements
- Advanced debugging information
5. Stop Detection
Click Stop Detection to end the session and review the recorded intensity patterns.
How to Use File Analysis Mode
1. Select Audio File Mode
Click the Audio File Upload button to switch to file analysis mode.
2. Upload Your Audio File
In the upload section:
- Click to select or drag and drop your audio file
- Supported formats: WAV, MP3, M4A
- The system will automatically analyze sound levels
- Analysis may take a few moments for longer files
3. View Analysis Results
Once uploaded, you'll see:
- Complete intensity timeline visualization
- Audio player with synchronized playback
- Dual-axis chart showing normalized levels and dBFS
- File status showing 📁 File Loaded
4. Interactive Playback
Use the audio player to navigate through your file while watching the intensity cursor move along the visualization in real-time, showing precise intensity values at each moment.
Understanding the Visualization
Chart Components
Shows elapsed time in seconds
Displays intensity from 0 to 1
Shows decibel levels from -96 to 0 dBFS
Green for normalized, red for dBFS measurements
Interactive Features
Tooltip Display
Enable the Show Tooltip checkbox to:
- See exact intensity values when hovering over the chart
- View precise time stamps for any point
- Access normalized level, dBFS, and raw RMS values
Debug Mode
- Toggle debug mode to see raw RMS measurements
- Additional blue line appears on chart for technical analysis
- Access advanced measurement data in tooltips
Sound Intensity Applications
Voice Training
- Monitor vocal volume consistency
- Practice dynamic range control
- Work on crescendo and diminuendo
- Develop breath support awareness
Speech Therapy
- Assess vocal intensity patterns
- Monitor voice volume disorders
- Track loudness therapy progress
- Analyze speech prominence patterns
Research & Analysis
- Study intensity variation in speech
- Analyze emotional expression levels
- Compare volume patterns over time
- Document vocal intensity changes
Audio Production
- Monitor recording levels
- Ensure consistent volume
- Avoid clipping and distortion
- Optimize dynamic range
Technical Specifications
0-1 normalized, -96 to 0 dBFS
Real-time with continuous monitoring
WAV, MP3, M4A (uncompressed preferred)
Normalized Level, dBFS, Raw RMS
Understanding Sound Level Measurements
Measurement Definitions
Relative intensity scale where 0 is silence and 1 is maximum detected level
Professional audio measurement where 0 dBFS is the maximum digital level
Technical measurement of signal power, useful for advanced analysis
How to Visualize Audio Waveforms
The Audio Waveform Visualizer provides detailed visual representation of audio signal amplitude over time. This tool displays the classic waveform view that shows the shape and structure of audio signals, making it ideal for analyzing speech patterns, audio quality, timing, and overall signal characteristics.
How to Access
- Log into your Vocametrix account
- Navigate to the Visualization section from the main dashboard
- Select Waveform Display from the available tools
- Or visit /visualization/waveform directly
Analysis Modes
Record audio directly from your microphone and visualize the waveform as it's captured
- Live audio recording
- Immediate waveform generation
- Perfect for speech analysis
- Interactive playback controls
Upload existing audio files to view and analyze their waveform patterns
- Supports WAV, MP3, M4A formats
- Complete waveform visualization
- Interactive audio playback
- Zoom and navigation controls
How to Use Real-time Recording Mode
1. Select Real-time Recording Mode
Click the Real-time Recording button to enable live audio capture and waveform generation.
2. Start Recording
Click Start Recording to begin:
- Grant microphone access when prompted
- The system will start capturing audio in real-time
- Recording status shows 🔴 Recording
- Audio is buffered for waveform generation
3. Monitor Recording Status
The interface displays current recording information:
Shows 🔴 Recording or ⏸️ Stopped
Displays 🎤 Real-time mode indicator
4. Stop Recording and View Waveform
Click Stop Recording to:
- Complete the recording session
- Generate the full waveform visualization
- Enable playback controls
- Access interactive waveform features
How to Use File Upload Mode
1. Select Audio File Upload Mode
Click the Audio File Upload button to switch to file analysis mode.
2. Upload Your Audio File
In the upload section, you have two options:
- Click to select or drag and drop your file
- Supports WAV, MP3, M4A formats
- File automatically loads into waveform display
- Use built-in audio recorder
- Maximum 60-second recordings
- Instantly creates waveform visualization
3. View Waveform Analysis
Once uploaded, you'll see:
- Complete audio waveform visualization
- Interactive audio player with playback controls
- Time and amplitude information
- Status showing ✅ Audio Loaded
4. Interactive Waveform Navigation
Use the interactive features to explore your audio:
- Click anywhere on the waveform to jump to that position
- Use playback controls to play/pause audio
- Visual cursor shows current playback position
- Time display shows current and total duration
Understanding the Waveform Display
Waveform Components
Shows audio duration from start to finish
Displays audio signal strength and volume
Visual representation of audio signal over time
Shows current playback position
Interactive Playback Features
Playback Controls
Start or stop audio playback
Return to beginning of audio
Visual volume indicator
Navigation Tips
- Press Space to play/pause audio
- Click anywhere on the waveform to seek to that position
- Use the time display to monitor current position
- Visual playback cursor follows audio progress
Waveform Analysis Applications
Speech Analysis
- Identify speech patterns and pauses
- Analyze speech rhythm and timing
- Detect speech segments and silences
- Study articulation clarity
Voice Training
- Monitor voice consistency
- Analyze breath control patterns
- Study vocal attack and release
- Compare different vocal exercises
Audio Quality Assessment
- Detect audio clipping and distortion
- Identify background noise patterns
- Analyze dynamic range
- Assess recording quality
Research & Documentation
- Document audio characteristics
- Compare different recordings
- Study temporal speech patterns
- Analyze conversation dynamics
Technical Specifications
WaveSurfer.js for interactive waveforms
WebAudio API for high-quality processing
WAV, MP3, M4A (WAV recommended)
60 seconds maximum for built-in recorder
Understanding Waveform Patterns
Reading Waveform Characteristics
Loud sounds, stressed syllables, or consonant bursts
Quiet sounds, vowels, or background noise
Silence, pauses, or gaps between words
Complex sounds, fricatives, or multiple speakers
How to Visualize Vowels on an IPA Chart
Accessing the IPA Voice Mapping Tool
- In the main navigation, click Clinical Tools.
- Select IPA Voice Mapping from the dropdown menu.
- Alternatively, you can go directly to /visualization/ipa-voice-mapping in your browser.
Mapping Your Vowel
- Select your language from the dropdown at the top of the page. This determines the reference vowels shown on the chart.
- Choose your gender (male/female) and age group (adult/child) using the settings controls. The chart and analysis will adapt to your selection.
- Record a new vowel sample or upload an audio file using the provided controls.
- After recording or uploading, click Analyze to start the formant analysis.
- The system will process your audio and display your vowel’s position on the IPA chart, based on the measured F1 and F2 values.
- Review the chart to see where your vowel falls, with coordinates representing tongue height (close/open) and tongue position (front/back).
Interpreting the Results
- The chart highlights your vowel’s location as a point within the IPA grid.
- Hover or tap on the point to view the measured F1 and F2 values.
- Use this visualization to compare your production to typical vowel targets for clinical or educational feedback.
How to Use the AI Assistant
The Vocametrix AI Assistant is your specialized speech therapy companion, providing expert guidance tailored to your specific role and needs. Whether you're a therapist, patient, or caregiver, the AI adapts its responses to give you the most relevant and helpful information.
How to Access
- Log into your Vocametrix account
- Navigate to the AI section from the main dashboard
- Select AI Assistant from the available tools
- Or visit /ai/chat directly
User Roles & Personalization
The AI Assistant automatically personalizes responses based on your account type. You don't need to specify your role - the system handles this automatically to give you the most relevant guidance:
Professional clinical support and advanced guidance (automatically detected)
- Clinical analysis and disorder identification
- Evidence-based intervention strategies
- Assessment recommendations and progress tracking
- Professional development guidance
- Case consultation and therapy planning
- Research insights and best practices
Accessible explanations and encouragement (automatically detected)
- Simple explanations of speech therapy concepts
- Home practice tips and motivation
- Basic information about your condition
- Questions to ask your therapist
- Accessible resources and tools
- Progress celebration and encouragement
Family support and home practice strategies (automatically detected)
- Home therapy support strategies
- Daily activity integration tips
- Speech development information
- Communication with therapy team
- Progress tracking guidance
- Age-appropriate practice activities
How to Use the AI Assistant
1. Simply Ask Your Question
Just type your question directly - no special prefixes needed. The system automatically personalizes responses based on your account type:
"slt asks: How do I assess voice mutation in adolescents?""patient asks: What exercises can I do at home for my /r/ sounds?""parent asks: How can I help my child practice speech therapy at home?"
2. Provide Context
Share relevant details about your situation for more accurate guidance:
- Age and developmental stage
- Specific symptoms or challenges
- Current therapy goals or objectives
- Cultural or linguistic background
3. Ask Specific Questions
Be clear about what guidance you need for the most helpful responses:
- Request specific intervention strategies
- Ask for explanation of concepts
- Seek home practice recommendations
- Request progress tracking advice
4. Build on Responses
Continue the conversation for deeper guidance and clarification:
- Ask follow-up questions for clarification
- Request modifications for specific situations
- Seek additional resources or strategies
- Provide feedback on suggested approaches
Key Features
Context-Aware Conversations
- Remembers previous messages
- Builds on earlier information
- Acknowledges user corrections
- Maintains conversation continuity
Platform Integration
- Works with other Vocametrix tools
- Can analyze platform files and reports
- Suggests relevant features
- Integrates with progress tracking
Evidence-Based Guidance
- Proven intervention strategies
- Professional guidelines
- Standardized clinical protocols
- Best practice recommendations
Cultural Sensitivity
- Adapts to different backgrounds
- Multi-language support
- Inclusive guidance
- Cultural considerations
Safety & Professional Guidelines
Important Limitations
- Does not provide medical diagnoses
- Emphasizes consulting healthcare professionals
- Maintains confidentiality standards
- Redirects emergencies to appropriate care
- Suggests types of resources rather than specific titles
- Recommends general categories of materials
- Directs to professional associations
- Avoids inventing specific references
Best Practices for Effective Conversations
✅ Do:
- Use the correct role prefix for personalized responses
- Provide specific context about age, symptoms, and goals
- Build on previous conversation points
- Ask follow-up questions for clarification
- Mention cultural or linguistic considerations
❌ Avoid:
- Asking for medical diagnoses or emergency care
- Expecting specific book or resource recommendations
- Repeating information already provided
- Using vague descriptions without context
- Asking about technical AI system details
How to Generate Word Lists
The AI Word Generator creates targeted speech therapy word lists with hints for specific IPA sounds and positions. This tool helps therapists and caregivers create engaging practice materials tailored to individual needs and developmental levels.
How to Access
- Log into your Vocametrix account
- Navigate to Generate Word Lists from the main dashboard
- Or visit the word generator page directly
Features & Capabilities
Multi-Language Support
Generate word lists in multiple languages:
- English - General American pronunciation
- French - Standard French pronunciation
- Spanish - Standard Spanish pronunciation
- German - Standard German pronunciation
IPA Sound Targeting
Select from comprehensive IPA sound categories:
- Consonants - Organized by articulation type (stops, fricatives, nasals, etc.)
- Vowels - Grouped by tongue position and mouth opening
- Diphthongs - Complex vowel combinations
- Position Control - Beginning, middle, end, or random placement
Age-Appropriate Content
Content automatically adapts to developmental levels:
- Young Children (Under 8) - Simple, concrete nouns and basic verbs
- Older Children/Adults - Complex vocabulary and advanced concepts
- Educational Focus - Hints promote learning and engagement
How to Generate Word Lists
1. Select Language
Choose your target language from the dropdown menu. All generated words and hints will be in the selected language.
2. Specify Age and Development Level
Enter the patient's age and developmental stage for appropriate word complexity:
"5 years old""8 years old, beginner level""Adult, intermediate"
3. Choose Target IPA Sound
Browse the comprehensive IPA sound library organized by categories:
- Each sound shows the IPA symbol, description, and example
- Select one position: beginning, middle, end, or random
- Only one sound and position can be selected at a time
4. Generate and Review
Click "Generate Word List" to create 20 targeted words with descriptive hints for therapy practice.
Understanding the Results
Word List Format
The generator provides two formats for maximum flexibility:
- Simple List - Semicolon-separated words for quick reference
- Words with Hints - Table format with descriptive clues for each word
Hint Quality
Each hint is crafted to be educational and engaging:
- 6-12 words forming complete sentences
- Specific details without being too obvious
- Age-appropriate language and concepts
- Perfect for word guessing games and exercises
Copy and Export Options
Use the copy buttons to easily transfer word lists to other applications or share with colleagues and families.
Sound Position Examples
Beginning Position (/s/)
Words like "sun", "sit", "soap" - target sound starts the word
Middle Position (/s/)
Words like "pencil", "lesson", "glasses" - target sound in middle syllables
End Position (/s/)
Words like "house", "dress", "bus" - target sound ends the word
Random Position (/s/)
Mix of all positions for varied practice opportunities
Therapeutic Applications
Practice Activities
- Articulation drilling with target sounds
- Word guessing games using hints
- Reading comprehension exercises
- Vocabulary building activities
- Home practice assignments
Assessment Tools
- Baseline pronunciation assessment
- Progress tracking with consistent word sets
- Difficulty level adjustments
- Cross-linguistic comparison studies
Credit Usage
Each word list generation uses 1 credit from your account balance. Monitor your remaining credits with the indicator at the top of the page.
- Check credit balance before generating
- Purchase additional credits as needed
- Each generation produces exactly 20 words with hints
Best Practices
✓ Specify exact age and development level for most appropriate word complexity
✓ Generate multiple lists for the same sound to provide variety in practice sessions
✓ Use hints creatively for games, comprehension activities, and engagement
✓ Copy lists to external apps for session planning and family communication
Technical Specifications
AI Model: Advanced language model trained on speech therapy methodologies
Output: Exactly 20 words per generation with matching hints
Languages: English, French, Spanish, German with native pronunciation standards
Validation: Automatic verification of sound position accuracy
Age Adaptation: Content complexity automatically adjusted for developmental stages
How to Generate Speech Exercises
The AI Exercise Generator creates personalized speech therapy exercises tailored to specific age groups, speech challenges, and languages. This tool helps therapists design targeted practice sessions and provides caregivers with structured activities for home practice.
How to Access
- Log into your Vocametrix account
- Navigate to the AI section from the main dashboard
- Select Exercise Generator from the available tools
- Or visit /ai/exercisegenerator directly
Features & Capabilities
Personalized Exercise Generation
Creates exactly 5 targeted exercises based on:
- Age/Development Level - From toddlers to adults
- Speech Challenge - Specific articulation, fluency, or voice issues
- Target Language - English, French, or Spanish support
- Progressive Difficulty - Exercises build from easy to challenging
Comprehensive Exercise Types
Generates various therapeutic exercise formats:
- Warm-up Exercises - Breathing and mouth positioning
- Sound Isolation - Focus on specific problematic sounds
- Word-Level Practice - Single words with target sounds
- Phrase/Sentence Practice - Connected speech development
- Functional Practice - Real-world application scenarios
Age-Appropriate Content
Automatically adapts content and complexity:
- Young Children (3-6) - Play-based exercises with fun themes
- School-Age (7-12) - Interactive activities and educational games
- Teenagers (13-18) - Practical social situations and confidence-building
- Adults - Professional contexts and functional communication
How to Generate Exercises
1. Specify Patient Information
Provide detailed information about the patient for personalized exercises:
- Age/Level:
"5 years old","Adult beginner" - Speech Challenge:
"R sound difficulties","Stuttering" - Language: English, French, or Spanish
2. Describe Specific Goals
Add any specific requirements or focus areas in your message:
"Focus on /th/ sounds at the beginning of words""Include exercises for home practice with parents""Need activities for group therapy sessions"
3. Generate and Review
Click "Generate Exercises" to receive 5 progressive therapy activities with detailed instructions, examples, and expected outcomes.
Understanding Exercise Structure
Exercise Components
Each generated exercise includes comprehensive guidance:
- Title - Clear, descriptive name for the exercise
- Instructions - Step-by-step directions for implementation
- Example - Concrete demonstrations or sample words/phrases
- Difficulty Level - Easy, Medium, or Hard classification
- Expected Outcomes - Therapeutic goals and progress indicators
Progressive Difficulty
Exercises are designed to build skills systematically:
- Exercise 1-2 - Foundation building and warm-up activities
- Exercise 3-4 - Skill development and practice consolidation
- Exercise 5 - Advanced application and real-world practice
Speech Challenge Categories
Articulation Challenges
R sounds, L sounds, TH sounds, S sounds, and other consonant/vowel difficulties
Fluency Issues
Stuttering, speech rhythm problems, pacing irregularities
Voice Disorders
Volume control, pitch variation, vocal clarity, breathiness
Language Development
Vocabulary building, sentence structure, comprehension skills
Language-Specific Features
English Exercises
- Focus on common English phonemes and sound patterns
- Age-appropriate vocabulary and cultural references
- Emphasis on problematic sounds like /r/, /l/, /th/
French Exercises
- Instructions and examples provided in French
- Focus on French-specific sounds (nasal vowels, uvular R)
- Cultural context appropriate for French-speaking patients
Spanish Exercises
- Native Spanish instructions and therapeutic content
- Focus on Spanish phonemes (rolled R, specific vowel sounds)
- Age-appropriate Spanish vocabulary and cultural references
Therapeutic Applications
Clinical Settings
- Individual therapy session planning
- Group therapy activity development
- Progress assessment and skill tracking
- Treatment plan documentation
Home Practice
- Parent/caregiver guided activities
- Independent practice routines
- Daily speech maintenance exercises
- Family engagement strategies
Educational Support
- School-based intervention activities
- Classroom communication practice
- Peer interaction exercises
- Academic presentation skills
Credit Usage & Threading
Each exercise generation uses 1 credit from your account. The system maintains conversation threads for follow-up requests and refinements.
- Monitor credit balance before generating exercises
- Threads preserve context for related follow-up requests
- Ask for modifications or additional exercises in the same session
- Each thread maintains conversation history for better personalization
Safety & Professional Guidelines
⚠ Professional Oversight - All generated exercises should be reviewed by qualified speech-language pathologists
⚠ Individual Assessment - Ensure exercises are appropriate for each patient's specific needs and abilities
⚠ Safety First - All exercises are designed to be safe, but supervision may be required for certain patients
Technical Specifications
AI Model: Specialized speech therapy exercise generation system
Output Format: Structured JSON with 5 comprehensive exercises
Languages: English, French, Spanish with native-speaker content
Threading: Conversation continuity for follow-up requests and modifications
Personalization: Age-appropriate content and difficulty adaptation
Therapeutic Focus: Evidence-based exercise types and progression patterns
How to Generate Therapeutic Images
The AI Image Generator creates customized therapeutic images for speech therapy sessions. Generate educational illustrations, patient exercises, and visual aids with professional formatting options including therapist information, borders, and print-ready layouts for clinical use.
How to Access
- Log into your Vocametrix account
- Navigate to the AI section from the main dashboard
- Select Image Generator from the available tools
- Or visit /ai/image-generator directly
Features & Capabilities
AI-Powered Image Generation
Create custom therapeutic images with advanced AI technology:
- Natural Language Prompts - Describe what you want in plain English
- Therapy-Focused Content - Optimized for speech therapy and educational use
- Quick Suggestions - Pre-built prompts for common therapeutic scenarios
- 500-Character Limit - Encourages clear, focused image descriptions
- Content Safety - Built-in content policy validation for appropriate images
Professional Customization
Personalize images for your practice and patients:
- Patient Information - Add patient names with show/hide toggle
- Exercise Titles - Label images with specific exercise names
- Instructions - Include detailed therapy instructions
- Date Stamps - Automatic or manual date inclusion
- Therapist Settings - Practice name, therapist name, contact info, and logo
Image Styles & Borders
Choose the perfect visual style for your therapeutic materials:
- Drawing Style 🎨 - Illustrated, cartoon-like images perfect for therapy
- Photo Realistic 📷 - Realistic, photograph-style images
- Border Options - None, Simple, Colorful, Waves, or Stars borders
- Professional Layout - Print-ready formatting with proper spacing
- Logo Integration - Upload and include your practice logo
How to Generate Images
1. Write Your Image Description
In the "Image Description" box, describe what you want to see in the image:
- Be specific: "A red apple on a wooden table" vs. "fruit"
- Include therapy context: "Speech therapy exercise with mouth positions"
- Mention style preferences: "Simple line drawing" or "colorful illustration"
- Use quick suggestions: Click pre-made prompts for common scenarios
2. Add Patient Information (Optional)
Personalize the image for specific patients:
- Patient Name: Enter the patient's name (can be hidden from image)
- Exercise Title: Name the specific therapy exercise
- Instructions: Add detailed instructions for the exercise
- Visibility Toggle: Use the eye icon to show/hide patient name on the image
3. Choose Visual Style
Select the appropriate style and formatting:
- Image Style: Drawing (recommended for therapy) or Photo Realistic
- Border Style: Choose from 5 border options (None, Simple, Colorful, Waves, Stars)
- Date Option: Include current date on the image
- Therapist Info: Include your practice information footer
4. Generate and Review
Generate your image and review the results:
- Generate Image: Click the "Generate Image" button (uses 1 credit)
- Review Preview: Check the generated image in the preview panel
- Regenerate if needed: Modify your description and try again
- Real-time Preview: See patient name, date, and title overlays
How to Set Up Language Learning Plans
The AI Language Learning Setup provides multilingual conversation coaching with real-time vocabulary assistance and pronunciation feedback. Configure your native and target languages, select conversation topics, and engage with specialized AI agents that provide vocabulary tips in your native language while maintaining immersion in your target language.
How to Access
- Log into your Vocametrix account
- Navigate to the AI section from the main dashboard
- Select Language Learning Setup from the available tools
- Or visit /ai/language-learning-setup directly
Features & Capabilities
Dual-Agent Language Learning System
Two specialized AI agents provide comprehensive language learning:
- Vocabulary Coach Agent - Real-time vocabulary improvement with native language explanations
- Pronunciation Coach Agent - Audio feedback and pronunciation guidance
Vocabulary Building Mode
Interactive conversations with real-time vocabulary enhancement:
- Native Language Explanations - Vocabulary tips provided in your first language
- Target Language Immersion - Main conversation conducted in learning language
- Alternative Expressions - Learn synonyms and natural phrasing
- Contextual Learning - New words introduced within conversation flow
- Cultural Context - Language use appropriate to cultural settings
Pronunciation Coaching Mode
Specialized pronunciation feedback and practice:
- Audio Assessment Integration - Works with Azure Speech Services for accuracy scoring
- Progress Tracking - Assessment history monitoring across sessions
- Language-Specific Challenges - Targeted practice for difficult sounds (French R, Spanish RR)
- Real-time Feedback - Immediate pronunciation guidance and corrections
- Fluency & Prosody Analysis - Beyond accuracy to natural speech patterns
Conversation Topics
Practice with 9 focused topic areas for targeted language learning:
- Family & Relationships 👨👩👧👦 - Personal connections and family matters
- Travel & Adventure ✈️ - Tourism, transportation, and exploration
- Food & Cooking 🍽️ - Cuisine, recipes, and dining experiences
- Work & Career 💼 - Professional contexts and workplace communication
- Hobbies & Interests 🎨 - Personal interests and recreational activities
- Health & Wellness 🏥 - Medical topics and health-related conversations
- Shopping & Services 🛍️ - Commerce, services, and transactions
- Education & Learning 📚 - Academic topics and learning experiences
- Free Conversation 💬 - Open-ended discussions on any topic
How to Set Up Your Learning Session
1. Choose Your Native Language
Select your first language - vocabulary tips and explanations will be provided in this language:
- English 🇺🇸 - Tips provided in English
- French 🇫🇷 - "Conseils de vocabulaire" et "Nouveaux mots pour vous"
- Spanish 🇪🇸 - "Consejos de vocabulario" y "Nuevas palabras para ti"
- German 🇩🇪 - "Vokabeltipps" und "Neue Wörter für Sie"
2. Select Your Target Language
Choose the language you want to learn - all main conversation will be in this language:
- Same 4 language options: English, French, Spanish, German
- AI coach responds primarily in this language for immersion
- Target language can be different from your native language
3. Set Age/Proficiency Level
Choose your skill level for appropriate content complexity:
- Child-Beginner: Simple words, basic grammar, fun examples
- Teen-Intermediate: Everyday vocabulary, modern expressions, social situations
- Adult-Advanced: Complex vocabulary, professional terms, cultural nuances
- Content adapts to your specified level automatically
4. Pick a Conversation Topic
Select a focused topic area for targeted vocabulary and context:
- 9 topic options: family, travel, food, work, hobbies, health, shopping, education, free
- Topic influences vocabulary focus and conversation scenarios
- Free conversation allows open-ended practice on any subject
5. Choose Learning Mode
Select your focus and corresponding AI agent:
- Vocabulary Building: Uses Vocabulary Coach Agent for real-time language tips
- Pronunciation Practice: Uses Pronunciation Coach Agent for audio-focused training
- Each mode has specialized training and response formats
How to Play Games
Vocametrix offers a collection of voice-responsive speech therapy games that react directly to voice and sounds, making practice engaging and effective while targeting specific speech skills with real-time or recorded audio feedback.
🎮 Game Categories
Speech Production Games
Practice basic sound production and volume control
Monster Dodge
Purpose: Practice making sounds to avoid obstacles
How it works: Use voice volume to move your character up/down
Balloon Game
Purpose: Control voice volume and breath support
How it works: Keep balloon inflated by maintaining sound above threshold
Pronunciation & Reading Games
Practice clear speech and word pronunciation
Repeat It
Purpose: Practice speech repetition with visual feedback
How it works: Listen to words/phrases and repeat them clearly
Explode the Words
Purpose: Practice pronunciation accuracy
How it works: Say words correctly to trigger explosions
Spell Master
Purpose: Practice spelling words letter by letter
How it works: Speak each letter clearly to spell words
Voice Control & Prosody Games
Master pitch, rhythm, and voice quality
Crystal Resonance
Purpose: Practice pitch matching and control
How it works: Match and maintain target pitch frequency
Match Rhythm
Purpose: Practice speech timing and rhythm
How it works: Speak in sync with visual rhythm patterns
Airplane Game
Purpose: Control voice level to navigate
How it works: Use voice level to fly airplane through corridors
Special & External Games
Creative activities and external tools
Treasure Hunt Designer
Purpose: Create interactive voice-based treasure hunts
How it works: Design custom riddles and voice challenges (External)
🚀 How to Get Started
Choose Your Game
Select a game from the Games section that matches your therapy goals
Configure Settings
Adjust game parameters to match your skill level and needs
Practice & Play
Follow game instructions and practice your speech skills
Track Progress
Review scores and improvements over time
⚙️ Common Game Features
Microphone Setup
All games require microphone access for voice input
Real-time Feedback
Immediate visual and audio feedback on your performance
Adaptive Difficulty
Games can be customized to your skill level
Session Management
Control game duration and track your practice time
How to Join the Therapist Waiting List
As a patient or parent, you can join the waiting list to be contacted by qualified speech therapists. This feature helps connect you with professional therapists who can provide personalized speech therapy services.
How to Access
- Log into your Vocametrix account
- You can access this feature in two ways:
- Via Profile: Go to Profile → Find Therapist tab
- Direct link: Navigate to
/find-therapist
Account Requirements: Only users with "Patient" or "Parent" account types can join the waiting list. Speech therapists and visitors cannot use this feature.
Step-by-Step Process
Step 1: Language & Problem Description
Select Your Language: Choose the language you speak from the dropdown menu. This helps match you with therapists who speak your language.
Describe Your Problem: Provide a detailed description of your speech difficulties, including:
- What specific speech issues you're experiencing
- How long you've had these difficulties
- Any treatments or therapies you've already tried
- Your goals for speech therapy
Privacy Note: Do NOT include your name, address, or any personally identifiable information. This description will be shared anonymously with therapists.
Step 2: Audio Samples (Optional but Recommended)
Providing audio samples helps therapists better understand your speech patterns. You can record three types of samples:
Reading Sample
Read aloud a provided text in your selected language (up to 60 seconds)
Conversation Sample
Speak freely about your likes, dislikes, and hobbies (up to 60 seconds)
Sustained Vowel Sample
Sustain the "ahhh" sound for about 5 seconds (up to 10 seconds recording)
Step 3: Geographic Information
Provide your location to help match you with local therapists:
- Country (required)
- State/Province (optional)
- City (optional)
This information helps connect you with therapists who may offer in-person sessions in your area.
Step 4: Consent & Review
Data Sharing Consent: You must agree to share your anonymous information with speech therapists.
Review Your Information: Check all details including language, problem description, audio samples, and location.
Searching Status: Confirm whether you are actively looking for a speech therapist.
How Contact Works
- Therapist Reviews: Qualified speech therapists can browse anonymous profiles of patients/parents seeking help
- Interest Expression: If a therapist believes they can help, they contact Vocametrix (not you directly)
- Credential Verification: Vocametrix verifies the therapist's credentials and qualifications
- Facilitated Introduction: Vocametrix facilitates a proper introduction via email, sharing both parties' contact information
- Direct Communication: You and the therapist can then communicate directly to arrange services
Privacy & Safety
What Therapists See
- Your problem description (anonymous)
- Your audio samples
- Your language preference
- Your general location (country/state/city)
What Therapists Don't See
- Your name, email address, or phone number
- Any personally identifiable information
- Your exact address or specific location details
Managing Your Information
Updating Your Profile
You can update your information at any time by returning to the Find Therapist section and modifying your details.
Changing Search Status
You can toggle "I am actively looking for a speech therapist" on or off if your situation changes.
Deleting Your Data
If you no longer need the service, you can delete all your therapist search data from your profile.
Important Notes
- No Guarantees: Being on the waiting list doesn't guarantee you'll be contacted by a therapist
- Therapist Availability: Contact depends on therapist availability and their assessment of whether they can help with your specific needs
- Professional Service: Any therapy services arranged are between you and the therapist directly; Vocametrix facilitates the introduction only
- Free Service: There is no cost to join the waiting list or be connected with therapists
- Account Required: You must have a Vocametrix account with "Patient" or "Parent" account type