Skip to content

PairaVoice Support

Find answers to common PairaVoice questions below. If you still need help, submit a support request and our team will follow up.

Search the FAQs to quickly find answers about PairaVoice features, settings, and common issues.

 

No matching FAQs found. Try a different keyword, or submit a support request using the form on this page.

General

What is PairaVoice?

PairaVoice is a real-time voice translation app built for healthcare, education, patients, families, and professionals who need secure multilingual communication. It supports live voice translation, text transcription, AI note-taking, and professional workflows such as SOAP note generation in PairaVoice Pro.

Who is PairaVoice for?

PairaVoice is designed for healthcare providers, educators, patients, students, families, and advocates. Common use cases include patient consultations, clinical rounds, EMS encounters, parent-teacher conferences, IEP meetings, ESL support, lectures, meetings, and multilingual family communication.

What makes PairaVoice different from a consumer translation app?

PairaVoice is built for professional conversations, not just casual travel translation. It combines live voice translation, multi-engine AI translation, transcription, voice playback, optional noise cancellation, hands-free mode, single or dual-device use, and security features designed for healthcare and education workflows. PairaVoice Pro also adds saved searchable transcripts, SOAP notes, and HIPAA / FERPA-focused capabilities.

Is PairaVoice available on mobile?

Yes. PairaVoice is a native iOS and Android translation app.

Does PairaVoice replace human interpreters?

PairaVoice can help reduce wait times and costs for many routine conversations, especially when instant translation is needed. For high-risk, legally sensitive, emergency, or highly nuanced conversations, organizations may still choose to involve a qualified human interpreter depending on their policies and regulatory requirements.

Best Practices for a Smooth Conversation

How should I speak for the best translation quality?

Speak clearly, use short complete sentences, pause between speakers, and avoid talking over the other person. Confirm important details by repeating them back, especially names, dates, medications, instructions, numbers, and follow-up steps.

How long should each speaking turn be?

Short turns usually work best. One to three sentences at a time is ideal for live conversations. For complex explanations, break the message into smaller parts and let PairaVoice translate each part before continuing.

What should I do before starting an important conversation?

Choose the source and target languages, select the best mode, test the microphone, reduce background noise, and tell both participants to speak one at a time. For sensitive healthcare or education conversations, confirm whether transcripts should be saved according to your organization’s policy.

How can I improve translation for technical topics?

Follow our suggested settings below, speak in plain language when possible, avoid idioms, and define specialized terms. For medical topics, say drug names, dosages, symptoms, and instructions slowly. For education topics, spell or repeat names of programs, forms, services, or acronyms.

What should I avoid?

Avoid side conversations, overlapping speech, slang, idioms, sarcasm, long monologues, and switching languages mid-sentence. These can reduce recognition and translation quality.

How should users handle names, addresses, and numbers?

Speak names, addresses, phone numbers, dates, times, dosages, and quantities slowly. Repeat them if they are important. For critical information, verify the translated text visually before moving on.

Languages and Accuracy

How accurate is PairaVoice?

PairaVoice uses three industry-leading AI engines and targets 95%+ accuracy.

Why can accuracy vary?

Accuracy can vary depending on the language pair, speaker accent, background noise, microphone quality, speaking speed, medical or technical terminology, and whether the conversation is casual or specialized. Short, clear turns usually produce better results than long, overlapping, or interrupted speech.

Which language pairs work best?

Common high-resource language pairs, such as English-Spanish, typically perform very well. For less common languages, regional dialects, or highly specialized topics, users may get better results by testing multiple engines and both modes available.

Translation Engines and Voice Recognition

What is the Voice Recognition Model?

The Voice Recognition Model is the technology that listens to the speaker and converts spoken words into text.

In PairaVoice, the available voice recognition options are: OpenAI, ElevenLabs, and Google.

This setting affects how well PairaVoice understands speech, accents, pronunciation, and audio quality.

Which Voice Recognition Model should I choose?

The best voice recognition model may depend on the language, accent, speaker, and environment.

As a general guide:

  • ElevenLabs delivered the best overall performance in our testing. It is the recommended option when high-quality voice recognition, natural audio output, and a smooth real-time conversation experience are priorities.
  • OpenAI is a good choice for natural conversations and context-heavy speech, particularly when understanding meaning across longer exchanges is important.
  • Google remains a good option for broad language support and both general and specialized speech recognition.

Although ElevenLabs performed best overall in our tests, results may still vary depending on the language pair, speaker, accent, audio quality, and conversation environment.

What is the Translation Model?

The Translation Model is the engine that translates the recognized text into the target language.

In PairaVoice, the available translation options are: OpenAI, DeepL, and Google.

This setting affects the translation’s accuracy, fluency, tone, and terminology.

Which Translation Model should I choose?

As a general guide:

  • DeepL produced the best overall translation quality in our PairaVoice testing. It delivered the most accurate and natural-sounding results across the scenarios we evaluated. However, performance may still vary depending on the language pair, terminology, accent, audio quality, and subject matter.
  • OpenAI is useful for natural, fluent, and context-aware speech translation. It can perform well in conversations where tone and meaning are especially important.
  • Google is a strong choice for fast, general-purpose speech translation across many language pairs.

For healthcare, education, technical, or other specialized conversations, we recommend testing the available models with real examples from your organization’s common use cases. This will help you determine which model provides the best balance of accuracy, speed, terminology, and natural conversational flow.

For further information or help selecting the right translation model for PairaVoice, contact the Pairaphrase team.

Can I switch engines?

Yes. Users can select their preferred AI model and switch among engines for optimal accuracy.

Should I use the same engine for every language pair?

Not necessarily. The best engine may vary by language pair, topic, accent, and desired output style. Organizations should test their most common use cases, such as English-Spanish clinical intake or English-Arabic school communication, and document the best setting for each scenario.

What is Voice Enhancement?

Voice Enhancement improves the quality and clarity of the speaker’s voice before the audio is processed. It helps PairaVoice better detect speech and can make the translated audio sound more natural.

Use Voice Enhancement ON for most conversations, especially when the speaker’s voice is soft, distant, or not perfectly clear.

Turn it OFF only if the voice sounds distorted or overly processed.

What is Noise Cancellation?

Noise Cancellation reduces background noise so PairaVoice can focus more clearly on the speaker’s voice.

Use Noise Cancellation ON when you are in a noisy place, such as a clinic, classroom, waiting room, hallway, conference room, or public space.

Use Noise Cancellation OFF when you are in a quiet room and the speaker is close to the microphone.

When should I turn Noise Cancellation on?

Turn Noise Cancellation ON when background sounds are making it difficult for PairaVoice to understand the speaker.

This is helpful when there are:

  • Other people talking nearby
  • Equipment sounds
  • Room echo
  • Traffic or public noise
  • Classroom or clinic background noise

For quiet conversations, it is usually better to leave it off.

What is Voice Enhancement Configuration?

Voice Enhancement Configuration allows users to adjust how voice enhancement works.

Most users should keep the default settings. Change these settings only if the voice sounds too quiet, too sharp, distorted, or unnatural.

A good practice is to test any changes with a short sample conversation before using them in an important session.

What is Stream mode?

The Stream mode delivers live translation while a participant is speaking. This setting is designed for a faster, more natural flow, reducing wait times by processing speech continuously rather than waiting for long pauses.

In our tests, this mode delivered the best overall results across various use cases, ranging from casual check-ins to highly technical discussions.

Use Stream mode for:

  • Fast-paced, routine back-and-forth
  • Basic greetings and uncomplicated queries
  • General conversational exchanges
  • Clinical, school-based, or legal meetings
  • Sessions involving specific directions or steps
  • Environments where responsive interaction is critical

Because it maintains context while providing immediate output, Stream mode is our recommended default for most user interactions.

What is Batch mode?

The Batch mode processes longer segments of speech, translating only after a person completes a full thought or phrase.

This creates a more deliberate, turn-based rhythm with visible pauses between the original speech and the audio translation. It is ideal for scenarios where participants prefer a structured approach where one person finishes entirely before the next begins.

Use Batch mode when:

  • Speakers want to pause after full statements
  • The meeting follows strict turn-taking
  • Translating large blocks of text as a single unit
  • A highly formal speaking process is desired
  • Stream mode does not align with the specific environment

Batch mode remains an effective alternative for users who favor a complete-phrase translation workflow.

Should I use Stream mode or Batch mode?

We suggest starting with Stream mode for the majority of sessions. Our testing confirms it provides the most fluid experience and highest responsiveness for professional communication.

Choose Batch mode if you prefer a traditional, turn-by-turn style where the app waits for a clear stop before translating.

Quick Reference:

  • Stream: Best for active, fast-moving conversations
  • Batch: Best for structured, deliberate exchanges

For additional assistance or guidance on selecting the right PairaVoice settings, reach out to the Pairaphrase support team.

What is the SMS setting?

The SMS setting allows users to receive a code by text message.

This may be used for verification, account access, or connecting another participant to a session.

Turn SMS ON if you want to receive codes by text message.

Turn SMS OFF if you do not want to receive codes by SMS or if your organization uses another access method.

What does “PairaVoice will recognize your voice” mean?

This screen introduces voice recognition or voice cloning setup.

Voice setup allows PairaVoice to recognize the user’s voice and may help create more natural translated audio.

Users can choose:

  • Let’s go to begin voice setup.
  • Skip to continue without setting it up.
Do I need to set up voice recognition?

You should set up voice recognition if you plan to use PairaVoice often and want a more personalized or natural voice experience.

You can skip it if you only need quick translation, text transcription, or do not want to use voice cloning features.

Recommended Settings by Scenario

What settings should I use for a normal conversation?

For most everyday conversations, use:

  • Voice Enhancement: ON
  • Noise Cancellation: OFF in quiet spaces, ON in noisy spaces
  • Speech Recognition Model: ElevenLabs, OpenAI, Google
  • Translation Model: OpenAI or DeepL
  • Translation Mode: Stream
What settings should I use in a noisy place?

For noisy environments, use:

  • Voice Enhancement: ON
  • Noise Cancellation: ON
  • Speech Recognition Model: ElevenLabs
  • Translation Model: OpenAI or DeepL
  • Translation Mode: Batch

Also try to keep the phone close to the speaker and avoid people speaking at the same time.

What settings should I use for medical or critical conversations?

For medical, safety, consent, or other critical conversations, use:

  • Voice Enhancement: ON
  • Noise Cancellation: ON if needed
  • Speech Recognition Model: Best-tested model for that language pair. ElevenLabs suggested
  • Translation Model: Best-tested model for that language pair. DeepL suggested
  • Translation Mode: Stream suggested

For important details, speak in short sentences and confirm names, numbers, dates, medications, and instructions before moving on.

What settings should I use for a parent-teacher conference?

For parent-teacher conferences, use:

  • Voice Enhancement: ON
  • Noise Cancellation: ON if the room is noisy
  • Translation Mode: Stream
What is the best combination of settings?

There is no single best combination for every situation. The best settings depend on the language pair, background noise, topic, and need for speed or accuracy.

A strong default setup is:

  • Voice Enhancement ON
  • Noise Cancellation OFF in quiet spaces / ON in noisy spaces
  • ElevenLabs for Voice Recognition
  • DeepL for Translation
  • Stream mode
Should I use one device or two devices?

Single-device mode is convenient when both people are nearby and can share the screen or listen to the same device. Dual-device mode is better when speakers need more personal space, when each person wants to read their own transcript, or when the conversation is easier with separate microphones.

When should I use hands-free mode?

Hands-free mode is useful when the user cannot hold the device, such as during clinical rounds, classroom walkthroughs, field visits, or other active environments. Hands-free translation can be used with earbuds.

Troubleshooting

Why is the translation delayed?

Delays can happen if the app is using batch mode, the internet connection is weak, the audio is noisy, or the speaker is using long sentences. For faster output, try streaming mode and shorter speaking turns.

Why did PairaVoice misunderstand a word?

Speech recognition can be affected by background noise, accents, overlapping speech, unfamiliar names, technical terms, or poor microphone placement. Repeat the phrase slowly, move closer to the microphone, or switch to batch mode.

What should I do if the translation seems wrong?

Pause the conversation, restate the idea in simpler language, and try a shorter sentence. For important information, verify the transcript and repeat the key point. If the issue persists, try another engine.

Why is medical or technical terminology sometimes inconsistent?

Specialized terminology is harder for speech recognition and machine translation systems. Use batch mode, speak clearly, and verify critical terms. For repeated professional use, organizations should test preferred engines by language pair and topic.

Security, Privacy, and Compliance

Is PairaVoice HIPAA compliant?

PairaVoice is designed to support healthcare workflows and HIPAA-aligned use cases. Organizations should follow their own privacy, consent, documentation, and interpreter-use policies when using PairaVoice in healthcare settings.

Is PairaVoice FERPA aligned?

PairaVoice is designed to support education workflows and FERPA-aligned use cases. Schools and organizations should follow their own policies for student information, consent, documentation, and data retention.

Are conversations stored?

Conversations are never stored on PairaVoice servers without explicit permission.

Are transcripts searchable?

Saved and searchable transcripts are listed as a PairaVoice Pro feature.

Can I use PairaVoice for sensitive conversations?

PairaVoice is positioned for healthcare and education conversations, including HIPAA and FERPA-sensitive environments. Organizations should still follow their own privacy, consent, documentation, and interpreter-use policies.

Plans, Trial, and Access

Can I try PairaVoice before buying?

Yes. The first five conversations are free with full feature access and no credit card required.

What is included in PairaVoice?

The PairaVoice plan includes real-time voice translation, unlimited conversations, personalized voice cloning, multi-engine AI translation, streaming and batch modes, background noise cancellation, hands-free mode, single or dual-device mode, and a chat-style transcription UI.

What is included in PairaVoice Pro?

PairaVoice Pro includes everything in PairaVoice, plus automatic SOAP note generation, saved and searchable transcripts, and HIPAA / FERPA-focused capabilities.

How much does PairaVoice cost?

PairaVoice costs $2.99/month and PairaVoice Pro costs $9.99/month.

Still need help?

If your question is not answered above, please submit a support request using the form below. Our team will review your message and follow up as soon as possible.