Live Captions for Talks and Conferences: A Setup Guide

ByAvoce·July 17, 2026
A talk becoming multilingual captions on a projector and audience phones

The biggest risk with live captions is an avoidable on-site failure: distorted audio, a missing display, or a mistranslated name. A 15-minute rehearsal prevents most of these problems.

Prepare three things before the doors open

Secure clean close-mic audio, decide whether captions appear on a screen or through a QR, and add speaker names and specialist terms to the glossary.

Signal path from speaker microphone and mixer to translation computer, screen, and phones

How do you get accurate audio?

  • Use a lapel or handheld mic close to the speaker.
  • Avoid speaker feedback and room noise.
  • For a PA venue, take a clean mixer feed into the computer instead of using its built-in mic.

How should the audience view captions?

MethodSetupBest for
Projected screenSend the Avoce caption window to the main or side screenOne shared language in a lecture hall
QR on personal phonesShow one QR and let viewers choose a languageMultilingual audiences and panels

Lock names and terms with a glossary

Enter speaker, company, product, and technical names before the event so recognition and translation use consistent wording.

Choose the right engine

Direct speech translation gives the lowest latency for general talks. Use the two-stage transcribe-then-translate engine when dense terminology calls for steadier recognition.

On-site backup checklist

  • Run the full audio, screen, QR, and caption flow 15 minutes early.
  • Prepare a phone hotspot.
  • Preload enough points: two hours of translation uses about 120.
  • Know how to stop and restart after a disconnect.

Before the live event, say one complete sentence through the real microphone and verify the entire path.

Frequently asked questions

Does the audience need an app?

No. They scan the QR in a phone browser and choose a language.

How many points for a two-hour talk?

About 120 points for live translation.

Can it work offline?

No. Recognition and translation require a connection, so prepare a hotspot.

Can I export the transcript?

Yes, as text, Markdown, or timestamped SRT.

Get started

Keep reading