Choosing an Input Mode
Atanzo AI supports voice, text, and camera input today, with smart glasses coming soon. Each one starts as its own session, so you pick the mode that fits the job when you begin — hands-free by voice, quietly by typing, or visually through the camera.
Voice Mode
Voice mode lets you talk to the system — advancing steps, asking questions, and confirming completions by speaking. The AI listens and responds in kind.
Best for: hands-free environments where gloves, machinery, or outdoor conditions make tapping a phone impractical.
Requirements: microphone permission must be granted in the browser. The browser will prompt for this on first use. On most devices you can review this permission under the site settings in your browser.
Smart Glasses (Coming Soon)
Smart glasses support isn't available yet — this section describes what's coming.
Smart glasses give you hands-free voice guidance with each step shown on a heads-up display, for jobs where you can't hold or keep glancing down at a phone.
Best for: work that needs both hands free the whole time — up a ladder, inside machinery, or anywhere pulling out a phone to check the next step isn't practical.
Requirements: a pair of smart glasses paired to your account (see Pairing Smart Glasses below).
Pairing Smart Glasses
Pairing links a physical pair of smart glasses to your account so they can be used to run a session.
- Turn on your glasses. They'll display a six-digit pairing code.
- In Atanzo AI, open the paired devices screen and enter the six-digit code shown on the glasses.
- Confirm the pairing. The glasses will show a success indicator, and the device will appear in your list of paired devices.
Only enter a code shown on glasses you are physically holding. The code approves that specific device to run sessions on your account — entering a code you were sent or read from someone else pairs their device, not yours.
If you lose your glasses or replace them, you can remove a paired device from the same screen at any time.
Text Mode
Text mode lets you type responses into the interface. The AI reads each entry and responds as it would in any other mode.
Best for: noisy environments where voice recognition is unreliable, or quiet areas such as office spaces, clean rooms, or occupied customer sites where speaking aloud is disruptive.
Requirements: any device with a keyboard or on-screen keyboard. No additional permissions are needed.
Camera Mode
Camera mode lets you point your phone at the work. Atanzo AI processes the camera frame and can confirm a step, identify a part or label, or surface relevant knowledge based on what it sees.
Best for: visual inspection, part identification, condition assessment, and any step where what you're looking at needs to carry into the guidance.
Requirements: camera permission must be granted in the browser. A rear-facing camera is required for most field use cases.
See Camera Input & Identification for the full capability, including how the system matches what the camera sees against your uploaded image library, and how to start a camera session.
Choosing a Mode
Mode is set when you start a session — each mode is its own session rather than a setting you flip mid-task. Pick whichever is practical for the environment you're in: voice for a loud shop floor or a job that needs both hands free, text for a quiet or noisy space where speaking isn't practical, camera for a step that hinges on identifying or checking something visually.
If conditions change partway through a job — you finish a hands-free stretch and want to check a part with the camera, for example — end the current session and start a new one in the mode you need. Your progress and knowledge base access carry over; only the input mode changes.
Device Requirements
- All modes: any modern smartphone running a current browser (Chrome, Safari, or equivalent)
- Voice mode: microphone access; works on both iOS and Android
- Smart glasses (coming soon): a paired pair of compatible smart glasses; a phone or browser session to complete pairing
- Text mode: on-screen or physical keyboard; no additional hardware required
- Camera mode: rear-facing camera; camera permission granted in the browser
Running Sessions from the Mobile App
Running a session in the Atanzo AI Console app is an input surface too — see The Atanzo AI Mobile App for getting the app and running voice, text, and camera sessions from your phone.
Related Articles
- Camera Input & Identification — visual identification, image matching, and how to prepare your image library
- The Atanzo AI Mobile App — running guided task and coaching sessions from your phone