Camera Input & Identification

COMMENT-> ACTIVATING CAMERA INPUT ISNT THE RIGHT PHRASE ; THE VIDEO AGENT USES CAMERA INPUT. RATHER THIS IS THE CAMERA INPUT AGENT. Camera mode lets workers use their phone camera as a direct input during a guided task or session. Rather than describing what they see, a worker can point the camera at the work and Bourne AI processes the frame — confirming a step, identifying a part or label, assessing a condition, or surfacing relevant knowledge from your content library.

What Camera Input Does

When a worker activates camera input, the system reads the camera frame in context with the current task. The AI can:

  • Confirm that a step has been completed correctly based on what it sees
  • Identify a part, component, or label against your uploaded reference images
  • Surface relevant documentation or instructions tied to the identified item
  • Assess a visible condition (e.g., wear, damage, orientation) and flag it if needed

The worker stays in the flow of the task. There is no separate lookup step — the camera input feeds directly into the guidance.

Visual Identification

When your Media Manager contains reference images — part photographs, diagrams, label photos, equipment placards — the system uses those images to recognise what the camera sees. There is no manual lookup step; the right reference is surfaced automatically during the task.

Example: a worker scans a serial plate on a piece of equipment. The system identifies the unit against your uploaded part images and immediately surfaces the relevant section of the service manual. The worker does not need to search; the match is surfaced in context.

To make visual identification work well:

  • Upload clear, well-lit reference images for each part or item workers are likely to encounter
  • Include multiple angles or variants where applicable (e.g., worn vs. new, different mounting orientations)
  • Use descriptive filenames and tags when uploading so the system has context alongside the image

See Media Manager for instructions on uploading and organizing reference images.

Object Segmentation

The camera can isolate specific objects within a cluttered frame. When multiple components are visible, the system uses segmentation to focus on the relevant item rather than treating the entire image as a single input. This is useful in environments where components are tightly packed or the background is visually complex.

Accessing Camera Mode

Camera input is available when the camera capability is enabled on your account. If camera mode does not appear as an option, contact your account administrator to confirm your access. Once available, camera mode is accessible from the dashboard at the start of a session, and can be used alongside voice or text input within a session.

Camera Permissions

The browser must be allowed to access the camera. On first use, the browser will display a permission prompt. If you deny it or need to re-enable it:

  • Chrome (Android/desktop): tap the lock icon in the address bar and set Camera to Allow
  • Safari (iOS): go to Settings > Safari > Camera and set to Allow or Ask
  • Other browsers: consult your browser's site permissions settings

Camera permission is per-site. Granting permission once is typically remembered for future sessions on the same device.