Skip to main content

Mazix Studios

MazixStudios

VLM · Task camera

Camera step reader

A vision-language check on frames from a task camera: what step is this, and does it match the procedure.

A camera on a small workbench

The problem

Labeling every frame by hand does not scale for a small team. The model had to read a frame and say the step.

What we built

  • Frames pulled from the working camera
  • A vision-language call against the expected step
  • Plain language back to the operator
  • A queue for frames the model would not commit to

Tell us what you need built.

A short note on the device, the user, and the deadline is enough to start.

Start a project