Desktop capture and OCR
Capture supported macOS windows and submit extracted context for analysis.
Supported targets
| Target | Behaviour |
|---|---|
| Chrome | Captures the target Chrome window. |
| PowerPoint | Captures the window or requires Slide Show mode for full-slide capture. |
| Word | Captures the supported Word window; CLI also exposes document automation. |
Capture
- 1Prepare the target
Open the supported app and keep it on the display where OpenSS expects to capture.
- 2Choose Capture
Select the target and full-slide behaviour where shown.
- 3Allow permission
Approve Screen Recording when macOS asks, then retry if the first permission change requires it.
- 4Review
OpenSS extracts text locally and sends the submitted prompt and relevant image/context through the selected AI route.
Image-first analysis
Screenshot requests always include the original PNG or JPEG pixels as well as OCR. The vision-capable model receives the image as primary evidence, with the user instruction and locally extracted OCR clearly separated as supporting text. Diagrams, charts, colours, selected states and spatial layouts are therefore available to the model even when OCR is empty or incomplete.