Vision Access - Image Description Add-on

This add-on enables NVDA users to describe graphical content on screen using AI. You can use either Ollama (local) or Google Gemini AI.

Table of Contents

Features

Keyboard Shortcuts

Shortcut Function
NVDA+Shift+E Describe the navigator object's visual
NVDA+Shift+G Describe the focused object's (graphic, image) visual

Settings

To access settings: NVDA Menu → Preferences → Settings → Vision Access

Provider Selection

Ollama Settings

Gemini Settings

Note: To get a Gemini API key, visit Google AI Studio.

Requirements

For Ollama

For Gemini

Usage

  1. Navigate to the visual you want to describe (image, graphic, screen area)
  2. Press NVDA+Shift+E or NVDA+Shift+G
  3. Listen to progress notifications (Percent 0, 25, 50, 75, 100)
  4. When you hear "Description ready", the result dialog opens
  5. Use "Ask More" field to ask follow-up questions

Troubleshooting

Ollama connection error: Make sure Ollama is running (ollama serve)
Gemini API error: Ensure your API key is valid and quota is not exceeded

Developer

Sarper Arıkan
Email: sarperarikan@gmail.com
Web: sarperarikan.net