Introduction
The AI Summarizer NVDA Add-on v2.5, developed by Sujan Rai at Team of Tech Visionary, is a cutting-edge accessibility tool that leverages the Google Gemini API to deliver concise, AI-generated summaries for a variety of content types, including code, audio, video, documents, and images. Designed specifically for NVDA users, this add-on enhances productivity and accessibility by providing a seamless, user-friendly interface for content summarization.
Description
The AI Summarizer enables NVDA users to upload files in supported formats and generate tailored summaries based on customizable prompts. With an accessible interface, intuitive keyboard shortcuts, and robust functionality, the add-on ensures a smooth user experience. Key features include threaded processing for responsiveness, clipboard and export options, and advanced capabilities like audio/video transcription and image text extraction.
Features
- Multiformat Support: Summarize a diverse range of file types:
- Video: MP4, MKV
- Audio: MP3, WAV
- Images: PNG, JPEG, JPG, ICO
- Documents: Text-based files (e.g., TXT, PDF)
- Customizable Prompts: Define specific instructions for AI to tailor summaries or transcriptions.
- Accessible Interface: Fully compatible with NVDA, featuring keyboard shortcuts and screen reader feedback.
- Transcription Capabilities: Transcribe audio and video files with high accuracy.
- Image Text Extraction: Extract text from images using advanced AI algorithms.
- Clipboard and Export Options: Copy summaries to the clipboard or export them as text files for easy sharing.
- Internet Connectivity Check: Verifies internet access to ensure reliable API communication.
- Threaded Processing: Runs summarization tasks in a separate thread to maintain UI responsiveness.
- Follow-Up Questions: Ask the AI follow-up questions based on generated summaries for deeper insights.
Key Highlights
- AI-Powered Summarization: Harnesses vision-pro s3 model by tech visionary AI for high-quality, concise content summaries.
- Enhanced Accessibility: Designed for seamless integration with NVDA, ensuring an inclusive user experience.
- Advanced Transcription: Transcribes audio and video content with precision.
- Image Text Extraction: Extracts text from images, broadening accessibility for visual content.
- Interactive Follow-Up: Supports iterative queries to refine or expand on summaries.
Keyboard Shortcuts
The AI Summarizer add-on provides intuitive keyboard gestures for efficient navigation:
- NVDA+Alt+N: Opens the AI Summarizer main dialog.
- Alt+A: Attaches a file in the main dialog.
- Alt+R: Removes the attached file in the main dialog or regenerates a summary in the response dialog.
- Alt+S: Initiates summarization in the main dialog or subscribes to the YouTube channel in the about dialog.
- Alt+B: Opens the about dialog or returns to the main dialog from the response dialog.
- Alt+C: Closes any dialog or copies the summary to the clipboard in the response dialog.
- Alt+E: Exports the summary as a text file in the response dialog.
- Alt+L: Closes the response dialog.
- Escape: Closes any active dialog.
- Backspace: Returns to the main dialog from the response dialog.
Usage Instructions
- Install the Add-on:
Download the AI Summarizer add-on from the official GitHub repository or the NVDA add-on store. Follow NVDA's standard add-on installation process.
- Access the Add-on:
Launch the AI Summarizer dialog using NVDA+Alt+N or navigate to the "AI Summarizer" option in NVDA's Tools menu.
- Summarize Content:
- In the main dialog, enter a prompt specifying how the AI should summarize or process the content.
- Press
Alt+A or click "Attach a File" to select a supported file.
- Press
Alt+S or click "Summarize" to process the file and generate a summary.
- View the AI-generated summary in the response dialog.
- Manage Summaries:
In the response dialog, you can:
- Copy the summary to the clipboard (
Alt+C).
- Export the summary as a text file (
Alt+E).
- Regenerate the summary with a new prompt (
Alt+R).
- Return to the main dialog (
Alt+B).
- Explore Additional Features:
Access the about dialog via Alt+B in the main dialog to learn more about the add-on or subscribe to the Team of Tech Visionary YouTube channel.
Update Information
The AI Summarizer NVDA Add-on has been updated to version 2.5, introducing significant enhancements and new features.
Changelog for v2.5
- Video File Support: Added support for summarizing video files (MP4, MKV), with fixes for issues present in previous versions.
- Follow-Up Questions: Introduced the ability to ask follow-up questions based on generated summaries, enabling iterative interaction with the AI.
- Transcription Support: Added transcription capabilities for audio and video files, allowing users to generate accurate text transcripts.
- Image Text Extraction: Enabled AI-driven text extraction from images (PNG, JPEG, JPG, ICO).
- Performance Improvements: Enhanced stability and reliability with bug fixes and optimized API integration.
- Upcoming Features: Additional capabilities are in development and will be released in future updates.
Important Notes
- Internet Requirement: An active internet connection is required to communicate with AISummarizer v2.5.
- Supported Formats: Only the listed file formats (MP4, MKV, MP3, WAV, PNG, JPEG, JPG, ICO, TXT, PDF) are supported for summarization and processing.
- Tutorials and Updates: Subscribe to the Team of Tech Visionary YouTube channel for tutorials, updates, and community support.
Contributions
We welcome contributions from the community! Visit our GitHub repository to report issues, suggest improvements, or contribute code.
Project Home Page
For more information, documentation, and updates, visit the official project page at https://github.com/s-toolkit/ai-summarizer-nvda-addon.