ScanSpeak Privacy Policy
1. What ScanSpeak is
ScanSpeak is an Android app for scanning book pages, recognizing text (OCR), storing content locally on the device, translating text into other languages, and reading text aloud.
2. What data the app processes
The app may process the following data saved or used by the user:
- images of scanned pages,
- text recognized using OCR and text added from the clipboard or from files,
- translations of scan text,
- scan titles and organizational metadata,
- language, voice, reading speed, and translation settings,
- saved resume position,
- audio files created with the "Export to audio" feature,
- downloaded translation language models,
- technical data required for system text-to-speech and notifications.
3. Where data is stored
Scan data, OCR text, translations, preferences, and resume state are stored locally on the user's device within the app's storage. In the current release build, user data backup is disabled so that book content and reading history are not copied to system backup services.
Translation language models are stored on the device by the Google ML Kit library.
Files created with "Export to audio" are saved in the public Music/ScanSpeak
folder so they can be played in other apps.
4. When data leaves the device
4.1. Text correction (LanguageTool)
To fix text recognition errors, the app sends the recognized text of pages to the external LanguageTool service (LanguageTooler GmbH), which returns spelling and grammar suggestions. This happens:
- automatically after saving a new camera scan, after creating a scan from a file, and after adding pages to a scan from images or files,
- when the "Improve text" feature is used manually.
Text added from the clipboard is not sent automatically. Only the text content of pages and the language code are sent — no images, titles, or user identifiers. When the device is offline, the app applies local text formatting only.
Endpoint used by the app: https://api.languagetool.org/v2/check.
LanguageTool's processing is described in the
LanguageTool privacy policy.
4.2. Text translation
Translation runs on the device using Google ML Kit — the content of scans and translations is not sent to any server. The first time a language is used, the app downloads its language model (about 30 MB) from Google servers. Users can limit model downloads to Wi-Fi and delete downloaded models in the app settings.
4.3. Google ML Kit and Google Play services
Text recognition (OCR), the document scanner, and translation use the Google ML Kit library and Google Play services. Images and text are processed on the device. The text recognition model is delivered through Google Play services.
ML Kit sends technical data to Google for diagnostics and usage analytics: device information (manufacturer, model, OS version), package name and app version, performance metrics, and the configuration of the features used, and for translation also the selected source and target languages. For translation, ML Kit uses Firebase Installations and Firebase Remote Config, which rely on an app installation identifier. This data is encrypted in transit (HTTPS). Page images and text content are not sent this way.
Details: Google's ML Kit data disclosure and the Google Privacy Policy.
4.4. Reading aloud
Reading aloud uses the text-to-speech engine installed on the system (for example Google or Samsung). Offline voices process text on the device. If the user selects a voice marked as online in the settings, the text-to-speech engine may send the text being read to its provider's servers under that provider's privacy policy.
5. Permissions and system features
- Document scanner (Google Play services) — to scan pages; the camera is handled by Google Play services and the app does not request the camera permission.
- Internet — for text correction via LanguageTool, downloading language models, and ML Kit diagnostic data.
- Network state — to check whether the device is on Wi-Fi when model downloads are limited to Wi-Fi.
- Notifications — to control background reading and report audio export progress.
- Foreground service (media playback) — so text-to-speech can continue after the app is minimized.
- Foreground service (data sync) — so audio export can finish in the background.
6. Ads, analytics, and user accounts
In the reviewed production build of ScanSpeak:
- it does not require creating an account,
- it does not display ads,
- it contains no in-house analytics or analytics SDKs that track user behavior; the only diagnostic data is the technical data sent by Google ML Kit, described in section 4.3.
7. How to delete data
- Deleting a scan in the app removes the related text, translations, and stored page files from the app's storage.
- Downloaded language models can be deleted in the app settings, in the "Translation" card.
- Audio export files in the
Music/ScanSpeakfolder can be deleted with a file manager. - Uninstalling the app removes all data stored in the app's storage.
8. Sharing data with third parties
Data is not sold. During normal operation, data leaves the device only as described in section 4:
- page text — to LanguageTool for text correction (section 4.1),
- technical data and language model downloads — to Google through ML Kit and Google Play services (sections 4.2 and 4.3),
- text being read — to the text-to-speech engine provider, only when the user selects an online voice (section 4.4).
9. Contact
For privacy or app-related questions, contact: tomasz.paciorek81@gmail.com
10. Changes to this policy
This policy may be updated as the app's features, data processing practices, or Google Play requirements change.