Skip to main content
Image recognition extracts Japanese text from screenshots, photos, or manga and fills it into the input box. After you confirm the text is correct, click Analyze.

Extract text from an image

  1. Click the Extract text from an image button next to the input box.
  2. Choose an image file.
  3. Wait for recognition to finish. The extracted text fills the input box.
The app only extracts text. It doesn’t translate or interpret the image content. Line breaks in the image are replaced with spaces, and the original text order is preserved. When recognition finishes, the page shows “Text extracted. Review it, then select ‘Analyze’.” Check and correct the recognized text, then click Analyze.
With streaming output on, the recognized text appears in the input box gradually.

Which model is used

Image recognition uses the text model you selected in Settings:

Image processing and limits

  • Before uploading, the app compresses the image in your browser. If the longer side exceeds 1600 pixels, the image is scaled down proportionally and re-encoded at 70% quality.
  • The compressed image data can’t exceed 8 MB.
  • Only image files are accepted. If you choose another type of file, the page shows “Please upload an image file.”

When recognition fails

The image is sent to the model provider through the app server only during recognition, and the app doesn’t store it. Umami analytics doesn’t record images or extracted text either. For details, see Keys and privacy.