Missing text extractions

This feature helps the administrator review documents whose text extraction failed, listing the error registered for each one: the document's UUID and name, the error message (e.g. an unsupported MIME type, per the Built-in Text extractor plugins), the error date, and whether the extraction file exists on disk.

  • To search, fill in the Filter by UUID, name or message field and click the search icon.

Act on a single document

  • Click the info icon to view the error log details.
  • Click the refresh icon to reset the document: it is removed from the error registry and queued for re-indexing (re-extraction).
  • Click the delete icon to delete the error log entry only, without touching the document.

Act on all documents sharing an error

  • Click the Error Summary button to see every distinct error message grouped, with a count of how many documents are affected by each one.
  • Click the refresh icon of a message to reset every document with that error at once.
  • Click the delete icon of a message to delete every error log entry with that message at once.
  • Both bulk actions require typing yes to confirm.

Resetting a document only queues it for re-extraction; if the underlying cause is not fixed (e.g. the MIME type is still not supported by any enabled text extractor), the same error will be logged again.