Use this page to set text extractor options.
Category | Option | Definition |
|---|---|---|
Text Extractor | Use the following fallback ANSI code page | Allows the administrator to select the fallback character set. The text extractor uses this character set to read input files when there is a problem identifying the correct code page. The default is to use the endpoint computer operating system native language. |
Maximum memory used (MB) per process | Sets the maximum amount used for text extraction. Default: 75 | |
Maximum input file size to scan (MB) | The maximum file size the text extractor can handle is 500. Default: 20 | |
Maximum output file size (MB) | The maximum file size the text extractor generates to be used by the Trellix DLP Endpoint client is 500. Default: 10 | |
Respect cell boundaries in spreadsheets | Prevents the incorrect identification of random 16-digit numbers. Default: Unchecked. | |
Content fingerprinting | Select which technologies to use for adding meta-data | Optimizes content fingerprinting performance. |
Content fingerprinting for Outlook | Preserve content fingerprints in attachments when sending or receiving an email | When selected, restores content fingerprints of email attachments if the recipient has Trellix DLP Outlook add-in installed. |
Ignored Processes | Process name | Specifies the original filename of the application in the text box to add it to ignored processes.
|
Folder | Specifies the folder name. Optional unless extensions are specified. | |
All files | Adds all files in a named folder. | |
Specific extensions | Adds the named extensions. | |
Empty extension | Adds a blank extension. | |
Web | When selected, dynamic fingerprints for web upload are not created. | |
App | When selected, files opened by the named application in the named location are not analyzed. | |
File | When selected, tags for files opened by the named application in the named location are not analyzed. | |
Actions | Allows editing or deletion of the ignored processes. |