TextExtractor extracts plain text from hundreds of different file types, storing the text extracted in suitably named text files.

TextExtractor 1.10 works in six different modes :-

Instant Mode - Just select any file and extract the text from it.
Batch Mode - Select a group of files and extract the text from all of them in one go.
Polling Mode - Watch a folder location, processing new files as they appear there.
Hierarchical Mode - Extract Text from files in a directory hierarchy.
File List Mode - Extract Text from files in a list.
File Viewer - Select individual files from a file tree to see their textual content.

Features

  • Reads PDF,DOC,DOX,XLS.XLSX,ODT,RTF and many other file types.
  • Also reads DLLs, EXE, COM and binary files.
  • Outputs plain text files, one for each file processed.
  • Extract text instantly, in batch mode, or poll a folder and process files as they appear there.
  • Fast, accurate text extraction.
  • Process multiple file types at the same time.
  • Process whole directory hierarchies.
  • View the text in individual files selected from a directory tree.
  • Make a single list of files, where ever they are located and extract text from them.

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow TextExtractor

TextExtractor Web Site

Other Useful Business Software
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of TextExtractor!

Additional Project Details

Operating Systems

Windows

Intended Audience

End Users/Desktop

User Interface

Java Swing

Registered

2022-11-16