We know that MS Word can create an HTML
document from a standard document. The images become gif or jpg and the text becomes HTML.
So I have a client with hundreds of Word documents that contain TIF hi res images and multi-column text. How can we extract the images and text into a data base of hi res images and unformatted text?