Google Makes All PDF Documents Copy & Paste Friendly with OCR

Oct 31, 2008 - 8:07 am 1 by
Filed Under Google

Google announced that they are now using OCR technology to index and show an HTML version of a scanned PDF document. In the past, Google only showed an HTML version of PDF's created with text enabled formatting. But now, if a document is scanned as an image, Google can create an HTML version using OCR.

For example, this PDF is a scan of a cooperative agreement between Google and Regents of the University of California. You cannot copy and paste the text from the PDF document. But now, with Google's OCR capabilities, you can view the HTML version and use this text in your own agreements, saving you the expense of starting from scratch on your own agreements.

Ever find that perfect document that you wanted to reuse for contracts, marketing material, how-tos, and so on? But you were unable to reuse it because it wasn't copy and paste friendly? Well, now you can use Google to get to it. From now on, when searching for documents like this, try filetype:pdf in the search box along with your search query.

Forum discussion at WebmasterWorld.

 

Popular Categories

The Pulse of the search community

Search Video Recaps

 
- YouTube
Video Details More Videos Subscribe to Videos

Most Recent Articles

Google

Google Christmas Decorations Are Live For 2024

Dec 21, 2024 - 6:55 pm
Search Forum Recap

Daily Search Forum Recap: December 20, 2024

Dec 20, 2024 - 10:00 am
Search Video Recaps

Search News Buzz Video Recap: Google December Core Update Done, Spam Update Starts, Google Ranking Exploit Leaked, Google Tests Double Serving Ads

Dec 20, 2024 - 8:01 am
Google Updates

Google December 2024 Spam Update 👾 Rollout Shocks Before Holidays

Dec 20, 2024 - 7:51 am
Google

Google Testing Shaded Button Sitelinks On Mobile

Dec 20, 2024 - 7:41 am
Google

Google Search To Gain AI Mode

Dec 20, 2024 - 7:31 am
Previous Story: Google AdWords To Take Ad Position Into Account For Quality Score