---
title: "More Google API musing -- \"404 Correction\" in a personal HTTP proxy via Google's cache"
date: 2002-04-15T00:49:41-04:00
url: https://blog.lmorchard.com/2002/04/15/oooohg/
author: Les Orchard
---

# More Google API musing -- "404 Correction" in a personal HTTP proxy via Google's cache

<p>Hmm... now that I finally stopped babbling and read the docs, I just noticed that the <a href="http://www.google.com/apis">Google APIs</a> has methods to access their cache.</p>
<p>Sounds like I need to write a personal HTTP proxy that includes "404 Correction" by consulting Google's cache whenever one encounters a 404.  Could be a new project, too, since someone I was talking to wanted searchable personal web browsing history and I think a personal HTTP proxy could help with that.<br />
</p>
<!--more-->
shortname=oooohg

<div id="comments" class="comments archived-comments">
            <h3>Archived Comments</h3>
            
        <ul class="comments">
            
        <li class="comment" id="comment-221090164">
            <div class="meta">
                <div class="author">
                    <a class="avatar image" rel="nofollow" 
                       href="http://webseitz.fluxent.com/wiki"><img src="http://www.gravatar.com/avatar.php?gravatar_id=2e83224d92ed7f1148f4dd3cdb0e4548&amp;size=32&amp;default=http://mediacdn.disqus.com/1320279820/images/noavatar32.png"/></a>
                    <a class="avatar name" rel="nofollow" 
                       href="http://webseitz.fluxent.com/wiki">Bill Seitz</a>
                </div>
                <a href="#comment-221090164" class="permalink"><time datetime="2002-04-23T20:55:44">2002-04-23T20:55:44</time></a>
            </div>
            <div class="content">The problem is that it wouldn't help the most common case, which is where someone like Time mag or the NYTimes moves their archives into a Paid category. I never find those things in Google.

(But I haven't looked that carefully, so I could be wrong...)

(Hmm, I wonder whether the Wayback machine holds such things? Quick check shows the NYTimes blocks robots from content.)

(Hmm, I wonder if the Wayback Machine has an API?)</div>
            
        </li>
    
        </ul>
    
        </div>
