<?xml version="1.0" encoding="UTF-8" standalone="yes"?><feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en-us"><id>https://blog.pgxn.org/tags/not-found/</id><title>Not Found</title><updated>2011-03-04T20:20:08Z</updated><link rel="self" type="application/atom+xml" href="https://blog.pgxn.org/tags/not-found/feed.xml"/><link rel="alternate" type="text/html" href="https://blog.pgxn.org/tags/not-found/"/><author><name>The PGXN Maintainers</name></author><generator uri="https://gohugo.io/" version="0.167.0">Hugo</generator><entry><id>https://blog.pgxn.org/post/3641963548</id><title type="html">Question: Ajax and Search Engine Indexing</title><link rel="alternate" type="text/html" href="https://blog.pgxn.org/2011/ajax-search-indexing/"/><updated>2026-10-07T16:13:48Z</updated><published>2011-03-04T20:20:08Z</published><author><name>David E. Wheeler</name></author><category scheme="https://blog.pgxn.org/tags" term="ajax" label="Ajax"/><category scheme="https://blog.pgxn.org/tags" term="api" label="API"/><category scheme="https://blog.pgxn.org/tags" term="search-engine" label="Search Engine"/><category scheme="https://blog.pgxn.org/tags" term="indexing" label="Indexing"/><category scheme="https://blog.pgxn.org/tags" term="indexability" label="Indexability"/><category scheme="https://blog.pgxn.org/tags" term="response-code" label="Response Code"/><category scheme="https://blog.pgxn.org/tags" term="not-found" label="Not Found"/><summary type="html"><![CDATA[I&rsquo;ve started working on the main (search) site in earnest now. The basic
layout is done, and I&rsquo;m working on the distribution view (<a href="https://theory.github.com/pgxn/pgtapdist.html">mockup</a>). My
thinking so far has been that I would simply serve a page that requested, say,
<code>/dist/pgTAP/</code>, and that page would use Ajax requests to fetch the data from
the <a href="https://api.pgxn.org/">API server</a> and display stuff. I think this will work pretty well except
for one thing: 404s.]]></summary><content type="html" xml:base="https://blog.pgxn.org/" xml:space="preserve"><![CDATA[<p>I&rsquo;ve started working on the main (search) site in earnest now. The basic
layout is done, and I&rsquo;m working on the distribution view (<a href="https://theory.github.com/pgxn/pgtapdist.html">mockup</a>). My
thinking so far has been that I would simply serve a page that requested, say,
<code>/dist/pgTAP/</code>, and that page would use Ajax requests to fetch the data from
the <a href="https://api.pgxn.org/">API server</a> and display stuff. I think this will work pretty well except
for one thing: 404s.</p>
<p>That is, if you request <code>/dist/nonexistent/</code>, then it will load a page with
the HTTP status code <code>200 OK</code>, but then, when the Ajax request 404s, it will
show a &ldquo;Not found&rdquo; error message. That&rsquo;s all well and good, but I&rsquo;m wondering
about the impact of two things:</p>
<ol>
<li>
<p>Since the page itself won&rsquo;t 404, search engines might index links to
nonexistent extensions. Of course, bad links won&rsquo;t be <em>that</em> common, but
of course they do happen and then tend to live forever.</p>
</li>
<li>
<p>If the search site uses Ajax to fetch the contents of a page via JSON (or,
for documentation, as an HTML document it will put into a div), will the
full content be properly indexed by search engines?</p>
</li>
</ol>
<p>So these are serious questions, in my mind. Do we loose good search engine
indelibility when we load content dynamically?</p>
<p>Of course, I can instead write it so that the back end fetches stuff from the
API server (and perhaps directly from the file system) and get &lsquo;round these
issues, but then it&rsquo;s less of a cool example of the use of the API server.</p>
<p>What do you think? Good advice much appreciated!</p>
]]></content></entry></feed>