<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Search Architecture Notes]]></title><description><![CDATA[Practical notes on technical SEO, content architecture, indexing, internal linking, website audits, and search visibility based on real-world projects.]]></description><link>https://search-architecture-notes.hashnode.dev</link><image><url>https://cdn.hashnode.com/res/hashnode/image/upload/v1593680282896/kNC7E8IR4.png</url><title>Search Architecture Notes</title><link>https://search-architecture-notes.hashnode.dev</link></image><generator>RSS for Node</generator><lastBuildDate>Wed, 02 Sep 2026 15:56:21 GMT</lastBuildDate><atom:link href="https://search-architecture-notes.hashnode.dev/rss.xml" rel="self" type="application/rss+xml"/><language><![CDATA[en]]></language><ttl>60</ttl><item><title><![CDATA[How I Audited the SEO Architecture of a Large Poetry Website]]></title><description><![CDATA[SEO problems on content-heavy websites are rarely caused by one broken setting.
Sometimes the sitemap is fine. Pages are indexable. Titles exist. Internal links are present. Yet search visibility keep]]></description><link>https://search-architecture-notes.hashnode.dev/seo-architecture-audit-large-poetry-website</link><guid isPermaLink="true">https://search-architecture-notes.hashnode.dev/seo-architecture-audit-large-poetry-website</guid><category><![CDATA[SEO]]></category><category><![CDATA[technical seo]]></category><category><![CDATA[#content strategy]]></category><category><![CDATA[Search engine optimization]]></category><category><![CDATA[Web Development]]></category><dc:creator><![CDATA[Muhammad Hanif]]></dc:creator><pubDate>Wed, 26 Aug 2026 11:36:03 GMT</pubDate><enclosure url="https://cdn.hashnode.com/uploads/covers/6a8ec91e72b9edbb4f126b26/c5cb3f90-b626-4661-8341-7cbc88322d47.png" length="0" type="image/jpeg"/><content:encoded><![CDATA[<p>SEO problems on content-heavy websites are rarely caused by one broken setting.</p>
<p>Sometimes the sitemap is fine. Pages are indexable. Titles exist. Internal links are present. Yet search visibility keeps falling.</p>
<p>I recently worked through this kind of problem on a poetry website with a large collection of closely related pages. The interesting part was that the site did not have one obvious technical failure. Instead, several smaller structural problems were working together.</p>
<p>This article explains the audit process I used, what I looked for, and why I eventually stopped thinking about the site as a collection of individual articles and started looking at it as an information architecture problem.</p>
<h2>The First Clue Wasn't Traffic</h2>
<p>Traffic loss gets attention first, but I did not begin by asking why traffic had fallen.</p>
<p>I started with a different question:</p>
<p><strong>What does a search engine see when it crawls this website?</strong></p>
<p>That changed the direction of the audit.</p>
<p>The site had many poetry collections targeting closely related emotional and situational searches. Some pages focused on romantic poetry, others on friendship, sadness, birthdays, family relationships, humor, and specific occasions.</p>
<p>Individually, many of those topics made sense.</p>
<p>The problem appeared when I looked at them together.</p>
<p>Several URLs were competing for overlapping ideas. Some contained very similar literary examples. Others were separated by only a small change in search intent.</p>
<p>That meant the important question was no longer whether each page was “optimized.”</p>
<p>The real question became whether every page had a sufficiently distinct reason to exist.</p>
<h2>I Mapped the Site Before Changing Anything</h2>
<p>Before editing content, I created a simple URL inventory.</p>
<p>For every important page, I recorded its primary topic, closest competing pages, ranking history, current impressions, internal links, and whether its intent was unique enough to justify a separate URL.</p>
<p>This immediately made patterns easier to see.</p>
<p>A poetry site is especially vulnerable to overlap because the same poem can naturally fit several categories. A Robert Burns poem, for example, might appear relevant to romance, relationships, anniversaries, or emotional poetry.</p>
<p>That is useful for readers, but if several search pages repeatedly use the same works, similar introductions, and similar interpretations, the distinction between those pages starts becoming weaker.</p>
<p>The audit therefore had to consider both <strong>technical duplication</strong> and <strong>semantic duplication</strong>.</p>
<p>Technical duplication is relatively easy to find.</p>
<p>Semantic duplication is harder because the pages can contain different words while still solving almost the same search intent.</p>
<h2>Search Performance Helped Me Prioritize</h2>
<p>I combined ranking-history data with Search Console performance instead of treating every URL equally.</p>
<p>Some pages had previously ranked for dozens or even hundreds of queries but had lost most of that visibility. Others were still receiving impressions despite sitting several pages deep in the search results.</p>
<p>Those URLs interested me more than completely inactive pages.</p>
<p>A URL with existing impressions tells me that Google still associates the page with relevant searches. It may not be ranking well, but the connection has not disappeared completely.</p>
<p>This was particularly useful while reviewing <a href="https://postpoetics.org/">PostPoetics</a>, because the site had enough historical visibility to reveal which sections Google had understood in the past.</p>
<p>Instead of spreading effort across the entire site, I created a smaller recovery group made up of pages with three characteristics:</p>
<ul>
<li><p>meaningful historical search visibility</p>
</li>
<li><p>current impressions or residual rankings</p>
</li>
<li><p>a clearly defensible search intent</p>
</li>
</ul>
<p>This prevented the audit from turning into a site-wide rewriting project.</p>
<h2>Pagination Needed Its Own Review</h2>
<p>Content sites often generate more URLs than editors realize.</p>
<p>Pagination is one of the easiest ways for this to happen.</p>
<p>A collection might have a main URL and then additional numbered pages. Those paginated URLs can sometimes begin receiving impressions independently.</p>
<p>That is not automatically a problem.</p>
<p>The problem begins when search engines are unsure whether the paginated URL is an independent destination, part of a sequence, or a weaker version of the main page.</p>
<p>I reviewed pagination separately from normal article URLs.</p>
<p>For every numbered URL I asked:</p>
<p>Does this page contain content users may specifically want?</p>
<p>Does it deserve to rank independently?</p>
<p>Is navigation between the pages clear?</p>
<p>Are canonical signals consistent with the actual purpose of the page?</p>
<p>Could the same material be presented more effectively on the primary collection?</p>
<p>The important lesson was not to automatically canonicalize everything to page one. Canonical decisions have to reflect the real structure.</p>
<p>But pagination should never be left to chance.</p>
<p>If numbered pages are appearing in search, they are part of the SEO architecture whether the content team intended that or not.</p>
<h2>Canonicals Are Not a Substitute for Structure</h2>
<p>Canonical tags are useful, but they cannot rescue unclear information architecture.</p>
<p>I often see canonicalization treated as a cleanup button.</p>
<p>Two pages look similar, so one is canonicalized to the other.</p>
<p>That may reduce duplication signals, but it does not answer the larger editorial question:</p>
<p><strong>Why were both pages created in the first place?</strong></p>
<p>If two URLs serve genuinely different reader needs, they should be differentiated more strongly.</p>
<p>If they do not, consolidation may be more appropriate.</p>
<p>The decision needs to happen at the content-model level before it happens in HTML.</p>
<p>This was one of the biggest changes in how I approached the poetry site.</p>
<p>I stopped asking, “Which canonical should this page use?”</p>
<p>I started asking, “Should these actually be two pages?”</p>
<p>That question produced better decisions.</p>
<h2>Internal Linking Became a Way to Explain the Site</h2>
<p>Internal links are often discussed as a method for passing authority.</p>
<p>That is only part of their value.</p>
<p>On a large topical site, internal links also explain relationships.</p>
<p>A broad poetry collection can act as a parent resource. More specific pages can address narrower situations. Related articles can reference each other without competing for the same primary intent.</p>
<p>When this structure is deliberate, internal links become a map.</p>
<p>When it is accidental, they become noise.</p>
<p>I reviewed internal links with three goals.</p>
<p>First, the strongest pages needed to point toward important recovery pages where the connection was contextually useful.</p>
<p>Second, related articles needed descriptive but natural anchor text rather than repeatedly using the exact same keyword.</p>
<p>Third, orphaned or weakly connected pages needed either stronger integration or reconsideration.</p>
<p>I did not try to maximize the number of internal links.</p>
<p>I tried to make each link explain why another page was worth visiting.</p>
<h2>Repeated Literary Examples Were More Important Than I Expected</h2>
<p>This was one of the more interesting parts of the audit.</p>
<p>On a poetry website, repeating a famous public-domain poem across different collections can seem completely reasonable.</p>
<p>From an editorial perspective, it often is.</p>
<p>But repetition at scale can create a different problem.</p>
<p>Imagine five pages with different titles but similar structures. Each has a short introduction, a collection of poems, short interpretations, and a conclusion. Several of the same famous poems appear on multiple pages.</p>
<p>The URLs are technically unique.</p>
<p>The copy may also be unique.</p>
<p>Yet the overall experience can still feel interchangeable.</p>
<p>That is the type of duplication that no canonical tag can properly solve.</p>
<p>The solution was not simply “never repeat a poem.”</p>
<p>Instead, every important page needed its own editorial perspective.</p>
<p>If the same literary work appeared in two places, the surrounding analysis needed to serve different purposes.</p>
<p>A romantic collection might discuss intimacy.</p>
<p>A long-distance collection might focus on separation.</p>
<p>An anniversary collection might emphasize continuity and shared history.</p>
<p>The poem can be the same.</p>
<p>The reason for including it should not be.</p>
<h2>I Separated Recovery Pages From Maintenance Pages</h2>
<p>One mistake in SEO recovery projects is treating every page as equally urgent.</p>
<p>That spreads effort too thin.</p>
<p>I divided the site into three practical groups.</p>
<p>The first contained URLs with meaningful historical performance and current recovery potential.</p>
<p>The second contained useful supporting pages that could strengthen topical structure but were not immediate priorities.</p>
<p>The third contained pages that needed consolidation, technical review, or simply no active promotion.</p>
<p>This made external promotion easier too.</p>
<p>If a page still has unclear intent, I do not want to send links to it.</p>
<p>Backlinks amplify whatever page already exists.</p>
<p>They do not repair its purpose.</p>
<p>So before thinking about external authority, I wanted the destination URLs to be stable enough that I would still be comfortable promoting them six months later.</p>
<h2>Content Pruning Did Not Mean Deleting Everything</h2>
<p>“Pruning” is often interpreted as deleting pages with low traffic.</p>
<p>I do not think that is a useful rule.</p>
<p>A page can have little traffic and still perform an important structural role.</p>
<p>Likewise, a page can have historical traffic and still be redundant.</p>
<p>For each weak URL, I considered whether to improve it, merge it, redirect it, leave it alone, or remove it from the active content strategy.</p>
<p>That decision depended on intent and usefulness rather than traffic alone.</p>
<p>If a page had a distinct reader need but weak execution, improvement made sense.</p>
<p>If two pages answered essentially the same question, consolidation became more attractive.</p>
<p>If a URL existed mainly because the CMS created it, technical cleanup was usually the better solution.</p>
<h2>The Audit Changed My View of Topical Authority</h2>
<p>Topical authority is often described as publishing more content around a subject.</p>
<p>That definition is incomplete.</p>
<p>A site with 300 pages can communicate its subject less clearly than a site with 50.</p>
<p>More URLs do not automatically mean more expertise.</p>
<p>A better signal is whether each page contributes something identifiable to the larger topic.</p>
<p>For a poetry website, that means a reader should understand why a friendship collection differs from a romantic collection, why a birthday page differs from a general celebration page, and why a sadness page exists separately from grief or heartbreak.</p>
<p>Those distinctions have to exist in the actual content.</p>
<p>They cannot live only in the keyword spreadsheet.</p>
<h2>What I Would Do Before Building More Backlinks</h2>
<p>If I were starting this project again, I would not begin with link building.</p>
<p>I would first decide which URLs deserve long-term authority.</p>
<p>Then I would improve those pages, strengthen their internal relationships, resolve overlapping intent, review pagination and canonical signals, and make sure important sections can be discovered naturally.</p>
<p>Only after that would I begin acquiring relevant external mentions.</p>
<p>A backlink is much more useful when it points to a page whose role in the site is already clear.</p>
<p>Otherwise, authority is being sent into an architecture that search engines may still struggle to interpret.</p>
<h2>Final Takeaway</h2>
<p>The most valuable result of this audit was not a single technical fix.</p>
<p>It was a clearer model of the website.</p>
<p>Large content sites usually accumulate complexity gradually. New categories are added. Similar keywords become separate articles. Pagination grows. Internal links are created inconsistently. Old content remains live because nobody has a strong reason to remove it.</p>
<p>Eventually, each page may look acceptable on its own while the site as a whole becomes difficult to understand.</p>
<p>That is why my SEO architecture audits now begin above the page level.</p>
<p>Before changing titles, building backlinks, or rewriting introductions, I want to understand the relationships between URLs.</p>
<p>Once those relationships become clear, many individual SEO decisions become much easier.</p>
]]></content:encoded></item></channel></rss>