Content inventory: list every page before you plan a single new one
A content inventory is a single list of every page on a site, with what each page is about and how it is doing. It is the first step of any content audit, and the step most often skipped. Without it, new pages are planned on top of old ones nobody remembers, and the site ends up competing with itself.
On this page — 4 sections
Why start with an inventory instead of new ideas?
Quick answer
Because most sites already have pages on the topics they plan to write about, often several. An inventory shows what exists, what works and what overlaps, so new work extends the site instead of duplicating it. It usually also finds quick wins in pages that only need an update.
A Dubai AC company planning ten new articles may find it already has three old posts on AC maintenance, a service page that barely mentions it, and a landing page from a summer campaign. Writing a fourth maintenance article would make things worse, not better.
Content sites feel this most. An Egyptian recipe site with five years of posts may hold four versions of the same molokhia recipe, each from a different writer, none of them clearly the main one.
Where do you get the list of pages?
Quick answer
Combine three sources: your XML sitemap, the pages Search Console reports in its Performance and Page indexing reports, and a crawl of the site. Each catches pages the others miss, such as old URLs still receiving clicks or pages missing from the sitemap.
- Sitemap: the pages you meant to publish. Often incomplete on older sites.
- Search Console Performance report: every page that received impressions, including ones you forgot about.
- Search Console Page indexing report: pages Google found and whether it indexed them.
- A crawl: what a visitor can actually reach by following links from the home page.
Pages that appear in Search Console but not in the crawl are often orphans: still known to Google, but no longer linked from the site. Mark them as you go; they get their own step later.
What should the inventory record for each page?
Quick answer
Keep it short enough to finish: the URL, the page type, the topic or target query, the hub it belongs to, clicks and impressions over the last few months, the number of internal links pointing to it, and the date it was last updated.
| URL | Type | Topic | Hub | Clicks (3 mo.) | Internal links in | Last updated |
|---|---|---|---|---|---|---|
| /sofas/corner-sofas/ | Category | Corner sofas | Sofas | 1,240 | 18 | 2026-06 |
| /blog/best-corner-sofas/ | Article | Corner sofas | None | 310 | 1 | 2023-02 |
| /offers/corner-sofa-sale/ | Landing | Corner sofas | None | 12 | 0 | 2024-11 |
The "Topic" and "Hub" columns are the ones that matter most, and the ones tools can't fill for you. They take judgement, and they are where overlaps become visible.
What patterns should you look for once the list is done?
Quick answer
Sort by topic and look for four patterns: several pages on one topic, pages with no hub, pages with impressions but almost no clicks, and pages no one has updated in years. Each pattern points to a different action in the next step of the audit.
- Several pages, one topic: candidates to merge.
- No hub: pages floating outside the structure, which need a parent and links.
- Impressions without clicks: a page Google shows but people skip; often the title or the opening fails to answer.
- Not updated in years: candidates for a refresh, especially pages with prices, laws or dates.
This article is part of the Content Operations series — Running content week to week: briefs that prevent rework, checks before publishing, fixing blockers first and refreshing on schedule.
About the author
Mohamed Youns
Semantic SEO Engineer · Author & system developer
Mohamed Youns writes about how search engines understand content — the same standards he applies when building semantic systems at Nut Hub. nut-hub.org