<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Search Engine Basics — new guides</title><description>Technical guides explaining how search engines crawl, index, retrieve and rank web pages, written for beginners and developers.</description><link>https://searchenginebasics.dev/</link><language>en</language><item><title>What Is a Web Crawler and How Does It Work?</title><link>https://searchenginebasics.dev/crawling/what-is-a-web-crawler/</link><guid isPermaLink="true">https://searchenginebasics.dev/crawling/what-is-a-web-crawler/</guid><description>A web crawler is a program that fetches URLs over HTTP and follows the links it finds. Here is the fetch loop, the rules that stop it, and how to verify one.</description><pubDate>Tue, 08 Sep 2026 00:00:00 GMT</pubDate><category>crawling</category><author>Hassan</author></item><item><title>HTTP Status Codes That Actually Matter for Search</title><link>https://searchenginebasics.dev/technical/http-status-codes-for-seo/</link><guid isPermaLink="true">https://searchenginebasics.dev/technical/http-status-codes-for-seo/</guid><description>The status code is the first thing a crawler reads, and it decides everything after it. Here is what each code means to a search engine, and the traps.</description><pubDate>Mon, 07 Sep 2026 00:00:00 GMT</pubDate><category>technical</category><author>Hassan</author></item><item><title>Crawled, Currently Not Indexed: What It Actually Means</title><link>https://searchenginebasics.dev/indexing/why-pages-are-crawled-but-not-indexed/</link><guid isPermaLink="true">https://searchenginebasics.dev/indexing/why-pages-are-crawled-but-not-indexed/</guid><description>Google fetched your page and chose not to store it. Here are the five real causes behind that Search Console status and how to tell which one applies.</description><pubDate>Sun, 06 Sep 2026 00:00:00 GMT</pubDate><category>indexing</category><author>Hassan</author></item><item><title>Why robots.txt Does Not Remove a Page From Google</title><link>https://searchenginebasics.dev/crawling/robots-txt-does-not-deindex/</link><guid isPermaLink="true">https://searchenginebasics.dev/crawling/robots-txt-does-not-deindex/</guid><description>Blocking a URL in robots.txt stops the fetch, not the listing. Here is why blocked pages still appear in results, and what actually removes them from the index.</description><pubDate>Sat, 05 Sep 2026 00:00:00 GMT</pubDate><category>crawling</category><author>Hassan</author></item></channel></rss>