- Follow all links within the domain to traverse the linked site.
- Only child pages of the provided URL to stay within a section, such as
/articles. - Set Page limit to
1to index only the starting page. - Set Depth of search to limit how many links away from the start the crawler may go.
Data sources
Website sources
A website source makes public pages available for search. Add it from the target Space’s Connections area and enter a public starting URL. The crawler follows links it finds on that page; it does not guess or discover pages that are not linked.
Choose the crawl boundary to fit the site: