# Semantic microdata and more content for crawlers

**URL:** https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211
**Category:** Feature
**Created:** [13.Февраль.2015 13:51:04 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211 "2015-02-13T13:51:04Z")
**Posts on this page:** 7
**Page:** 1

<div class="post-metadata">

### Author: ![fantasticfears](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/fantasticfears/32/119608_2.png) [@fantasticfears](https://meta.discourse.org/u/fantasticfears)
#### Post date: [13.Февраль.2015 13:51:04 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/1 "2015-02-13T13:51:04Z")

</div>

[Breadcrumb](http://schema.org/BreadcrumbList) is one of many semantic markup described in [http://schema.org](http://schema.org). Search engines rely on these markups to improve the display of search result.

> [@Breadcrumbs microformat for SEO](https://meta.discourse.org/t/breadcrumbs-microformat-for-seo/14634):
>
> Given the url structure Discourse uses, Google has little hints about the hierarchy of the forum. How about adopting the microdata for breadcrumbs? Discourse already has the categories in the right order rendered in the posts page. I don’t have any articles to indicate if ranking improves. One of the advantages is to allow the user to get to a particular category from the search results. The breadcrumbs replace the link, just bellow the title. [Implementation](https://support.google.com/webmasters/answer/185417?hl=en) with microdata is trivia…

This feature request is to add more things for crawler.

- **More microdata** can be added into topic view, lists, category list and faq page. Maybe it can help for displaying topic list in the search engines.

![](https://global.discourse-cdn.com/meta/original/3X/a/5/a5d601393acc61bc521cd0a6755b36040c3aedb0.png)

Beyond that, add some content for crawler in some views.

- **Add category description** in the category list.
- **Add link to category list** in topic list page. There is no link to get there.

BTW, can I add `360Spider` to crawler ua list? It’s the second search engine in China now. (appr. Baidu 56% - 360 30%. bye to the third search engine for the last 10 😰 )

---

<div class="post-metadata">

### Author: ![codinghorror](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/codinghorror/32/110067_2.png) [@codinghorror](https://meta.discourse.org/u/codinghorror)
#### Post date: [14.Февраль.2015 03:24:30 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/2 "2015-02-14T03:24:30Z")

</div>

Sure add 360 UA that is fine.

---

<div class="post-metadata">

### Author: ![fantasticfears](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/fantasticfears/32/119608_2.png) [@fantasticfears](https://meta.discourse.org/u/fantasticfears)
#### Post date: [14.Февраль.2015 14:34:28 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/3 "2015-02-14T14:34:28Z")

</div>

Does this looks fine to you? @codinghorror

[https://github.com/fantasticfears/discourse/commit/4ede84758e5d4536b6fc50c4440bce374fa6e115](https://github.com/fantasticfears/discourse/commit/4ede84758e5d4536b6fc50c4440bce374fa6e115)

---

<div class="post-metadata">

### Author: ![codinghorror](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/codinghorror/32/110067_2.png) [@codinghorror](https://meta.discourse.org/u/codinghorror)
#### Post date: [15.Февраль.2015 01:00:04 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/4 "2015-02-15T01:00:04Z")

</div>

I would not link Top, that is entirely duplicate links from the crawler’s perspective. The rest seems fine.

---

<div class="post-metadata">

### Author: ![fantasticfears](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/fantasticfears/32/119608_2.png) [@fantasticfears](https://meta.discourse.org/u/fantasticfears)
#### Post date: [15.Февраль.2015 05:58:49 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/5 "2015-02-15T05:58:49Z")

</div>

What about `<meta name="robots" content="nofollow">` for top page? Then the top page can be accessible for `noscript`, eg Lynx.

---

<div class="post-metadata">

### Author: ![codinghorror](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/codinghorror/32/110067_2.png) [@codinghorror](https://meta.discourse.org/u/codinghorror)
#### Post date: [15.Февраль.2015 06:29:30 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/6 "2015-02-15T06:29:30Z")

</div>

I seriously doubt anyone is using those crawler-only pages.

---

<div class="post-metadata">

### Author: ![chapoi](https://sea3.discourse-cdn.com/meta/user_avatar/meta.discourse.org/chapoi/32/537252_2.png) [@chapoi](https://meta.discourse.org/u/chapoi)
#### Post date: [04.Декабрь.2025 11:33:33 UTC](https://meta.discourse.org/t/semantic-microdata-and-more-content-for-crawlers/25211/7 "2025-12-04T11:33:33Z")

</div>


