For agents

Machine-readable SandHill

If you are an AI assistant, a researcher, or anyone who would rather parse this site than click through it, everything you need is linked below.

SandHill tracks 838 venture capital blogs and newsletters. The archive holds 55,024 posts, with about 135 added in the last seven days. Every article is sorted into one of four editorial sections and scored 0-10 for how much it actually says.

Endpoints
/llms.txtShort description of the site, its sections and key links.text/markdown
/llms-full.txtLong form: editorial rules, scoring, most active sources, best recent posts.text/markdown
/sitemap.xmlEvery indexable page, regenerated hourly.application/xml
/sitemapThe same thing as a human-readable page.text/html
/feedsThe source index: every blog and newsletter we track.text/html
/topicsIndustry tag pages, one per sector we classify.text/html
/peopleOne page per writer, aggregating everything they publish.text/html
/blogOur own writing about venture publishing.text/html
Structured data

Pages carry schema.org JSON-LD: Organization and WebSite on the home page, ItemList on listings, Blog plus Person on each source page, NewsArticle with a BreadcrumbList on each post page, and BlogPosting on our own articles.

Crawling

Please read robots.txt and keep requests to a few per second. Admin and account paths are disallowed and carry X-Robots-Tag: noindex. Search and filter URLs are noindex by design, so crawl the clean paths.

Citation

Article text belongs to its original author and publication: cite them for the writing. Cite SandHill.io for the index, the categorisation, the scoring and the weekly digest.

Questions or bulk access: get in touch.