{"id":41911,"date":"2026-07-28T22:27:04","date_gmt":"2026-07-28T20:27:04","guid":{"rendered":"https:\/\/www.graviton.at\/letterswaplibrary\/self-promotion-paid-podcast-sponsorship-dataset-which-brands-sponsor-which-shows-with-the-verbatim-evidence-line-for-every-record-free-tier-available\/"},"modified":"2026-07-28T22:27:04","modified_gmt":"2026-07-28T20:27:04","slug":"self-promotion-paid-podcast-sponsorship-dataset-which-brands-sponsor-which-shows-with-the-verbatim-evidence-line-for-every-record-free-tier-available","status":"publish","type":"post","link":"https:\/\/www.graviton.at\/letterswaplibrary\/self-promotion-paid-podcast-sponsorship-dataset-which-brands-sponsor-which-shows-with-the-verbatim-evidence-line-for-every-record-free-tier-available\/","title":{"rendered":"[self-promotion] [PAID] Podcast Sponsorship Dataset: Which Brands Sponsor Which Shows, With The Verbatim Evidence Line For Every Record (free Tier Available)"},"content":{"rendered":"<p><!-- SC_OFF --><\/p>\n<div class=\"md\">\n<p>Disclosure: I built this, it&#8217;s my project, and paid tiers exist. There&#8217;s a free tier and everything shown below is viewable without signing up.<\/p>\n<p>What it is: structured sponsorship records extracted from public podcast RSS show notes. One row per (brand, episode):<\/p>\n<p>brand (canonically resolved) | show | episode | publish date | promo code | promo URL + registrable domain | sponsor type (paid \/ affiliate \/ house ad) | confidence | confidence tier | first seen | last seen | the verbatim sentence the claim came from<\/p>\n<p>Sample rows straight out of the DB:<\/p>\n<p>&#8211; AG1 on Huberman Lab, 2026-07-27, evidence: &#8220;AG1: <a href=\"https:\/\/drinkag1.com\/huberman\">https:\/\/drinkag1.com\/huberman<\/a>&#8220;<\/p>\n<p>&#8211; Visible on Good Hang with Amy Poehler, code HANG, 2 episodes, 14-day span<\/p>\n<p>&#8211; Saily on Machtwechsel (German news podcast), code &#8220;Machtwechsel&#8221;, 3 episodes over 18 days<\/p>\n<p>Method, since this sub cares about it: LLM extraction over the show-notes text, then a deterministic brand-resolution layer on top. Domain evidence merges entities first (drinkag1.com and athleticgreens.com collapse into one AG1 entity), exact normalized-name match second, and anything that is merely name-similar goes to an adjudication queue and is never auto-merged. That last rule is what keeps Dove the soap separate from Dove the chocolate. Every record retains its source sentence so any claim can be audited by hand.<\/p>\n<p>Honest limits, up front:<\/p>\n<p>&#8211; Show notes only. Ads that exist purely in audio and never appear in the notes are invisible to this. Transcript coverage is not built yet.<\/p>\n<p>&#8211; The corpus is small right now: 314 episodes across 93 shows, US + DE + FR. It grows daily but this is not a historical archive.<\/p>\n<p>&#8211; I am deliberately not publishing an accuracy percentage. I ran a held-out evaluation, then used its failures to fix the extractor, which burns that holdout. Any number I quoted today would be inflated. A fresh untouched holdout is the next task. Until then every record carries a confidence tier and only the CONFIRMED tier is presented as fact.<\/p>\n<p>&#8211; No spend or impression estimates. This answers who advertises where, not how much they paid.<\/p>\n<p>Free tier is 200 requests\/month, paid is $49\/$199\/$499. Keys are not self-serve yet, so the page is an early-access list rather than a checkout.<\/p>\n<p>Two things I would actually like this sub&#8217;s read on: is a per-record evidence string useful to you, or is it dead weight next to a confidence score? And what would you want joined onto this that is missing (show category, audience estimates, historical backfill)?<\/p>\n<p><a href=\"https:\/\/podintel.github.io\/?src=datasets\">https:\/\/podintel.github.io\/?src=datasets<\/a><\/p>\n<\/div>\n<p><!-- SC_ON -->   submitted by   <a href=\"https:\/\/www.reddit.com\/user\/Smart-Farmer1966\"> \/u\/Smart-Farmer1966 <\/a> <br \/> <span><a href=\"https:\/\/www.reddit.com\/r\/datasets\/comments\/1v99tdb\/selfpromotion_paid_podcast_sponsorship_dataset\/\">[link]<\/a><\/span>   <span><a href=\"https:\/\/www.reddit.com\/r\/datasets\/comments\/1v99tdb\/selfpromotion_paid_podcast_sponsorship_dataset\/\">[comments]<\/a><\/span><\/p><div class='watch-action'><div class='watch-position align-right'><div class='action-like'><a class='lbg-style1 like-41911 jlk' href='javascript:void(0)' data-task='like' data-post_id='41911' data-nonce='ddd07821da' rel='nofollow'><img class='wti-pixel' src='https:\/\/www.graviton.at\/letterswaplibrary\/wp-content\/plugins\/wti-like-post\/images\/pixel.gif' title='Like' \/><span class='lc-41911 lc'>0<\/span><\/a><\/div><\/div> <div class='status-41911 status align-right'><\/div><\/div><div class='wti-clear'><\/div>","protected":false},"excerpt":{"rendered":"<p>Disclosure: I built this, it&#8217;s my project, and paid tiers exist. There&#8217;s a free tier and everything&#8230;<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[85],"tags":[],"class_list":["post-41911","post","type-post","status-publish","format-standard","hentry","category-datatards","wpcat-85-id"],"_links":{"self":[{"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/posts\/41911","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/comments?post=41911"}],"version-history":[{"count":0,"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/posts\/41911\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/media?parent=41911"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/categories?post=41911"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.graviton.at\/letterswaplibrary\/wp-json\/wp\/v2\/tags?post=41911"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}