DONE - today's run finished 09:46:06; beehiiv post created
2026-09-03 23:40 PDT · json · log tail
300 row(s) attempted, 130 article(s) read (43%)
why the fetches failed:
299 network (HTTPError)
8 network (TimeoutError)
8 too short after HTML strip
4 looks_like_code heuristic
15 article(s) found by searching the headline (Google link was unusable)
1 fetched article(s) REJECTED - the text never mentioned this deal's companies
hosts attempted:
171 news.google.com
9 www.prnewswire.com
6 www.pehub.com
5 finance.yahoo.com
5 pulse2.com
4 www.citybiz.co
4 www.globenewswire.com
3 dallasinnovates.com
10 row(s) updated, 10 person record(s) added, 3 source URL(s) de-Googled
fields filled: deal_value_type 3, geography 3, deal_value 2, deal_leads 2
originals kept as output/deals_*.csv.bak.articles
Then rebuild the site so the new people get pages:
python3 build_pages.py --all
===== backfill done 22:25:36 =====