Ran the CSV through our deduper first. Categorisation was granular enough to segment by sub-vertical without regex. One column I'd add next refresh: secondary email.
Thanks for flagging the small gaps, noted on the minor gaps — they're tracked for the next refresh. Drop a note via Help → Contact Support and we'll line up a fresh export for you.