Ran the CSV through our deduper first. Categorisation was granular enough to segment by sub-vertical without regex. One column I'd add next refresh: secondary email.
Appreciate the constructive note, we're actively working on the gaps you called out. Open a ticket via the Support tab and we'll prioritise the columns you flagged for the next refresh.