How Pioneer Finds Leads
Pioneer explores the open web the way a diligent researcher would, iteratively, with every candidate identity-checked and qualified before you see it.
Last updated August 11, 2026
You don't need this page to use Pioneer. But once you know how discovery actually works, you'll write sharper descriptions and criteria, you'll understand why leads keep arriving hours after you start a search, and you'll know exactly what kind of lead Pioneer can find that no database can.
Pioneer searches the web, not a database
Most lead-generation tools are databases: a fixed set of records you filter. If your ideal lead isn't in the database, or the signal you care about isn't a column, you're out of luck.
Pioneer searches the open web. There's no pre-existing dataset limiting what you can discover: if the information exists publicly (a press release, a job posting, a case study, a conference agenda), Pioneer can find it and use it. That's why criteria like "should have open compliance roles" or "must have published technical papers" work in Pioneer and don't work anywhere else: no database has that column, but the web has the answer.
One instruction becomes many searches
Take a single criterion: "Must have raised Series A funding." Pioneer doesn't just append "Series A" to a query. It looks for funding announcements, checks press coverage, cross-references what it finds, and lets each result suggest where to look next.
That's the pattern for the whole search. Pioneer:
- Runs multiple search strategies, not one query
- Explores promising results deeper
- Follows related links and references
- Adjusts its approach based on what it finds
- Continues until it's confident it has covered the space
This is why discovery takes time, and why that's a feature: initial leads usually arrive quickly, but thorough coverage of a space takes hours, not seconds. A tool that answers instantly is querying a database; a tool that keeps delivering all afternoon is doing research.
Every candidate has to prove its identity
Before anything becomes a lead on your board, it passes three checks, each one there because its absence would fill your board with junk:
- It must be the right kind of thing. If your pipeline looks for companies, a listicle about companies doesn't qualify, and neither does a person. Pioneer checks that each candidate is actually the kind of entity you're looking for.
- It must resolve to its own web address. Search results usually point at pages about a company: a LinkedIn page, a directory entry, an aggregator. Pioneer resolves each candidate to its own primary website, because all further research starts from that address, and research aimed at a directory page would research the directory.
- It must be new. Candidates are deduplicated against leads already in your pipeline, so the same company doesn't land twice. Dedup checks more than the URL: a company that enters through its own website and through a directory profile (like a Crunchbase or an accelerator listing) resolves to two different URLs, but Pioneer recognizes them as the same entity and keeps only one. The fingerprint is multi-faceted (normalized name, domain, URL) and includes fuzzy matching, so similar names that refer to the same company are caught.
Leads arrive qualified, not raw
Discovery and qualification aren't separate steps you manage. Every candidate that passes the identity checks is researched against your criteria before it reaches you: evidence gathered per criterion, a summary written, a verdict recorded with its reasoning. What lands in your Leads column is a researched answer, ready to judge. How to read one: Read a Lead's Research.
What Pioneer can find, and what it can't
If it's public on the internet, Pioneer can find it:
- Companies and startups: websites, Crunchbase profiles, press coverage
- People: LinkedIn profiles, personal websites, speaker bios
- Organizations: nonprofits, agencies, institutions
- Investors: fund websites, portfolio pages, news coverage
- Anything else: if it has a web presence and matches your criteria
What it can't:
- Private information: data that isn't publicly available online
- Paywalled content: information behind subscriptions or logins
- Very recent information: there's a lag between publication and searchability
This boundary is worth internalizing when you write criteria: a rule about private information comes back Inconclusive everywhere. See Write rules the web can answer.
Discovery and import get the same treatment
Pioneer offers two ways to add leads, and they converge:
Discovery (automatic): describe what you're looking for; Pioneer finds candidates you don't know about yet.
Import (manual): paste a list of names, links, or emails; Pioneer researches and qualifies each one exactly as if it had discovered them, with the same enrichment, the same criteria verdicts, and the same board.
Use both in the same pipeline: discovery to explore the space, import to add referrals, conference contacts, or an existing list. A lead's origin never changes its quality.
Help the search help you
Be specific in your description. "Fintech startups" is okay. "B2B fintech startups building payment infrastructure for marketplaces" is better; every extra qualifier redirects real search effort.
Write criteria as search hints. Criteria don't just filter. They aim the exploration. "Must have raised Series A" tells Pioneer where to look, not just what to accept.
Give it time. Check back after a few hours for the full picture. And you don't have to wait around: when a run finishes, Pioneer emails you the results. See Export Your Leads.
Next steps
- Teach Pioneer What a Great Lead Looks Like: write the rules that steer all of this
- Import Your Own Leads: when you already have the list
- Refine Your Pipeline: make each search better than the last
Need help?
If you have questions, reach out to us at support@pioneerclimate.com