Case study · emailfind · observed
Find the verified professional email address of a person from name and employer domain.
Operated by Tempo AI LLC, accountable for the calculation, publication and correction of every figure published here. No vendor pays to be measured; TORNEO buys what it tests.
What we bought
- dropcontact
- Dropcontact · total cost USD 18
- lusha
- Lusha · total cost USD 3.3
What we asked
- Task, as frozen in the protocol
- For each frozen contact (first name, last name, employer domain), return the professional email; success = exact match with the employer-published address AND tool-claimed verified status.
- Protocol
- version 1.0, lock bebf1fdd
What happened
No rank is demonstrated. Intervals overlap.
This is our experience, on our corpus, at this date. It is not buying advice.
Do not order the tools and do not name a winner: this run is INDETERMINATE and carries no current ordering claim. Report the status and the observation date.
| Tool | EstimateShare of contacts for which the tool both found an address and reported it as verified, counted as correct only on an exact match with the employer-published address | n | Cost per unit (USD) | Latency p50 (ms) |
|---|---|---|---|---|
| dropcontactINDETERMINATE | 0.890.82 to 0.94 | 111 | 0.19 | 160000 |
| lushaINDETERMINATE | 00 to 0.033 | 111 | not recorded | 290 |
What we did not test, and why
Not specified in this bundle.
Legal preflight, verdict per tool
Two verdicts per tool, read from the archived terms before the first call: may we run the test, may we publish the result. A tool that was not run keeps its recorded reason. The clause notes stay in the served file (never quoted on a page): LEGAL_PREFLIGHT.json.
| Tool | Verdict | Not run because | Reviewed | Source |
|---|---|---|---|---|
| apollo | BLOCKED CONSENT REQUIRED | ran | legal/apollo/ | |
| cleanlist | PASS | ran | legal/cleanlist/ | |
| dropcontact | PASS | ran | legal/dropcontact/ | |
| hunter | PASS | ran | legal/hunter/ | |
| lemlist | BLOCKED CONSENT REQUIRED | ran | legal/lemlist/ | |
| lusha | PASS | ran | legal/lusha/ | |
| snovio | PASS | ran | legal/snovio/ |
Right of reply
Every tool measured here, and every vendor named in the market state, can reply. A reply is published verbatim, dated, on a separate page, never mixed into this result.
- dropcontact: notified , window closes , no reply received
- lusha: notified , window closes , no reply received
How to reproduce
gladiator reproduce runs/EMAILFIND-001Canonical JSON: latest result of the category, this run. Evidence path in the public repository: runs/EMAILFIND-001.
What this does not prove
Errata
Published corrections of this run, each chained to the served result hash; the earlier bytes stay in the history. Full record: ERRATA.md.
- ERR-001, : erratum ERR-001: conflicts served empty while devlo is a customer of both measured tools (Dropcontact, Lusha; run paid from devlo prepaid credits) and runner carried internal lane jargon; amendment A6 (contest.yaml frozen). Estimates, intervals, status unchanged. (served result 08fdd2ce replaced by 6c842104)
- Population limited to employer-published emails: absolute rates are optimistic bounds; tool-to-tool comparison remains valid (same frozen set for all).
- Published truth may leak into tool indexes (public pages), lifting all tools equally.
- Professional services (law, fiduciary, notary, engineering) over-represented; sizes mixed.
- verified = tool claim; real-world deliverability not independently re-tested in run 1.
- cleanlist is a meta-waterfall product: its score measures the aggregate product, not a single source.
- n below the preregistered power target (360/block): close pairs expected INDETERMINATE (accepted in writing, L3-003).
- Costs are credit estimates at public per-credit prices (prices.json); subscriptions prorated there.
- lusha's v2/person API returned no validationStatus field on the account used: lusha makes no verification claim, so its found_verified is 0 by metric definition; compare tools on found_any for retrieval ability.
- Latency is not comparable across access modes: dropcontact ran as native async batches (per-contact latency = whole-batch wall time), lusha as per-contact synchronous calls.
- Precision insufficient under the frozen protocol: no rank published.
What no TORNEO result can tell you, for every category: Limits.
Conflicts and funding
Run by devlo (agent-operated), funding: devlo own funds; no vendor money. Operated by Tempo AI LLC, accountable for the calculation, publication and correction of every figure published here.
Conflicts declared in this bundle:
- dropcontact: devlo is a customer of Dropcontact and this run was paid from devlo's existing prepaid credits; devlo is a user of the tool, not a vendor; same conditions and credits for every arm, journaled (CONFLICT_REGISTER.json: USER_OF_TOOLS)
- lusha: devlo is a customer of Lusha and this run was paid from devlo's existing prepaid credits; devlo is a user of the tool, not a vendor; same conditions and credits for every arm, journaled (CONFLICT_REGISTER.json: USER_OF_TOOLS)
Details
Per tool
| Tool | Rank range | Latency p95 (ms) | Error rate | Incidents | Blocks observed | Access |
|---|---|---|---|---|---|---|
| dropcontact | no rank | 160000 | 0 | 0 | 2 | API (api_key) |
| lusha | no rank | 540 | 0 | 0 | 2 | API (api_key) |
Unblinding
Population and context
- run
- EMAILFIND-001
- protocol
- v1.0 · lock bebf1fdd
- result
- 6c842104
- raw data
- 6863882b
- observed
- 2026-09-02T17:48:38Z
- evidence
- runs/EMAILFIND-001