World Equivalent Names
Overview
Global Empirical Name Equivalents Registry: Geographic Variant Database
Standard algorithmic name matching frequently fails because it relies on rigid, theoretical rules rather than how names evolve in the real world. Our Global Empirical Name Equivalents Registry bridges this gap by indexing actual orthographic and phonetic variations based on documented real-world usage across highly specific geographic zones.
Engineered for localized precision, this dataset maps the real-world mutations, colloquial shifts, and script-transcription behaviors unique to target regions. It provides global enterprise software, risk-compliance engines, and identity management systems with the exact data infrastructure required to resolve cross-border identities with flawless regional context. Other potential applications include:
- Combating the Financing of Terrorism
- Culture-sensitive CRM applications
- Compliance and Governance
- Customer Data Management
- Employment Diversity
- Ethnic-based marketing campaigns
- Fraud Detection
- Entity Matching
- Identity Resolution
- Immigration Control
- Intelligence Analysis
- Name Screening
- KYC and Due Diligence
- Law Enforcement
- Name Screening
- PEP and sanctions screening
- Voters list correction
Related Datasets & Solutions
Technical Specifications
1. Geographically Anchor Data Architecture
- Empirical Usage Mapping: Contains millions of verified orthographic and phonetic variations based strictly on real-world document registries and field-tested usage.
- Geographic Stratification: Every name variation is tagged with specific geographic metadata, linking spelling preferences to exact countries, sub-regions, or cultural corridors.
- Master Script Parallelization: Maps colloquial and regional Latinized variations directly back to their native, localized script equivalents (e.g., Arabic, Cyrillic, Indic).
2. Advanced Phonetic & Orthographic Intelligence
- Dialectal Drift Processing: Identifies and links variants that emerge from regional vocalization patterns, ensuring colloquial adaptations are caught by text-matching systems.
- Bilingual Data Validation: Features fully vocalized and unvocalized reference standards to ensure flawless downstream tokenization in text-parsing tools.
- Fuzzy Engine Optimization: Designed to seamlessly replace or augment standard rule-based algorithms with real-world, deterministic matching tables to minimize search missed-hits.
3. Primary High-Impact Applications
- Hyper-Localized KYC & Compliance: Empowers compliance platforms to filter out regional noise and target variations used within specific geographies, preventing false matches.
- SaaS Enterprise Search & Indexing: Maximizes search engine retrieval rates by natively indexing names exactly how local populations spell them in everyday travel documents and forms.
- Data Harmonization & MDM: Cleans, matches, and deduplicates multi-national user registries by grouping regional spelling mutations under a single, globally mapped master identity record.
Sample downloads:
- TXT: larger text sample Database of World Equivalent Names
- Other: related samples samples page
Reference: DBEQ
Entries: 300,000+
Last updated: 11/6/2026
Field Definitions
VID: name variant ID
Variant: name in basic Latin script
Native: name variant in native writing script
Gender: M: male, F: female, U: unisex
Type: C: common, F: first/given name, L: last name/surname
Language: variant original language
Locale: country of origin, ISO country code
Sample
| ID | VID | Variant | Native | Gender | Type | Language | Locale |
|---|---|---|---|---|---|---|---|
| 1 | N0001 | Abraham | M | F | English | GBR | |
| 2 | N0001 | Avram | M | F | Serbian | Serbia | |
| 3 | N0002 | Andre | André | M | F | French | France |
| 4 | N0002 | Andrea | M | F | Bulgarian | Bulgaria | |
| 5 | N0002 | Andreas | M | F | Danish | Denmark | |
| 6 | N0002 | Andries | M | F | Dutch | Netherland | |
| 7 | N0002 | Andrija | M | F | Serbian | Serbia | |
| 8 | N0003 | Anthony | M | F | English | GBR | |
| 9 | N0003 | Antoaneta | F | F | Bulgaria | Bulgaria | |
| 10 | N0003 | Antoine | M | F | French | France | |
| 11 | N0003 | Anton | M | F | German | Austria | |
| 12 | N0003 | Antonella | F | F | Italian | Italy | |
| 13 | N0003 | Antonello | M | F | Italian | Italy | |
| 14 | N0003 | Antonina | F | F | Polish | Poland | |
| 15 | N0003 | Antonino | M | F | Italian | Italy | |
| 16 | N0003 | Antonio | M | F | Italian | Italy | |
| 17 | N0004 | Elisabeth | F | F | English | ||
| 18 | N0004 | Lisbeth | F | F | Danish | Denmark | |
| 19 | N0004 | Elisaveta | F | F | Bulgarian | Bulgaria | |
| 20 | N0004 | Elizabeth | F | F | English | GBR | |
| 21 | N0005 | Christian | M | F | Danish | Denmark | |
| 22 | N0005 | Christiane | M | F | German | Austria | |
| 23 | N0005 | Christina | F | F | Bulgarian | Bulgaria | |
| 24 | N0005 | Christo | M | F | Bulgarian | Bulgaria | |
| 25 | N0005 | Kristijan | M | F | Serbian | Serbia | |
| 26 | N0005 | Cristina | F | F | Dutch | Netherland | |
| 27 | N0005 | Hristo | M | F | Bulgarian | Bulgaria | |
| 28 | N0005 | Khristo | M | F | Bulgarian | Bulgaria | |
| 29 | N0006 | Vilhelm | M | F | Danish | Denmark | |
| 30 | N0006 | Willem | M | F | Dutch | Netherland | |
| 31 | N0006 | William | M | F | English | GBR | |
| 32 | N0007 | Josefina | F | F | Spanish | Spain | |
| 33 | N0007 | Joseph | M | F | German | Austria | |
| 34 | N0007 | Giuseppe | M | F | Italian | Italy | |
| 35 | N0008 | George | M | F | English | ||
| 36 | N0008 | Georges | M | F | French | France | |
| 37 | N0008 | Georgi | M | F | Bulgarian | Bulgaria | |
| 38 | N0008 | Giorgia | F | F | Italian | Italy | |
| 39 | N0008 | Giorgio | M | F | Italian | Italy | |
| 40 | N0008 | Gregorio | M | F | Spanish | Spain | |
| 41 | N0008 | Jorge | M | F | Spanish | Spain | |
| 42 | N0008 | Dorde | Đorđe | M | F | Serbian | Serbia |
| 43 | N0009 | Johnson | M | L | English | GBR | |
| 44 | N0009 | Johnston | M | L | |||
| 45 | N0009 | Jonassen | M | L | Danish | Denmark | |
| 46 | N0010 | Willemsz | M | L | Polish | Poland | |
| 47 | N0010 | Williams | M | L | English | GBR | |
| 48 | N0011 | Jacobowitz | M | L | Polish | Poland | |
| 49 | N0011 | Jacobsen | M | L | Danish | Denmark | |
| 50 | N0011 | Jacobson | M | L | Irish | Ireland |