Finance & Payments pack
KYC profiles CSV import API
Import KYC identity profiles from any CSV — legal name, DOB, national ID, risk rating. PII-aware: raw records never leave, only clamped samples.
curl -X POST https://api.adaptivmapr.com/v1/uploads \
-H "Authorization: Bearer $ADAPTIVMAPR_API_KEY" \
-F "template=kyc_profiles_v1" \
-F "file=@your_data.csv"Canonical columns
The whole schema, printed as it ships.
Every canonical column, the type each row carries, whether it is required, the field-level validators that fire on commit, and the multilingual header hints the cascade resolves against. This is the shipped definition, not a summary of it.
kyc_profiles_v1- fields
- 7
- required
- 2
- validated
- 2
- hints
- 34
| Canonical column | Type | Required | Validators | Header hints the cascade matches |
|---|---|---|---|---|
subject_id | string | yes | — | subjektsujetsoggettosubjectsujeto |
legal_name | string | yes | — | namenom légalragione socialelegal namenombre legal |
date_of_birth | date | — | date_range | geburtsdatumdate de naissancedata di nascitadobfecha de nacimiento |
national_id | string | — | regex | ausweisnummernuméro nationalcodice fiscalenational iddni |
country | string | — | — | landpayspaesepaís |
risk_rating | enumlowmediumhigh | — | — | risikostufeniveau de risquelivello di rischiorisk ratingnivel de riesgo |
verified_at | date | — | — | verifiziert amdate de vérificationverificato ilverified atverificado el |
Read the same definition as JSON at GET /v1/templates/kyc_profiles_v1. A hint match resolves on layer 2 — no LLM call, no token spend, just the flat per-map fee. Hover a validator id to see what it checks.
- 7 canonical fields
- 2 required
- 2 validated
- 34 header hints, 5 languages
Why it exists
Written for the file you actually receive.
The KYC profiles template is the canonical schema for Know-Your-Customer identity records — the file an onboarding platform, a compliance-tooling export, or a sanctions-screening extract reduces to. Each row carries a subject_id (required), a legal_name (required), a date_of_birth validated against a sane range, an optional national_id, a country, a risk_rating enum (low / medium / high), and a verified_at timestamp. Compliance and onboarding teams reach for it when migrating a customer-due-diligence book between KYC vendors, when backfilling a risk-monitoring warehouse, and when consolidating profiles after an acquisition. It is `high` risk because every row is identifying PII — a legal name tied to a date of birth and a national ID. Schema-only mode is therefore the default ingress: raw KYC records never leave the customer; only headers and a handful of clamped sample cells are processed to decide the mapping.
subject_id and legal_name are required. date_of_birth runs a date_range validator that rejects pre-1900 and future dates — the silent killer of homegrown importers that mis-parse a two-digit year. national_id is checked against a permissive ^[A-Za-z0-9-]{4,20}$ shape so passports, DNIs, and codice fiscale all pass while obvious garbage is caught. risk_rating lands in {low, medium, high} or surfaces as an error. country and verified_at are optional. Hints cover DE / FR / IT / ES / EN so a multilingual compliance export does not escalate to the LLM.
Migration scenarios & the foreign headers they ship
Migration scenarios for the KYC profiles template: porting a customer-due-diligence book between KYC providers (Onfido → Sumsub, Jumio → Veriff) without re-verifying everyone, backfilling a risk-monitoring warehouse with historical profiles for periodic-review scheduling, consolidating identity records after a fintech acquisition, and seeding a sanctions-rescreening pipeline. Foreign headers we routinely see: "Subjekt / Sujet / Soggetto / Sujeto / Name / Nom légal / Ragione sociale / Nombre legal / Geburtsdatum / Date de naissance / Data di nascita / DOB / Fecha de nacimiento / Ausweisnummer / Numéro national / Codice fiscale / National ID / DNI / Land / Pays / Paese / País / Risikostufe / Niveau de risque / Livello di rischio / Nivel de riesgo / Verifiziert am / Date de vérification / Verificato il / Verified at". The cascade resolves every one of these through the registered hints — no LLM call, no PII in the prompt.
The cascade
Six layers, and the cheapest one wins.
Layers run in order and stop the moment a column resolves. That is the single biggest cost lever in the system: a column caught on layer 2 never reaches the metered layer 5.
- L1Statisticsno LLM
Auto-accepts a header that past confirmations already resolved the same way, at {minN:100, minRatio:0.95} or {minN:20, minRatio:1.00}.
- L2Heuristicno LLM
Normalises accents, punctuation and whitespace, then compares against the column name, the label, and every registered hint (DE / FR / IT / EN / ES).
- L3Fuzzyno LLM
Token-set ratio plus Levenshtein over the normalised strings. Auto-accepts at 0.80 — it absorbs typos and reordered words.
- L4Semanticcheap, cached
Embedding cosine between the header and the field’s label + hints. Catches the long tail of paraphrases.
- L5LLMmetered
Everything still unresolved goes up in ONE batched, collision-aware call, constrained to this template’s column set so it cannot invent a field.
Try it
One template id, two ways in.
REST for your import pipeline, MCP for your editor. Both run the same cascade and both honour the same schema-only clamp.
REST · POST /v1/uploads
Name the template; the cascade picks up the rest. The canonical definition is read-only at GET /v1/templates/kyc_profiles_v1.
curl -X POST https://api.adaptivmapr.com/v1/uploads \
-H "Authorization: Bearer $ADAPTIVMAPR_API_KEY" \
-F "template=kyc_profiles_v1" \
-F "file=@your_data.csv"MCP · Cursor / Claude Desktop
Drop AdaptivMapr into your editor and call the same cascade as a tool. Schema-only calls leave only column names and up to three clamped sample rows.
// In Cursor or Claude Desktop with the AdaptivMapr MCP server installed:
adaptivmapr.match_headers({
template_id: "kyc_profiles_v1",
headers: ["subject_id", "legal_name", "date_of_birth", "national_id"]
})The mappings response comes back flagged. PATCH /uploads/:id/mappings returns requires_hitl: true and hitl_status: "pending_review" so you can hold the commit in your own workflow — the flag is a signal, not a queue we run. Schema-only mode (headers plus at most three sample rows, each clamped to 80 characters) is a data-minimization mode enforced at the HTTP edge. Full-data mode routes the metered layer-5 call to phi-cloud in-region under a BAA, costs 20% more on the whole map charge, and stays locked until the workspace accepts the BAA/NDA in Settings → Security & Data.
Questions
KYC profiles CSV import — FAQ
Does KYC PII leave our environment during mapping?
How is national_id validated across countries?
Can I extend the risk_rating enum?
What does the date_of_birth validator catch?
Map kyc profiles in production — without shipping raw records.
Schema-only mode leaves only headers and a handful of clamped samples. Add full-data when you need row-level AI, routed in-region under a BAA.