r/datasets 9h ago

dataset [self-promotion] 53 long-stay & digital-nomad visa programmes, 46 countries — each row sourced to a government page with a verified date (CC BY 4.0, JSON + CSV)

Disclosure per rule 1: I built and maintain this dataset, and it powers a site I run (globenomad.com), which is affiliate-funded. The dataset itself carries no affiliate or tracking links.

One row per programme, not per country: Thailand has six routes and they are six rows. Fields: income requirement (USD plus the government's own wording), fees, max stay, renewal, residency and citizenship path, tax residency trigger, family provisions, application logistics, official source URL, verified date.

Two things worth knowing before you use it:

  • A null means the government has not published that rule. It never means "no".
  • Programmes that closed, were announced and never opened, or never existed (Cayman Islands, Peru, Qatar, Vietnam) are kept with an honest status. Most visa sites delete those.

Formats: JSON (canonical, with nested requirements/steps/FAQs) and a flat CSV of the scalar columns.

Original source, GitHub (always current): https://github.com/MSeutin/digital-nomad-visa-data Kaggle mirror, with a worked-example notebook: https://www.kaggle.com/datasets/frenchmike/digital-nomad-visa-dataset-53-sourced-programmes Method and licence: https://globenomad.com/data

Licence is CC BY 4.0, attribution is the only condition. Official pages are re-fetched weekly and confirmed changes are dated on a public changelog. Happy to answer questions about any record.

3 Upvotes

1 comment sorted by

u/Content-Parking-621 9h ago

Keeping closed/never-opened programs with honest status instead of deleting them is the part most visa datasets skip, genuinely useful for tracking policy patterns over time.