About this site 한국어

Unofficial explanatory translation of the Korean AI-Ready Data (AIRD) draft standard. The Korean text prevails.

Preparing data › Procedure

Choosing persistent identifiers

Type · ProcedureReading time · about 3 minData providers

ContentsContents
  1. Deciding on a persistent identifier
  2. Construction rules
  3. Examples
  4. Keeping identifiers stable once assigned
  5. Points to confirm with the responsible department
Persistent identifier decision flow

How to give a dataset a persistent identifier that does not change. Check your organization’s identifier policy first; if there is none, follow the interim rules.

What the standard requires — a persistent identifier in URI form. A resolvable URI is the principle. [Part 3 Annex A.1] Required

If your organization (public institution, company, or association) has an issuance policy, that policy takes precedence. This page gives interim rules to use until a policy is in place. Recommended

Deciding on a persistent identifier

SituationDecision
Organization has an issuance policyApply the organization’s policy (takes precedence over the interim rules)
No policy · organization has a domain1st choice · set a fixed path under the organization’s domain — https://data.<org-domain>/id/dataset/<name>
No policy · no organization domain2nd choice · check whether a parent organization (the supervising body of a public institution, or a company’s group) or a shared namespace can be used
No policy, domain, or shared namespaceAssign an identifier whose uniqueness is guaranteed · record the fact that it does not resolve as a limitation

Never leave the identifier blank. Without an identifier, the dataset does not pass the stage 1 check.

Construction rules

An identifier has two parts: namespace + dataset name.

https://data.<org-domain>/id/dataset/<dataset-name>
└─────────── namespace ───────────┘└─ dataset name ─┘
   fixed area the organization controls   never changed once set
PartIncludeExclude
NamespaceA fixed address representing the organizationDepartment name — changes on reorganization
Dataset nameA fixed name for the dataset · a meaningless serial number is also allowedVersion · year — change on update
File format — invalid when the format changes
Portal listing key — changes when re-registered on the portal

Key rule: do not put anything that changes into the identifier. The data provider records version, reference time, and format in separate elements: version in dcat:version, reference time in the temporal coverage, and format in dct:format.

Examples

The example identifier on this site is https://data.example.go.kr/id/dataset/biz-registry.

Examples — expand the 8-row table
JudgmentValueReason
Suitablehttps://data.example.go.kr/id/dataset/biz-registryFixed namespace · unchanging name
Suitablehttps://data.example.go.kr/id/dataset/d-00417A meaningless serial number is also allowed
Suitablehttps://data.example.com/id/dataset/sales-dailyFixed path under a company domain · unchanging name
Unsuitablehttps://data.example.go.kr/file/biz_2025_v3.csvContains year, version, and format. Record the file address as the access URL (dcat:accessURL)
UnsuitableBIZ-2025-001Not a URI · contains a year
Unsuitable15029008Portal listing key · changes when re-registered on the portal
Unsuitablehttps://www.data.go.kr/data/15029008/fileData.doPortal detail page address · changes on re-registration. Record it as the landing page (dcat:landingPage)
Unsuitablehttps://market.example.com/products/48213Data marketplace product page address · changes when the product is re-registered. Record it as the landing page (dcat:landingPage)

Do not use the portal listing key or the portal detail page address as the persistent identifier. Both values change when the data is re-registered on the portal. The same applies to a data marketplace’s product page address.

Record the portal detail page address as the landing page (dcat:landingPage). Record the file download address as the access URL (dcat:accessURL). The persistent identifier, landing page, and access URL are different values.

Keeping identifiers stable once assigned

SituationIdentifier handling
Values cleanedKeep — same dataset
New version publishedKeep — versions are distinguished by version number
File format addedKeep — a distribution is added
Purpose-specific operation file createdKeep — use the source dataset’s identifier
Portal re-registration · migrationKeep — update only the landing page and access URL
Datasets merged · splitAssign a new one — a different dataset

The purpose of an identifier is not to change. Changing an identifier breaks every reference that pointed to the data.

Points to confirm with the responsible department

When requesting a policy, the data provider passes on the following points.

Point to confirmPurpose
Whether the organization controls a domainDeciding the namespace
Whether the domain’s retention period is guaranteedEnsuring identifier persistence
Whether an existing dataset numbering scheme existsReflecting the existing numbering scheme in the URI path
Who manages identifier issuance and retirementPreventing duplicate issuance

The rules on this page are interim rules. Once an organizational policy or higher-level guideline is in place, that policy or guideline takes precedence. Identifiers already assigned are kept.

Completion criteria

  • The “dataset identifier” is in URI form.
  • The identifier contains no year, version, file format, department name, or portal listing key.
  • The identifier, landing page, and access URL are different values.

Measuring quality (stage 2)

Last updated · 2026-10-07Report an error