ISO 24619:2011
Language resource management — Persistent identification and sustainable access (PISA)
Language resource management — Persistent identification and sustainable access (PISA)
- Статус документа:
- Действующий
- Формат:
- Электронный (PDF)
- Количество страниц:
- 34
- Дата публикации:
- 12 мая 2011 г.
- Издание:
- ISO IS 24619 edition 1 version 1
- ICS:
- 01.140.20
ISO 24619:2011 specifies requirements for the persistent identifier (PID) framework and for using PIDs as references and citations of language resources in documents as well as in language resources themselves. In this context, examples of language resources include such works as digital dictionaries, language-purposed terminological resources, machine-translation lexica, annotated multimedia/multimodal corpora, text corpora that have been annotated with, for example, morpho-syntactic information, and the like. Computational and applied linguists and information specialists create such resources. ISO 24619:2011 also addresses issues of persistence and granularity of references to resources, first by requiring that persistent references be implemented by using a PID framework and further by imposing requirements on any PID frameworks used for this purpose. PID frameworks also allow the association of general metadata with the identifier, which can also contain citation information. ISO 24619:2011 specifies minimum requirements for effective use of PIDs in language resources and cites the use of several possible existing standards and de-facto standards.
Abstract
ISO 24619:2011 (PISA) - Overview
ISO 24619:2011, titled Language resource management - Persistent identification and sustainable access (PISA), defines requirements for using persistent identifiers (PIDs) to reference, cite and access language resources reliably over time. The standard targets digital language resources - for example digital dictionaries, terminological resources, machine‑translation lexica, annotated corpora and multimedia/multimodal language data - and specifies minimum requirements for PID frameworks and PID usage to ensure persistence, resolvability and reproducible citation.
Key topics and technical requirements
- PID framework requirements: mandates that persistent references be implemented through an established PID framework and imposes minimum functional and operational requirements on such frameworks (resolver services, metadata association, persistence guarantees).
- PID usage and citation: describes how PIDs can be used as references and citations both in documents and embedded inside language resources; supports inclusion of citation metadata tied to identifiers.
- Granularity and parts identification: addresses how to identify and access resource parts, snapshots and versions; advises on granularity decisions for complex, federated or multi‑layered language resources.
- Collections and virtual incarnations: covers published collections and virtual assemblies (resource collection incarnations) that are referenced as a whole while enabling access to constituent parts.
- Complementary and practical constraints: discusses issues of persistence, sustainability (avoiding fragile dependencies), and best practices to enable machine and human resolvability.
- Terminology and definitions: provides clear definitions for resources, collections, archives, repositories, fragments, snapshots, versions and other core concepts relevant to language data management.
Practical applications and users
ISO 24619:2011 is intended for:
- Computational and applied linguists who create or cite corpora, lexica and annotated datasets.
- Information specialists, digital librarians and repository managers responsible for long‑term access and citation of language resources.
- Research projects and archives that need stable linking between distributed resources, parts, and multimedia elements.
- Developers of language technology platforms that must resolve identifiers reliably for automated processing (e.g., NLP pipelines, e‑Science workflows).
Typical applications:
- Assigning PIDs to corpora, datasets or lexical entries to enable reproducible research and machine‑actionable citations.
- Using resolver services to keep links valid when resources move locations or are mirrored.
- Referencing specific parts or snapshots of complex or federated resources in publications and metadata records.
Related standards (examples cited in ISO 24619)
- Citation formats: ISO 690, APA, MLA
- Fragment/part identification: ISO/IEC 21000‑17 (MPEG‑21), XPointer, IETF RFC 5147
- PID systems and implementations: DOI, Handle System, ARK, PURL
- Web and metadata technologies referenced for PID use and metadata attachment
ISO 24619:2011 helps organizations and researchers ensure persistent identification, sustainable access, and reliable citation of language resources - critical for reproducibility, long‑term preservation and machine‑readable scholarly infrastructures.
Технические детали
- Технический комитет
- ISO/TC 37/SC 4 - Language resource management
- SKU
- ISO 24619:2011
Похожие стандарты
Стандарты, упомянутые в описании
BS ISO 24619:2011
ДействующийLanguage resource management. Persistent identification and sustainable access (PISA).
ISO 6903:1984
ОтменёнCinematography — Motion-picture camera cartridge, 8 mm Type S, Model 1 (capacity 60 m) — Cartridge-camera int…
ISO/IEC 21000-17:2006
ДействующийInformation technology — Multimedia framework (MPEG-21) — Part 17: Fragment Identification of MPEG Resources
Overview ISO/IEC 21000-17:2006 - part of the MPEG-21 multimedia framework - defines a normative syntax for URI Fragment Identifiers to address parts of MPEG resources. It applies to resources whose I…