“ISO 11940:2007 Information and documentation – Transliteration of Thai characters into Latin characters – Part 2: Simplified transcription of Thai language” has been added to your cart. View cart

ISO 24614:2011

Name: ISO 24614:2011 Language resource management - Word segmentation of written texts - Part 2: Word segmentation for Chinese, Japanese and Korean
SKU: 33967f531121
Availability: InStock

ISO 24614:2011 Language resource management – Word segmentation of written texts – Part 2: Word segmentation for Chinese, Japanese and Korean

CDN $312.00

SKU: 33967f531121 Category: ICS:01.140.10

Description

Description

The basic concepts and general principles of word segmentation as defined in ISO 24614-1 apply to Chinese, Japanese and Korean. Text needs to be segmented into tokens, words, phrases or some other types of smaller textual units in order to perform certain computational applications on language resources, such as natural language processing, information retrieval and machine translation. ISO 24614-2:2011 is restricted to the segmentation of a text into words or other word segmentation units (WSUs). This task is distinct from morphological or syntactic analysis per se, although it greatly depends on morphosyntactic analysis. It is also different from the task of laying out a framework for constructing a lexicon and identifying its lexical entries, namely lemmas and lexemes. The frameworks for the latter tasks are provided by ISO 24611, ISO 24613 and ISO 24615.

ISO 24614-2:2011 specifies rules for delineating WSUs for Chinese, Japanese and Korean. Some rules are common to all three languages, though each language also has its own distinct rules for identifying WSUs. The common features are discussed, then the distinct rules are laid out for Chinese, for Japanese and for Korean.

Edition

Published Date

2011-08-25

Status

PUBLISHED

Language

English

Abstract

Previous Editions

Can’t find what you are looking for?

Please contact us at:

[email protected]

Customer Care

Customer Care

Quick Links

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

Placeholder headline

ISO 24614:2011

ISO 24614:2011 Language resource management – Word segmentation of written texts – Part 2: Word segmentation for Chinese, Japanese and Korean

Description

Edition

Published Date

Status

Pages

Language

Format

Secure – PDF details

Abstract

Previous Editions

Can’t find what you are looking for?

Please contact us at:

Related Documents

ISO 259:1994 Information and documentation – Transliteration of Hebrew characters into Latin characters – Part 2: Simplified transliteration

ISO 233:2023 Information and documentation – Transliteration of Arabic characters into Latin characters – Part 3: Persian language – Transliteration

ISO 11940:1998 Information and documentation – Transliteration of Thai

ISO 9:1995 Information and documentation – Transliteration of Cyrillic characters into Latin characters – Slavic and non-Slavic languages

Need Help?

Customer Care

Need Help?

Customer Care

Quick Links