Building the foundation text for Nanyang Technological University : multilingual corpus (NTU-MC)

The NTU-MC is a multilingual corpus that taps on the availability of multilingual text available in Singapore. The current version of NTU-MC contains a total of ~375,000 words (15,096 sentences) for the NTU-MC in 6 languages (English, Chinese, Japanese, Korean, Indonesian and Vietnamese) from 6 lang...

Full description

Bibliographic Details
Main Author: Tan, Li Ling.
Other Authors: Francis Bond
Format: Final Year Project (FYP)
Language:English
Published: 2012
Subjects:
Online Access:https://hdl.handle.net/10356/92348
http://hdl.handle.net/10220/7790