OLAC Record
oai:www.ldc.upenn.edu:LDC2007T02

Metadata
Title:English Chinese Translation Treebank v 1.0
Access Rights:Licensing Instructions for Subscription & Standard Members, and Non-Members: http://www.ldc.upenn.edu/language-resources/data/obtaining
Bibliographic Citation:Bies, Ann, et al. English Chinese Translation Treebank v 1.0 LDC2007T02. Web Download. Philadelphia: Linguistic Data Consortium, 2007
Contributor:Bies, Ann
Palmer, Martha
Mott, Justin
Warner, Colin
Date (W3CDTF):2007
Date Issued (W3CDTF):2007-01-22
Description:*Description* This release of English Chinese Translation Treebank v. 1.0 consists of 146,300 words in 325 files of individual news stories from Xinhua News Agency (corresponding to the Xinhua data in Chinese Treebank 5.0 LDC2005T01) that are translated into English, part-of-speech tagged and treebanked. The files were compressed using gzip. The source files for the treebank annotation contain the final updated translation of these files. Translation errors that prevented complete treebank annotation have been corrected. This translation and annotation were completed in October 2004 and supersede any earlier translation. This publication was compiled under National Science Foundation Grant #IIS-0325646. *Samples* For an example of the data in this publication, please view this sample.
Extent:Corpus size: 7680 KB
Identifier:LDC2007T02
https://catalog.ldc.upenn.edu/LDC2007T02
ISBN: 1-58563-408-5
ISLRN: 877-578-293-641-1
DOI: 10.35111/jn2g-zd52
Language:English
Language (ISO639):eng
License:LDC User Agreement for Non-Members: https://catalog.ldc.upenn.edu/license/ldc-non-members-agreement.pdf
Medium:Distribution: Web Download
Publisher:Linguistic Data Consortium
Publisher (URI):https://www.ldc.upenn.edu
Relation (URI):https://catalog.ldc.upenn.edu/docs/LDC2007T02
Rights Holder:Portions © 1994-1998 Xinhua News Agency, © 2004, 2007 Trustees of the University of Pennsylvania
Type (DCMI):Text
Type (OLAC):primary_text

OLAC Info

Archive:  The LDC Corpus Catalog
Description:  http://www.language-archives.org/archive/www.ldc.upenn.edu
GetRecord:  OAI-PMH request for OLAC format
GetRecord:  Pre-generated XML file

OAI Info

OaiIdentifier:  oai:www.ldc.upenn.edu:LDC2007T02
DateStamp:  2020-11-30
GetRecord:  OAI-PMH request for simple DC format

Search Info

Citation: Bies, Ann; Palmer, Martha; Mott, Justin; Warner, Colin. 2007. Linguistic Data Consortium.
Terms: area_Europe country_GB dcmi_Text iso639_eng olac_primary_text


http://www.language-archives.org/item.php/oai:www.ldc.upenn.edu:LDC2007T02
Up-to-date as of: Mon Mar 25 7:20:12 EDT 2024