MADCAT phase 3 training set.
(/)

Title MADCAT phase 3 training set.
Title Original /
Publication Place - [Philadelphia, PA] : Linguistic Data Consortium, c2013.
Subject Arabic language -- Written Arabic -- Data processing. Arabic language -- Machine translating. Arabic language -- Translating into English. Machine translating. Machine translating. Arabic language -- Translating into English. Arabic language -- Machine translating.
Type Other
Language Arabic
Digital Yes
Manuscript No
Physical Dimensions |
Library: University of Chicago
Record ID 9347803
Notes Title from disc label.Data type: Text.Data sources: Newsgroups, newswire, weblogs.Applications: Handwriting recognition, machine translation."LDC2013T16".Authors: David Lee, Safa Ismael, Dave Doermann, Stephanie Strassel, Zhiyi Song, Stephen Grimes.Arabic. 1 DVD ; 4 3/4 in.. "MADCAT (Multilingual Automatic Document Classification Analysis and Translation) Phase 3 Training Set contains all training data created by the Linguistic Data Consortium (LDC) to support Phase 3 of the DARPA MADCAT Program. The data in this release consists of handwritten Arabic documents, scanned at high resolution and annotated for the physical coordinates of each line and token. Digital transcripts and English translations of each document are also provided, with the various content and annotation layers integrated in a single MADCAT XML output. The goal of the MADCAT program is to automatically convert foreign text images into English transcripts." -- LDC online catalogue.
Başlığın Farklı Biçimleri Multilingual automatic document classification analysis and translation phase 3 training set
Diğer yazarlar / katkıda bulunanlar Lee, David. Linguistic Data Consortium.
ISBN 15856365179781585636518
View in source University of Chicago University of Chicago - Historical works, archives, and periodicals search engine
University of Chicago - Historical works, archives, and periodicals search engine University of Chicago

MADCAT phase 3 training set.

(/)
Publication Place - [Philadelphia, PA] : Linguistic Data Consortium, c2013.
Subject Arabic language -- Written Arabic -- Data processing. Arabic language -- Machine translating. Arabic language -- Translating into English. Machine translating. Machine translating. Arabic language -- Translating into English. Arabic language -- Machine translating.
Type Other
Language Arabic
Digital Yes
Manuscript No
Physical Dimensions |
Library University of Chicago
Record ID 9347803
Notes Title from disc label.Data type: Text.Data sources: Newsgroups, newswire, weblogs.Applications: Handwriting recognition, machine translation."LDC2013T16".Authors: David Lee, Safa Ismael, Dave Doermann, Stephanie Strassel, Zhiyi Song, Stephen Grimes.Arabic. 1 DVD ; 4 3/4 in.. "MADCAT (Multilingual Automatic Document Classification Analysis and Translation) Phase 3 Training Set contains all training data created by the Linguistic Data Consortium (LDC) to support Phase 3 of the DARPA MADCAT Program. The data in this release consists of handwritten Arabic documents, scanned at high resolution and annotated for the physical coordinates of each line and token. Digital transcripts and English translations of each document are also provided, with the various content and annotation layers integrated in a single MADCAT XML output. The goal of the MADCAT program is to automatically convert foreign text images into English transcripts." -- LDC online catalogue.
Başlığın Farklı Biçimleri Multilingual automatic document classification analysis and translation phase 3 training set
Diğer yazarlar / katkıda bulunanlar Lee, David. Linguistic Data Consortium.
ISBN 15856365179781585636518
University of Chicago - Historical works, archives, and periodicals search engine
University of Chicago You are being redirected...

Please wait