ISLRN

Romanian - English literature corpus (Processed)

Full Official Name: Romanian - English literature corpus (Processed)

Submission date: March 9, 2020, 12:27 p.m.

This dataset has been created within the framework of the European Language Resource Coordination (ELRC) Connecting Europe Facility - Automated Translation (CEF.AT) action. For further information on the project: http://lr-coordination.eu. Bilingual Romanian – English literature corpus built from a small set of freely available literature books (drama, sci-fi, etc.). The texts are positionally aligned, i.e. the sentence on line i in the English text is aligned with the sentence on line i in the Romanian text. Alignment was manually validated.

Creator(s)

Distributor(s)