Resources for Romanian: Difference between revisions

From ACL Wiki
Jump to navigation Jump to search
No edit summary
 
Zeman (talk | contribs)
HamleDT
 
(8 intermediate revisions by 5 users not shown)
Line 5: Line 5:
===Proprietary===
===Proprietary===


*


==Lexical resources==
==Lexical resources==
Line 12: Line 11:


==Corpora==
==Corpora==
===Free===
* [http://www.statmt.org/europarl Europarl corpus], sentence aligned with English
* [http://ufal.mff.cuni.cz/hamledt HamleDT], harmonized dependency treebanks of many languages, common annotation style.
* [http://www.cs.unt.edu/~rada/downloads.html Romanian NLP]
* [http://www.statmt.org/setimes/ Southeast European Times] (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)
===Proprietary===


* [http://consilr.info.uaic.ro/en/index.php?showpage=060103 Corpora] (Monolingual, POS tagged and bilingual English/French<->Romanian).
* [http://consilr.info.uaic.ro/en/index.php?showpage=060103 Corpora] (Monolingual, POS tagged and bilingual English/French<->Romanian).
Line 17: Line 24:
==Bibliography==
==Bibliography==


*


==External links==
==External links==

Latest revision as of 15:50, 26 May 2014

Machine translation systems

Free software

Proprietary

Lexical resources

Corpora

Free

  • Europarl corpus, sentence aligned with English
  • HamleDT, harmonized dependency treebanks of many languages, common annotation style.
  • Romanian NLP
  • Southeast European Times (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)

Proprietary

  • Corpora (Monolingual, POS tagged and bilingual English/French<->Romanian).

Bibliography

External links