Resources for Bulgarian: Difference between revisions

From ACL Wiki
Jump to navigation Jump to search
Free: +Europarl corpus
Zeman (talk | contribs)
HamleDT
 
Line 31: Line 31:
* [http://www.statmt.org/setimes/ Southeast European Times], sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language
* [http://www.statmt.org/setimes/ Southeast European Times], sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language
* [http://www.statmt.org/europarl Europarl corpus], sentence aligned with English
* [http://www.statmt.org/europarl Europarl corpus], sentence aligned with English
* [http://ufal.mff.cuni.cz/hamledt HamleDT], harmonized dependency treebanks of many languages, common annotation style.


===Proprietary===
===Proprietary===

Latest revision as of 15:36, 26 May 2014

Machine translation systems

Free software

Proprietary

Lexical resources

Morphological analysis

Free software

Proprietary

Grammars

Proprietary

Corpora

Free

  • Southeast European Times, sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language
  • Europarl corpus, sentence aligned with English
  • HamleDT, harmonized dependency treebanks of many languages, common annotation style.

Proprietary

Bibliography

External links