Difference between revisions of "Resources for Serbian"

From ACL Wiki
Jump to navigation Jump to search
 
 
Line 1: Line 1:
 
 
==Corpora==
 
==Corpora==
  
 
===Free===
 
===Free===
  
* [http://xixona.dlsi.ua.es/~fran/setimes/ Southeast European Times] (paragraph aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — 9,678 paragraphs, 92,450— 122,912 words per language)
+
* [http://www.statmt.org/setimes/ Southeast European Times] (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)
  
 
[[Category:Resources by language|Serbian]]
 
[[Category:Resources by language|Serbian]]

Latest revision as of 13:06, 25 March 2010

Corpora

Free

  • Southeast European Times (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)