1. What contributions have the authors mentioned in the paper "Exploring methods and resources for discriminating similar languages" ?
In this paper, the authors describe the submissions made by team UniMelb-NLP, which took part in both the closed and open categories.. The authors present the text representations and modeling techniques used, including cross-lingual POS tagging as well as fine-grained tags extracted from a deep grammar of English, and discuss additional data they collected for the open submissions, utilizing custombuilt web corpora based on top-level domains as well as existing corpora.
read more
2. What future works have the authors mentioned in the paper "Exploring methods and resources for discriminating similar languages" ?
Overall, these results highlight the need for further research into discriminating between varieties of English.
read more
3. What is the main assumption for the.pt TLD?
For instance, the authors assume that Portuguese text found when crawling the .pt TLD will primarily be European Portuguese, while the Portuguese found in .br will be primarily Brazilian Portuguese.
read more
4. What was the first use of de-lexicalized text representations?
De-lexicalized text representations through POS tagging were first considered for native language identification (NLI), where they were used as a proxy for syntax in order to capture certain types of grammatical errors (Wong and Dras, 2009).
read more




