25.77, Calls: Computational Linguistics/Iceland

linguist at linguistlist.org linguist at linguistlist.org
Wed Jan 8 20:26:35 UTC 2014


LINGUIST List: Vol-25-77. Wed Jan 08 2014. ISSN: 1069 - 4875.

Subject: 25.77, Calls: Computational Linguistics/Iceland

Moderator: Damir Cavar, Eastern Michigan U <damir at linguistlist.org>

Reviews: 
Monica Macaulay, U of Wisconsin Madison
Rajiv Rao, U of Wisconsin Madison
Joseph Salmons, U of Wisconsin Madison
Mateja Schuck, U of Wisconsin Madison
Anja Wanner, U of Wisconsin Madison
       <reviews at linguistlist.org>

Homepage: http://linguistlist.org

Do you want to donate to LINGUIST without spending an extra penny? Bookmark
the Amazon link for your country below; then use it whenever you buy from
Amazon!

USA: http://www.amazon.com/?_encoding=UTF8&tag=linguistlist-20
Britain: http://www.amazon.co.uk/?_encoding=UTF8&tag=linguistlist-21
Germany: http://www.amazon.de/?_encoding=UTF8&tag=linguistlistd-21
Japan: http://www.amazon.co.jp/?_encoding=UTF8&tag=linguistlist-22
Canada: http://www.amazon.ca/?_encoding=UTF8&tag=linguistlistc-20
France: http://www.amazon.fr/?_encoding=UTF8&tag=linguistlistf-21

For more information on the LINGUIST Amazon store please visit our
FAQ at http://linguistlist.org/amazon-faq.cfm.

Editor for this issue: Bryn Hauk <bryn at linguistlist.org>
================================================================  


Date: Wed, 08 Jan 2014 15:26:16
From: Reinhard Rapp [reinhardrapp at gmx.de]
Subject: 7th Workshop on Building and Using Comparable Corpora

E-mail this message to a friend:
http://linguistlist.org/issues/emailmessage/verification.cfm?iss=25-77.html&submissionid=25100269&topicid=3&msgnumber=1
 
Full Title: 7th Workshop on Building and Using Comparable Corpora 
Short Title: BUCC 2014 

Date: 27-May-2014 - 27-May-2014
Location: Reykjavik, Iceland 
Contact Person: Pierre Zweigenbaum
Meeting Email: pz at limsi.fr
Web Site: http://comparable.limsi.fr/bucc2014/ 

Linguistic Field(s): Computational Linguistics 

Call Deadline: 10-Feb-2014 

Meeting Description:

Motivation:

In the language engineering and the linguistics communities, research in comparable corpora has been motivated by two main reasons. In language engineering, on the one hand, it is chiefly motivated by the need to use comparable corpora as training data for statistical Natural Language Processing applications such as statistical machine translation or cross-lingual retrieval. In linguistics, on the other hand, comparable corpora are of interest in themselves by making possible inter-linguistic discoveries and comparisons. It is generally accepted in both communities that comparable corpora are documents in one or several languages that are comparable in content and form in various degrees and dimensions. We believe that the linguistic definitions and observations related to comparable corpora can improve methods to mine such corpora for applications of statistical NLP. As such, it is of great interest to bring together builders and users of such corpora.

Call for Papers:

7th Workshop on Building and Using Comparable Corpora
Building Resources for Machine Translation Research
http://comparable.limsi.fr/bucc2014/

May 27, 2014
Co-located with LREC 2014
Harpa Conference Centre, Reykjavik (Iceland) 

Deadline for Papers: February 10, 2014
https://www.softconf.com/lrec2014/BUCC2014/

Parallel corpora are a key resource as training data for statistical machine translation, and for building or extending bilingual lexicons and terminologies. However, beyond a few language pairs such as English- French or English-Chinese and a few contexts such as parliamentary debates or legal texts, they remain a scarce resource, despite the creation of automated methods to collect parallel corpora from the Web. To exemplify such issues in a practical setting, this year's special focus will be on building resources for machine translation research.

This special topic aims to address the need for:

(1) Machine Translation training and testing data such as spoken or written monolingual, comparable or parallel data collections
(2) Methods and tools used for collecting, annotating, and verifying MT data such as Web crawling, crowdsourcing, tools for language experts and for finding MT data in comparable corpora

Important Dates:
 
February 10, 2014: Deadline for submission of full papers
March 10, 2014: Notification of acceptance
March 27, 2014: Camera-ready papers due
May 27, 2014: Workshop date

Organisers:
 
Pierre Zweigenbaum, LIMSI, CNRS, Orsay (France)
Ahmet Aker, University of Sheffield (UK)
Serge Sharoff, University of Leeds (UK)
Stephan Vogel, QCRI (Qatar)
Reinhard Rapp, Universities of Mainz (Germany) and Aix-Marseille (France)







----------------------------------------------------------
LINGUIST List: Vol-25-77	
----------------------------------------------------------



More information about the LINGUIST mailing list