<html><head>

<meta http-equiv="content-type" content="text/html; charset=UTF-8"></head><body

 bgcolor="#FFFFFF" text="#000000">

SHARED TASK ON THE LEXICAL ACCESS PROBLEM (COMPUTING ASSOCIATIONS WHEN 

GIVEN MULTIPLE STIMULI)<br>

  <br>

In the framework of the 4th Workshop on Cognitive Aspects of the Lexicon

 (CogALex) to be held at COLING 2014, we invite participation in a 

shared task devoted to the problem of lexical access in language 

production, with the aim of providing a quantitative comparison between 

different systems.<br>

  <br>

 <br>

MOTIVATION<br>

  <br>

The quality of a dictionary depends not only on coverage, but also on 

the accessibility of the information. That is a crucial point is 

dictionary access. Access strategies vary with the task (text 

understanding vs. text production) and the knowledge available at the 

very moment of consultation (words, concepts, speech sounds). Unlike 

readers who look for meanings, writers start from them, searching for 

the corresponding words. While paper dictionaries are static, permitting

 only limited strategies for accessing information, their electronic 

counterparts promise dynamic, proactive search via multiple criteria 

(meaning, sound, related words) and via diverse access routes. 

Navigation takes place in a huge conceptual lexical space, and the 

results are displayable in a multitude of forms (e.g. as trees, as 

lists, as graphs, or sorted alphabetically, by topic, by frequency).<br>

  <br>

To bring some structure into this multitude of possibilities, the shared

 task will concentrate on a crucial subtask, namely multiword 

association.  What we mean by this in the context of this workshop is 

the following. Suppose, we were looking for a word expressing the 

following ideas: 'superior dark coffee made of beans from Arabia', but 

could not remember the intended word 'mocha' due to the 

tip-of-the-tongue problem. Since people always remember something 

concerning the elusive word, it would be nice to have a system accepting

 this kind of input, to propose then a number of candidates for the 

target word. Given the above example, we might enter 'dark', 'coffee', 

'beans', and 'Arabia', and the system would be supposed to come up with 

one or several associated words such as 'mocha', 'espresso', or 

'cappuccino'.<br>

  <br>

 <br>

TASK DEFINITION<br>

  <br>

The participants will receive lists of five given words (primes) such as

 'circus', 'funny', 'nose', 'fool', and 'fun' and are supposed to 

compute the word which is most closely associated to all of them. In 

this case, the word 'clown' would be the expected response. Here are 

some more examples:<br>

  <br>

   given words:  gin, drink, scotch, bottle, soda<br>

   target word:  whisky<br>

  <br>

   given words:  wheel, driver, bus, drive, lorry<br>

   target word:  car<br>

  <br>

   given words:  neck, animal, zoo, long, tall<br>

   target word:  giraffe<br>

  <br>

   given words:  holiday, work, sun, summer, abroad<br>

   target word:  vacation<br>

  <br>

   given words:  home, garden, door, boat, chimney<br>

   target word:  house<br>

  <br>

   given words:  blue, cloud, stars, night, high<br>

   target word:  sky<br>

  <br>

We will provide a training set of 2000 sets of five input words 

(multiword stimuli), together with the expected target words 

(associative responses). The participants will have about five weeks to 

train their systems on this data. After the training phase, we will 

release a test set containing another 2000 sets of five input words, but

 without providing the expected target words. <br>

  <br>

Participants will have five days to run their systems on the test data, 

thereby predicting the target words. For each system, we will compare 

the results to the expected target words and compute an accuracy. The 

participants will be invited to submit a paper describing their approach

 and their results.<br>

  <br>

For the participating systems, we will distinguish two categories: <br>

  <br>

(1) Unrestricted systems. They can use any kind of data to compute their

 results. <br>

(2) Restricted systems: These systems are only allowed to draw on the 

freely available ukWaC corpus in order to extract information on word 

associations. The ukWaC corpus comprises about 2 billion words and is 

can be downloaded from <a class="moz-txt-link-freetext" href="http://wacky.sslmit.unibo.it/doku.php?id=corpora">http://wacky.sslmit.unibo.it/doku.php?id=corpora</a>.

 <br>

  <br>

Participants are allowed to compete in either category or in both.<br>

  <br>

  <br>

VENUE<br>

  <br>

The shared task will take place as part of the CogALex workshop which is

 co-located with COLING 2014 (Dublin). The workshop date is August 23, 

2014. Shared task participants who wish to have a paper published in the

 workshop proceedings will be required to present their work at the 

workshop.<br>

  <br>

  <br>

SHARED TASK SCHEDULE<br>

  <br>

Training data release:  March 27, 2014<br>

Test data release:  May 5, 2014<br>

Final results due:  May 9, 2014<br>

Deadline for paper submission: May 25, 2014  <br>

Reviewers' feedback:  June, 15, 2014<br>

Camera-ready version:  July 7, 2014<br>

Workshop date:  August 23, 2014<br>

  <br>

  <br>

FURTHER INFORMATION<br>

  <br>

CogALex workshop website: 

<a class="moz-txt-link-freetext" href="http://pageperso.lif.univ-mrs.fr/~michael.zock/CogALex-IV/cogalex-webpage/index.html">http://pageperso.lif.univ-mrs.fr/~michael.zock/CogALex-IV/cogalex-webpage/index.html</a><br>

Data releases: To be found on the above workshop website from the dates 

given in the schedule.<br>

Registration for the shared task: Send e-mail to Michael Zock, with 

Reinhard Rapp in copy.<br>

  <br>

  <br>

WORKSHOP ORGANIZERS<br>

  <br>

Michael Zock (LIF-CNRS, Marseille, France), michael.zock AT 

lif.univ-mrs.fr<br>

Reinhard Rapp (University of Aix Marseille (France) and Mainz (Germany),

 reinhardrapp AT gmx.de<br>

Chu-Ren Huang (The Hong Kong Polytechnic University, Hong Kong), 

churen.huang AT inet.polyu.edu.hk<br>

  <br>

  <br>

  <br>

  <div class="moz-signature">-- <br>

------------------------------------------------------<br>

    <span>Michael ZOCK<br>

<span></span><span>

<link rel="File-List" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0/clip_filelist.xml">

<link rel="themeData" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0/clip_themedata.xml">

<p class="MsoNormal"><span style="mso-fareast-font-family:"Times 

New Roman";

mso-bidi-font-family:"Times New Roman"" lang="FR">Aix-Marseille

 Université, <br>

CNRS & LIF, UMR 7279, <br>

163 Avenue de Luminy<br>

F-13288 <u>Marseille</u> / France</span><span lang="FR"><o:p></o:p></span></p>

 </span><br>

<span style="color: rgb(51, 0, 153);">Mail</span>:    <span 

style="color: rgb(0, 0, 153);"><a class="moz-txt-link-abbreviated" href="mailto:michael.zock@lif.univ-mrs.fr">michael.zock@lif.univ-mrs.fr</a></span><br>

Tel.:    +33 (0) 4 91 82 94 88<br>

<span>

<link rel="File-List" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0/clip_filelist.xml">

<link rel="themeData" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0/clip_themedata.xml">

<p class="MsoNormal"><span style="mso-fareast-font-family:"Times 

New Roman";

mso-bidi-font-family:"Times New Roman"" lang="FR">Secr.:       

+33

 (0) 4 91 82 90 70<br>

Fax:          +33 (0) 4 91 82 92 75</span><span lang="FR"><o:p></o:p></span></p>

 </span><br>

  <span style="color: rgb(51, 0, 153);">Web</span>:   <span 

style="color: rgb(0, 0, 153);"><a class="moz-txt-link-freetext" href="http://pageperso.lif.univ-mrs.fr/~michael.zock/">http://pageperso.lif.univ-mrs.fr/~michael.zock/</a></span><span>

<link rel="File-List" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0clip_filelist.xml">

<link rel="themeData" 

href="file://localhost/Users/MIKA/Library/Caches/TemporaryItems/msoclip/0clip_themedata.xml">

<span style="font-size:12.0pt;font-family:"Times New Roman";

mso-fareast-font-family:"Times New 

Roman";mso-ansi-language:FR;mso-fareast-language:

FR;mso-bidi-language:AR-SA" lang="FR"></span></span> </span><br>

------------------------------------------------------

  </div>

</body>

</html>