<?xml version="1.0"?><!DOCTYPE article SYSTEM "/project/take/software/searchbench_offline_processing/paperxml_generator/aclextractor/src/python/../resource/dtd/paperxml.dtd"><article><header><firstpageheader><page local="1" global="75"/><title>SemEval-2010 Task 17: All-Words Word Sense Disambiguation on a Specific Domain</title><pubinfo>Proceedings of the 5th International Workshop on Semantic Evaluation, ACL 2010,pages 75-80, Uppsala, Sweden, 15-16 July 2010. ©2010 Association for Computational Linguistics</pubinfo><author surname="Agirre" givenname="Eneko"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author><author surname="López de Lacalle" givenname="Oier"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author><author surname="Hsieh" givenname="Shu-Kai"><org  name="Hsieh" city="Shu-Kai"/></author><author surname="Monachini" givenname="Monica"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author><author surname="Fellbaum" givenname="Christiane"><org  name="Princeton University" country="USA" city="Princeton"/></author><author surname="Tesconi" givenname="Maurizio"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author><author surname="Vossen" givenname="Piek"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author><author surname="Segers" givenname="Roxanne"><org  name="Vrije University" country="The Netherlands" city="Amsterdam"/></author></firstpageheader><frontmatter><p><b>SemEval-2010 Task 17: All-words Word Sense Disambiguation</b></p><p><b>on a Specific Domain</b></p><p><b>Eneko Agirre, Oier Lopez de Lacalle</b></p><p>IXA NLP group UBC</p><p>Donostia, Basque Country</p><p>{e.agirre,oier.lopezdelacalle}@ehu.es</p><p><b>Shu-Kai Hsieh</b></p><p>Department of English National Taiwan Normal University Taipei, Taiwan</p><p>shukai@ntnu.edu.tw</p><p><b>Monica Monachini</b></p><p>ILC CNR Pisa, Italy</p><p>monica.monachini@ilc.cnr.it p.^</p><p><b>Christiane Fellbaum</b></p><p>Department of Computer Science Princeton University Princeton, USA</p><p>fellbaum@princeton.edu</p><p><b>Maurizio Tesconi</b></p><p>IIT CNR Pisa, Italy</p><p>maurizio.tesconi@iit.cnr.it</p><p><b>Piek Vossen, Roxanne Segers</b></p><p>Faculteit der Letteren Vrije Universiteit Amsterdam Amsterdam, Netherlands</p><p>ssen@let.vu.ni,roxane.segers@gmail.com</p></frontmatter><abstract>Domain portability and adaptation of NLP components and Word Sense Disambigua­tion systems present new challenges. The difficulties found by supervised systems to adapt might change the way we assess the strengths and weaknesses of supervised and knowledge-based WSD systems. Un­fortunately, all existing evaluation datasets for specific domains are lexical-sample corpora. This task presented all-words datasets on the environment domain for WSD in four languages (Chinese, Dutch, English, Italian). 11 teams participated, with supervised and knowledge-based sys­tems, mainly in the English dataset. The results show that in all languages the par­ticipants where able to beat the most fre­quent sense heuristic as estimated from general corpora. The most successful ap­proaches used some sort of supervision in the form of hand-tagged examples from the domain. </abstract></header><body><section number="1" title="Introduction"><p>Word Sense Disambiguation (WSD) competitions have focused on general domain texts, as attested in previous Senseval and SemEval competitions (Kilgarriff, 2001; Mihalcea et al., 2004; Snyder and Palmer, 2004; Pradhan et al., 2007). Specific domains pose fresh challenges to WSD sys­tems: the context in which the senses occur might change, different domains involve different sense distributions and predominant senses, some words tend to occur in fewer senses in specific domains, the context of the senses might change, and new senses and terms might be involved. Both super­vised and knowledge-based systems are affected by these issues: while the first suffer from differ­ent context and sense priors, the later suffer from lack of coverage of domain-related words and in­formation.</p><p>The main goal of this task is to provide a mul­tilingual testbed to evaluate WSD systems when faced with full-texts from a specific domain. All datasets and related information are publicly avail­able from the task websites<footnote anchor="1"/>.</p><p>This task was designed in the context of Ky­oto (Vossen et al., 2008)<footnote anchor="2"/>, an Asian-European project that develops a community platform for modeling knowledge and finding facts across lan­guages and cultures. The platform operates as a Wiki system with an ontological support that so­cial communities can use to agree on the mean­ing of terms in specific domains of their interest. Kyoto focuses on the environmental domain be­cause it poses interesting challenges for informa­tion sharing, but the techniques and platforms are</p><p>!http://xmlgroup.iit.cnr.it/SemEval2 010/ and http://semeval2.fbk.eu/<page local="2" global="76"/></p><footnote label="2">http://www.kyoto-project.eu/</footnote><p>independent of the application domain.</p><p>The paper is structured as follows. We first present the preparation of the data. Section 3 re­views participant systems and Section 4 the re­sults. Finally, Section 5 presents the conclusions.</p></section><section number="2" title="Data preparation"><p>The data made available to the participants in­cluded the test set proper, and background texts. Participants had one week to work on the test set, but the background texts where provided months earlier.</p><subsection number="2.1" title="Test datasets"><p>The WSD-domain comprises comparable all-words test corpora on the environment domain. Three texts were compiled for each language by the European Center for Nature Conservation<footnote anchor="3"/> and Worldwide Wildlife Forum<footnote anchor="4"/>. They are documents written for a general but interested public and in­volve specific terms from the domain. The docu­ment content is comparable across languages. Ta­ble 1 shows the numbers for the datasets.</p><p>Although the original plan was to annotate mul­tiword terms, and domain terminology, due to time constraints we focused on single-word nouns and verbs. The test set clearly marked which were the words to be annotated. In the case of Dutch, we also marked components of single-word com­pounds. The format of the test set followed that of previous all-word exercises, which we extended to accommodate Dutch compounds. For further de­tails check the datasets in the task website.</p><p>The sense inventory was based on publicly available wordnets of the respective languages (see task website for details). The annotation pro­cedure involved double-blind annotation by ex­perts plus adjudication, which allowed us to also provide Inter Annotator Agreement (IAA) figures for the dataset. The procedure was carried out us­ing KAFnotator tool (Tesconi et al., 2010). Due to limitations in resources and time, the English dataset was annotated by a single expert annota­tor. For the rest of languages, the agreement was very good, as reported in Table 1.</p><p>Table 1 includes the results of the random base­line, as an indication of the polysemy in each dataset. Average polysemy is highest for English, and lowest for Dutch.</p><footnote label="3">http://www.ecnc.org http :// www.wwf.org</footnote><p>Table 1: Dataset numbers, including number of tokens, nouns and verbs to be tagged, Inter-Annotator Agreement (IAA) and precision of ran­dom baseline.</p><p>In addition to the test datasets proper, we also pro­vided additional documents on related subjects, kindly provided by ECNC and WWF. Table 2 shows the number of documents and words made available for each language. The full list with the urls of the documents are available from the task website, together with the background documents.</p></subsection></section><section number="3" title="Participants"><p>Eleven participants submitted more than thirty runs (cf. Table 3). The authors classified their runs into supervised (S in the tables, three runs), weakly supervised (WS, four runs), unsupervised (no runs) and knowledge-based (KB, the rest of runs)<footnote anchor="5"/>. Only one group used hand-tagged data from the domain, which they produced on their own. We will briefly review each of the participant groups, ordered fol­lowing the rank obtained for English. They all par­ticipated on the English task, with one exception as noted below, so we report their rank in the En­glish task. Please refer to their respective paper in these proceedings for more details.</p><p><b>CFILT: </b>They participated with a domain-specific knowledge-based method based on Hop-field networks (Khapra et al., 2010). They first identify domain-dependant words using the back­ground texts, use a graph based on hyponyms in WordNet, and a breadth-first search to select the most representative synsets within domain. In ad­dition they added manually disambiguated around one hundred examples from the domain as seeds.</p><footnote label="5">Note that boundaries are slippery. We show the classifi­cations as reported by the authors.</footnote><table class="main" frame="box" rules="all" border="1" regular="False"><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Total</p></td><td class="cell"><p>Noun</p></td><td class="cell"><p>Verb</p></td><td class="cell"><p>IAA</p></td><td class="cell"><p>Random</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Chinese</p></td><td class="cell"><p>3989</p></td><td class="cell"><p>754</p></td><td class="cell"><p>450</p></td><td class="cell"><p>0.96</p></td><td class="cell"><p>0.321</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Dutch</p></td><td class="cell"><p>8157</p></td><td class="cell"><p>997</p></td><td class="cell"><p>635</p></td><td class="cell"><p>0.90</p></td><td class="cell"><p>0.328</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>English</p></td><td class="cell"><p>5342</p></td><td class="cell"><p>1032</p></td><td class="cell"><p>366</p></td><td class="cell"><p>n/a</p></td><td class="cell"><p>0.232</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Italian</p></td><td class="cell"><p>8560</p></td><td class="cell"><p>1340</p></td><td class="cell"><p>513</p></td><td class="cell"><p>0.72</p></td><td class="cell"><p>0.294</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr></table><table caption="Table 2: Size of the background data.2.2   Background data" class="main" frame="box" rules="all" border="1" regular="False"><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p></p></td><td class="cell"><p>Documents</p></td><td class="cell"><p>Words</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Chinese</p></td><td class="cell"><p>58</p></td><td class="cell"><p>455359</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Dutch</p></td><td class="cell"><p>98</p></td><td class="cell"><p>21089</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>English</p></td><td class="cell"><p>113</p></td><td class="cell"><p>2737202</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Italian</p></td><td class="cell"><p>27</p></td><td class="cell"><p>240158</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr></table><page local="3" global="77"/><doubt alpha="100.0" length="7" tooSmall="False" monospace="0.0">Chinese</doubt><doubt alpha="100.0" length="5" tooSmall="False" monospace="0.0">Dutch</doubt><doubt alpha="100.0" length="7" tooSmall="False" monospace="0.0">English</doubt><p>This is the only group using hand-tagged data from the target domain. Their best run ranked 1st.</p><p><b>IIITTH: </b>They presented a personalized PageR-ank algorithm over a graph constructed from WordNet similar to   (Agirre and Soroa, 2009),</p><p>with two variants. In the first (IIITH1), the vertices of the graph are initialized following the rank­ing scores obtained from predominant senses as in (McCarthy et al., 2007). In the second (IIITH2), the graph is initialized with keyness values as in<page local="4" global="78"/></p><table class="main" frame="box" rules="all" border="1" regular="False"><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Rank</p></td><td class="cell"><p>Participant</p></td><td class="cell"><p>System ID</p></td><td class="cell"><p>Type</p></td><td class="cell"><p>P</p></td><td class="cell"><p>R</p></td><td class="cell"><p>R nouns</p></td><td class="cell"><p>R verbs</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>1</p></td><td class="cell"><p>Anup Kulkarni</p></td><td class="cell"><p>CFILT-2</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.570</p></td><td class="cell"><p>0.555</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>024</p></td><td class="cell"><p>0.594 ±0.028</p></td><td class="cell"><p>0.445 ±0.047</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>2</p></td><td class="cell"><p>Anup Kulkarni</p></td><td class="cell"><p>CFILT-1</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.554</p></td><td class="cell"><p>0.540</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>021</p></td><td class="cell"><p>0.580 ±0.025</p></td><td class="cell"><p>0.426 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>3</p></td><td class="cell"><p>Siva Reddy</p></td><td class="cell"><p>IHTHl-d.l.ppr.05</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.534</p></td><td class="cell"><p>0.528</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>027</p></td><td class="cell"><p>0.553 ±0.023</p></td><td class="cell"><p>0.456 ±0.041</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>4</p></td><td class="cell"><p>Abhilash Inumella</p></td><td class="cell"><p>IIITH2-d.r.l.ppr.05</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.522</p></td><td class="cell"><p>0.516</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>023</p></td><td class="cell"><p>0.529 ±0.027</p></td><td class="cell"><p>0.478 ±0.041</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>5</p></td><td class="cell"><p>Ruben Izquierdo</p></td><td class="cell"><p>B LC 20 S emcorB ackground</p></td><td class="cell"><p>s</p></td><td class="cell"><p>0.513</p></td><td class="cell"><p>0.513</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.534 ±0.026</p></td><td class="cell"><p>0.454 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Most Frequent Sense</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.505</p></td><td class="cell"><p>0.505</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>023</p></td><td class="cell"><p>0.519 ±0.026</p></td><td class="cell"><p>0.464 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>6</p></td><td class="cell"><p>Ruben Izquierdo</p></td><td class="cell"><p>BLC20Semcor</p></td><td class="cell"><p>s</p></td><td class="cell"><p>0.505</p></td><td class="cell"><p>0.505</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>025</p></td><td class="cell"><p>0.527 ±0.031</p></td><td class="cell"><p>0.443 ±0.045</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>7</p></td><td class="cell"><p>Anup Kulkarni</p></td><td class="cell"><p>CFILT-3</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.512</p></td><td class="cell"><p>0.495</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>023</p></td><td class="cell"><p>0.516 ±0.027</p></td><td class="cell"><p>0.434 ±0.048</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>8</p></td><td class="cell"><p>Andrew Tran</p></td><td class="cell"><p>Treematch</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.506</p></td><td class="cell"><p>0.493</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>021</p></td><td class="cell"><p>0.516 ±0.028</p></td><td class="cell"><p>0.426 ±0.046</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>9</p></td><td class="cell"><p>Andrew Tran</p></td><td class="cell"><p>Treematch-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.504</p></td><td class="cell"><p>0.491</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>021</p></td><td class="cell"><p>0.515 ±0.030</p></td><td class="cell"><p>0.425 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>10</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.481</p></td><td class="cell"><p>0.481</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.487 ±0.025</p></td><td class="cell"><p>0.462 ±0.039</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>11</p></td><td class="cell"><p>Andrew Tran</p></td><td class="cell"><p>Treematch-3</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.492</p></td><td class="cell"><p>0.479</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.494 ±0.028</p></td><td class="cell"><p>0.434 ±0.039</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>12</p></td><td class="cell"><p>Radu Ion</p></td><td class="cell"><p>RACAI-MFS</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.461</p></td><td class="cell"><p>0.460</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.458 ±0.025</p></td><td class="cell"><p>0.464 ±0.046</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>13</p></td><td class="cell"><p>Hansen A. Schwartz</p></td><td class="cell"><p>UCF-WS</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.447</p></td><td class="cell"><p>0.441</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.440 ±0.025</p></td><td class="cell"><p>0.445 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>14</p></td><td class="cell"><p>Yuhang Guo</p></td><td class="cell"><p>HIT-CIR-DMFS-l.ans</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.436</p></td><td class="cell"><p>0.435</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>023</p></td><td class="cell"><p>0.428 ±0.027</p></td><td class="cell"><p>0.454 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>15</p></td><td class="cell"><p>Hansen A. Schwartz</p></td><td class="cell"><p>UCF-WS-domain</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.440</p></td><td class="cell"><p>0.434</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>024</p></td><td class="cell"><p>0.434 ±0.029</p></td><td class="cell"><p>0.434 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>16</p></td><td class="cell"><p>Abhilash Inumella</p></td><td class="cell"><p>IIITH2-d.r.l.baseline.05</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.496</p></td><td class="cell"><p>0.433</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>024</p></td><td class="cell"><p>0.452 ±0.023</p></td><td class="cell"><p>0.390 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>17</p></td><td class="cell"><p>Siva Reddy</p></td><td class="cell"><p>IHTHl-d.l.baseline.05</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.498</p></td><td class="cell"><p>0.432</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>021</p></td><td class="cell"><p>0.463 ±0.026</p></td><td class="cell"><p>0.344 ±0.038</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>18</p></td><td class="cell"><p>Radu Ion</p></td><td class="cell"><p>RACAI-2MFS</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.433</p></td><td class="cell"><p>0.431</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.434 ±0.027</p></td><td class="cell"><p>0.399 ±0.049</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>19</p></td><td class="cell"><p>Siva Reddy</p></td><td class="cell"><p>IHTHl-d.l.ppv.05</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.426</p></td><td class="cell"><p>0.425</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>026</p></td><td class="cell"><p>0.434 ±0.028</p></td><td class="cell"><p>0.399 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>20</p></td><td class="cell"><p>Abhilash Inumella</p></td><td class="cell"><p>IIITH2-d.r.l.ppv.05</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.424</p></td><td class="cell"><p>0.422</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>023</p></td><td class="cell"><p>0.456 ±0.025</p></td><td class="cell"><p>0.325 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>21</p></td><td class="cell"><p>Hansen A. Schwartz</p></td><td class="cell"><p>UCF-WS-domain.noPropers</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.437</p></td><td class="cell"><p>0.392</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>025</p></td><td class="cell"><p>0.377 ±0.025</p></td><td class="cell"><p>0.434 ±0.043</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>22</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-1</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.384</p></td><td class="cell"><p>0.384</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.382 ±0.024</p></td><td class="cell"><p>0.391 ±0.047</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>23</p></td><td class="cell"><p>Ruben Izquierdo</p></td><td class="cell"><p>BLC20Background</p></td><td class="cell"><p>S</p></td><td class="cell"><p>0.380</p></td><td class="cell"><p>0.380</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.385 ±0.026</p></td><td class="cell"><p>0.366 ±0.037</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>24</p></td><td class="cell"><p>Davide Buscaldi</p></td><td class="cell"><p>NLEL-WSD-PDB</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.381</p></td><td class="cell"><p>0.356</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.357 ±0.027</p></td><td class="cell"><p>0.352 ±0.049</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>25</p></td><td class="cell"><p>Radu Ion</p></td><td class="cell"><p>RACAI-Lexical-Chains</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.351</p></td><td class="cell"><p>0.350</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>015</p></td><td class="cell"><p>0.344 ±0.017</p></td><td class="cell"><p>0.368 ±0.030</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>26</p></td><td class="cell"><p>Davide Buscaldi</p></td><td class="cell"><p>NLEL-WSD</p></td><td class="cell"><p>WS</p></td><td class="cell"><p>0.370</p></td><td class="cell"><p>0.345</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.352 ±0.027</p></td><td class="cell"><p>0.328 ±0.037</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>27</p></td><td class="cell"><p>Yoan Gutierrez</p></td><td class="cell"><p>Relevant Semantic Trees</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.328</p></td><td class="cell"><p>0.322</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.335 ±0.026</p></td><td class="cell"><p>0.284 ±0.044</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>28</p></td><td class="cell"><p>Yoan Gutierrez</p></td><td class="cell"><p>Relevant Semantic Trees-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.321</p></td><td class="cell"><p>0.315</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>022</p></td><td class="cell"><p>0.327 ±0.024</p></td><td class="cell"><p>0.281 ±0.040</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>29</p></td><td class="cell"><p>Yoan Gutierrez</p></td><td class="cell"><p>Relevant Cliques</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.312</p></td><td class="cell"><p>0.303</p></td><td class="cell"><p>±0</p></td><td class="cell"><p>021</p></td><td class="cell"><p>0.304 ±0.024</p></td><td class="cell"><p>0.301 ±0.041</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Random baseline</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.232</p></td><td class="cell"><p>0.232</p></td><td class="cell"><p></p></td><td class="cell"><p></p></td><td class="cell"><p>0.253</p></td><td class="cell"><p>0.172</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr></table><table class="main" frame="box" rules="all" border="1" regular="False"><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Rank</p></td><td class="cell"><p>Participant</p></td><td class="cell"><p>System ID</p></td><td class="cell"><p>Type</p></td><td class="cell"><p>P</p></td><td class="cell"><p>R</p></td><td class="cell"><p>R nouns</p></td><td class="cell"><p>R verbs</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Most Frequent Sense</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.562</p></td><td class="cell"><p>0.562 ±0.026</p></td><td class="cell"><p>0.589 ±0.027</p></td><td class="cell"><p>0.518 ±0.039</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>1</p></td><td class="cell"><p>Meng-Hsien Shih</p></td><td class="cell"><p>HR</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.559</p></td><td class="cell"><p>0.559 ±0.024</p></td><td class="cell"><p>0.615 ±0.026</p></td><td class="cell"><p>0.464 ±0.039</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>2</p></td><td class="cell"><p>Meng-Hsien Shih</p></td><td class="cell"><p>GHR</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.517</p></td><td class="cell"><p>0.517 ±0.024</p></td><td class="cell"><p>0.533 ±0.035</p></td><td class="cell"><p>0.491 ±0.038</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Random baseline</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.321</p></td><td class="cell"><p>0.321</p></td><td class="cell"><p>0.326</p></td><td class="cell"><p>0.312</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>4</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-3</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.322</p></td><td class="cell"><p>0.296 ±0.022</p></td><td class="cell"><p>0.257 ±0.027</p></td><td class="cell"><p>0.360 ±0.038</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>3</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.342</p></td><td class="cell"><p>0.285 ±0.021</p></td><td class="cell"><p>0.251 ±0.026</p></td><td class="cell"><p>0.342 ±0.040</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>5</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-1</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.310</p></td><td class="cell"><p>0.258 ±0.023</p></td><td class="cell"><p>0.256 ±0.029</p></td><td class="cell"><p>0.261 ±0.031</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr></table><table caption="Table 3: Overall results for the domain WSD datasets, ordered by recall." class="main" frame="box" rules="all" border="1" regular="False"><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Rank</p></td><td class="cell"><p>Participant</p></td><td class="cell"><p>System ID</p></td><td class="cell"><p>Type</p></td><td class="cell"><p>P</p></td><td class="cell"><p>R</p></td><td class="cell"><p>R nouns</p></td><td class="cell"><p>R verbs</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>1</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-3</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.526</p></td><td class="cell"><p>0.526 ±0.022</p></td><td class="cell"><p>0.575 ±0.029</p></td><td class="cell"><p>0.450 ±0.034</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>2</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.519</p></td><td class="cell"><p>0.519 ±0.022</p></td><td class="cell"><p>0.561 ±0.027</p></td><td class="cell"><p>0.454 ±0.034</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Most Frequent Sense</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.480</p></td><td class="cell"><p>0.480 ±0.022</p></td><td class="cell"><p>0.600 ±0.027</p></td><td class="cell"><p>0.291 ±0.025</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>3</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-1</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.465</p></td><td class="cell"><p>0.465 ±0.021</p></td><td class="cell"><p>0.505 ±0.026</p></td><td class="cell"><p>0.403 ±0.033</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Random baseline</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.328</p></td><td class="cell"><p>0.328</p></td><td class="cell"><p>0.350</p></td><td class="cell"><p>0.293</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p></p></td><td class="cell"><p></p></td><td class="cell"><p></p></td><td class="cell"><p>Italian</p></td><td class="cell"><p></p></td><td class="cell"><p></p></td><td class="cell"><p></p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>Rank</p></td><td class="cell"><p>Participant</p></td><td class="cell"><p>System ID</p></td><td class="cell"><p>Type</p></td><td class="cell"><p>P</p></td><td class="cell"><p>R</p></td><td class="cell"><p>R nouns</p></td><td class="cell"><p>R verbs</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>1</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-3</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.529</p></td><td class="cell"><p>0.529 ±0.021</p></td><td class="cell"><p>0.530 ±0.024</p></td><td class="cell"><p>0.528 ±0.038</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>2</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-2</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.521</p></td><td class="cell"><p>0.521 ±0.018</p></td><td class="cell"><p>0.522 ±0.023</p></td><td class="cell"><p>0.519 ±0.035</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>3</p></td><td class="cell"><p>Aitor Soroa</p></td><td class="cell"><p>kyoto-1</p></td><td class="cell"><p>KB</p></td><td class="cell"><p>0.496</p></td><td class="cell"><p>0.496 ±0.019</p></td><td class="cell"><p>0.507 ±0.020</p></td><td class="cell"><p>0.468 ±0.037</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Most Frequent Sense</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.462</p></td><td class="cell"><p>0.462 ±0.020</p></td><td class="cell"><p>0.472 ±0.024</p></td><td class="cell"><p>0.437 ±0.035</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"><p>-</p></td><td class="cell"><p>-</p></td><td class="cell"><p><i>Random baseline</i></p></td><td class="cell"><p>-</p></td><td class="cell"><p>0.294</p></td><td class="cell"><p>0.294</p></td><td class="cell"><p>0.308</p></td><td class="cell"><p>0.257</p></td><td class="cell"></td></tr><tr class="row"><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td><td class="cell"></td></tr></table><doubt alpha="68.2" length="400" tooSmall="True" monospace="0.0">CFILT-2 CFILT-1 IIITH1 -d.l.ppr.05 IIITH2-d.l.ppr.05 BLC20SCBG BLC20SC CFILT-3 Treematch Treematch-2 Kyoto-2 Treematch-3 RACAI-MFS UCF-WS HIT-CIR-DMFS UCF-WS-domain IIITH2-d.r.l.baseline.05 IIITH1 -d.i. baseline. 05 RACAI-2MFS-BOW IIITH1 -d.l.ppv.05 IIITH2-d.r.l.ppv.05 UCF-WS-domain.no Propers Kyoto-1 BLC20BG NLEL-WSD-PDB RACAI-Lexical-Chains NLEL-WSD Rel. Sem. Trees Rel. Sem. Trees-2 Rel. Cliques</doubt><doubt alpha="100.0" length="3" tooSmall="False" monospace="0.0">MFS</doubt><doubt alpha="0.0" length="3" tooSmall="False" monospace="0.0">0.3</doubt><doubt alpha="0.0" length="4" tooSmall="False" monospace="0.0">0.35</doubt><doubt alpha="0.0" length="3" tooSmall="False" monospace="0.0">0.4</doubt><doubt alpha="0.0" length="4" tooSmall="False" monospace="0.0">0.45</doubt><doubt alpha="0.0" length="3" tooSmall="False" monospace="0.0">0.5</doubt><doubt alpha="0.0" length="4" tooSmall="False" monospace="0.0">0.55</doubt><p>Figure 1 : Plot for all the systems which participated in English domain WSD. Each point correspond to one system (denoted in axis <i>y) </i>according each recall and confidence interval (axis <i>X). </i>Systems are ordered depending on their rank.</p><p>(Rayson and Garside, 2000). Some of the runs use sense statistics from SemCor, and have been classified as weakly supervised. They submitted a total of six runs, with the best run ranking 3rd.</p><p><b>BLC20(SC/BG/SCBG): </b>This system is super­vised. A Support Vector Machine was trained us­ing the usual set of features extracted from con­text and the most frequent class of the target word. Semantic class-based classifiers were built from SemCor (Izquierdo et al., 2009), where the classes were automatically obtained exploiting the struc­tural properties of WordNet. Their best run ranked 5th.</p><p><b>Treematch: </b>This system uses a knowledge-based disambiguation method that requires a dic­tionary and untagged text as input. A previously developed system (Chen et al., 2009) was adapted to handle domain specific WSD. They built a domain-specific corpus using words mined from relevant web sites (e.g. WWF and ECNC) as seeds. Once parsed the corpus, the used the de­pendency knowledge to build a nodeset that was used for WSD. The background documents pro­vided by the organizers were only used to test how exhaustive the initial seeds were. Their best run ranked 8th.</p><p><b>Kyoto: </b>This system participated in all four languages, with a free reimplementation of the domain-specific knowledge-based method for WSD presented in (Agirre et al., 2009). It uses a module to construct a distributional the­saurus, which was run on the background text, and a disambiguation module based on Personalized PageRank over wordnet graphs. Different Word-Net were used as the LKB depending on the lan­guage. Their best run ranked 10th. Note that this team includes some of the organizers of the task. A strict separation was kept, in order to keep the test dataset hidden from the actual developers of the system.</p><p><b>RACAI: </b>This participant submitted three differ­ent knowledge-based systems. In the first, they use the mapping to domains of WordNet (version 2.0) in order to constraint the domains of the content words of the test text. In the second, they choose among senses using lexical chains (Ion and Ste-fanescu, 2009). The third system combines the previous two. Their best system ranked 12th.</p><p><b>HIT-CIR: </b>They presented a knowledge-based system which estimates predominant sense from raw test. The predominant senses were calculated with the frequency information in the provided background text, and automatically constructed<page local="5" global="79"/></p><p>thesauri from bilingual parallel corpora. The sys­tem ranked 14.</p><p><b>UCFWS: </b>This knowledge-based WSD system was based on an algorithm originally described in (Schwartz and Gomez, 2008), in which selectors are acquired from the Web via searching with lo­cal context of a given word. The sense is cho­sen based on the similarity or relatedness between the senses of the target word and various types of selectors. In some runs they include predom­inant senses(McCarthy et al., 2007). The best run ranked 13th.</p><p><b>NLEL-WSD(-PDB) </b>: The system used for the participation is based on an ensemble of different methods using fuzzy-Borda voting. A similar sys­tem was proposed in SemEval-2007 task-7 (Bus­caldi and Rosso, 2007). In this case, the com­ponent method used where the following ones: 1) Most Frequent Sense from SemCor; 2) Con­ceptual Density ; 3) Supervised Domain Relative Entropy classifier based on WordNet Domains; 4) Supervised Bayesian classifier based on Word-Net Domains probabilities; and 5) Unsupervised Knownet-20 classifiers. The best run ranked 24th.</p><p><b>UMCC-DLSI (Relevant): </b>The team submitted three different runs using a knowledge-based sys­tem. The first two runs use domain vectors and the third is based on cliques, which measure how much a concept is correlated to the sentence by obtaining Relevant Semantic Trees. Their best run ranked 27th.</p><p><b>(G)HR: </b>They presented a Knowledge-based WSD system, which make use of two heuristic rules (Li et al., 1995). The system enriched the Chinese WordNet by adding semantic relations for English domain specific words (e.g. ecology, en­vironment). When in-domain senses are not avail­able, the system relies on the first sense in the Chi­nese WordNet. In addition, they also use sense definitions. They only participated in the Chinese task, with their best system ranking 1st.</p></section><section number="4" title="Results"><p>The evaluation has been carried out using the stan­dard Senseval/SemEval scorer s cor er 2 as in­cluded in the trial dataset, which computes preci­sion and recall. Table 3 shows the results in each dataset. Note that the main evaluation measure is recall (R). In addition we also report precision (P) and the recall for nouns and verbs. Recall mea­sures are accompanied by a 95% confidence interval calculated using bootstrap resampling pro­cedure (Noreen, 1989). The difference between two systems is deemed to be statistically signifi­cant if there is no overlap between the confidence intervals. We show graphically the results in Fig­ure 1. For instance, the differences between the highest scoring system and the following four sys­tems are not statistically significant. Note that this method of estimating statistical significance might be more strict than other pairwise methods.</p><p>We also include the results of two baselines. The random baseline was calculated analytically. The first sense baseline for each language was taken from each wordnet. The first sense baseline in English and Chinese corresponds to the most frequent sense, as estimated from out-of-domain corpora. In Dutch and Italian, it followed the in­tuitions of the lexicographer. Note that we don't have the most frequent sense baseline from the do­main texts, which would surely show higher re­sults (Koeling et al., 2005).</p></section><section number="5" title="Conclusions"><p>Domain portability and adaptation of NLP com­ponents and Word Sense Disambiguation systems present new challenges. The difficulties found by supervised systems to adapt might change the way we assess the strengths and weaknesses of super­vised and knowledge-based WSD systems. With this paper we have motivated the creation of an all-words test dataset for WSD on the environ­ment domain in several languages, and presented the overall design of this SemEval task.</p><p>One of the goals of the exercise was to show that WSD systems could make use of unannotated background corpora to adapt to the domain and improve their results. Although it's early to reach hard conclusions, the results show that in each of the datasets, knowledge-based systems are able to improve their results using background text, and in two datasets the adaptation of knowledge-based systems leads to results over the MFS baseline. The evidence of domain adaptation of supervised systems is weaker, as only one team tried, and the differences with respect to MFS are very small. The best results for English are obtained by a sys­tem that combines a knowledge-based system with some targeted hand-tagging. Regarding the tech­niques used, graph-based methods over WordNet and distributional thesaurus acquisition methods have been used by several teams.</p><page local="6" global="80"/><p>All datasets and related information are publicly available from the task websites<footnote anchor="6"/>.</p></section><section title="Acknowledgments"><p>We thank the collaboration of Lawrence Jones-Walters, Amor Torre-Marin (ECNC) and Karin de Boom (WWF), com­piling the test and background documents. This work task is partially funded by the European Commission (KY­OTO ICT-2007-211423), the Spanish Research Department (KNOW-2 TIN2009-14715-C04-01) and the Basque Govern­ment (BERBATEK IE09-262).</p></section><references><p>Eneko Agirre and Aitor Soroa. 2009. Personalizing pager-ank for word sense disambiguation. In <i>Proceedings of the 12th Conference of the European Chapter of the Associa­tion for Computational Linguistics (EACL09), </i>pages 33-41. Association for Computational Linguistics.</p><p>Eneko Agirre, Oier Lopez de Lacalle, and Aitor Soroa. 2009. Knowledge-based wsd on specific domains: Performing better than generic supervised wsd. In <i>Proceedigns ofU-CAI.pp. 1501-1506.".</i></p><p>Davide Buscaldi and Paolo Rosso. 2007. Upv-wsd : Com­bining different wsd methods by means of fuzzy borda voting. In <i>Proceedings of the Fourth International Work­shop on Semantic Evaluations (SemEval-2007), </i>pages 434-437.</p><p>P. Chen, W. Ding, and D. Brown. 2009. A fully unsupervised word sense disambiguation method and its evaluation on coarse-grained all-words task. In <i>Proceeding of the North American Chapter of the Association for Computational Linguistics (NAACL09).</i></p><p>Radu Ion and Dan Stefanescu. 2009. Unsupervised word sense disambiguation with lexical chains and graph-based context formalization. In <i>Proceedings of the 4th Language and Technology Conference: Human Language Technolo­gies as a Challenge for Computer Science and Linguistics, </i>pages 190-194.</p><p>Rubén Izquierdo, Armando Suârez, and German Rigau. 2009. An empirical study on class-based word sense dis­ambiguation. In <i>EACL '09: Proceedings of the 12th Con­ference of the European Chapter of the Association for Computational Linguistics, </i>pages 389-397, Morristown, NJ, USA. Association for Computational Linguistics.</p><p>Mitesh Khapra, Sapan Shah, Piyush Kedia, and Pushpak Bhattacharyya. 2010. Domain-specific word sense dis­ambiguation combining corpus based and wordnet based parameters. In <i>Proceedings of the 5th International Con­ference on Global Wordnet (GWC2010).</i></p><p>A. Kilgarriff. 2001. English Lexical Sample Task Descrip­tion. In <i>Proceedings of the Second International Work­shop on evaluating Word Sense Disambiguation Systems, </i>Toulouse, France.</p><p>R. Koeling, D. McCarthy, and J. Carroll. 2005. Domain-specific sense distributions and predominant sense acqui­sition. In <i>Proceedings of the Human Language Technol­ogy Conference and Conference on Empirical Methods in</i> <i>Natural Language Processing.</i><i> HLT/EMNLP, </i>pages 419-426, Ann Arbor, Michigan.</p><footnote label="6">http://xmlgroup.iit.cnr.it/SemEval2 010/ and http ://semeval2.fbk.eu/</footnote><p>Xiaobin Li, Stan Szpakowicz, and Stan Matwin. 1995. A wordnet-based algorithm for word sense disambiguation. In <i>Proceedings of The 14th International Joint Conference on Artificial Intelligence (IJCAI95).</i></p><p>Diana McCarthy, Rob Koeling, Julie Weeds, and John Car­roll. 2007. Unsupervised acquisition of predominant word senses. <i>Computational Linguistics, </i>33(4).</p><p>R. Mihalcea, T. Chklovski, and Adam Killgariff. 2004. The Senseval-3 English lexical sample task. In <i>Proceedings of the 3rd ACL workshop on the Evaluation of Systems for the Semantic Analysis ofText(SENSEVAL), </i>Barcelona, Spain.</p><p>Eric W. Noreen. 1989. <i>Computer-Intensive Methods for Test­ing Hypotheses. </i>John Wiley &amp; Sons.</p><p>Sameer Pradhan, Edward Loper, Dmitriy Dligach, and Martha Palmer. 2007. Semeval-2007 task-17: English lexical sample, srl and all words. In <i>Proceedings of the Fourth International Workshop on Semantic Evaluations (SemEval-2007), </i>pages 87-92, Prague, Czech Republic.</p><p>Paul Ray son and Roger Garside. 2000. Comparing corpora using frequency profiling. In <i>Proceedings of the workshop on Comparing corpora, </i>pages 1-6.</p><p>Hansen A. Schwartz and Fernando Gomez. 2008. Acquir­ing knowledge from the web to be used as selectors for noun sense disambiguation. In <i>Proceedings of the Twelfth Conference on Computational Natural Language Learn­ing (CONLL08).</i></p><p>B. Snyder and M. Palmer. 2004. The English all-words task. In <i>Proceedings of the 3rd ACL workshop on the Evalua­tion of Systems for the Semantic Analysis of Text (SENSE-VAL), </i>Barcelona, Spain.</p><p>M. Tesconi, F. Ronzano, S. Minutoli, C. Aliprandi, and A. Marchetti. 2010. Kafnotator: a multilingual seman­tic text annotation tool. In <i>In Proceedings of the Second International Conference on Global Interoperability for Language Resources.</i></p><p>Piek Vossen, Eneko Agirre, Nicoletta Calzolari, Christiane Fellbaum, Shu kai Hsieh, Chu-Ren Huang, Hitoshi Isa-hara, Kyoko Kanzaki, Andrea Marchetti, Monica Mona-chini, Federico Neri, Remo Raffaelli, German Rigau, Maurizio Tescon, and Joop VanGent. 2008. Kyoto: a system for mining, structuring and distributing knowl­edge across languages and cultures. In <i>Proceedings of the Sixth International Language Resources and Evaluation (LREC'08), </i>Marrakech, Morocco, may. European Lan­guage Resources Association (ELRA). http://www.lrec-conf.org/proceedings/lrec2008/.</p></references></body></article>