<?xml version="1.0" encoding="UTF-8"?><?xml-stylesheet type="text/xsl" href="static/style.xsl"?><OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd"><responseDate>2026-09-18T17:58:15Z</responseDate><request verb="GetRecord" identifier="oai:gupea.ub.gu.se:2077/60509" metadataPrefix="dim">https://gupea.ub.gu.se/server/oai/request</request><GetRecord><record><header><identifier>oai:gupea.ub.gu.se:2077/60509</identifier><datestamp>2019-08-24T01:35:44Z</datestamp><setSpec>com_2077_18176</setSpec><setSpec>com_2077_19</setSpec><setSpec>com_2077_10556</setSpec><setSpec>col_2077_19517</setSpec><setSpec>col_2077_10557</setSpec></header><metadata><dim:dim xmlns:dim="http://www.dspace.org/xmlns/dspace/dim" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xmlns:doc="http://www.lyncode.com/xoai" xsi:schemaLocation="http://www.dspace.org/xmlns/dspace/dim http://www.dspace.org/schema/dim.xsd">
   <dim:field mdschema="dc" element="contributor" qualifier="author">Nieto Piña, Luis</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="accessioned">2019-08-23T08:03:45Z</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="available">2019-08-23T08:03:45Z</dim:field>
   <dim:field mdschema="dc" element="date" qualifier="issued">2019-08-23</dim:field>
   <dim:field mdschema="dc" element="identifier" qualifier="isbn">978-91-87850-75-2</dim:field>
   <dim:field mdschema="dc" element="identifier" qualifier="issn">0347-948X</dim:field>
   <dim:field mdschema="dc" element="identifier" qualifier="uri">http://hdl.handle.net/2077/60509</dim:field>
   <dim:field mdschema="dc" element="description" qualifier="abstract" lang="sv">The representation of written language semantics is a central problem of language technology and a crucial component of many natural language processing applications, from part-of-speech tagging to text summarization. These representations of linguistic units, such as words or sentences, allow computer applications that work with language to process and manipulate the meaning of text. In particular, a family of models has been successfully developed based on automatically learning semantics from large collections of text and embedding them into a vector space, where semantic or lexical similarity is a function of geometric distance. Co-occurrence information of words in context is the main source of data used to learn these representations.&#xd;
&#xd;
Such models have typically been applied to learning representations for word forms, which have been widely applied, and proven to be highly successful, as characterizations of semantics at the word level. However, a word-level approach to meaning representation implies that the different meanings, or senses, of any polysemic word share one single representation. This might be problematic when individual word senses are of interest and explicit access to their specific representations is required. For instance, in cases such as an application that needs to deal with word senses rather than word forms, or when a digital lexicon&amp;apos;s sense inventory has to be mapped to a set of learned semantic representations.&#xd;
&#xd;
In this thesis, we present a number of models that try to tackle this problem by automatically learning representations for word senses instead of for words. In particular, we try to achieve this by using two separate sources of information: corpora and lexica for the Swedish language. Throughout the five publications compiled in this thesis, we demonstrate that it is possible to generate word sense representations from these sources of data individually and in conjunction, and we observe that combining them yields superior results in terms of accuracy and sense inventory coverage. Furthermore, in our evaluation of the different representational models proposed here, we showcase the applicability of word sense representations both to downstream natural language processing applications and to the development of existing linguistic resources.</dim:field>
   <dim:field mdschema="dc" element="language" qualifier="iso" lang="sv">eng</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="ispartofseries" lang="sv">Data Linguistica</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="ispartofseries" lang="sv">30</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="haspart" lang="sv">Luis Nieto Piña and Richard Johansson 2015. A simple and efficient method to generate word sense representations. Proceedings of the International Conference Recent Advances in Natural Language Processing, 465–472. Hissar, Bulgaria.</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="haspart" lang="sv">Luis Nieto Piña and Richard Johansson 2016. Embedding senses for efficient graph-based word sense disambiguation. Proceedings of TextGraphs-10: the Workshop on Graph-based Methods for Natural Language Processing, NAACL-HLT 2016, 1–5. San Diego, USA.</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="haspart" lang="sv">Luis Nieto Piña and Richard Johansson 2017. Training word sense embeddings with lexicon-based regularization. Proceedings of the Eighth International Joint Conference on Natural Language Processing (Volume 1: Long Papers). Asian Federation of Natural Language Processing. Taipei, Taiwan.</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="haspart" lang="sv">Luis Nieto Piña and Richard Johansson 2018. Automatically Linking Lexical Resources with Word Sense Embedding Models. Proceedings of the third workshop on semantic deep learning (SemDeep-3), COLING 2018, 23–29. Association for Computational Linguistics. Santa Fe, USA.</dim:field>
   <dim:field mdschema="dc" element="relation" qualifier="haspart" lang="sv">Lars Borin, Luis Nieto Piña and Richard Johansson 2015. Here be dragons? The perils and promises of inter-resource lexical-semantic mapping. Proceedings of the workshop on semantic resources and semantic annotation for natural language processing and the digital humanities at NODALIDA 2015, 1–11. Vilnius, Lithuania.</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">language technology</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">natural language processing</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">distributional models</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">semantic representations</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">distributed representations</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">word senses</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">embeddings</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">word sense disambiguation</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">linguistic resources</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">corpus</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">lexicon</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">machine learning</dim:field>
   <dim:field mdschema="dc" element="subject" lang="sv">neural networks</dim:field>
   <dim:field mdschema="dc" element="title" lang="sv">Splitting rocks: Learning word sense representations from corpora and lexica</dim:field>
   <dim:field mdschema="dc" element="type">Text</dim:field>
   <dim:field mdschema="dc" element="type" qualifier="svep" lang="eng">Doctoral thesis</dim:field>
   <dim:field mdschema="dc" element="type" qualifier="degree" lang="sv">Doctor of Philosophy</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="origin" lang="swe">Göteborgs universitet. Humanistiska fakulteten</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="origin" lang="eng">University of Gothenburg. Faculty of Arts</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="department" lang="sv">Department of Swedish ; Institutionen för svenska språket</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="defenceplace" lang="sv">Fredagen den 13 september 2019, kl. 13.15, Lilla hörsalen, Humanisten, Lundgrensgatan 1B</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="defencedate">2019-09-13</dim:field>
   <dim:field mdschema="dc" element="gup" qualifier="dissdb-fakultet">HF</dim:field>
   <dim:field mdschema="others" element="access-status">open.access</dim:field>
</dim:dim>
</metadata></record></GetRecord></OAI-PMH>