{"id":{"repo_id":"wlv","oai_identifier":"oai:wlv.openrepository.com:2436/624530"},"canonical_url":"https://search.dev.ndltd.org/etd/wlv/oai:wlv.openrepository.com:2436/624530","repository":{"repo_id":"wlv","name":"University of Wolverhampton","base_url":"https://wlv.openrepository.com/server/oai/request"},"display":{"title":"Deep learning based semantic textual similarity for applications in translation technology","abstract":"Semantic Textual Similarity (STS) measures the equivalence of meanings between two textual segments. It is a fundamental task for many natural language processing applications. In this study, we focus on employing STS in the context of translation technology. We start by developing models to estimate STS. We propose a new unsupervised vector aggregation-based STS method which relies on contextual word embeddings. We also propose a novel Siamese neural network based on efficient recurrent neural network units. We empirically evaluate various unsupervised and supervised STS methods, including these newly proposed methods in three different English STS datasets, two non- English datasets and a bio-medical STS dataset to list the best supervised and unsupervised STS methods. We then embed these STS methods in translation technology applications. Firstly we experiment with Translation Memory (TM) systems. We propose a novel TM matching and retrieval method based on STS methods that outperform current TM systems. We then utilise the developed STS architectures in translation Quality Estimation (QE). We show that the proposed methods are simple but outperform complex QE architectures and improve the state-of-theart results. The implementations of these methods have been released as open source.","abstract_html":"Semantic Textual Similarity (STS) measures the equivalence of meanings between two textual segments. It is a fundamental task for many natural language processing applications. In this study, we focus on employing STS in the context of translation technology. We start by developing models to estimate STS. We propose a new unsupervised vector aggregation-based STS method which relies on contextual word embeddings. We also propose a novel Siamese neural network based on efficient recurrent neural network units. We empirically evaluate various unsupervised and supervised STS methods, including these newly proposed methods in three different English STS datasets, two non- English datasets and a bio-medical STS dataset to list the best supervised and unsupervised STS methods. We then embed these STS methods in translation technology applications. Firstly we experiment with Translation Memory (TM) systems. We propose a novel TM matching and retrieval method based on STS methods that outperform current TM systems. We then utilise the developed STS architectures in translation Quality Estimation (QE). We show that the proposed methods are simple but outperform complex QE architectures and improve the state-of-theart results. The implementations of these methods have been released as open source.","abstract_has_math":false,"creators":["Ranasinghe, Tharindu"],"institution":"University of Wolverhampton","degree_name":"PhD","degree_level":"Doctoral","degree_discipline":null,"degree_department":null,"school":null,"contributors":[],"advisors":["Mitkov, Ruslan"],"committee_chairs":[],"committee_members":[],"year":2021,"date_issued":"2021","date_published":"2021","updated_at":"2026-07-24T06:09:06Z","subjects":["semantic textual similarity","translation memories","translation quality estimation","deep learning","transformers","sentence encoders","TransQuest"],"languages":[],"rights":["Attribution-NonCommercial-NoDerivatives 4.0 International"],"rights_urls":["https://wlv.dspace7.openrepository.com/bitstreams/96331ba7-f58f-495d-acbe-38f99671693c/download"],"identifier_entries":[]},"links":{"outbound_url":null,"outbound_label":null,"outbound_source":null},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Mitkov, Ruslan"]},{"key":"dc:creator","label":"Author","values":["Ranasinghe, Tharindu"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.issued","label":"Date","values":["2021"]},{"key":"dc:publisher.institution","label":"Dc Publisher Institution","values":["University of Wolverhampton"]},{"key":"dc:relation.isreferencedby","label":"Dc Relation Isreferencedby","values":["http://hdl.handle.net/2436/624530"]},{"key":"dc:type","label":"Dc Type","values":["Thesis or dissertation"]},{"key":"dc:type.qualificationlevel","label":"Dc Type Qualificationlevel","values":["Doctoral"]},{"key":"dc:type.qualificationname","label":"Dc Type Qualificationname","values":["PhD"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["semantic textual similarity","translation memories","translation quality estimation","deep learning","transformers","sentence encoders","TransQuest"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:rights","label":"Dc Rights","values":["https://wlv.dspace7.openrepository.com/bitstreams/96331ba7-f58f-495d-acbe-38f99671693c/download","Attribution-NonCommercial-NoDerivatives 4.0 International"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://wlv.dspace7.openrepository.com/bitstreams/ebe63d04-ad4d-476d-8020-b0fa91283691/download"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["Semantic Textual Similarity (STS) measures the equivalence of meanings between two textual segments. It is a fundamental task for many natural language processing applications. In this study, we focus on employing STS in the context of translation technology. We start by developing models to estimate STS. We propose a new unsupervised vector aggregation-based STS method which relies on contextual word embeddings. We also propose a novel Siamese neural network based on efficient recurrent neural network units. We empirically evaluate various unsupervised and supervised STS methods, including these newly proposed methods in three different English STS datasets, two non- English datasets and a bio-medical STS dataset to list the best supervised and unsupervised STS methods. We then embed these STS methods in translation technology applications. Firstly we experiment with Translation Memory (TM) systems. We propose a novel TM matching and retrieval method based on STS methods that outperform current TM systems. We then utilise the developed STS architectures in translation Quality Estimation (QE). We show that the proposed methods are simple but outperform complex QE architectures and improve the state-of-theart results. The implementations of these methods have been released as open source."]},{"key":"dc:format.checksum.md5","label":"Dc Format Checksum Md5","values":["341f9be91f1fad98b74b9e82d6eb2030","9a8bca7caa7f8d947bb968f887038807","8a4605be74aa9ea9d79846c1fba20a33"]},{"key":"dc:title","label":"Title","values":["Deep learning based semantic textual similarity for applications in translation technology"]}]}],"canonical_facts":{"dc:contributor.advisor":["Mitkov, Ruslan"],"dc:creator":["Ranasinghe, Tharindu"],"dc:date.issued":["2021"],"dc:description.abstract":["Semantic Textual Similarity (STS) measures the equivalence of meanings between two textual segments. It is a fundamental task for many natural language processing applications. In this study, we focus on employing STS in the context of translation technology. We start by developing models to estimate STS. We propose a new unsupervised vector aggregation-based STS method which relies on contextual word embeddings. We also propose a novel Siamese neural network based on efficient recurrent neural network units. We empirically evaluate various unsupervised and supervised STS methods, including these newly proposed methods in three different English STS datasets, two non- English datasets and a bio-medical STS dataset to list the best supervised and unsupervised STS methods. We then embed these STS methods in translation technology applications. Firstly we experiment with Translation Memory (TM) systems. We propose a novel TM matching and retrieval method based on STS methods that outperform current TM systems. We then utilise the developed STS architectures in translation Quality Estimation (QE). We show that the proposed methods are simple but outperform complex QE architectures and improve the state-of-theart results. The implementations of these methods have been released as open source."],"dc:format.checksum.md5":["341f9be91f1fad98b74b9e82d6eb2030","9a8bca7caa7f8d947bb968f887038807","8a4605be74aa9ea9d79846c1fba20a33"],"dc:identifier.uri":["https://wlv.dspace7.openrepository.com/bitstreams/ebe63d04-ad4d-476d-8020-b0fa91283691/download"],"dc:publisher.institution":["University of Wolverhampton"],"dc:relation.isreferencedby":["http://hdl.handle.net/2436/624530"],"dc:rights":["https://wlv.dspace7.openrepository.com/bitstreams/96331ba7-f58f-495d-acbe-38f99671693c/download","Attribution-NonCommercial-NoDerivatives 4.0 International"],"dc:subject":["semantic textual similarity","translation memories","translation quality estimation","deep learning","transformers","sentence encoders","TransQuest"],"dc:title":["Deep learning based semantic textual similarity for applications in translation technology"],"dc:type":["Thesis or dissertation"],"dc:type.qualificationlevel":["Doctoral"],"dc:type.qualificationname":["PhD"]},"updated_at":"2026-07-24T06:09:06Z"}