Results 21 to 30 of about 266 (115)
Adapting Off-the-Shelf Speech Recognition Systems for Novel Words
Current speech recognition systems with fixed vocabularies have difficulties recognizing Out-of-Vocabulary words (OOVs) such as proper nouns and new words. This leads to misunderstandings or even failures in dialog systems.
Wiam Fadel +3 more
doaj +1 more source
Harmonizing taxon names in biodiversity data: A review of tools, databases and best practices
Abstract The process of standardizing taxon names, taxonomic name harmonization, is necessary to properly merge data indexed by taxon names. The large variety of taxonomic databases and related tools are often not well described. It is often unclear which databases are actively maintained or what is the original source of taxonomic information.
Matthias Grenié +5 more
wiley +1 more source
Abstract An app‐based clinical trial enrolment process can contribute to duplicated records, carrying data management implications. Our objective was to identify duplicated records in real time in the Apple Heart Study (AHS). We leveraged personal identifiable information (PII) to develop a dissimilarity score (DS) using the Damerau–Levenshtein ...
Ariadna Garcia +23 more
wiley +1 more source
A systematic approach to reviewing the literature can reduce bias in the types of datasets identified to understand conservation issues. Here, we present a community‐driven evidence synthesis framework that can be adapted to gather and assess evidence from many broad topics in conservation and apply the approach to gather datasets documenting long‐term
Eliza M. Grames +13 more
wiley +1 more source
Real Word Spelling Error Detection and Correction for Urdu Language
Non-word and real-word errors are generally two types of spelling errors. Non-word errors are misspelled words that are nonexistent in the lexicon while real-word errors are misspelled words that exist in the lexicon but are used out of context in a ...
Romila Aziz +7 more
doaj +1 more source
A multi‐agent system for itinerary suggestion in smart environments
Abstract Modern smart environments pose several challenges, among which the design of intelligent algorithms aimed to assist the users. When a variety of points of interest are available, for instance, trajectory recommendations are needed to suggest users the most suitable itineraries based on their interests and contextual constraints. Unfortunately,
Alessandra De Paola +4 more
wiley +1 more source
Distance Measure for Controlled Random Tests
The problem of constructing characteristics of the difference between test sequences is investigated. Its relevance for generating controlled random tests and the complexity of finding difference measures for symbolic tests are substantiated.
V. N. Yarmolik +2 more
doaj +1 more source
Conceptual Similarity and Communicative Need Shape Colexification: An Experimental Study
Abstract Colexification refers to the phenomenon of multiple meanings sharing one word in a language. Cross‐linguistic lexification patterns have been shown to be largely predictable, as similar concepts are often colexified. We test a recent claim that, beyond this general tendency, communicative needs play an important role in shaping colexification ...
Andres Karjus +4 more
wiley +1 more source
Damerau Levenshtain Distance dengan Metode Empiris untuk Koreksi Ejaan Bahasa Indonesia
Damerau Levenshtein Distance (DLD) adalah algoritma untuk koreksi kesalahan penulisan. Kesalahan terjadi karena penyisipan, penghapusan, pertukaran, dan penggantian alfabet dalam sebuah kata. Ini mungkin terjadi karena hilangnya spasi di antara dua kata.
Aji Prasetya Wibawa +4 more
doaj +1 more source
Abstract The recognition of named entities in Spanish medieval texts presents great complexity, involving specific challenges: First, the complex morphosyntactic characteristics in proper‐noun use in medieval texts. Second, the lack of strict orthographic standards.
Mª Luisa Díez Platas +4 more
wiley +1 more source

