Results 151 to 160 of about 307,200 (226)
Lexical Simplification: From a Japanese Dataset to Multilingual Methods Leveraging Large Language Models and a YouTube Subtitle Corpus [PDF]
Nohejl Adam
openalex +1 more source
Introduction: Address Terms in World Englishes
ABSTRACT This special issue explores variations in English address terms across world Englishes, highlighting how globalization and digital communication continue to shape their use. While Western Englishes like British and American English show a decline in hierarchy‐stressing terms (e.g.
Anke Lensch +2 more
wiley +1 more source
Abstract This paper provides ethnographic analysis of interactions between volunteers in an open‐science AI project entrusted with preparing multilingual text for use in fine‐tuning a pretrained LLM. Volunteers were encouraged to entextualize translated text into structured corpora and to write and edit prompt‐completion pairings in over 100 languages.
Joseph Wilson
wiley +1 more source
Automatic lexical simplification is a task to substitute lexical items that may be unfamiliar and difficult to understand with easier and more common words.
Salazar, Martin Solis +4 more
core
Image Encoding With Blipped Gradients in an Echo Train
ABSTRACT Purpose An MRI scanner was designed and built to encode k‐space points in the spin echoes occurring between 180∘ pulses in an echo train by using blipped B0 gradient pulses applied just prior to those spin echoes. Methods The proposed MRI scanner is a modification of an original design that used TRansmit Array Spatial Encoding (TRASE), a built‐
Logi Vidarsson, Gordon E. Sarty
wiley +1 more source
ArGemma: A Multi‐Task Fine‐Tuning Framework for Adapting Gemma to Arabic
ABSTRACT Open‐source large language models (LLMs) have significantly advanced natural language processing (NLP), particularly for English. However, their performance in Arabic has remained limited due to the scarcity of high‐quality datasets and the high computational cost of full fine‐tuning.
Taha Alselwi +2 more
wiley +1 more source
Simplified Chinese lexicon project: A lexical decision database with 8105 characters and 4864 pseudocharacters [PDF]
Yixia Wang +5 more
openalex +1 more source
The proposed model takes an imbalanced English dataset as input and balances it by Generating Synthetic Tweets (GST) using GPT‐based paraphrasing. These new tweets are then annotated with one of 10 categorical labels representing their primary meaning, using another large language model like GPT.
Hossein Nekouei +1 more
wiley +1 more source

