Results 111 to 120 of about 27,036 (297)
Segmenting DNA sequence into words based on statistical language model
This paper presents a novel method to segment/decode DNA sequences based on n-gram statistical language model. Firstly, we find the length of most DNA “words” is 12 to 15 bps by analyzing the genomes of 12 model species.
Wang Liang
core +1 more source
Word Segmentation in Indo-China Languages for Digital Libraries
This chapter introduces word segmentation methods for Indo-China languages. It describes six different word segmentation methods developed for the Thai, Vietnamese, and Myanmar languages and compare different approaches in terms of their algorithms and ...
Schubert Foo +4 more
core +1 more source
Microstructure Evolution of a VMnFeCoNi High‐Entropy Alloy After Synthesis, Swaging, and Annealing
The synthesis and processing (rotary swaging and annealing) of the novel VMnFeCoNi alloy is investigated, alongside the estimation of the grain size effect on hardness. Analysis of a wide grain size range of recrystallized microstructures (12–210 µm) reveals a low annealing twin density.
Aditya Srinivasan Tirunilai +6 more
wiley +1 more source
A Lightweight Procedural Layer for Hybrid Experimental–Computational Workflows in Materials Science
We unveil a prototype hybrid‐workflow framework that fuses automatedcomputation with hands‐on experiments. Built atop pyiron, a lightweight, parameterized layer translates procedure descriptions into executable manual steps, syncing instrument settings, human interventions, and data capture in real‐time today.
Steffen Brinckmann +8 more
wiley +1 more source
Unsupervised Training for Overlapping Ambiguity Resolution in
This paper proposes an unsupervised training approach to resolving overlapping ambiguities in Chinese word ...
Changning Huang +3 more
core
In this paper, a novel approach for word extraction and character segmentation from the handwritten Bangla document images is reported. At first, a modified Run Length Smoothing Algorithm (RLSA), called Spiral Run Length Smearing Algorithm (SRLSA), is ...
Sarkar Ram +5 more
doaj +1 more source
Reproduction of stacking fault energy calculations from literature with a semi‐automated large language model‐assisted extraction procedure: extraction of simulation protocol, atomistic structures, computational parameters, and reported results, ontology alignment, knowledge graph construction and, finally, recomputation forvalidation.
Sepideh Baghaee Ravari +5 more
wiley +1 more source
Unlike alphabetic language, Chinese is an ideographic language that does not contain spaces between words. Chinese readers must develop unique segmentation strategies for word recognition and reading comprehension.
Mingjing Chen +3 more
doaj +1 more source
Coarse‐grained (left) and atomistic (right) models of the shape memory polymer ESTANE ETE 75DT3 are shown schematically. The two representations bridge molecular detail and mesoscopic description. Both models capture shape memory behavior, linking segmental mobility and conformational relaxation of anisotropic chains to macroscopic recovery, and ...
Fathollah Varnik
wiley +1 more source
Grain boundary triple junctions are an essential ingredient of the microstructure of polycrystalline materials. In this study, a triple junction is observed using atomic‐resolution scanning transmission electron microscopy and characterized. Computer simulations reveal that the junction has a dislocation character that is determined by the joining ...
Tobias Brink +4 more
wiley +1 more source

