Results 241 to 250 of about 152,555 (310)
Some of the next articles are maybe not open access.
An Image is Worth 32 Tokens for Reconstruction and Generation
Neural Information Processing SystemsRecent advancements in generative models have highlighted the crucial role of image tokenization in the efficient synthesis of high-resolution images.
Qihang Yu +5 more
semanticscholar +1 more source
Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens
arXiv.orgVision-language models (VLMs) excel at multimodal understanding, yet their text-only decoding forces them to verbalize visual reasoning, limiting performance on tasks that demand visual imagination.
Zeyuan Yang +4 more
semanticscholar +1 more source
Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
Annual Meeting of the Association for Computational LinguisticsInformation about pretraining corpora used to train the current best-performing language models is seldom discussed: commercial models rarely detail their data, and even open models are often released without accompanying training data or recipes to ...
Luca Soldaini +35 more
semanticscholar +1 more source
Token Assorted: Mixing Latent and Text Tokens for Improved Language Model Reasoning
International Conference on Machine LearningLarge Language Models (LLMs) excel at reasoning and planning when trained on chainof-thought (CoT) data, where the step-by-step thought process is explicitly outlined by text tokens.
DiJia Su +5 more
semanticscholar +1 more source
Software: Practice and Experience, 2001
AbstractA common problem when writing compilers for programming languages or little, domain‐specific languages is that an input token may have several interpretations, depending on context. Solutions to this problem demand programmer intervention, obfuscate the language's grammar, and may introduce subtle bugs.
John Aycock, R. Nigel Horspool
openaire +2 more sources
AbstractA common problem when writing compilers for programming languages or little, domain‐specific languages is that an input token may have several interpretations, depending on context. Solutions to this problem demand programmer intervention, obfuscate the language's grammar, and may introduce subtle bugs.
John Aycock, R. Nigel Horspool
openaire +2 more sources
Recent Advances in Discrete Speech Tokens: A Review
IEEE Transactions on Pattern Analysis and Machine IntelligenceThe rapid advancement of speech generation technologies in the era of large language models (LLMs) has established discrete speech tokens as a foundational paradigm for speech representation.
Yiwei Guo +9 more
semanticscholar +1 more source
Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency
Conference on Empirical Methods in Natural Language ProcessingRecent advances in large reasoning models have enabled complex, step-by-step reasoning but often introduce significant overthinking, resulting in verbose and redundant outputs that hinder efficiency.
Chenlong Wang +5 more
semanticscholar +1 more source
∞Bench: Extending Long Context Evaluation Beyond 100K Tokens
Volume 1Processing and reasoning over long contexts is crucial for many practical applications of Large Language Models (LLMs), such as document comprehension and agent construction.
Xinrong Zhang +10 more
semanticscholar +1 more source
TimeChat-Online: 80% Visual Tokens are Naturally Redundant in Streaming Videos
ACM MultimediaThe rapid growth of online video platforms, particularly live streaming services, has created an urgent need for real-time video understanding systems.
Linli Yao +13 more
semanticscholar +1 more source

