Hot-Start from Pixels: Low-Resolution Visual Tokens for Chinese Language Modeling
arXiv:2601.09566v2 Announce Type: replace-cross Abstract: Large language models typically represent Chinese characters as discrete index-based tokens, largely ignoring their visual…
AI Survival Stories: a Taxonomic Analysis of AI Existential Risk
arXiv:2601.09765v1 Announce Type: new Abstract: Since the release of ChatGPT, there has been a lot of debate about whether AI…
Bridging Semantic Understanding and Popularity Bias with LLMs
arXiv:2601.09478v2 Announce Type: replace-cross Abstract: Semantic understanding of popularity bias is a crucial yet underexplored challenge in recommender systems, where…
Who Owns the Text? Design Patterns for Preserving Authorship in AI-Assisted Writing
arXiv:2601.10236v1 Announce Type: cross Abstract: AI writing assistants can reduce effort and improve fluency, but they may also weaken writers’…
Introduction to optimization methods for training SciML models
arXiv:2601.10222v1 Announce Type: cross Abstract: Optimization is central to both modern machine learning (ML) and scientific machine learning (SciML), yet…
Reward Learning through Ranking Mean Squared Error
arXiv:2601.09236v2 Announce Type: replace-cross Abstract: Reward design remains a significant bottleneck in applying reinforcement learning (RL) to real-world problems. A…
