Structured Matrix Scaling for Multi-Class Calibration
arXiv:2511.03685v1 Announce Type: cross Abstract: Post-hoc recalibration methods are widely used to ensure that classifiers provide faithful probability estimates. We…
Whisper Leak: a side-channel attack on Large Language Models
arXiv:2511.03675v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in sensitive domains including healthcare, legal services, and…
In Situ Training of Implicit Neural Compressors for Scientific Simulations via Sketch-Based Regularization
arXiv:2511.02659v2 Announce Type: replace-cross Abstract: Focusing on implicit neural representations, we present a novel in situ training protocol that employs…
Multimodal Detection of Fake Reviews using BERT and ResNet-50
arXiv:2511.00020v1 Announce Type: new Abstract: In the current digital commerce landscape, user-generated reviews play a critical role in shaping consumer…
Best Practices for Biorisk Evaluations on Open-Weight Bio-Foundation Models
arXiv:2510.27629v3 Announce Type: replace-cross Abstract: Open-weight bio-foundation models present a dual-use dilemma. While holding great promise for accelerating scientific research…
CMI-MTL: Cross-Mamba interaction based multi-task learning for medical visual question answering
arXiv:2511.01357v1 Announce Type: cross Abstract: Medical visual question answering (Med-VQA) is a crucial multimodal task in clinical decision support and…
Thinking with DistilQwen: A Tale of Four Distilled Reasoning and Reward Model Series
arXiv:2511.01354v1 Announce Type: cross Abstract: Recently, the demand for small and efficient reasoning models to support real-world applications has driven…
How Similar Are Grokipedia and Wikipedia? A Multi-Dimensional Textual and Structural Comparison
arXiv:2510.26899v2 Announce Type: replace-cross Abstract: The launch of Grokipedia, an AI-generated encyclopedia developed by Elon Musk’s xAI, was presented as…
CATArena: Evaluation of LLM Agents through Iterative Tournament Competitions
arXiv:2510.26852v1 Announce Type: new Abstract: Large Language Model (LLM) agents have evolved from basic text generation to autonomously completing complex…
Faithful and Fast Influence Function via Advanced Sampling
arXiv:2510.26776v2 Announce Type: replace-cross Abstract: How can we explain the influence of training data on black-box models? Influence functions (IFs)…
