DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
arXiv:2510.15015v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have advanced rapidly, yet they remain vulnerable to semantic leakage, the unintended…
ClapperText: A Benchmark for Text Recognition in Low-Resource Archival Documents
arXiv:2510.15557v1 Announce Type: cross Abstract: This paper presents ClapperText, a benchmark dataset for handwritten and printed text recognition in visually…
PAD: Phase-Amplitude Decoupling Fusion for Multi-Modal Land Cover Classification
arXiv:2504.19136v3 Announce Type: replace-cross Abstract: The fusion of Synthetic Aperture Radar (SAR) and RGB imagery for land cover classification remains…
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
arXiv:2510.14588v1 Announce Type: cross Abstract: Video generation has recently made striking visual progress, but maintaining coherent object motion and interactions…
Does FLUX Already Know How to Perform Physically Plausible Image Composition?
arXiv:2509.21278v2 Announce Type: replace-cross Abstract: Image composition aims to seamlessly insert a user-specified object into a new scene, but existing…
From Craft to Constitution: A Governance-First Paradigm for Principled Agent Engineering
arXiv:2510.13857v1 Announce Type: cross Abstract: The advent of powerful Large Language Models (LLMs) has ushered in an “Age of the…
LiRA: Linguistic Robust Anchoring for Cross-lingual Large Language Models
arXiv:2510.14466v1 Announce Type: cross Abstract: As large language models (LLMs) rapidly advance, performance on high-resource languages (e.g., English, Chinese) is…
Attention-Aided MMSE for OFDM Channel Estimation: Learning Linear Filters with Attention
arXiv:2506.00452v2 Announce Type: replace-cross Abstract: In orthogonal frequency division multiplexing (OFDM), accurate channel estimation is crucial. Classical signal processing based…
Autonomous ultrasound with UltraBot
Post Content
