Qwen Councils
arXiv is taking too long to respond. Please try again or narrow your search.
Showing downloaded papers while arXiv is unavailable.

All arXiv

arXiv preprints from January 1, 2026 through September 13, 2026 — 08:13:00 EST

0

Posted in cs.CV · 2026-01-21 · Gautom Das, Vincent La, Ethan Lau, Abhinav Shrivastava, Matthew Gwilliam

Towards Understanding Best Practices for Quantization of Vision-Language Models

Large language models (LLMs) deliver impressive results for a variety of tasks, but state-of-the-art systems require fast GPUs with large amounts of memory. To reduce both the memory and latency of these systems, practitioners quantize their learned parameters, typically at half precision. A growing body of research focuses on...

💬 0 commentsarXiv:2601.15287v1PDF
0

Posted in cs.CV · 2026-01-21 · Shantanu Jaiswal, Mihir Prabhudesai, Nikash Bhardwaj, Zheyang Qin, Amir Zadeh, Chuan Li, Katerina Fragkiadaki, Deepak Pathak

Iterative Refinement Improves Compositional Image Generation

Text-to-image (T2I) models have achieved remarkable progress, yet they continue to struggle with complex prompts that require simultaneously handling multiple objects, relations, and attributes. Existing inference-time strategies, such as parallel sampling with verifiers or simply increasing denoising steps, can improve prompt...

💬 0 commentsarXiv:2601.15286v1PDF
0

Posted in cs.CV · 2026-01-21 · Anurag Bagchi, Zhipeng Bao, Homanga Bharadhwaj, Yu-Xiong Wang, Pavel Tokmakov, Martial Hebert

Walk through Paintings: Egocentric World Models from Internet Priors

What if a video generation model could not only imagine a plausible future, but the correct one, accurately reflecting how the world changes with each action? We address this question by presenting the Egocentric World Model (EgoWM), a simple, architecture-agnostic method that transforms any pretrained video diffusion model into an...

💬 0 commentsarXiv:2601.15284v1PDF
0

Posted in astro-ph.EP · 2026-01-21 · Yinuo Han, Mark Wyatt, Kate Y. L. Su, Antranik A. Sefilian, Joshua B. Lovell, Carlos del Burgo, Jonathan P. Marshall, Sebastian Marino, David J. Wilner, Brenda C. Matthews, Max Sommer, A. Meredith Hughes, John M. Carpenter, Meredith A. MacGregor, Nicole Pawellek, Thomas Henning

A radially broad collisional cascade in the debris disk of $γ$ Ophiuchi observed by JWST

The A1V star $γ$ Oph, at a distance of 29.7 pc, is known from Spitzer imaging to host a debris disk with a large radial extent and from its spectral energy distribution to host inner warm dust. We imaged $γ$ Oph with JWST/MIRI at 15 and 25.5 $μ$m, revealing smooth and radially broad emission that extends to a radius of at least 250 au...

💬 0 commentsarXiv:2601.15285v2PDF
0

Posted in cs.CV · 2026-01-21 · Ruofan Liang, Norman Müller, Ethan Weber, Duncan Zauss, Nandita Vijaykumar, Peter Kontschieder, Christian Richardt

LuxRemix: Lighting Decomposition and Remixing for Indoor Scenes

We present a novel approach for interactive light editing in indoor scenes from a single multi-view scene capture. Our method leverages a generative image-based light decomposition model that factorizes complex indoor scene illumination into its constituent light sources. This factorization enables independent manipulation of...

💬 0 commentsarXiv:2601.15283v2PDF
0

Posted in cs.CV · 2026-01-21 · Yufan Deng, Zilin Pan, Hongyu Zhang, Xiaojie Li, Ruoqing Hu, Yufei Ding, Yiming Zou, Yan Zeng, Daquan Zhou

Rethinking Video Generation Model for the Embodied World

Video generation models have significantly advanced embodied intelligence, unlocking new possibilities for generating diverse robot data that capture perception, reasoning, and action in the physical world. However, synthesizing high-quality videos that accurately reflect real-world robotic interactions remains challenging, and the...

💬 0 commentsarXiv:2601.15282v1PDF
0

Posted in cs.CV · 2026-01-21 · Ying Yang, Zhengyao Lv, Tianlin Pan, Haofan Wang, Binxin Yang, Hubery Yin, Chen Li, Ziwei Liu, Chenyang Si

StableWorld: Towards Stable and Consistent Long Interactive Video Generation

In this paper, we explore the overlooked challenge of stability and temporal consistency in interactive video generation, which synthesizes dynamic and controllable video worlds through interactive behaviors such as camera movements and text prompts. Despite remarkable progress in world modeling, current methods still suffer from...

💬 0 commentsarXiv:2601.15281v1PDF
0

Posted in cs.HC · 2026-01-21 · Chloe Qianhui Zhao, Jie Cao, Jionghao Lin, Kenneth R. Koedinger

LLM-based Multimodal Feedback Produces Equivalent Learning and Better Student Perceptions than Educator Feedback

Providing timely, targeted, and multimodal feedback helps students quickly correct errors, build deep understanding and stay motivated, yet making it at scale remains a challenge. This study introduces a real-time AI-facilitated multimodal feedback system that integrates structured textual explanations with dynamic multimedia...

💬 0 commentsarXiv:2601.15280v1PDF
0

Posted in cs.LG · 2026-01-21 · Christoph Bartmann, Johannes Schimunek, Mykyta Ielanskyi, Philipp Seidl, Günter Klambauer, Sohvi Luukkonen

MolecularIQ: Characterizing Chemical Reasoning Capabilities Through Symbolic Verification on Molecular Graphs

A molecule's properties are fundamentally determined by its composition and structure encoded in its molecular graph. Thus, reasoning about molecular properties requires the ability to parse and understand the molecular graph. Large Language Models (LLMs) are increasingly applied to chemistry, tackling tasks such as molecular name...

💬 0 commentsarXiv:2601.15279v1PDF
0

Posted in cs.MM · 2026-01-21 · Mingyue Zha, Ho-Chun Herbert Chang

Interpreting Multimodal Communication at Scale in Short-Form Video: Visual, Audio, and Textual Mental Health Discourse on TikTok

Short-form video platforms integrate text, visuals, and audio into complex communicative acts, yet existing research analyzes these modalities in isolation, lacking scalable frameworks to interpret their joint contributions. This study introduces a pipeline combining automated multimodal feature extraction with Shapley value-based...

💬 0 commentsarXiv:2601.15278v1PDF
0

Posted in cs.CL · 2026-01-21 · Sahar Tahmasebi, Eric Müller-Budack, Ralph Ewerth

Robust Fake News Detection using Large Language Models under Adversarial Sentiment Attacks

Misinformation and fake news have become a pressing societal challenge, driving the need for reliable automated detection methods. Prior research has highlighted sentiment as an important signal in fake news detection, either by analyzing which sentiments are associated with fake news or by using sentiment and emotion features for...

💬 0 commentsarXiv:2601.15277v1PDF
0

Posted in math.CO · 2026-01-21 · Ruben Carpenter, Colin Defant, Noah Kravitz

On the number of permutation-twisted dot products

Let $\mathbb{K}$ be a field of characteristic $0$. For each choice of distinct $a_1, \ldots, a_n\in \mathbb{K}$ and distinct $b_1, \ldots, b_n\in \mathbb{K}$, consider the sum $S=\sum_{i=1}^n a_i b_{π(i)}$ as $π$ ranges over the permutations of $[n]$. We show that this sum always assumes at least $Ω(n^3)$ distinct values. This...

💬 0 commentsarXiv:2601.15276v3PDF
0

Posted in cs.CV · 2026-01-21 · Yu Wu, Minsik Jeon, Jen-Hao Rick Chang, Oncel Tuzel, Shubham Tulsiani

RayRoPE: Projective Ray Positional Encoding for Multi-view Attention

We study positional encodings for multi-view transformers that process tokens from a set of posed input images, and seek a mechanism that encodes patches uniquely, allows SE(3)-invariant attention with multi-frequency similarity, and can adapt to the geometry of the underlying 3D scene. We find that prior (absolute or relative)...

💬 0 commentsarXiv:2601.15275v3PDF
0

Posted in astro-ph.GA · 2026-01-21 · Zichen Hua, Lelli Federico, Enrico Di Teodoro, Stacy McGaugh, James Schombert

The baryonic mass-size relation of galaxies. II. Implications for the evolutionary paths between star-forming and passive galaxies

The baryonic mass-size relation of galaxies links the total baryonic mass (stars plus gas) to the baryonic half-mass radius. In the first paper of this series, we showed that star-forming galaxies from the SPARC sample follow two distinct relations in the baryonic mass-size plane: one defined by high-surface-density (HSD),...

💬 0 commentsarXiv:2601.15274v1PDF
0

Posted in q-bio.QM · 2026-01-21 · Paul Van Liedekerke, Jiří Pešek, Kevin Alessandri, Dirk Drasdo

How high-resolution agent-based models can improve fundamental insights in tissue development and cell culturing methods

The fundamental understanding of how cells physically interact with each other and their environment is key to understanding their organisation in living tissues. Over the past decades several computational methods have been developed to decipher emergent multi-cellular behaviors. In particular agent-based (or cell-based) models that...

💬 0 commentsarXiv:2601.15273v1PDF
0

Posted in cs.LG · 2026-01-21 · Maciej Kilian, Oleg Mkrtchyan, Luke Zettlemoyer, Akshat Shrivastava, Armen Aghajanyan

Improving MoE Compute Efficiency by Composing Weight and Data Sparsity

Mixture-of-Experts layers achieve compute efficiency through weight sparsity: each token activates only a subset of experts. Data sparsity, where each expert processes only a subset of tokens, offers a complementary axis. Expert-choice routing implements data sparsity directly but violates causality in autoregressive models, creating...

💬 0 commentsarXiv:2601.15370v1PDF
0

Posted in math.CO · 2026-01-21 · Ronald Orozco López

Lucas-Pantograph Type Exponential, Trigonometric, and Hyperbolic Functions

In this paper, we include some new results for the Lucas calculus. A Lucas-Pantograph type exponential function is introduced. Additionally, we define Lucas-Pantograph type trigonometric functions, and some of their most notable identities are given: parity, sum and difference formulas, Pythagorean identities, double-angle identities,...

💬 0 commentsarXiv:2601.15272v1PDF
0

Posted in math.AG · 2026-01-21 · Jun-Yong Park

Height moduli of elliptic surfaces: Motivic height zeta rationality and Kudla-Millson modularity of Mordell-Weil rank jumps

Let $k$ be a perfect field with $\mathrm{char}(k)\neq 2,3$, set $K=k(t)$, and let $\mathcal{W}_n^{\min}$ be the moduli stack of minimal elliptic curves over $K$ of Faltings height $n$, constructed via the height-moduli framework of Bejleri-Park-Satriano applied to $\overline{\mathcal{M}}_{1,1}\simeq\mathcal{P}(4,6)$. The Shioda-Tate...

💬 0 commentsarXiv:2601.15543v5PDF
0

Posted in astro-ph.CO · 2026-01-21 · Sandro D. P. Vitenti, Nelson Pinto-Neto, Patrick Peter, Luiz Felipe Demétrio

Two Fluid Quantum Bouncing Cosmology I: Theoretical Model

Bouncing cosmologies offer an alternative to inflation by resolving the initial singularity through a contracting phase followed by a bounce into expansion. In many such models, the contracting phase is dominated by a single matter component, typically pressureless dust, which leads to an almost scale-invariant spectrum of scalar...

💬 0 commentsarXiv:2601.15542v2PDF
0

Posted in cs.RO · 2026-01-21 · Heng Zhang, Wei-Hsing Huang, Qiyi Tong, Gokhan Solak, Puze Liu, Kaidi Zhang, Sheng Liu, Jan Peters, Yu She, Arash Ajoudani

CompliantVLA-adaptor: VLM-Guided Variable Impedance Action for Safe Contact-Rich Manipulation

We propose a CompliantVLA-adaptor that augments the state-of-the-art Vision-Language-Action (VLA) models with vision-language model (VLM)-informed context-aware variable impedance control (VIC) to improve the safety and effectiveness of contact-rich robotic manipulation tasks. Existing VLA systems (e.g., RDT, Pi0.5, OpenVLA-oft)...

💬 0 commentsarXiv:2601.15541v2PDF
0

Posted in cs.LG · 2026-01-21 · Dongchen Huang

PRISM: Deriving a White-Box Transformer as a Signal-Noise Decomposition Operator via Maximum Coding Rate Reduction

Deep learning models, particularly Transformers, are often criticized as "black boxes" and lack interpretability. We propose Prism, a white-box attention-based architecture derived from the principles of Maximizing Coding Rate Reduction ($\text{MCR}^2$). By modeling the attention mechanism as a gradient ascent process on a distinct...

💬 0 commentsarXiv:2601.15540v2PDF
0

Posted in eess.IV · 2026-01-21 · Ali Khreis, Ro'Yah Radaideh, Quinn McGill

A Machine Vision Approach to Preliminary Skin Lesion Assessments

Early detection of malignant skin lesions is critical for improving patient outcomes in aggressive, metastatic skin cancers. This study evaluates a comprehensive system for preliminary skin lesion assessment that combines the clinically established ABCD rule of dermoscopy (analyzing Asymmetry, Borders, Color, and Dermoscopic...

💬 0 commentsarXiv:2601.15539v1PDF
0

Posted in cs.LG · 2026-01-21 · Himanshu Mishra, Kanwal Mehreen

QUAIL: Quantization Aware Unlearning for Mitigating Misinformation in LLMs

Machine unlearning aims to remove specific knowledge (e.g., copyrighted or private data) from a trained model without full retraining. In practice, models are often quantized (e.g., 4-bit) for deployment, but we find that quantization can catastrophically restore forgotten information [1]. In this paper, we (1) analyze why low-bit...

💬 0 commentsarXiv:2601.15538v1PDF
0

Posted in eess.AS · 2026-01-21 · Jiaheng Dong, Hong Jia, Ting Dang

Test-Time Adaptation for Speech Emotion Recognition

The practical utility of Speech Emotion Recognition (SER) systems is undermined by their fragility to domain shifts, such as speaker variability, the distinction between acted and naturalistic emotions, and cross-corpus variations. While domain adaptation and fine-tuning are widely studied, they require either source data or labelled...

💬 0 commentsarXiv:2601.16240v1PDF
0

Posted in physics.soc-ph · 2026-01-21 · Jhordan Silveira de Borba, Celia Anteneodo, Sebastian Gonçalves

Can Rising Consumption Deepen Inequality?

The impact of rising consumption on wealth inequality remains an open question. Here we revisit and extend the Social Architecture of Capitalism agent-based model proposed by Ian Wright, which reproduces stylized facts of wealth and income distributions. In a previous study, we demonstrated that the macroscopic behavior of the model...

💬 0 commentsarXiv:2601.15537v2PDF