Qwen Councils

Computer Science

arXiv preprints from January 1, 2026 through September 11, 2026 — 04:44:15 EST

0

Posted in cs.IT · 2026-01-14 · Mingyu Hu, Nan Liu, Wei Kang

Movable Antenna Assisted Dual-Polarized Multi-Cell Cooperative AirComp: An Alternating Optimization Approach

Over-the-air computation (AirComp) is a key enabler for distributed optimization, since it leverages analog waveform superposition to perform aggregation and thereby mitigates the communication bottleneck caused by iterative information exchange. However, AirComp is sensitive to wireless environment and conventional systems with fixed...

💬 0 commentsarXiv:2601.09137v1PDF
0

Posted in cs.CV · 2026-01-14 · Lijun Liu, Linwei Chen, Zhishou Zhang, Meng Tian, Hengfu Cui, Ruiyang Li, Zhaocheng Liu, Qiang Ju, Qianxi Li, Hong-Yu Zhou

SkinFlow: Efficient Information Transmission for Open Dermatological Diagnosis via Dynamic Visual Encoding and Staged RL

General-purpose Large Vision-Language Models (LVLMs), despite their massive scale, often falter in dermatology due to "diffuse attention" - the inability to disentangle subtle pathological lesions from background noise. In this paper, we challenge the assumption that parameter scaling is the only path to medical precision. We...

💬 0 commentsarXiv:2601.09136v1PDF
0

Posted in cs.CL · 2026-01-14 · Jongha Kim, Byungoh Ko, Jeehye Na, Jinsung Yoon, Hyunwoo J. Kim

Relevance-aware Multi-context Contrastive Decoding for Retrieval-augmented Visual Question Answering

Despite the remarkable capabilities of Large Vision Language Models (LVLMs), they still lack detailed knowledge about specific entities. Retrieval-augmented Generation (RAG) is a widely adopted solution that enhances LVLMs by providing additional contexts from an external Knowledge Base. However, we observe that previous decoding...

💬 0 commentsarXiv:2602.06050v1PDF
0

Posted in cs.CR · 2026-01-14 · Xiaonan Liu, Zhihao Li, Xiao Lan, Hao Ren, Haizhou Wang, Xingshu Chen

KryptoPilot: An Open-World Knowledge-Augmented LLM Agent for Automated Cryptographic Exploitation

Capture-the-Flag (CTF) competitions play a central role in modern cybersecurity as a platform for training practitioners and evaluating offensive and defensive techniques derived from real-world vulnerabilities. Despite recent advances in large language models (LLMs), existing LLM-based agents remain ineffective on high-difficulty...

💬 0 commentsarXiv:2601.09129v1PDF
0

Posted in cs.CV · 2026-01-14 · Yurun Song, Jiong Yin, Rongjunchen Zhang, Ian G. Harris

Compress to Focus: Efficient Coordinate Compression for Policy Optimization in Multi-Turn GUI Agents

Multi-turn GUI agents enable complex task completion through sequential decision-making, but suffer from severe context inflation as interaction history accumulates. Existing strategies either sacrifice long-term context via truncation or compromise spatial structure through token pruning. In this paper, we propose Coordinate...

💬 0 commentsarXiv:2601.11631v1PDF
0

Posted in cs.IR · 2026-01-14 · Anh Nguyen Van, Huy Ngo Hoang, Khoi Ngo Nguyen, Ngoc Pham Thi, Khanh Ngo Mai Bao, Quyen Nguyen Van

Consensus-Driven Group Recommendation on Sparse Explicit Feedback: A Collaborative Filtering and Choquet-Borda Aggregation Framework

Group Recommender Systems (GRS) play an essential role in supporting collective decision-making among users with diverse and potentially conflicting preferences. However, achieving stable intra-group consensus becomes particularly challenging when only sparse userID-itemID-rating data are available and no demographic, contextual, or...

💬 0 commentsarXiv:2603.21012v1PDF
0

Posted in cs.IT · 2026-01-14 · Rachel St. Clair, John Austin Cook, Peter Sutor, Victor Cavero, Garrett Mindt

The .serva Standard: One Primitive for All AI Cost Reduced, Barriers Removed

Artificial Intelligence (AI) infrastructure faces two compounding crises. Compute payload - the unsustainable energy and capital costs of training and inference - threatens to outpace grid capacity and concentrate capability among a handful of organizations. Data chaos - the 80% of project effort consumed by preparation, conversion,...

💬 0 commentsarXiv:2601.09124v1PDF
0

Posted in cs.CV · 2026-01-14 · Xin Yuan, Meiqi Wan, Wei Liu, Xin Xu, Zheng Wang

Beyond Seen Bounds: Class-Centric Polarization for Single-Domain Generalized Deep Metric Learning

Single-domain generalized deep metric learning (SDG-DML) faces the dual challenge of both category and domain shifts during testing, limiting real-world applications. Therefore, aiming to learn better generalization ability on both unseen categories and domains is a realistic goal for the SDG-DML task. To deliver the aspiration,...

💬 0 commentsarXiv:2601.09121v1PDF
0

Posted in cs.SE · 2026-01-14 · Jiali Cheng, Rui Pan, Hadi Amiri

Investigating Tool-Memory Conflicts in Tool-Augmented LLMs

Tool-augmented large language models (LLMs) have powered many applications. However, they are likely to suffer from knowledge conflict. In this paper, we propose a new type of knowledge conflict -- Tool-Memory Conflict (TMC), where the internal parametric knowledge contradicts with the external tool knowledge for tool-augmented LLMs....

💬 0 commentsarXiv:2601.09760v1PDF
0

Posted in cs.CL · 2026-01-14 · Chen-Wei Liang, Bin Guo, Zhen-Yuan Wei, Mu-Jiang-Shan Wang

Adaptive Multi-Stage Patent Claim Generation with Unified Quality Assessment

Current patent claim generation systems face three fundamental limitations: poor cross-jurisdictional generalization, inadequate semantic relationship modeling between claims and prior art, and unreliable quality assessment. We introduce a novel three-stage framework that addresses these challenges through relationship-aware...

💬 0 commentsarXiv:2601.09120v1PDF
0

Posted in cs.CL · 2026-01-14 · Yongming Sun

Contrastive Bi-Encoder Models for Multi-Label Skill Extraction: Enhancing ESCO Ontology Matching with BERT and Attention Mechanisms

Fine-grained labor market analysis increasingly relies on mapping unstructured job advertisements to standardized skill taxonomies such as ESCO. This mapping is naturally formulated as an Extreme Multi-Label Classification (XMLC) problem, but supervised solutions are constrained by the scarcity and cost of large-scale,...

💬 0 commentsarXiv:2601.09119v1PDF
0

Posted in cs.CV · 2026-01-14 · Jackie Alex, Guoqiang Huan

LPCAN: Lightweight Pyramid Cross-Attention Network for Rail Surface Defect Detection Using RGB-D Data

This paper addresses the limitations of current vision-based rail defect detection methods, including high computational complexity, excessive parameter counts, and suboptimal accuracy. We propose a Lightweight Pyramid Cross-Attention Network (LPCANet) that leverages RGB-D data for efficient and accurate defect identification. The...

💬 0 commentsarXiv:2601.09118v2PDF
0

Posted in cs.CY · 2026-01-14 · Shalmoli Ghosh, Matthew R. DeVerna, Filippo Menczer

A Marketplace for AI-Generated Adult Content and Deepfakes

Generative AI systems increasingly enable the production of highly realistic synthetic media. Civitai, a popular community-driven platform for AI-generated content, operates a monetized feature called Bounties, which allows users to commission the generation of content in exchange for payment. To examine how this mechanism is used and...

💬 0 commentsarXiv:2601.09117v3PDF
0

Posted in cs.CV · 2026-01-14 · Haoyan Gong, Hongbin Liu

LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models

Real-world License Plate Recognition (LPR) faces significant challenges from severe degradations such as motion blur, low resolution, and complex illumination. The prevailing "restoration-then-recognition" two-stage paradigm suffers from a fundamental flaw: the pixel-level optimization objectives of image restoration models are...

💬 0 commentsarXiv:2601.09116v1PDF
0

Posted in cs.CR · 2026-01-14 · Fengchao Chen, Tingmin Wu, Van Nguyen, Surya. Nepal, Carsten Rudolph

Agents at Risk: How Users Unwittingly Undermine LLM Safety

Large language model (LLM)-based agents are increasingly deployed in applications, such as trip-planning agents and web-use agents, to perform complex planning and execution tasks. Prior work has shown that LLM-based agents are vulnerable to context confusion, where external adversarial content incorporated into the agent's reasoning...

💬 0 commentsarXiv:2601.10758v3PDF
0

Posted in cs.DC · 2026-01-14 · Yufan Xia, Marco De La Pierre, Amanda S. Barnard, Giuseppe Maria Junior Barca

A Machine Learning Approach Towards Runtime Optimisation of Matrix Multiplication

The GEneral Matrix Multiplication (GEMM) is one of the essential algorithms in scientific computing. Single-thread GEMM implementations are well-optimised with techniques like blocking and autotuning. However, due to the complexity of modern multi-core shared memory systems, it is challenging to determine the number of threads that...

💬 0 commentsarXiv:2601.09114v1PDF
0

Posted in cs.AI · 2026-01-14 · Zixia Jia, Jiaqi Li, Yipeng Kang, Yuxuan Wang, Tong Wu, Quansen Wang, Xiaobo Wang, Shuyi Zhang, Junzhe Shen, Qing Li, Siyuan Qi, Yitao Liang, Di He, Zilong Zheng, Song-Chun Zhu

The AI Hippocampus: How Far are We From Human Memory?

Memory plays a foundational role in augmenting the reasoning, adaptability, and contextual fidelity of modern Large Language Models and Multi-Modal LLMs. As these models transition from static predictors to interactive systems capable of continual learning and personalized inference, the incorporation of memory mechanisms has emerged...

💬 0 commentsarXiv:2601.09113v1PDF
0

Posted in cs.CY · 2026-01-14 · Ying He, Baiyang Li, Yule Cao, Huirun Xu, Qiuxian Chen, Shu Chen, Shangsheng Ren

Seeking Human Security Consensus: A Unified Value Scale for Generative AI Value Safety

The rapid development of generative AI has brought value- and ethics-related risks to the forefront, making value safety a critical concern while a unified consensus remains lacking. In this work, we propose an internationally inclusive and resilient unified value framework, the GenAI Value Safety Scale (GVS-Scale): Grounded in a...

💬 0 commentsarXiv:2601.09112v1PDF
0

Posted in cs.CV · 2026-01-14 · Yang Li, Aming Wu, Zihao Zhang, Yahong Han

Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive Reasoning

Vision-Language Navigation (VLN) aims to enable agents to navigate to a target location based on language instructions. Traditional VLN often follows a close-set assumption, i.e., training and test data share the same style of the input images and instructions. However, the real world is open and filled with various unseen...

💬 0 commentsarXiv:2601.09111v2PDF
0

Posted in cs.CL · 2026-01-14 · Ziyang Zhou, Ziqi Liu, Yan Wang, Yiming Lin, Yangbin Chen

RAM-SD: Retrieval-Augmented Multi-agent framework for Sarcasm Detection

Sarcasm detection remains a significant challenge due to its reliance on nuanced contextual understanding, world knowledge, and multi-faceted linguistic cues that vary substantially across different sarcastic expressions. Existing approaches, from fine-tuned transformers to large language models, apply a uniform reasoning strategy to...

💬 0 commentsarXiv:2601.17002v1PDF
0

Posted in cs.CV · 2026-01-14 · Kai Hu, Yaozu Feng, Vladimir Lysenko, Ya Guo, Huayi Wu

SAM-Aug: Leveraging SAM Priors for Few-Shot Parcel Segmentation in Satellite Time Series

Few-shot semantic segmentation of time-series remote sensing images remains a critical challenge, particularly in regions where labeled data is scarce or costly to obtain. While state-of-the-art models perform well under full supervision, their performance degrades significantly under limited labeling, limiting their real-world...

💬 0 commentsarXiv:2601.09110v2PDF
0

Posted in cs.CV · 2026-01-14 · Yanguang Sun, Chao Wang, Jian Yang, Lei Luo

Small but Mighty: Dynamic Wavelet Expert-Guided Fine-Tuning of Large-Scale Models for Optical Remote Sensing Object Segmentation

Accurately localizing and segmenting relevant objects from optical remote sensing images (ORSIs) is critical for advancing remote sensing applications. Existing methods are typically built upon moderate-scale pre-trained models and employ diverse optimization strategies to achieve promising performance under full-parameter...

💬 0 commentsarXiv:2601.09108v1PDF
0

Posted in cs.CV · 2026-01-14 · Lachlan Holden, Feras Dayoub, Alberto Candela, David Harvey, Tat-Jun Chin

Vision Foundation Models for Domain Generalisable Cross-View Localisation in Planetary Ground-Aerial Robotic Teams

Accurate localisation in planetary robotics enables the advanced autonomy required to support the increased scale and scope of future missions. The successes of the Ingenuity helicopter and multiple planetary orbiters lay the groundwork for future missions that use ground-aerial robotic teams. In this paper, we consider rovers using...

💬 0 commentsarXiv:2601.09107v1PDF
0

Posted in cs.AI · 2026-01-14 · Wenbin Li, Jingling Wu, Xiaoyong Lin. Jing Chen, Cong Chen

AviationLMM: A Large Multimodal Foundation Model for Civil Aviation

Civil aviation is a cornerstone of global transportation and commerce, and ensuring its safety, efficiency and customer satisfaction is paramount. Yet conventional Artificial Intelligence (AI) solutions in aviation remain siloed and narrow, focusing on isolated tasks or single modalities. They struggle to integrate heterogeneous data...

💬 0 commentsarXiv:2601.09105v2PDF
0

Posted in cs.RO · 2026-01-14 · Ko Yamamoto, Kyosuke Ishibashi, Hiroki Ishikawa, Osamu Azami

Design Methodology of Hydraulically-driven Soft Robotic Gripper for a Large and Heavy Object

This paper presents a design methodology of a hydraulically-driven soft robotic gripper for grasping a large and heavy object -- approximately 10 - 20 kg with 20 - 30 cm diameter. Most existing soft grippers are pneumatically actuated with several hundred kPa pressure, and cannot generate output force sufficient for such a large and...

💬 0 commentsarXiv:2601.09104v1PDF