Expertini Research Research

Browse Research Papers

346,661+ open-access research outputs.

โœ• Clear
๐Ÿ” avoidance learning
Showing 346661 results for "avoidance learning"
AI & Data Science Preprint PDF DOI

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

Shouzheng Huang, Meishan Zhang, Baotian Hu, Min Zhang ยท 2026

Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evolving tool repositories, existing methods relyinโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Zero-shot Evaluation of Deep Learning for Java Code Clone Detection

Thomas S. Heinze ยท 2026

Deep Learning (DL) is becoming more and more widespread in clone detection, motivated by achieving near-perfect performance for this task. In particular in case of semantic code clones, which share onโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Soft $Q(\lambda)$: A multi-step off-policy method for entropy regularised reinforcement learning using eligibility traces

Pranav Mahajan, Ben Seymour ยท 2026

Soft Q-learning has emerged as a versatile model-free method for entropy-regularised reinforcement learning, optimising for returns augmented with a penalty on the divergence from a reference policy. โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

From Anchors to Supervision: Memory-Graph Guided Corpus-Free Unlearning for Large Language Models

Wenxuan Li, Zhenfei Zhang, Mi Zhang, Geng Hong, Mi Wen, Xiaoyu You, Min Yang ยท 2026

Large language models (LLMs) may memorize sensitive or copyrighted content, raising significant privacy and legal concerns. While machine unlearning has emerged as a potential remedy, prevailing paradโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

OffloadFS: Leveraging Disaggregated Storage for Computation Offloading

Sungho Moon, Daegyu Han, Hera Koo, Sangeun Chae, Duck-Ho Bae, Euiseong Seo, Beomseok Nam ยท 2026

Disaggregated storage systems improve resource utilization and enable independent scaling of storage and compute resources by separating storage resources from computing resources in data centers. NVMโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Online learning with noisy side observations

Tomas Kocak, Gergely Neu, Michal Valko ยท 2026

We propose a new partial-observability model for online learning problems where the learner, besides its own loss, also observes some noisy feedback about the other actions, depending on the underlyinโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Jump-Start Reinforcement Learning with Vision-Language-Action Regularization

Angelo Moroncelli, Roberto Zanetti, Marco Maccarini, Loris Roveda ยท 2026

Reinforcement learning (RL) enables high-frequency, closed-loop control for robotic manipulation, but scaling to long-horizon tasks with sparse or imperfect rewards remains difficult due to inefficienโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Doc-V*:Coarse-to-Fine Interactive Visual Reasoning for Multi-Page Document VQA

Yuanlei Zheng, Pei Fu, Hang Li, Ziyang Wang, Yuyi Zhang, Wenyu Ruan, Xiaojin Zhang, Zhongyu Wei, Zhenbo Luo, Jian Luan, Wei Chen, Xiang Bai ยท 2026

Multi-page Document Visual Question Answering requires reasoning over semantics, layouts, and visual elements in long, visually dense documents. Existing OCR-free methods face a trade-off between capaโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

ReConText3D: Replay-based Continual Text-to-3D Generation

Muhammad Ahmed Ullah Khan, Muhammad Haris Bin Amir, Didier Stricker, Muhammad Zeshan Afzal ยท 2026

Continual learning enables models to acquire new knowledge over time while retaining previously learned capabilities. However, its application to text-to-3D generation remains unexplored. We present Rโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

Yanfeng Shi, Pengfei Cai, Jun Liu, Qing Gu, Nan Jiang, Lirong Dai, Ian McLoughlin, Yan Song ยท 2026

Large Audio-Language Models (LALMs) enable general audio understanding and demonstrate remarkable performance across various audio tasks. However, these models still face challenges in temporal percepโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

Sinan Kurtyigit, Sabine Schulte im Walde, Alexander Fraser ยท 2026

Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memorization. To address this, we analyze generalizaโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration

Sayan Kumar Chaki, Antoine Gourru, Julien Velcin ยท 2026

Fairness in language models is typically studied as a property of a single, centrally optimized model. As large language models become increasingly agentic, we propose that fairness emerges through inโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

MIND: AI Co-Scientist for Material Research

Geonhee Ahn, Donghyun Lee, Hayoung Doo, Jonggeol Na, Hyunsoo Cho, Sookyung Kim ยท 2026

Large language models (LLMs) have enabled agentic AI systems for scientific discovery, but most approaches remain limited to textbased reasoning without automated experimental verification. We proposeโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

Pirzada Suhail, Aditya Anand, Amit Sethi ยท 2026

Reliable and interpretable decision-making is essential in medical imaging, where diagnostic outcomes directly influence patient care. Despite advances in deep learning, most medical AI systems operatโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Beyond Voxel 3D Editing: Learning from 3D Masks and Self-Constructed Data

Yizhao Xu, Hongyuan Zhu, Caiyun Liu, Tianfu Wang, Keyu Chen, Sicheng Xu, Jiaolong Yang, Nicholas Jing Yuan, Qi Zhang ยท 2026

3D editing refers to the ability to apply local or global modifications to 3D assets. Effective 3D editing requires maintaining semantic consistency by performing localized changes according to promptโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

VIGILant: an automatic classification pipeline for glitches in the Virgo detector

Tiago Fernandes, Francesco Di Renzo, Antonio Onofre, Alejandro Torres-Forne, Jose A. Font ยท 2026

Glitches frequently contaminate data in gravitational-wave detectors, complicating the observation and analysis of astrophysical signals. This work introduces VIGILant, an automatic pipeline for classโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

EMGFlow: Robust and Efficient Surface Electromyography Synthesis via Flow Matching

Boxuan Jiang, Chenyun Dai, Can Han ยท 2026

Deep learning-based surface electromyography (sEMG) gesture recognition is frequently bottlenecked by data scarcity and limited subject diversity. While synthetic data generation via Generative Adversโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Scalable Design for RIS-Assisted Multi-User Downlink System Empowered by RSMA under Partial CSI

Yifan Fang, Bile Peng, Yingyang Chen, Qiang Li, Marwa Chafii, Eduard A. Jorswieck ยท 2026

In large-scale reconfigurable intelligent surface (RIS) communication systems, the precise acquisition of channel state information (CSI) is challenging. Consider a practical RIS configuration where oโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Empirical Prediction of Pedestrian Comfort in Mobile Robot Pedestrian Encounters

Alireza Jafari, Hong-Son Nguyen, Yen-Chen Liu ยท 2026

Mobile robots joining public spaces like sidewalks must care for pedestrian comfort. Many studies consider pedestrians' objective safety, for example, by developing collision avoidance algorithms, butโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Behavioral Systems Theory Meets Machine Learning: Control-Aware Learning of the Intrinsic Behavior from Big Data

Yitao Yan, Yu Tong, Jie Bao, Wei Wang ยท 2026

The abundance of process operating data in modern industries, along with the rapid advancement of learning techniques, has led to a paradigm shift towards data-centric analysis and control. However, iโ€ฆ

Read Paper โ†’
โ† Prev Page 123 of 17334 Next โ†’