Expertini Research Research

Browse Research Papers

346,661+ open-access research outputs.

โœ• Clear
๐Ÿ” avoidance learning
Showing 346661 results for "avoidance learning"
Chemistry Preprint PDF DOI

Transferable excited-state dynamics enable screening of fluorescent protein chromophores

Rhyan Barrett, Sophia Wesely, Julia Westermayr ยท 2026

Transferable excited-state dynamics offer a route to efficient screening of photophysical behavior across molecular systems, but conventional nonadiabatic simulations remain prohibitively expensive. Hโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Risk-Calibrated Learning: Minimizing Fatal Errors in Medical AI

Abolfazl Mohammadi-Seif, Ricardo Baeza-Yates ยท 2026

Deep learning models often achieve expert-level accuracy in medical image classification but suffer from a critical flaw: semantic incoherence. These high-confidence mistakes that are semantically incโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

BID-LoRA: A Parameter-Efficient Framework for Continual Learning and Unlearning

Jagadeesh Rachapudi, Ritali Vatsi, Praful Hambarde, Amit Shukla ยท 2026

Recent advances in deep learning underscore the need for systems that can not only acquire new knowledge through Continual Learning (CL) but also remove outdated, sensitive, or private information thrโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Intelligent resource prediction for SAP HANA continuous integration build workloads

Torsten Mandel, Jonathan Bader, Hanyoung Yoo, Stephan Kraft ยท 2026

Large enterprises often operate extensive Continuous Integration (CI) pipelines on large, heterogeneous compute clusters, where conservative, statically defined resource requirements are used to ensurโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

A hierarchical spatial-aware algorithm with efficient reinforcement learning for human-robot task planning and allocation in production

Jintao Xue, Xiao Li, Nianmin Zhang ยท 2026

In advanced manufacturing systems, humans and robots collaborate to conduct the production process. Effective task planning and allocation (TPA) is crucial for achieving high production efficiency, yeโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Safe reinforcement learning with online filtering for fatigue-predictive human-robot task planning and allocation in production

Jintao Xue, Xiao Li, Nianmin Zhang ยท 2026

Human-robot collaborative manufacturing, a core aspect of Industry 5.0, emphasizes ergonomics to enhance worker well-being. This paper addresses the dynamic human-robot task planning and allocation (Hโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

Chuang Peng, Wei Zhang, Renshuai Tao, Xinhao Zhang, Jian Yang ยท 2026

Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy and heterogeneous nature of real-world HTML. Standโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

Rui Wang, Yi Zheng, Dongxin Wang, Haiping Huang, Yuanzhi Yao, Yuxiang Zhou, Jialin Yu, Philip Torr ยท 2026

Existing topic modeling methods, from LDA to recent neural and LLM-based approaches, which focus mainly on statistical coherence, often produce redundant or off-target topics that miss the user's undeโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Multi-Agent Digital Twins for Strategic Decision-Making using Active Inference

Francesco Maria Mancinelli, Matteo Torzoni, Domenico Maisto, Francesco Donnarumma, Alberto Corigliano, Giovanni Pezzulo, Andrea Manzoni ยท 2026

Active Inference is an emerging framework providing a quantitative account of behavioral processes in neuroscience and a principled approach to decision-making under uncertainty. Its application to agโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Robust Semi-Supervised Temporal Intrusion Detection for Adversarial Cloud Networks

Anasuya Chattopadhyay, Daniel Reti, Hans D. Schotten ยท 2026

Cloud networks increasingly rely on machine learning based Network Intrusion Detection Systems to defend against evolving cyber threats. However, real-world deployments are challenged by limited labelโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

PromptEcho: Annotation-Free Reward from Vision-Language Models for Text-to-Image Reinforcement Learning

Jinlong Liu, Wanggui He, Peng Zhang, Mushui Liu, Hao Jiang, Pipei Huang ยท 2026

Reinforcement learning (RL) can improve the prompt following capability of text-to-image (T2I) models, yet obtaining high-quality reward signals remains challenging: CLIP Score is too coarse-grained, โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Learning Chain Of Thoughts Prompts for Predicting Entities, Relations, and even Literals on Knowledge Graphs

Alkid Baci, Luke Friedrichs, Caglar Demir, N'Dah Jean Kouagou, Axel-Cyrille Ngonga Ngomo ยท 2026

Knowledge graph embedding (KGE) models perform well on link prediction but struggle with unseen entities, relations, and especially literals, limiting their use in dynamic, heterogeneous graphs. In coโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

TimeSAF: Towards LLM-Guided Semantic Asynchronous Fusion for Time Series Forecasting

Fan Zhang, Shiming Fan, Hua Wang ยท 2026

Despite the recent success of large language models (LLMs) in time-series forecasting, most existing methods still adopt a Deep Synchronous Fusion strategy, where dense interactions between textual anโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Contextual Multi-Task Reinforcement Learning for Autonomous Reef Monitoring

Melvin Laux, Yi-Ling Liu, Rina Alo, Soren Topper, Mariela De Lucas Alvarez, Frank Kirchner, Rebecca Adam ยท 2026

Although autonomous underwater vehicles promise the capability of marine ecosystem monitoring, their deployment is fundamentally limited by the difficulty of controlling vehicles under highly uncertaiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

LLMs Are Not a Silver Bullet: A Case Study on Software Fairness

Xinyue Li, Sixuan Li, Ying Xiao, Jie M. Zhang, Zhou Yang, Xuanzhe Liu, Zhenpeng Chen ยท 2026

Fairness is a critical requirement for human-related, high-stakes software systems, motivating extensive research on bias mitigation. Prior work has largely focused on tabular data settings using tradโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Calibration-Aware Policy Optimization for Reasoning LLMs

Ziqi Wang, Xingzhou Lou, Meiqi Wu, Zhengqi Wen, Junge Zhang ยท 2026

Group Relative Policy Optimization (GRPO) enhances LLM reasoning but often induces overconfidence, where incorrect responses yield lower perplexity than correct ones, degrading relative calibration asโ€ฆ

Read Paper โ†’
Mathematics Preprint PDF DOI

A Comparison of Reinforcement Learning and Optimal Control Methods for Path Planning

Qiang Le, Yaguang Yang, Isaac E. Weintraub ยท 2026

Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find ideal paths, the computational time is often too slow โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance

Linhao Yu, Tianmeng Yang, Siyu Ding, Renren Jin, Naibin Gu, Xiangzhao Hao, Shuaiyi Nie, Deyi Xiong, Weichong Yin, Yu Sun, Hua Wu ยท 2026

RLVR improves reasoning in large language models, but its effectiveness is often limited by severe reward sparsity on hard problems. Recent hint-based RL methods mitigate sparsity by injecting partialโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Joint Activity Detection and Channel Estimation for Massive Random Access Using SBL and SCA

Esa Ollila, Majdoddin Esfandiari, Daniel P. Palomar ยท 2026

In massive machine-type communication (mMTC) applications, a key challenge is joint device activity detection and channel estimation (JADCE) under grant-free random access, as a massive number of deviโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models

You Qin, Linqing Wang, Hao Fei, Roger Zimmermann, Liefeng Bo, Qinglin Lu, Chunyu Wang ยท 2026

The post-training pipeline for diffusion models currently has two stages: supervised fine-tuning (SFT) on curated data and reinforcement learning (RL) with reward models. A fundamental gap separates tโ€ฆ

Read Paper โ†’
โ† Prev Page 133 of 17334 Next โ†’