Expertini Research Research

Browse Research Papers

346,661+ open-access research outputs.

โœ• Clear
๐Ÿ” avoidance learning
Showing 346661 results for "avoidance learning"
Engineering Preprint PDF DOI

Discovery of unobservable parameters via physical embedding

Le Cheng, Xiaoran Liu, Lingjin Kong, Haitao Zhao, Jun Xiong, Fanglin Gu, Xiaoying Zhang, Baoquan Ren, Jibo Wei, Hao Yin ยท 2026

Recovering a source signal from indirect measurements often requires estimating latent parameters, such as wireless channel states or MRI coil sensitivities, that cannot be directly observed. Here, weโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Flexible Empowerment at Reasoning with Extended Best-of-N Sampling

Taisuke Kobayashi ยท 2026

This paper proposes a novel method that incorporates empowerment when reasoning actions in reinforcement learning (RL), thereby achieving the flexibility of exploration-exploitation dilemma (EED). In โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models

Yunbei Zhang, Shuaicheng Niu, Chengyi Cai, Feng Liu, Jihun Hamm ยท 2026

Test-Time Adaptation (TTA) for black-box models accessible only via APIs remains a largely unexplored challenge. Existing approaches such as post-hoc output refinement offer limited adaptive capacity,โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Privacy, Prediction, and Allocation

Ben Jacobsen, Nitin Kohli ยท 2026

Algorithmic predictions are increasingly used to inform the allocation of scarce resources. The promise of these methods is that, through machine learning, they can better identify the people who woulโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

BioHiCL: Hierarchical Multi-Label Contrastive Learning for Biomedical Retrieval with MeSH Labels

Mengfei Lan, Lecheng Zheng, Halil Kilicoglu ยท 2026

Effective biomedical information retrieval requires modeling domain semantics and hierarchical relationships among biomedical texts. Existing biomedical generative retrievers build on coarse binary reโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

CSLE: A Reinforcement Learning Platform for Autonomous Security Management

Kim Hammar ยท 2026

Reinforcement learning is a promising approach to autonomous and adaptive security management in networked systems. However, current reinforcement learning solutions for security management are mostlyโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Classifying Supermassive Black Hole Growth Regimes to Observables Across Cosmological Simulations with Forecasts for LSST

Hitaishi Chillara ยท 2026

The possibility of over-massive black holes suggested by James Webb Space Telescope photometric discoveries of 'little red dots', may disfavor light supermassive black hole (SMBH) seeds. However, whatโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

"Excuse me, may I say something..." CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

Yang Wu, Jinhong Yu, Jingwei Xiong, Zhimin Tao, Xiaozhong Liu ยท 2026

The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However, the reactive nature of LLMs, which respond only wโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Learning Behaviorally Grounded Item Embeddings via Personalized Temporal Contexts

Rafael T. Sereicikas, Pedro R. Pires, Gregorio F. Azevedo, Tiago A. Almeida ยท 2026

Effective user modeling requires distinguishing between short-term and long-term preference evolution. While item embeddings have become a key component of recommender systems, standard approaches likโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Reward Weighted Classifier-Free Guidance as Policy Improvement in Autoregressive Models

Alexander Peysakhovich, William Berman ยท 2026

Consider an auto-regressive model that produces outputs x (e.g., answers to questions, molecules) each of which can be summarized by an attribute vector y (e.g., helpfulness vs. harmlessness, or bio-aโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Why Fine-Tuning Encourages Hallucinations and How to Fix It

Guy Kaplan, Zorik Gekhman, Zhen Zhu, Lotem Rozner, Yuval Reif, Swabha Swayamdipta, Derek Hoiem, Roy Schwartz ยท 2026

Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information through supervised fine-tuning (SFT), which can incโ€ฆ

Read Paper โ†’
Mathematics Preprint PDF DOI

A bi-level priority sorting framework for flexible AGV service scheduling in smart warehouses

Xiaozhu Sun, Bilal Farooq ยท 2026

This paper proposes a bi-level optimization framework to coordinate Automated Guided Vehicle (AGV) flexible operations in smart independent warehouses, addressing the critical challenge of balancing hโ€ฆ

Read Paper โ†’
Earth & Environmental Sciences Preprint PDF DOI

ExoNet: Calibrated Multimodal Deep Learning for TESS Exoplanet Candidate Vetting using Phase-Folded Light Curves, Stellar Parameters, and Multi-Head Attention

Md.Rashadul Islam ยท 2026

The discovery of exoplanets at scale has become one of the defining data science challenges in modern astrophysics. NASA's Transiting Exoplanet Survey Satellite (TESS) had catalogued over 7,800 planetโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

Jacob Dang, Brian Y. Xie, Omar G. Younis ยท 2026

Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those traits. However, it remains unclear whether behavโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Learning Affine-Equivariant Proximal Operators

Oriel Savir, Zhenghan Fang, Jeremias Sulam ยท 2026

Proximal operators are fundamental across many applications in signal processing and machine learning, including solving ill-posed inverse problems. Recent work has introduced Learned Proximal Networkโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Natural gradient descent with momentum

Anthony Nouy, Agustin Somacal ยท 2026

We consider the problem of approximating a function by an element of a nonlinear manifold which admits a differentiable parametrization, typical examples being neural networks with differentiable actiโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Feature-level analysis and adversarial transfer in rotationally equivariant quantum machine learning

Maureen Krumtunger, Martin Sevior, Muhammad Usman ยท 2026

Group-equivariant quantum models are designed to exploit symmetry and can improve trainability, but it remains unclear how symmetry constraints shape their adversarial robustness. We study this questiโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Optimizing Stochastic Gradient Push under Broadcast Communications

Tuan Nguyen, Ting He ยท 2026

We consider the problem of minimizing the convergence time for decentralized federated learning (DFL) in wireless networks under broadcast communications, with focus on mixing matrix design. The mixinโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

UA-Net: Uncertainty-Aware Network for TRISO Image Semantic Segmentation

Kyle Lucke, Zuzanna Krajewska-Travar, Shoukun Sun, Lu Cai, John D. Stempien, Min Xian ยท 2026

Tristructural isotropic (TRISO)-coated particle fuels undergo dimensional changes and chemical reactions during high-temperature neutron irradiation. Post-irradiation materialography helps understand โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Verification Modulo Tested Library Contracts

Abhishek Uppar, Omar Muhammad, Sumanth Prabhu, Deepak D'Souza, Madhusudan P, Adithya Murali ยท 2026

We consider the problem of \emph{verification modulo tested library contracts} as a step towards automating the verification of client programs that use complex libraries. We formulate this problem โ€ฆ

Read Paper โ†’
โ† Prev Page 108 of 17334 Next โ†’