Expertini Research Research

Browse Research Papers

346,661+ open-access research outputs.

โœ• Clear
๐Ÿ” avoidance learning
Showing 346661 results for "avoidance learning"
AI & Data Science Preprint PDF DOI

Positive-Only Drifting Policy Optimization

Qi Zhang ยท 2026

In the field of online reinforcement learning (RL), traditional Gaussian policies and flow-based methods are often constrained by their unimodal expressiveness, complex gradient clipping, or stringentโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

TIP: Token Importance in On-Policy Distillation

Yuanda Xu, Hejian Sang, Zhengze Zhou, Ran He, Zhipeng Wang, Alborz Geramifard ยท 2026

On-policy knowledge distillation (OPD) trains a student on its own rollouts under token-level supervision from a teacher. Not all token positions matter equally, but existing views of token importanceโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Learning ultra-compressible hyperelasticity with splines: Constitutive asymmetries and non-unique representations

Miguel Angel Moreno-Mateos, Simon Wiesheier, Paul Steinmann, Ellen Kuhl ยท 2026

Highly compressible solids, such as foams, exhibit complex responses, including pronounced tension-compression asymmetry. Capturing such behaviors within unified hyperelastic frameworks remains challeโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Distributional Inverse Homogenization

Arnaud Vadeboncoeur, Mark Girolami, Kaushik Bhattacharya, Andrew M. Stuart ยท 2026

For many materials, macroscopic mechanical behavior is determined by an intricate microstructure. Understanding the relation between these two scales helps scientists and engineers design better materโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Generative design of inorganic materials

Jose Recatala-Gomez, Haiwen Dai, Zhu Ruiming, Nikita Kazeev, Nong Wei, Gang Wu, Maciej Koperski, Tan Teck Leong, Andrey Ustyuzhanin, Gerbrand Ceder, Kostya Novoselov, Kedar Hippalgaonkar ยท 2026

Materials discovery is fundamental to advance next-generation technologies as well as for sustainable and circular economy. Beyond computational screening, generative models are efficient at finding mโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Natural Language Embeddings of Synthesis and Testing conditions Enhance Glass Dissolution Prediction

Sajid Mannan, K. Sidharth Nambudiripad, Indrajeet Mandal, Nitya Nand Gosvami, N. M. Anoop Krishnan ยท 2026

Long-term chemical durability of glass, crucial for immobilizing nuclear waste, is governed by glass properties such as composition, surface geometry, as well as external factors like thermodynamic coโ€ฆ

Read Paper โ†’
Biology & Life Sciences Preprint PDF DOI

A deep learning framework for glomeruli segmentation with boundary attention

Behnaz Elhaminia, Catherine King, Jiaqi Lv, Lorraine Harper, Paul Moss, Owen Cain, Dimitrios Chanouzas, Shan E Ahmed Raza ยท 2026

Accurate detection and segmentation of glomeruli in kidney tissue are essential for diagnostic applications. Traditional deep learning methods primarily rely on semantic segmentation, which often failโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Neural architectures for resolving references in program code

Gergo Szalay, Gergely Zsolt Kovacs, Sandor Teleki, Balazs Pinter, Tibor Gregorics ยท 2026

Resolving and rewriting references is fundamental in programming languages. Motivated by a real-world decompilation task, we abstract reference rewriting into the problems of direct and indirect indexโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Deep Neural Network-guided PSO for Tracking a Global Optimal Position in Complex Dynamic Environment

Stephen Raharja, Toshiharu Sugawara ยท 2026

We propose novel particle swarm optimization (PSO) variants incorporated with deep neural networks (DNNs) for particles to pursue globally optimal positions in dynamic environments. PSO is a heuristicโ€ฆ

Read Paper โ†’
Economics & Finance Preprint PDF DOI

A Comparative Study of Dynamic Programming and Reinforcement Learning in Finite Horizon Dynamic Pricing

Lev Razumovskiy, Nikolay Karenin ยท 2026

This paper provides a systematic comparison between Fitted Dynamic Programming (DP), where demand is estimated from data, and Reinforcement Learning (RL) methods in finite-horizon dynamic pricing probโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

$\pi$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data

Yaocheng Zhang, Yuanheng Zhu, Wenyue Chong, Songjun Tu, Qichao Zhang, Jiajun Chai, Xiaohan Wang, Wei Lin, Guojun Yin, Dongbin Zhao ยท 2026

Deep search agents have emerged as a promising paradigm for addressing complex information-seeking tasks, but their training remains challenging due to sparse rewards, weak credit assignment, and limiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Enhancing Local Life Service Recommendation with Agentic Reasoning in Large Language Model

Shiteng Cao, Xiaochong Lan, Yuwei Du, Jie Feng, Yinxing Liu, Xinlei Shi, Yong Li ยท 2026

Local life service recommendation is distinct from general recommendation scenarios due to its strong living need-driven nature. Fundamentally, accurately identifying a user's immediate living need anโ€ฆ

Read Paper โ†’
Economics & Finance Preprint PDF DOI

How do you know you won't like it if you've (never) tried it? Preference discovery and data design

Sebastiano Della Lena, Alessio Muscillo, Paolo Pin ยท 2026

Consumers discover their preferences through experience, yet the sequence and composition of those experiences are often designed by firms, digital platforms, or policymakers. We introduce a ``data-deโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

On-Orbit Space AI: Federated, Multi-Agent, and Collaborative Algorithms for Satellite Constellations

Ziyang Wang ยท 2026

Satellite constellations are transforming space systems from isolated spacecraft into networked, software-defined platforms capable of on-orbit perception, decision making, and adaptation. Yet much ofโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Hierarchical Reinforcement Learning with Runtime Safety Shielding for Power Grid Operation

Gitesh Malik ยท 2026

Reinforcement learning has shown promise for automating power-grid operation tasks such as topology control and congestion management. However, its deployment in real-world power systems remains limitโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Dual-Enhancement Product Bundling: Bridging Interactive Graph and Large Language Model

Zhe Huang, Peng Wang, Yan Zheng, Sen Song, Longjun Cai ยท 2026

Product bundling boosts e-commerce revenue by recommending complementary item combinations. However, existing methods face two critical challenges: (1) collaborative filtering approaches struggle withโ€ฆ

Read Paper โ†’
Biology & Life Sciences Preprint PDF DOI

Continual Learning for fMRI-Based Brain Disorder Diagnosis via Functional Connectivity Matrices Generative Replay

Qianyu Chen, Shujian Yu ยท 2026

Functional magnetic resonance imaging (fMRI) is widely used for studying and diagnosing brain disorders, with functional connectivity (FC) matrices providing powerful representations of large-scale neโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

InfoChess: A Game of Adversarial Inference and a Laboratory for Quantifiable Information Control

Kieran A. Murphy ยท 2026

We propose InfoChess, a symmetric adversarial game that elevates competitive information acquisition to the primary objective. There is no piece capture, removing material incentives that would otherwโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective

Weijie Wang, Qihang Cao, Sensen Gao, Donny Y. Chen, Haofei Xu, Wenjing Bian, Songyou Peng, Tat-Jen Cham, Chuanxia Zheng, Andreas Geiger, Jianfei Cai, Jia-Wang Bian, Bohan Zhuang ยท 2026

Reconstructing 3D representations from 2D inputs is a fundamental task in computer vision and graphics, serving as a cornerstone for understanding and interacting with the physical world. While traditโ€ฆ

Read Paper โ†’
Mathematics Preprint PDF DOI

Stochastic Trust-Region Methods for Over-parameterized Models

Aike Yang, Hao Wang ยท 2026

Under interpolation-type assumptions such as the strong growth condition, stochastic optimization methods can attain convergence rates comparable to full-batch methods, but their performance, particulโ€ฆ

Read Paper โ†’
โ† Prev Page 120 of 17334 Next โ†’