Expertini Research Research

Browse Research Papers

378,930+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation
Showing 378930 results for "program evaluation"
AI & Data Science Preprint PDF DOI

Benchmarking the Safety of Large Language Models for Robotic Health Attendant Control

Mahiro Nakao, Kazuhiro Takemoto ยท 2026

Large language models (LLMs) are increasingly considered for deployment as the control component of robotic health attendants, yet their safety in this context remains poorly characterized. We introduโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Cosmological evolution of fast radio bursts and its rapid decline relative to star formation rate

X. D. Jia, D. H. Gao, J. H. Chen, Q. Wu, S. X. Yi, F. Y. Wang (NJU) ยท 2026

Fast radio bursts (FRBs) are enigmatic millisecond-duration radio transients whose physical origins remain debated. To shed light on this, we analyze the CHIME/FRB Catalog 2. By using the probability โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Advancing multi-site emission control: A physics-informed transfer learning framework with mixture of experts for carbon-pollutant synergy

Yuxuan Ying, Hanqing Yang, Kaige Wang, Yu Hu, Zhiming Zheng, Yunliang Jiang, Xiaoqing Lin, Xiaodong Li, Jun Chen ยท 2026

Municipal solid waste incineration is increasingly central to urban waste management, yet its sustainability benefit depends on controlling carbon emissions and multiple air pollutants under highly heโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision

Xiaoya Cheng, Rouwan Wu, Xinyi Liu, Zeyu Cui, Yan Liu, Na Zhao, Yu Liu, Maojun Zhang, Shen Yan ยท 2026

Despite the rapid progress in data-driven 3D vision, aerial geometric 3D vision remains a formidable challenge due to the severe scarcity of large-scale, high-fidelity training data. Existing benchmarโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

DenseStep2M: A Scalable, Training-Free Pipeline for Dense Instructional Video Annotation

Mingji Ge, Qirui Chen, Zeqian Li, Weidi Xie ยท 2026

Long-term video understanding requires interpreting complex temporal events and reasoning over procedural activities. While instructional video corpora, like HowTo100M, offer rich resources for model โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Preserving Disagreement: Architectural Heterogeneity and Coherence Validation in Multi-Agent Policy Simulation

Ariel Sela ยท 2026

Multi-agent deliberation systems using large language models (LLMs) are increasingly proposed for policy simulation, yet they suffer from artificial consensus: evaluator agents converge on the same opโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference

Bodon Jeong, Hongsu Byun, Youngjae Kim, Weikuan Yu, Kyungkeun Lee, Jihoon Yang, Sungyong Park ยท 2026

The increasing deployment of Large Language Model (LLM) inference on edge AI systems demands efficient execution under tight memory budgets. A key challenge arises from Key-Value (KV) caches, which ofโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Fidelity, Diversity, and Privacy: A Multi-Dimensional LLM Evaluation for Clinical Data Augmentation

Guillermo Iglesias, Gema Bello-Orgaz, Maria Navas-Loro, Cristian Ramirez-Atencia, Merce Salvador Robert, Enrique Baca-Garcia ยท 2026

The scarcity of high-quality annotated medical data, particularly in mental health, poses a significant bottleneck for training robust machine learning models. Privacy regulations restrict data sharinโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Fractional Cosmic String Loops In Expanding Universe

Pankaj Chaturvedi, Bikram Nath ยท 2026

We study the dynamics of circular cosmic string loops in a spatially flat Friedmann Lema\^itre Robertson Walker universe within a fractional Polyakov framework that incorporates nonlocal memory effectโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Verification and Validation (V&V)-in-the-Loop for RISC-V Design: The Holistic Vision of BZL

Sajjad Ahmed, Alexander Kropotov, Roberto Ignacio Genovese, Bernat Homs, Eloi Merino, Francesco Urbani, Henrique Yano, Ivan Diaz, Joan Gracia Fernandez, Matteo Toselli, Muhammad Imran, Muhammad Abu Bakar Umar Haider Iqbal, Nadeem Yaseen, Quswar Abid, Shaista Cheema, Samuel Sanchez, Daniel Garcia, Joan Cabre, Mostafa Elyasi, Fernando Ayats, Miquel Moreto, Teresa Cervero, Oscar Palomar, Behzad Salami ยท 2026

The Barcelona Zetascale Lab (BZL) project aims to strengthening Europe's capacity in the design and manufacture of RISC-V based high-performance computing chips. In this context, we present a holisticโ€ฆ

Read Paper โ†’
Mathematics Preprint PDF DOI

A stellated tetrahedron that is probably not Rupert

Tony Zeng ยท 2026

A convex polyhedron is Rupert if a hole can be cut into it (making its genus $1$) such that an identical copy of the polyhedron can pass through the hole. Resolving a conjecture of Jerrard-Wetzel-Yuanโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Fundamental Physics, Existential Risks and Human Futures

Adrian Kent (Centre for Quantum Information, Foundations, DAMTP, University of Cambridge, Perimeter Institute for Theoretical Physics) ยท 2026

Over the past 25 years, I have been involved in some intriguing developments in the foundations of physics, exploring the quantum reality problem, the relationship between quantum theory and gravity aโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

AGEL-Comp: A Neuro-Symbolic Framework for Compositional Generalization in Interactive Agents

Mahnoor Shahid, Hannes Rothe ยท 2026

Large Language Model (LLM)-based agents exhibit systemic failures in compositional generalization, limiting their robustness in interactive environments. This work introduces AGEL-Comp, a neuro-symbolโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

GIFGuard: Proactive Forensics against Deepfakes in Facial GIFs via Spatiotemporal Watermarking

Shupeng Che, Zhiqing Guo, Changtao Miao, Dan Ma, Gaobo Yang ยท 2026

The rapid evolution of deepfake technology poses an unprecedented threat to the authenticity of Graphics Interchange Format (GIF) imagery, which serves as a representative of short-loop temporal mediaโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

GMT: A Geometric Multigrid Transformer Solver for Microstructure Homogenization

Yu Xing, Yang Liu, Tianyang Xue, Lin Lu ยท 2026

Lattice metamaterials enable lightweight, multifunctional structures, yet homogenization-based evaluation of their effective properties remains computationally expensive. Neural surrogates offer speedโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Automatic Causal Fairness Analysis with LLM-Generated Reporting

Alessia Berarducci, Eric Rossetto, Alessandro Antonucci, Marco Zaffalon ยท 2026

AutoML, intended as the process of automating the application of machine learning to real-world problems, is a key step for AI popularisation. Most AutoML frameworks are not accounting for the potentiโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Microsecond-resolved electro-optic dual-comb spectroscopy in the 10~12.5 $\mu$m fingerprint region for radical kinetics

Pei-Ling Luo, I-Yun Chen ยท 2026

Dual-comb spectroscopy enables broadband, high-resolution measurements with microsecond temporal resolution, but extending this capability to the 10~12.5 $\mu$m molecular fingerprint region remains teโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

Matteo Leonesi, Francesco Belardinelli, Flavio Corradini, Marco Piangerelli ยท 2026

Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences once monitoring is lifted. Current detection methodโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

3D Generation for Embodied AI and Robotic Simulation: A Survey

Tianwei Ye, Yifan Mao, Minwen Liao, Jian Liu, Chunchao Guo, Dazhao Du, Quanxin Shou, Fangqi Zhu, Song Guo ยท 2026

Embodied AI and robotic systems increasingly depend on scalable, diverse, and physically grounded 3D content for simulation-based training and real-world deployment. While 3D generative modeling has aโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Prompts

Yuan Xin, Yixuan Weng, Minjun Zhu, Ying Ling, Chengwei Qin, Michael Hahn, Michael Backes, Yue Zhang, Linyi Yang ยท 2026

As Large Language Models (LLMs) are increasingly integrated into academic peer review, their vulnerability to adversarial prompts -- adversarial instructions embedded in submissions to manipulate outcโ€ฆ

Read Paper โ†’
โ† Prev Page 18 of 18947 Next โ†’