Expertini Research Research

Browse Research Papers

378,930+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation
Showing 378930 results for "program evaluation"
Computer Science Preprint PDF DOI

Exploring Sparse Matrix Multiplication Kernels on the Cerebras CS-3

Milan Shah, Sheng Di, Michela Becchi ยท 2026

In recent years, novel AI accelerators have emerged as promising alternatives to GPU for AI model training and inference tasks. One such accelerator, the Cerebras CS-3, achieves strong performance on โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery

Hanane Nour Moussa, Yifei Li, Zhuoyang Li, Yankai Yang, Cheng Tang, Tianshu Zhang, Nesreen K. Ahmed, Ali Payani, Ziru Chen, Huan Sun ยท 2026

Despite recent progress in language models and agents for scientific data-driven discovery, further advancing their capabilities is held back by the absence of verifiable environments representing reaโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Near--extremal gravitational collapse in 4+1 dimensions: Schwarzschild--de--Sitter space

Maciej Dunajski, Sebastian J. Szybka ยท 2026

We numerically study a formation of near extremal horizons from a gravitational collapse of radially symmetric gravitational waves in $4+1$ dimensions within the framework of pure Einstein gravity witโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

FineState-Bench: Benchmarking State-Conditioned Grounding for Fine-grained GUI State Setting

Fengxian Ji, Jingpu Yang, Zirui Song, Yuanxi Wang, Zhexuan Cui, Yuke Li, Qian Jiang, Xiuying Chen ยท 2026

Despite the rapid progress of large vision-language models (LVLMs), fine-grained, state-conditioned GUI interaction remains challenging. Current evaluations offer limited coverage, imprecise target-stโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

From LLM-Driven Trading Card Generation to Procedural Relatedness: A Pok\'emon Case Study

Johannes Pfau, Panagiotis Vrettis ยท 2026

Since the dawn of Trading Card Games, the genre has grown into a multi-billion-dollar industry engaging millions of analog and digital players worldwide. Popular TCGs rely on regular updates, balance โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

ClimateVID -- Social Media Videos Analysis and Challenges Involved

Shiqi Xu, Moritz Burmester, Katharina Prasse, Isaac Bravo, Stefanie Walter, Margret Keuper ยท 2026

The pervasive growth of digital content, specifically short videos on social media platforms, has significantly altered how topics are discussed and understood in public discourse. In this work, we adโ€ฆ

Read Paper โ†’
Sociology & Anthropology Preprint PDF DOI

Clustering in co-evolving opinion dynamics: reduced SPDE models

Sebastian Zimper, Natasa Djurdjevac Conrad, Federico Cornalba, Ana Djurdjevac ยท 2026

Clustering is a fundamental collective phenomenon in agent-based models (ABMs) of opinion dynamics. To study clustering in systems with co-evolving social and opinion variables, we derive stochastic pโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

LLMs as ASP Programmers: Self-Correction Enables Task-Agnostic Nonmonotonic Reasoning

Adam Ishay, Joohyung Lee ยท 2026

Recent large language models (LLMs) have achieved impressive reasoning milestones but continue to struggle with high computational costs, logical inconsistencies, and sharp performance degradation on โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On

Dingbao Shao, Song Wu, Shenyi Wang, Ye Wang, Ziheng Tang, Fei Liu, Jiang Lin, Xinyu Chen, Qian Wang, Ying Tai, Jian Yang, Zili Yi ยท 2026

Due to the scarcity of large-scale in-the-wild triplet data and the improper use of masks, the performance of video virtual try-on models remains limited. In this paper, we first introduce **TripVVT-1โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Real-Time Control of a Virtual Orchestra by Recognition of Conducting Gestures

Mert Mermerci, Emile Pascoe, Fredrik Edstrom, Hedvig Kjellstrom ยท 2026

We present a museum installation in a 180{\deg} dome theater, which gives the museum visitor the experience of conducting a symphony orchestra. We have pre-recorded a short music piece performed by a โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

The Effects of Visual Priming on Cooperative Behavior in Vision-Language Models

Kenneth J. K. Ong ยท 2026

As Vision-Language Models (VLMs) become increasingly integrated into decision-making systems, it is essential to understand how visual inputs influence their behavior. This paper investigates the effeโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

CRS-LLM: Cooperative Beam Prediction with a GPT-Style Backbone and Switch-Gated Fusion

Fangzhi Li, Cunhua Pan, Hong Ren, Dongming Wang, Jiangzhou Wang ยท 2026

Millimeter-wave (mmWave) communication depends on highly directional beamforming, while fast mobility, blockage, and rapid geometry changes in vehicle-to-everything (V2X) scenarios make beam tracking โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Calibrating Attribution Proxies for Reward Allocation in Participatory Weather Sensing

Mark C. Ballandies, Michael T. C. Chiu, Claudio J. Tessone ยท 2026

Large-scale IoT weather sensing networks require incentive mechanisms to sustain participation, yet determining how much value individual data contributions bring to the network remains an open probleโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Constrained Symplectic and Contact Hamiltonian Systems: A Review

Callum Bell, David Sloan ยท 2026

Singular theories, characterised by the presence of degeneracies in their Lagrangian or Hamiltonian descriptions, require the systematic implementation of constraints in order to obtain well-defined dโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Enhancing multimodal affect recognition in healthcare: the robustness of appraisal dimensions over labels within age groups and in cross-age generalisation

Hippolyte Fournier, Sina Alisamir, Safaa Azzakhnini, Isabella Zsoldos, Eleonore Tran, Gerard Bailly, Frederic Elisei, Beatrice Bouchot, Brice Varini, Patrick Constant, Joan Fruitet, Franck Tarpin-Bernard, Solange Rossato, Francois Portet, Olivier Koenig, Hanna Chainay, Fabien Ringeval ยท 2026

The integration of artificial intelligence (AI) into healthcare has advanced significantly, yet affect recognition remains a major challenge, particularly in AI-assisted interventions such as Computerโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Can AI Be a Good Peer Reviewer? A Survey of Peer Review Process, Evaluation, and the Future

Sihong Wu, Owen Jiang, Yilun Zhao, Tiansheng Hu, Yiling Ma, Kaiyan Zhang, Manasi Patwardhan, Arman Cohan ยท 2026

Peer review is a multi-stage process involving reviews, rebuttals, meta-reviews, final decisions, and subsequent manuscript revisions. Recent advances in large language models (LLMs) have motivated meโ€ฆ

Read Paper โ†’
Mathematics Preprint PDF DOI

Data-Driven Continuous-Time Linear Quadratic Regulator via Closed-Loop and Reinforcement Learning Parameterizations

Armin Gie{ss}ler, Felix Thommes, Soren Hohmann ยท 2026

This paper studies data-driven approaches to the continuous-time linear quadratic regulator (LQR) problem based on two existing parameterizations, namely a closed-loop (CL) parameterization from behavโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Beyond Semantics: Measuring Fine-Grained Emotion Preservation in Small Language Model-Based Machine Translation

Dawid Wisniewski, Igor Czudy ยท 2026

Preserving affective nuance remains a challenge in Machine Translation (MT), where semantic equivalence often takes precedence over emotional fidelity. This paper evaluates the performance of three stโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Geometry-Calibrated Conformal Abstention for Language Models

Rui Xu, Yi Chen, Sihong Xie, Hui Xiong ยท 2026

When language models lack relevant knowledge for a given query, they frequently generate plausible responses that can be hallucinations, rather than admitting being agnostic about the answer. Retrainiโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

HiMix: Hierarchical Artifact-aware Mixup for Generalized Synthetic Image Detection

Shuchang Zhou, Kaiwen Shen, Jiwei Wei, Yuyang Zhou, Peng Wang, Yang Yang ยท 2026

The rapid evolution of generative models has enabled the creation of highly realistic and diverse synthetic images, posing significant challenges to reliable and generalizable Synthetic Image Detectioโ€ฆ

Read Paper โ†’
โ† Prev Page 3 of 18947 Next โ†’