Expertini Research Research

Browse Research Papers

378,930+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation
Showing 378930 results for "program evaluation"
AI & Data Science Preprint PDF DOI

Zero-to-CAD: Agentic Synthesis of Interpretable CAD Programs at Million-Scale Without Real Data

Mohammadmehdi Ataei, Farzaneh Askari, Kamal Rahimi Malekshan, Pradeep Kumar Jayaraman ยท 2026

Computer-Aided Design (CAD) models are defined by their construction history: a parametric recipe that encodes design intent. However, existing large-scale 3D datasets predominantly consist of boundarโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems

Pablo Mateo-Torrejon, Alfonso Sanchez-Macian ยท 2026

The rapid integration of Large Language Models (LLMs) into Multi-Agent Systems (MAS) has significantly enhanced their collaborative problem-solving capabilities, but it has also expanded their attack โ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus

Johannes Moll, Jannik Lubberstedt, Christoph Nuernbergk, Jacob Stroh, Luisa Mertens, Anna Purcarea, Christopher Zirn, Zeineb Benchaaben, Fabian Drexel, Hartmut Hantze, Anirudh Narayanan, Friedrich Puttkammer, Andrei Zhukov, Jacqueline Lammert, Sebastian Ziegelmayer, Markus Graf, Marion Hogner, Marcus Makowski, Florian Bassermann, Lisa C. Adams, Jiazhen Pan, Daniel Rueckert, Krischan Braitsch, Keno K. Bressem ยท 2026

Multiple myeloma is managed through sequential lines of therapy over years to decades, with each decision depending on cumulative disease history distributed across dozens to hundreds of heterogeneousโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Machine learning technique for morphological classification of galaxies from SDSS. IV. Visual inspection vs CNN for merging, irregular, edge-on, barred, ringed, and with dust lanes galaxies at 0.02<z<0.1

Dobrycheva D.V., Vavilova I.B., Kompaniiets O.V., Khramtsov V., Vasylenko M.Yu., Hetmantsev O.O., Melnyk O.V., Karachentseva V.E ยท 2026

Context. Convolutional neural networks (CNNs) are widely used for automated galaxy morphological classification in large surveys. However, projection effects, image artefacts, and intrinsic degeneraciโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

Zero-shot Large Language Models for Automatic Readability Assessment

Riley Grossman, Yi Chen ยท 2026

Unsupervised automatic readability assessment (ARA) methods have important practical and research applications (e.g., ensuring medical or educational materials are suitable for their target audiences)โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Geometric Analysis of Self-Supervised Vision Representations for Semantic Image Retrieval

Esteban Rodriguez-Betancourt, Edgar Casasola-Murillo ยท 2026

Content-based image retrieval (CBIR) systems enable users to search images based on visual content instead of relying on metadata. The text domain has benefited from vector search of representations cโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Incisor: Ex Ante Cloud Instance Selection for HPC Jobs

Michael A. Laurenzano, Shihan Cheng, David A. B. Hyde ยท 2026

We present Incisor, a cloud HPC job submission system for the ex ante instance selection problem: choosing suitable hardware in the challenging but common setting where only the executable, inputs, anโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Measuring Successful Cooperation in Human-AI Teamwork: Development and Validation of the Perceived Cooperativity and Teaming Perception Scales

Christiane Attig, Christiane Wiebel-Herboth, Patricia Wollstadt, Tim Schrills, Mourad Zoubir, Thomas Franke ยท 2026

As human-AI cooperation becomes increasingly prevalent, reliable instruments for assessing the subjective quality of cooperative human-AI interaction are needed. We introduce two theoretically groundeโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

TextGround4M: A Prompt-Aligned Dataset for Layout-Aware Text Rendering

Dongxing Mao, Yilin Wang, Linjie Li, Zhengyuan Yang, Alex Jinpeng Wang ยท 2026

Despite recent advances in text-to-image generation, models still struggle to accurately render prompt-specified text with correct spatial layout -- especially in multi-span, structured settings. Thisโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

NL-COMM-Sat: Breaking the Direct Device-to-Satellite Communication Barrier via "Aggressive" Non-Orthogonal Transmissions and Non-Linear Processing

Konstantinos Nikitopoulos, Chathura Jayawardena ยท 2026

Direct Device-to-Satellite (D2S) communications, which enable direct satellite connectivity with unmodified user equipment (UE), not only expand global coverage but also reshape the evolution of futurโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

Kaijun Zhou, Qiwei Chen, Da Peng, Zhiyang Li, Xijun Li, Jinyu Gu ยท 2026

Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under tight cost and energy budgets. Most prior evaluatioโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark

Hongxin Li, Xiping Wang, Jingran Su, Zheng Ju, Yuntao Chen, Qing Li, Zhaoxiang Zhang ยท 2026

Autonomous agents capable of navigating Graphical User Interfaces (GUIs) hold the potential to revolutionize digital productivity. However, achieving true digital autonomy extends beyond reactive elemโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

CSST Preparations: Galaxy Completeness and S\'ersic Profile Fitting across the Wide, Deep, and Extreme Fields

Ziqi Ma, Si-Yue Yu, Taotao Fang, Jinyi Shangguan, Zhao-Yu Li, Luis C. Ho ยท 2026

The upcoming imaging survey of the Chinese Space-station Survey Telescope (CSST) will deliver high-resolution imaging of an unprecedented number of galaxies for galaxy studies. To understand CSST's caโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

SWE-QA: A Dataset and Benchmark for Complex Code Understanding

Laila Elkoussy (LRE, EPITA), Julien Perez (EPITA, LRE) ยท 2026

In this paper, we introduce SWE-QA, a text and code corpus aimed at benchmarking multi-hop code comprehension, addressing the gap between simplified evaluation tasks and the complex reasoning requiredโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Phase-Separated Complex Hilbert PCA on Markerless 3D Pose Estimation Data: A Global Phase Network and Its Extension to a Continuous Field on the Body Surface

Hiromitsu Goto, Tao Tao, Zheng-Lin Chia ยท 2026

Quantitative analysis of the kinematic chain in sports motion is essential for performance evaluation and injury prevention. Conventional methods such as the kinematic-sequence (KS) and continuous relโ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Gravitational waves of extreme-mass-ratio inspirals in a rotating black hole with Dehnen dark matter halo

Kun Meng, Shao-Jun Zhang, Nan Yang ยท 2026

Extreme Mass Ratio Inspirals (EMRIs) are among the key targe sources for the space-based gravitational wave (GW) detectors. The waveforms of the EMRIs are highly sensitive to the types of the central โ€ฆ

Read Paper โ†’
Physics Preprint PDF DOI

Impact of thermal and dissipative effects in a periodically-kicked quantum battery

Sebastian V. Romero, Xi Chen, Yue Ban ยท 2026

Quantum batteries (QBs) have emerged as a promising route for fast energy storage and on-chip power supply in quantum devices. Given the limited analytical understanding of open Floquet QBs, we employโ€ฆ

Read Paper โ†’
AI & Data Science Preprint PDF DOI

AD-Relight: Training-Free Banner Relighting via Illumination Translation with Diffusion Priors

Rameshwar Mishra, A V Subramanyam ยท 2026

The recent surge in content consumption through streaming services has driven a growing demand for personalized content. Personalized advertisements (ads) play a crucial role in enhancing both user enโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

How Personal Characteristics Shape User Exploration of Diverse Movie Recommendations with a LLM-Based Multi-Agent System

Yufan Zhou, Yirui Huang, Zhao Wang, Yucheng Jin ยท 2026

Diversity is an important evaluation criterion for recommender systems beyond accuracy, yet users differ in their willingness to engage with novel and diverse content. In this work, we investigate howโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

All That Glitters Is Not Audio: Rethinking Text Priors and Audio Reliance in Audio-Language Evaluation

Leonardo Haw-Yang Foo, Chih-Kai Yang, Chen-An Li, Ke-Han Lu, Hung-yi Lee ยท 2026

Large Audio-Language Models show consistent performance gains across speech and audio benchmarks, yet high scores may not reflect true auditory perception. If a model can answer questions without procโ€ฆ

Read Paper โ†’
โ† Prev Page 40 of 18947 Next โ†’