Expertini Research Research

Browse Research Papers

57,149+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation ๐Ÿ“‚ Computer Science
Showing 57149 results for "program evaluation" in Computer Science
Computer Science Preprint PDF DOI

K-CARE: Knowledge-driven Symmetrical Contextual Anchoring and Analogical Prototype Reasoning for E-commerce Relevance

Chen Yifei, Tian Zhixing, Wang Chenyang, Cheng Ziguang ยท 2026

This paper targets e-commerce search relevance. While Large Language Models (LLMs) have demonstrated significant potential in this field, they often encounter performance bottlenecks in persistent 'coโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Using Large Language Models for Black-Box Testing of FMU-Based Simulations

Abdullah Mughees, Gaadha Sudheerbabu, Tanwir Ahmad, Dragos Truscan, Mikael Manng{aa}rd, Kristian Klemets ยท 2026

We propose a human in the loop approach for black-box testing of Functional Mock-up Units (FMUs) using Large Language Models (LLMs). The goal is to reduce the manual effort in defining test scenarios โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

The Surprising Universality of LLM Outputs: A Real-Time Verification Primitive

Alex Bogdan, Adrian de Valois-Franklin ยท 2026

We report a striking statistical regularity in frontier LLM outputs that enables a CPU-only scoring primitive running at 2.6 microseconds per token, with estimated latency up to 100,000$\times$ (five โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Health System Scale Semantic Search Across Unstructured Clinical Notes

Faith Wavinya Mutinda, Spandana Makeneni, Anna Lin, Shivaji Dutta, Irit R. Rasooly, Patrick Dibussolo, Shivani Kamath Belman, Hessam Shahriari, Kevin Murphy, Alex B. Ruan, Barbara H. Chaiyachati, Sanjay Chainani, Robert W. Grundmeier, Scott M. Haag, Jeffrey M. Miller, Heather M. Griffis, Ian M. Campbell ยท 2026

Introduction: Semantic search, which retrieves documents based on conceptual similarity rather than keyword matching, offers substantial advantages for retrieval of clinical information. However, deplโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Generating Synthetic Citation Networks with Communities

{L}ukasz Brzozowski, Marek Gagolewski, Grzegorz Siudem ยท 2026

Generating realistic synthetic citation, patent, or component dependency networks is essential for benchmarking community detection, graph visualisation, and network data mining algorithms. We presentโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

An Empirical Analysis of Mobile Energy Consumption Across User Configurations

Wellington Oliveira ยท 2026

Mobile devices have become ubiquitous tools for communication, entertainment, and productivity, yet battery autonomy remains a constraint. While energy-saving tips exist, they are often generic, anecdโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

The Attention Market: Interpreting Online Fair Re-ranking as Manifold Optimization under Walrasian Equilibrium

Chen Xu, Wei Chu, Wenyu Hu, Fengran Mo, Jun Xu, Maarten de Rijke ยท 2026

Fair re-ranking aims to promote long-tail items and enhance diversity within groups in information retrieval. While previous research on online fairness-aware re-ranking has shown promising outcomes, โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Design Insights into Partition Placement and Routing for DNN Inference in Multi-Hop Edge Networks

Jinkun Zhang, Poonam Yadav ยท 2026

Partitioned DNN inference is a promising approach for latency-sensitive intelligent services in edge networks, since it allows different parts of a model to be executed across end devices, edge serverโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

SnapGuard: Lightweight Prompt Injection Detection for Screenshot-Based Web Agents

Mengyao Du, Han Fang, Haokai Ma, Jiahao Chen, Kai Xu, Quanjun Yin, Ee-Chien Chang ยท 2026

Web agents have emerged as an effective paradigm for automating interactions with complex web environments, yet remain vulnerable to prompt injection attacks that embed malicious instructions into webโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

From CRUD to Autonomous Agents: Formal Validation and Zero-Trust Security for Semantic Gateways in AI-Native Enterprise Systems

Ignacio Peyrano ยท 2026

Enterprise software engineering is shifting away from deterministic CRUD/REST architectures toward AI-native systems where large language models act as cognitive orchestrators. This transition introduโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

On the degradations of Binary-Input Discrete Memoryless Channels

Yadong Jiao, Xiaoyan Cheng, Yuansheng Tang, Ming Xu ยท 2026

For the polar codes introduced by Arikan in 2009, the first code family achieving the capacity of binary-input discrete memoryless channels (BIDMCs) with low-complexity encoding and decoding, it is crโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

SymphonyGen: 3D Hierarchical Orchestral Generation with Controllable Harmony Skeleton

Xuzheng He, Nan Nan, Zhilin Wang, Ziyue Kang, Zhuoru Mo, Ao Li, Yu Pan, Xiaobing Li, Feng Yu, Xiaohong Guan ยท 2026

Generating symphonic music requires simultaneously managing high-level structural form and dense, multi-track orchestration. Existing symbolic models often struggle with a "complexity-control imbalancโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Practical Insights into Fair Comparison and Evaluation Frame for Neutral-Atom Compilers

Emil Khusainov, Yanbin Chen, Jonas Winklmann, Helmut Seidl, Christian B. Mendl ยท 2026

Neutral-atom quantum computing is among the most promising platforms for scalable quantum computation, and compilation toolchains are crucial for leveraging capabilities such as qubit shuttling and paโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

PSP: An Interpretable Per-Dimension Accent Benchmark for Indic Text-to-Speech

Venkata Pushpak Teja Menta ยท 2026

Standard text-to-speech (TTS) evaluation measures intelligibility (WER, CER) and overall naturalness (MOS, UTMOS) but does not quantify accent. A synthesiser may score well on all four yet sound non-nโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Generative UI as an Accessibility Bridge: Lessons from C2C E-Commerce

Bektur Ryskeldiev ยท 2026

Web accessibility rests on static standards and developer compliance. That model frays in platforms where content is user-generated: photos arrive blurry or off-frame, descriptions skip size and condiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Central Limit Theorem for Mutation Systems

Liav Koram, Ohad Elishco ยท 2026

DNA-based storage has emerged as a promising alternative to traditional data storage methods, offering unmatched advantages in data density, longevity, and sustainability. Two main approaches have devโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Recommending Usability Improvements with Multimodal Large Language Models

Sebastian Lubos, Alexander Felfernig, Damian Garber, Viet-Man Le, Manuel Henrich ยท 2026

Usability describes quality attributes of application user interfaces that determine how effectively users can interact with them. Traditional usability evaluation methods require considerable expertiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

CoRE: A Fine-Grained Code Reasoning Benchmark Beyond Output Prediction

Jun Gao, Yun Peng, Qian Qiao, Changhai Zhou, Yuhua Zhou, Shiyang Zhang, Shichao Weng, Zhenchang Xing, Xiaoxue Ren ยท 2026

Despite strong performance on code generation tasks, it remains unclear whether large language models (LLMs) genuinely reason about code execution. Existing code reasoning benchmarks primarily evaluatโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

GeoSearch: Augmenting Worldwide Geolocalization with Web-Scale Reverse Image Search and Image Matching

Tung-Duong Le-Duc, Hoang-Quoc Nguyen-Son, Minh-Son Dao ยท 2026

Worldwide image geolocalization, which aims to predict the GPS coordinates of any image on Earth, remains challenging due to global visual diversity. Recent generative approaches based on Retrieval-Auโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Stop Using the Wilcoxon Test: Myth, Misconception and Misuse in IR Research

Julian Urbano ยท 2026

In benchmarking of Information Retrieval systems, the Wilcoxon signed-rank test is often treated as a safer alternative to the t-test. This belief is fueled by textbooks and recommendations that portrโ€ฆ

Read Paper โ†’
โ† Prev Page 7 of 2858 Next โ†’