Expertini Research Research

Browse Research Papers

57,149+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation ๐Ÿ“‚ Computer Science
Showing 57149 results for "program evaluation" in Computer Science
Computer Science Preprint PDF DOI

Hypencoder Revisited: Reproducibility and Analysis of Non-Linear Scoring for First-Stage Retrieval

Arne Eichholtz, Yongkang Li, Jutte Vijverberg, Tobias Groot, Mohammad Aliannejadi ยท 2026

The Hypencoder, proposed by Killingback et al., is a retrieval framework that replaces the fixed inner-product scoring function used in standard bi-encoders with a query-specific neural network (the $โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A 3GPP Perspective on Spectrum Sharing for the 5G-to-6G Migration: From DSS to MRSS

Xingqin Lin ยท 2026

Dynamic spectrum sharing (DSS) played an important role in the 4G-to-5G transition by allowing 5G new radio (NR) to enter valuable legacy spectrum without immediate static refarming. Yet practical depโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Resume-ing Control: (Mis)Perceptions of Agency Around GenAI Use in Recruiting Workflows

Sajel Surati, Rosanna Bellini, Emily Black ยท 2026

When generative AI (genAI) systems are used in high-stakes decision-making, its recommended role is to aid, rather than replace, human decision-making. However, there is little empirical exploration oโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Solving Positive Linear Programs with Differential Privacy

Alina Ene, Huy Le Nguyen, Ta Duy Nguyen, Adrian Vladu ยท 2026

We study differentially private approximation algorithms for positive linear programs (LPs with nonnegative coefficients and variables), focusing on the fundamental families of packing, covering, and โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Full Definability in a Profunctorial Model

Takeshi Tsukada, Kazuyuki Asada, Kengo Hirata ยท 2026

A semantic model enjoys full definability if every semantic element in the model is a denotation of some proof or program. Full definability indicates that the model captures programs and proofs in a โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A Semantic Quantum Circuit Cache for Scalable and Distributed Quantum-Classical Workflows

Mar Tejedor, Javier Conejero, Rosa M. Badia ยท 2026

Hybrid quantum--classical workflows often execute large ensembles of circuits that differ syntactically but implement identical operations, leading to substantial redundant computation. To address thiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

On the Complexity of Robust Markov Decision Processes and Bisimulation Metrics

Marnix Suilen, Guillermo A. Perez ยท 2026

Robust Markov decision processes (RMDPs) extend standard Markov decision processes (MDPs) to account for uncertainty in the transition probabilities. RMDPs have an uncertainty set that defines a set oโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Catching the Fly: Practical Challenges in Making Blockchain FlyClient Real

Pericle Perazzo, Dario Capecchi ยท 2026

FlyClient is a lightweight blockchain verification protocol that enables proof-of-work validation using minimal data, making it ideal for resource-constrained environments like mobile wallets, Interneโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

COPUS: Co-adaptive Parallelism and Batch Size Selection in Large Language Model Training

Akhmed Sakip, Erland Hilman Fuadi, Omar Sayedelahl, Zonghang Li, Jianshu She, Alham Fikri Aji, Steve Liu, Eric Xing, Qirong Ho ยท 2026

Training large language models requires jointly configuring two interdependent aspects of the system: the global batch size, which governs statistical efficiency, and the 3D parallelism strategy, whicโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

When Model Editing Meets Service Evolution: A Knowledge-Update Perspective for Service Recommendation

Guodong Fan, Cuiyun Gao, Chun Yong Chong, Lu Zhang, Jing Li, Jinglin Zhang, Shizhan Chen ยท 2026

The rapid evolution of software services poses substantial challenges to the design and implementation of effective recommendation systems. Traditional service recommendation approaches often rely on โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

MultEval: Supporting Collaborative Alignment for LLM-as-a-Judge Evaluation Criteria

Charles Chiang, Simret Gebreegziabher, Annalisa Szymanski, Yukun Yang, Hyo Jin Do, Zahra Ashktorab, Werner Geyer, Toby Li, Diego Gomez-Zara ยท 2026

LLM-as-a-judge approaches have emerged as a scalable solution for evaluating model behaviors, yet they rely on evaluation criteria often created by a single individual, embedding that person's assumptโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Reproducible Automated Program Repair Is Hard -- Experiences With the Defects4J Dataset

Adam Krafczyk, Klaus Schmid ยท 2026

In the research of automated program repair (APR), benchmark datasets consisting of known defects in combination with test suites that indicate the defects are of high importance. They allow for an evโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Which Types of Heterogeneity Matter for Root Cause Localization in Microservice Systems ?

Runzhou Wang, Shenglin Zhang, Wenwei Gu, Yongxin Zhao, Chenyu Zhao, Dan Pei, Yuxuan Chen, Yangyuxin Huang ยท 2026

Microservice root cause localization is fundamentally challenged by the inherent heterogeneity of cloud-native systems, which encompasses diverse observability data and multiple system entities. Existโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

FACT: Compositional Kernel Synthesis with a Three-Stage Agentic Workflow

Sina Heidari, Dimitrios S. Nikolopoulos ยท 2026

Deep learning compilers and vendor libraries deliver strong baseline performance but are bounded by finite, engineer-curated catalogs. When these omit needed optimizations, practitioners substitute haโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

The Bandit's Blind Spot: The Critical Role of User State Representation in Recommender Systems

Pedro R. Pires, Gregorio F. Azevedo, Rafael T. Sereicikas, Pietro L. Campos, Tiago A. Almeida ยท 2026

With the increasing availability of online information, recommender systems have become an important tool for many web-based systems. Due to the continuous aspect of recommendation environments, theseโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Impact of Attitude and Bounded Rationality on Collective Behavioral Transitions

Chen Song, Vladimir Cvetkovic, Angela Fontan, Rong Su, Karl H. Johansson ยท 2026

The theory of planned behavior (TPB) is one of the most influential frameworks in social psychology, stating that a person's behavior is driven by intention, which is primarily shaped by attitude, subโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Recommendations for Efficient and Responsible LLM Adoption within Industrial Software Development

Krishna Ronanki, Beatriz Cabrero-Daniel, Tomas Herda, Stefan Sitkovich, Jennifer Horkoff, Christian Berger ยท 2026

Context: Large language models (LLMs) are observed to have a significant positive impact on various software engineering (SE) activities. With improved accessibility, the adoption of powerful LLMs in โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Graph Construction and Matching for Imperative Programs using Neural and Structural Methods

Arshad Beg, Diarmuid O'Donoghue, Rosemary Monahan ยท 2026

Reusing verification artefacts requires identifying structural and semantic similarities across programs and their specifications. In this paper, we focus on graph construction as a foundational step โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

DUAL-BLADE: Dual-Path NVMe-Direct KV-Cache Offloading for Edge LLM Inference

Bodon Jeong, Hongsu Byun, Youngjae Kim, Weikuan Yu, Kyungkeun Lee, Jihoon Yang, Sungyong Park ยท 2026

The increasing deployment of Large Language Model (LLM) inference on edge AI systems demands efficient execution under tight memory budgets. A key challenge arises from Key-Value (KV) caches, which ofโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Verification and Validation (V&V)-in-the-Loop for RISC-V Design: The Holistic Vision of BZL

Sajjad Ahmed, Alexander Kropotov, Roberto Ignacio Genovese, Bernat Homs, Eloi Merino, Francesco Urbani, Henrique Yano, Ivan Diaz, Joan Gracia Fernandez, Matteo Toselli, Muhammad Imran, Muhammad Abu Bakar Umar Haider Iqbal, Nadeem Yaseen, Quswar Abid, Shaista Cheema, Samuel Sanchez, Daniel Garcia, Joan Cabre, Mostafa Elyasi, Fernando Ayats, Miquel Moreto, Teresa Cervero, Oscar Palomar, Behzad Salami ยท 2026

The Barcelona Zetascale Lab (BZL) project aims to strengthening Europe's capacity in the design and manufacture of RISC-V based high-performance computing chips. In this context, we present a holisticโ€ฆ

Read Paper โ†’
โ† Prev Page 4 of 2858 Next โ†’