Expertini Research Research

Browse Research Papers

57,149+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation ๐Ÿ“‚ Computer Science
Showing 57149 results for "program evaluation" in Computer Science
Computer Science Preprint PDF DOI

Entrywise Low-Rank Approximation and Matrix $p \rightarrow q$ Norms via Global Correlation Rounding

Prashanti Anderson, Ainesh Bakshi, Samuel B. Hopkins ยท 2026

Given a matrix $A$, the goal of the entrywise low-rank approximation problem is to find $\operatorname{argmin} \|A-B\|_p$ over all rank-$k$ matrices $B$, where $\| \cdot \|_p$ is the entrywise $\ell_pโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

How Supply Chain Dependencies Complicate Bias Measurement and Accountability Attribution in AI Hiring Applications

Gauri Sharma, Maryam Molamohammadi ยท 2026

The increasing adoption of AI systems in hiring has raised concerns about algorithmic bias and accountability, prompting regulatory responses including the EU AI Act, NYC Local Law 144, and Colorado'sโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Measuring Epistemic Unfairness for Algorithmic Decision-Making

Camilla Quaresmini, Lisa Piccinin, Valentina Breschi ยท 2026

Algorithmic systems increasingly function as epistemic infrastructures that govern the conditions of interpretative access and social belief. Yet, mainstream auditing strategies operationalize fairnesโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Cuts and Gauges for Submodular Width

Matthias Lanzinger ยท 2026

Submodular width is a central structural measure governing the complexity of conjunctive query evaluation. In this paper we recast submodular width in geometric terms. We how that submodular width canโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines

Negar Arabzadeh, Andrew Drozdov, Michael Bendersky, Matei Zaharia ยท 2026

Large Language Models (LLMs) have made query reformulation ubiquitous in modern retrieval and Retrieval-Augmented Generation (RAG) pipelines, enabling the generation of multiple semantically equivalenโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

RealBench: A Repo-Level Code Generation Benchmark Aligned with Real-World Software Development Practices

Jia Li, Hongyi Deng, Yiran Zhang, Kechi Zhang, Tianqi Shao, Tiankuo Zhao, Weinan Wang, Zhi Jin, Ge Li, Yang Liu, Yingtao Fang, Yihong Dong ยท 2026

Writing code requires significant time and effort in software development. To automate this process, researchers have made substantial progress using Large Language Models (LLMs) for code generation. โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Verifier Warnings Do Not Improve Comprehensibility Prediction

Nadeeshan De Silva, Martin Kellogg, Oscar Chaparro ยท 2026

Proponents of software verification suggest that code simplicity is linked to the effort to verify code, hypothesizing that formal verifiers produce fewer false positive warnings and require less manuโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Cloud to Edge: Benchmarking LLM Inference On Hardware-Accelerated Single-Board Computers

Harri Renney, Fouad Trad, Michael Mattarock, Zena Wood ยท 2026

Large language models (LLMs) are becoming increasingly capable at small parameter scales. At the same time, conventional cloud-centric deployment introduces challenges around data privacy, latency, anโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A Comparison of ROS 2 and AUTOSAR Adaptive Platform Against Industry-Elicited Automotive Middleware Requirements

Lucas Hegerath, David Philipp Kluner, Philipp Pelcz, Viswanatha Reddy Batchu, Marius Molz, Julius Kahle, Thomas Schulik, Stefan Kowalewski, Alexandru Kampmann ยท 2026

In software-defined vehicles, automotive middleware plays a fundamental role in enabling efficient communication, integration, and coordination among software components. This paper examines how well โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Adversarial Co-Evolution of Malware and Detection Models: A Bilevel Optimization Perspective

Olha Jureckova, Martin Jurecek, Matous Kozak, Robert Lorencz ยท 2026

Machine learning-based malware detectors are increasingly vulnerable to adversarial examples. Traditional defenses, such as one-shot adversarial training, often fail against adaptive attackers who useโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Module Lattice Security (Part II): Module Lattice Reduction via Optimal Sign Selection

Ming-Xing Luo ยท 2026

We extend the CDPR lattice reduction algorithm from ideal to module lattices, leveraging the trace orthogonality of the power basis to decompose the module into rank-1 submodules and applying CDPR indโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

DEKL 2.0: Trace-Indexed Knowledge Evolution in Dependent Type Theory

Chen Peng ยท 2026

DEKL 2.0 is a dependent type-theoretic framework for trace-indexed knowledge evolution. Its central claim is that the proof calculus remains monotone under standard structural rules, while non-monotonโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Benchmarking LLM-Driven Network Configuration Repair

Ioannis Protogeros, Rufat Asadli, Benjamin Hoffman, Laurent Vanbever ยท 2026

There is a rapidly growing interest in using Large Language Models (LLMs) to automate complex network operations, but their reliable adoption requires rigorous assessment of their effectiveness and saโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Gamifying Architectural Governance to Reduce Organizational Coupling in Microservice Systems

Xiaozhou Li ยท 2026

Microservice is a popular software architecture that relies on decentralized teams and clear service ownership to support modularity and scalability. However, in practice, developers frequently contriโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A comprehensive evaluation of spatial co-execution on GPUs using MPS and MIG technologies

Jorge Villarrubia, Luis Costero, Francisco D. Igual, Katzalin Olcoz ยท 2026

To mitigate the increasingly common underutilization of computational resources in modern GPUs, spatial sharing methods enable multiple applications to use them simultaneously. This work presents a coโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Automation-Exploit: A Multi-Agent LLM Framework for Adaptive Offensive Security with Digital Twin-Based Risk-Mitigated Exploitation

Biagio Andreucci, Arcangelo Castiglione ยท 2026

The offensive security landscape is highly fragmented: enterprise platforms avoid memory-corruption vulnerabilities due to Denial of Service (DoS) risks, Automatic Exploit Generation (AEG) systems sufโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

How Hard is it to Decide if a Fact is Relevant to a Query?

Meghyn Bienvenu, Diego Figueira, Pierre Lafourcade ยท 2026

We consider the following fundamental problem: given a database D, Boolean conjunctive query (CQ) q, and fact f in D, decide whether f is relevant to q wrt. D, i.e., does f belong to a minimal subset โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Trust as a Situated User State in Social LLM-Based Chatbots: A Longitudinal Study of Snapchat's My AI

Annie Landerberg, Kari Flatmo, Alan Said ยท 2026

Social chatbots based on large language models are increasingly embedded in everyday platforms, yet how users develop trust in these systems over time remains unclear. We present a four-week longitudiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A Model-Driven Approach to Database Migration with a Unified Data Model

Maria J. Ortin, Jose R. Hoyos, Jesus Garcia-Molina ยท 2026

Database migration is a key task in software modernization, increasingly involving transformations across heterogeneous data models such as relational and NoSQL systems. Existing approaches are typicaโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Ownership Refinement Types for Pointer Arithmetic and Nested Arrays

Yusuke Fujiwara, Yusuke Matsushita, Kohei Suenaga, Atsushi Igarashi ยท 2026

Tanaka et al. proposed a type system for verifying functional correctness properties of programs that use arrays and pointer arithmetic. Their system extends ConSORT -- a type system combining fractioโ€ฆ

Read Paper โ†’
โ† Prev Page 14 of 2858 Next โ†’