Expertini Research Research

Browse Research Papers

57,149+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation ๐Ÿ“‚ Computer Science
Showing 57149 results for "program evaluation" in Computer Science
Computer Science Preprint PDF DOI

BLAST: Benchmarking LLMs with ASP-based Structured Testing

Manuel Alejandro Borroto Santana, Erica Coppolillo, Francesco Calimeri, Giuseppe Manco, Simona Perri, Francesco Ricca ยท 2026

Large Language Models (LLMs) have demonstrated remarkable performance across a broad spectrum of tasks, including natural language understanding, dialogue systems, and code generation. Despite evidentโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Resource-Aware Layered Intrusion Detection Allocation Model

Ioan Padurean, Bela Genge, Roland Bolboaca ยท 2026

This paper proposes a resource-aware allocation model for layered intrusion detection in het erogeneous networks. Monitoring traffic at higher protocol layers improves the ability to detect sophisticaโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Exploiting pre-optimized kernels with polyhedral transformations for CGRA compilation

Yuxuan Wang, Maria Jose Belda, Fernando Castro, Katzalin Olcoz, David Atienza, Giovanni Ansaloni ยท 2026

Modern computing workloads commonly involve matrix-matrix multiplication (mmul) as a core computing pattern. Coarse-Grained Reconfigurable Arrays (CGRAs) can flexibly and efficiently support it, sinceโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Transformer-Based Rhythm Quantization of Performance MIDI Using Beat Annotations

Maximilian Wachter, Sebastian Murgul, Michael Heizmann ยท 2026

Rhythm transcription is a key subtask of notation-level Automatic Music Transcription (AMT). While deep learning models have been extensively used for detecting the metrical grid in audio and MIDI perโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

TRUST-SC: Truthful Multi-Task Double Auction for Quality-Aware Spatial Crowdsourcing in Strategic Environment

Chattu Bhargavi, Vikash Kumar Singh, Alok Kumar Shukla ยท 2026

Spatial crowdsourcing (SC) enables the assignment of location-based tasks to mobile users who must travel to specific locations to perform sensing or service activities. However, SC systems often operโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

GR-Evolve: Design-Adaptive Global Routing via LLM-Driven Algorithm Evolution

Taizun Jafri, Vidya A. Chhabria ยท 2026

Modern ASIC design is becoming increasingly complex, driving up design costs while limiting productivity gains from existing EDA tools. Despite decades of progress, current tools rely on fixed heuristโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies and Their Limitations

Anna Arnaudo, Riccardo Coppola, Maurizio Morisio, Flavio Giobergia, Andrea Bioddo, Angelo Bongiorno, Luca Dadone ยท 2026

Due to the textual and repetitive nature of many Requirements Engineering (RE) artefacts, Large Language Models (LLMs) have proven useful to automate their generation and processing. In this paper, weโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

FixV2W: Correcting Invalid CVE-CWE Mappings with Knowledge Graph Embeddings

Sevval Simsek, Varsha Athreya, David Starobinski ยท 2026

Accurate mapping between Common Vulnerabilities and Exposures (CVE) and Common Weakness Enumeration (CWE) entries is critical for effective vulnerability management and risk assessment. However, publiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Same Project, Different Start: How Contribution Events Shape Activity and Retention in Open Source

Mohamed Ouf, Mariam Guizani ยท 2026

Open source projects depend on newcomers who stay, yet most leave after a single contribution. Contribution events such as Google Summer of Code, LFX Mentorship, Hacktoberfest, and 24 Pull Requests atโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Implementation and Privacy Guarantees for Scalable Keyword Search on SOLID-based Decentralized Data with Granular Visibility Constraints

Mohamed Ragab, Faria Ferooz, Mohammad Bahrani, Helen Oliver, Thanassis Tiropanis, Alexandra Poulovassilis, Adriane Chapman, George Roussos ยท 2026

In decentralized personal data ecosystems grounded in architectures such as Solid, users retain sovereignty over their data via personal online data stores (pods), hosted on Solid-compliant server infโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Who Audits the Auditor? Tamper-Proof Fraud Detection with Blockchain-Anchored Explainable ML

Zhaohui Wang ยท 2026

In enterprise fraud detection, model accuracy alone is insufficient when insiders can tamper with audit logs or bypass approval workflows. Real-world incidents show that fraud often persists not becauโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

FlashSpread: IO-Aware GPU Simulation of Non-Markovian Epidemic Dynamics via Kernel Fusion

Heman Shakeri, Behnaz Moradi-Jamei, Aram Vajdi, Ehsan Ardjmand ยท 2026

Non-Markovian (renewal) epidemic simulation on multi-million-node contact networks is essential for realistic forecasting under general age-dependent holding-time distributions (log-normal, Weibull, Eโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

TRACE: Topology-aware Reconstruction of Accidents in CARLA for AV Evaluation

Nahian Salsabil, Sebastian Elbaum ยท 2026

Validating Autonomous Vehicles (AVs) requires exposure to rare, safety-critical scenarios, infrequent in routine driving data. Existing benchmarks address this by generating synthetic conflicts or mapโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Turnstile Streaming Algorithms Might (Still) as Well Be Linear Sketches, for Polynomial-Length Streams

Cheng Jiang, Yinchen Liu, Huacheng Yu ยท 2026

A fundamental question in streaming complexity is whether every space-efficient turnstile algorithm is implicitly a linear sketch. The landmark work of Li, Nguyen, and Woodruff [LNW14] established an โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Probabilistic Epistemic Dynamic Agentive Logic

Shay Allen Logan ยท 2026

I introduce PEDAL -- a probabilistic epistemic logic meant to capture, in propositional dynamic terms, the epistemic state of an agent engaged in checking whether a program meets its specification. Seโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Causality and Semantic Separation

Anna Zhang, Qinglan Luo, London Bielicke, Eunice Jun, Adam Chlipala ยท 2026

The design of scientific experiments deserves its own variation of formal verification to catch cases where scientists made important mistakes, such as forgetting to take confounding variables into acโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

FlyCatcher: Neural Inference of Runtime Checkers from Tests

Beatriz Souza, Chang Lou, Suman Nath, Michael Pradel ยท 2026

Complex software systems often suffer from silent failures, i.e., violations of the intended semantics that do not cause explicit errors. A promising approach to detect such errors is to use system-spโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models

Tanmay Gautam, Alireza Bahramali, Sandeep Atluri ยท 2026

Automated red-teaming methods for large language models typically optimize attack prompts within a fixed, human-designed strategy, leaving the attack strategy itself unchanged. We instead optimize theโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Emergent Technology, Emergent Critique: Students and Teachers Developing Critical AI Literacy through Participatory Design around Generative AI

Santiago Ojeda-Ramirez, Eva Durall Gazulla, Kylie Peppler ยท 2026

Who gets to decide how generative AI tools enter students' classrooms? We report on a five-week participatory design program in which three 11th-grade Latinx students and three high school teachers inโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

CrossCommitVuln-Bench: A Dataset of Multi-Commit Python Vulnerabilities Invisible to Per-Commit Static Analysis

Arunabh Majumdar ยท 2026

We present CrossCommitVuln-Bench, a curated benchmark of 15 real-world Python vulnerabilities (CVEs) in which the exploitable condition was introduced across multiple commits - each individually benigโ€ฆ

Read Paper โ†’
โ† Prev Page 15 of 2858 Next โ†’