Expertini Research Research

Browse Research Papers

54,648+ open-access research outputs.

✕ Clear
🔍 program evaluation 📂 Computer Science 📄 Preprint
Showing 54648 results for "program evaluation" in Computer Science · Preprint
Computer Science Preprint PDF DOI

RuC: HDL-Agnostic Rule Completion Benchmark Generation

Arnau Ayguade Domingo, Miquel Alberti-Binimelis, Cristian Gutierrez-Gomez, Emanuele Parisi, Razine Moundir Ghorab, Miquel Moreto, Gokcen Kestor, Dario Garcia-Gasulla · 2026

Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RTL) development increasingly attractive. Mimicking …

Read Paper →
Computer Science Preprint PDF DOI

LLM-as-a-Judge for Human-AI Co-Creation: A Reliability-Aware Evaluation Framework for Coding

Md Faizul Ibne Amin, Yutaka Watanobe, Daniel M. Muepu, Haruto Suzuki, Kenta Nanaumi, Md Mostafizer Rahman · 2026

LLMs are increasingly employed both as judges for evaluating open-ended outputs and as co-creation partners in AI-assisted programming; yet rigorous evaluation in human-AI co-creation settings remains…

Read Paper →
Computer Science Preprint PDF DOI

AgentEconomist: An End-to-end Agentic System Translating Economic Intuitions into Executable Computational Experiments

Jiaju Chen, Jinghua Piao, Xia Xu, Songwei Li, Tong Xia, Xiangnan He, Yong Li · 2026

A long-standing challenge in economics lies not in the lack of intuition, but in the difficulty of translating intuitive insights into verifiable research. To address this challenge, we introduce Agen…

Read Paper →
Computer Science Preprint PDF DOI

Towards an Ethical AI Curriculum: A Pan-African, Culturally Contextualized Framework for Primary and Secondary Education

Abidemi Kuburat Adedeji, Franklin Tchakounte, Sulaiman Oluwasegun Yusuff · 2026

Artificial intelligence (AI) is now embedded in educational, civic, and economic systems worldwide. For African primary and secondary education, this creates a double imperative: to prepare a young po…

Read Paper →
Computer Science Preprint PDF DOI

PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion Models

Haocheng Huang, Yuchen Chen, Weisong Sun, Peizhuo Lv, Yuan Xiao, Chunrong Fang, Yang Liu, Xiaofang Zhang · 2026

Constructing and curating high-quality code datasets requires significant resources, making them valuable intellectual property. Unfortunately, these datasets currently face severe risks of unauthoriz…

Read Paper →
Computer Science Preprint PDF DOI

VOW: Verifiable and Oblivious Watermark Detection for Large Language Models

Xiaokun Luan, Yihao Zhang, Pengcheng Su, Feiran Lei, Meng Sun · 2026

Large Language Model (LLM) watermarking is crucial for establishing the provenance of machine-generated text, but most existing methods rely on a centralized trust model. This model forces users to re…

Read Paper →
Computer Science Preprint PDF DOI

Back to the Future: Rethinking Endorsement in Order-Execute Blockchains

Rongji Huang, Yifeng Ye, Gerui Wang, Mingchao Wan, Yuxing Duan, Jingjing Zhang, Guangtao Xue, Shengyun Liu · 2026

Due to regulatory compliance and governance management, modern (permissioned) blockchains require flexible endorsement, which allows the endorsement policy for each contract or state object to be indi…

Read Paper →
Computer Science Preprint PDF DOI

An Exact 56-Addition, Rank-23 Scheme for General 3*3 Matrix Multiplication

Yinqi Sun · 2026

We present a rank-$23$ algorithm for general $3\times3$ matrix multiplication that uses $56$ additions/subtractions and $23$ multiplications, for a total of $79$ scalar operations in the standard bili…

Read Paper →
Computer Science Preprint PDF DOI

SecGoal: A Benchmark for Security Goal Extraction and Formalization from Protocol Documents

Dawei Huang, Hui Li, Haonan Feng, Jingjing Guan, Yueshuang Jiao, Bo Jia (Beijing University of Posts, Telecommunications) · 2026

Formal verification provides rigorous guarantees for cryptographic security, yet automating the extraction and formalization of security goals from natural language protocol documents remains a major …

Read Paper →
Computer Science Preprint PDF DOI

Thinking like a business: Reconfiguring relationships to sustain open data infrastructures

Kathleen Gregory, Dorothea Strecker · 2026

Sustaining open data infrastructures over time is a complex puzzle, involving dynamic funding models and relationships with customers, collaborators, and competitors. Despite their importance, these m…

Read Paper →
Computer Science Preprint PDF DOI

I'm Fine, But My Voice Isn't: Cross-Modal Affective Dissonance Detection for Reflective Journaling

Sumin Lee · 2026

Digital journaling creates an authenticity gap: users consciously translate raw emotions into text, often sanitizing narratives even in private writing. We formalize this as Cross-Modal Affective Diss…

Read Paper →
Computer Science Preprint PDF DOI

ScaleBox: Enabling High-Fidelity and Scalable Code Verification for Large Language Models

Jiasheng Zheng, Xin Zheng, Boxi Cao, Pengbo Wang, Zhengzhao Ma, Qiming Zhu, Jiazhen Jiang, Yaojie Lu, Hongyu Lin, Xianpei Han, Le Sun · 2026

Code sandboxes have emerged as a critical infrastructure for advancing the coding capabilities of large language models, providing verifiable feedback for both RL training and evaluation. However, exi…

Read Paper →
Computer Science Preprint PDF DOI

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

Luyao Xu, Xiang Chen · 2026

Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, introducing security risks beyond traditional prompt…

Read Paper →
Computer Science Preprint PDF DOI

ReVo: A Cross-Layer Reliable Volumetric Videoconferencing System

Ankur Aditya, Diptyaroop Maji, Lingdong Wang, Bhavya Ramakrishna, Ramesh Sitaraman, Prashant Shenoy · 2026

Volumetric videoconferencing enables immersive six Degrees of Freedom interactions by jointly transmitting visual appearance and 3D geometry. However, delivering volumetric video over today's networks…

Read Paper →
Computer Science Preprint PDF DOI

A Study on the Performance of Distributed Training of Data-driven CFD Simulations

Sergio Iserte, Alejandro Gonzalez-Barbera, Paloma Barreda, Krzysztof Rojek · 2026

Data-driven methods for computer simulations are blooming in many scientific areas. The traditional approach to simulating physical behaviors relies on solving partial differential equations (PDE). Si…

Read Paper →
Computer Science Preprint PDF DOI

A Reproducibility Study of LLM-Based Query Reformulation

Amin Bigdeli, Radin Hamidi Rad, Hai Son Le, Mert Incesu, Negar Arabzadeh, Charles L. A. Clarke, Ebrahim Bagheri · 2026

Large Language Models (LLMs) are now widely used for query reformulation and expansion in Information Retrieval, with many studies reporting substantial effectiveness gains. However, these results are…

Read Paper →
Computer Science Preprint PDF DOI

Exploring the Adoption Intention in Using AI-Enabled Educational Tools Among Preservice Teachers in the Philippines: A Partial-Least Square Modeling

Vanessa B. Sibug, Emerson Q. Fernando, Almer B. Gamboa, Roque Francis B. Dianelo, Agnes R. Regala, Joseph Alexander Bansil, Jan Henry B. Sunga, Vernon Grace M. Maniago, John Paul P. Miranda · 2026

This study examines the factors influencing pre-service teachers' behavioral intention to use AI-enabled educational tools during their practicum, using the Unified Theory of Acceptance and Use of Tec…

Read Paper →
Computer Science Preprint PDF DOI

One Size Fits All? An Empirical Comparison of ADR Templates regarding Comprehension, Usability, and Ease of Adoption

Fernando Nogueira, Nabson Silva, Tayana Conte · 2026

Context: Documenting Architectural Design Decisions (ADDs) is a critical factor in the software lifecycle, essential for efficient system maintenance, developer onboarding, and preventing knowledge va…

Read Paper →
Computer Science Preprint PDF DOI

SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields

Youkang Kong, Yang Liu, Yue Dong, Xin Tong, Heung-Yeung Shum · 2026

3D shapes from scanning, reconstruction, or AI-generated content often lack simple quad mesh layouts -- critical for efficient editing and modeling. Existing quad-remeshing techniques typically produc…

Read Paper →
Computer Science Preprint PDF DOI

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

Jun Yeon Won, Xin Jin, Shiqing Ma, Zhiqiang Lin · 2026

Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including computer security. In reverse engineering, LLMs are inc…

Read Paper →
← Prev Page 2 of 2733 Next →