Expertini Research Research

Browse Research Papers

57,149+ open-access research outputs.

โœ• Clear
๐Ÿ” program evaluation ๐Ÿ“‚ Computer Science
Showing 57149 results for "program evaluation" in Computer Science
Computer Science Preprint PDF DOI

RuC: HDL-Agnostic Rule Completion Benchmark Generation

Arnau Ayguade Domingo, Miquel Alberti-Binimelis, Cristian Gutierrez-Gomez, Emanuele Parisi, Razine Moundir Ghorab, Miquel Moreto, Gokcen Kestor, Dario Garcia-Gasulla ยท 2026

Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RTL) development increasingly attractive. Mimicking โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

LLM-as-a-Judge for Human-AI Co-Creation: A Reliability-Aware Evaluation Framework for Coding

Md Faizul Ibne Amin, Yutaka Watanobe, Daniel M. Muepu, Haruto Suzuki, Kenta Nanaumi, Md Mostafizer Rahman ยท 2026

LLMs are increasingly employed both as judges for evaluating open-ended outputs and as co-creation partners in AI-assisted programming; yet rigorous evaluation in human-AI co-creation settings remainsโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

AgentEconomist: An End-to-end Agentic System Translating Economic Intuitions into Executable Computational Experiments

Jiaju Chen, Jinghua Piao, Xia Xu, Songwei Li, Tong Xia, Xiangnan He, Yong Li ยท 2026

A long-standing challenge in economics lies not in the lack of intuition, but in the difficulty of translating intuitive insights into verifiable research. To address this challenge, we introduce Agenโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Towards an Ethical AI Curriculum: A Pan-African, Culturally Contextualized Framework for Primary and Secondary Education

Abidemi Kuburat Adedeji, Franklin Tchakounte, Sulaiman Oluwasegun Yusuff ยท 2026

Artificial intelligence (AI) is now embedded in educational, civic, and economic systems worldwide. For African primary and secondary education, this creates a double imperative: to prepare a young poโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion Models

Haocheng Huang, Yuchen Chen, Weisong Sun, Peizhuo Lv, Yuan Xiao, Chunrong Fang, Yang Liu, Xiaofang Zhang ยท 2026

Constructing and curating high-quality code datasets requires significant resources, making them valuable intellectual property. Unfortunately, these datasets currently face severe risks of unauthorizโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

VOW: Verifiable and Oblivious Watermark Detection for Large Language Models

Xiaokun Luan, Yihao Zhang, Pengcheng Su, Feiran Lei, Meng Sun ยท 2026

Large Language Model (LLM) watermarking is crucial for establishing the provenance of machine-generated text, but most existing methods rely on a centralized trust model. This model forces users to reโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Back to the Future: Rethinking Endorsement in Order-Execute Blockchains

Rongji Huang, Yifeng Ye, Gerui Wang, Mingchao Wan, Yuxing Duan, Jingjing Zhang, Guangtao Xue, Shengyun Liu ยท 2026

Due to regulatory compliance and governance management, modern (permissioned) blockchains require flexible endorsement, which allows the endorsement policy for each contract or state object to be indiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

An Exact 56-Addition, Rank-23 Scheme for General 3*3 Matrix Multiplication

Yinqi Sun ยท 2026

We present a rank-$23$ algorithm for general $3\times3$ matrix multiplication that uses $56$ additions/subtractions and $23$ multiplications, for a total of $79$ scalar operations in the standard biliโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

SecGoal: A Benchmark for Security Goal Extraction and Formalization from Protocol Documents

Dawei Huang, Hui Li, Haonan Feng, Jingjing Guan, Yueshuang Jiao, Bo Jia (Beijing University of Posts, Telecommunications) ยท 2026

Formal verification provides rigorous guarantees for cryptographic security, yet automating the extraction and formalization of security goals from natural language protocol documents remains a major โ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Thinking like a business: Reconfiguring relationships to sustain open data infrastructures

Kathleen Gregory, Dorothea Strecker ยท 2026

Sustaining open data infrastructures over time is a complex puzzle, involving dynamic funding models and relationships with customers, collaborators, and competitors. Despite their importance, these mโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

I'm Fine, But My Voice Isn't: Cross-Modal Affective Dissonance Detection for Reflective Journaling

Sumin Lee ยท 2026

Digital journaling creates an authenticity gap: users consciously translate raw emotions into text, often sanitizing narratives even in private writing. We formalize this as Cross-Modal Affective Dissโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

ScaleBox: Enabling High-Fidelity and Scalable Code Verification for Large Language Models

Jiasheng Zheng, Xin Zheng, Boxi Cao, Pengbo Wang, Zhengzhao Ma, Qiming Zhu, Jiazhen Jiang, Yaojie Lu, Hongyu Lin, Xianpei Han, Le Sun ยท 2026

Code sandboxes have emerged as a critical infrastructure for advancing the coding capabilities of large language models, providing verifiable feedback for both RL training and evaluation. However, exiโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

Luyao Xu, Xiang Chen ยท 2026

Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, introducing security risks beyond traditional promptโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

ReVo: A Cross-Layer Reliable Volumetric Videoconferencing System

Ankur Aditya, Diptyaroop Maji, Lingdong Wang, Bhavya Ramakrishna, Ramesh Sitaraman, Prashant Shenoy ยท 2026

Volumetric videoconferencing enables immersive six Degrees of Freedom interactions by jointly transmitting visual appearance and 3D geometry. However, delivering volumetric video over today's networksโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A Study on the Performance of Distributed Training of Data-driven CFD Simulations

Sergio Iserte, Alejandro Gonzalez-Barbera, Paloma Barreda, Krzysztof Rojek ยท 2026

Data-driven methods for computer simulations are blooming in many scientific areas. The traditional approach to simulating physical behaviors relies on solving partial differential equations (PDE). Siโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

A Reproducibility Study of LLM-Based Query Reformulation

Amin Bigdeli, Radin Hamidi Rad, Hai Son Le, Mert Incesu, Negar Arabzadeh, Charles L. A. Clarke, Ebrahim Bagheri ยท 2026

Large Language Models (LLMs) are now widely used for query reformulation and expansion in Information Retrieval, with many studies reporting substantial effectiveness gains. However, these results areโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

Exploring the Adoption Intention in Using AI-Enabled Educational Tools Among Preservice Teachers in the Philippines: A Partial-Least Square Modeling

Vanessa B. Sibug, Emerson Q. Fernando, Almer B. Gamboa, Roque Francis B. Dianelo, Agnes R. Regala, Joseph Alexander Bansil, Jan Henry B. Sunga, Vernon Grace M. Maniago, John Paul P. Miranda ยท 2026

This study examines the factors influencing pre-service teachers' behavioral intention to use AI-enabled educational tools during their practicum, using the Unified Theory of Acceptance and Use of Tecโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

One Size Fits All? An Empirical Comparison of ADR Templates regarding Comprehension, Usability, and Ease of Adoption

Fernando Nogueira, Nabson Silva, Tayana Conte ยท 2026

Context: Documenting Architectural Design Decisions (ADDs) is a critical factor in the software lifecycle, essential for efficient system maintenance, developer onboarding, and preventing knowledge vaโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

SQuadGen: Generating Simple Quad Layouts via Chart Distance Fields

Youkang Kong, Yang Liu, Yue Dong, Xin Tong, Heung-Yeung Shum ยท 2026

3D shapes from scanning, reconstruction, or AI-generated content often lack simple quad mesh layouts -- critical for efficient editing and modeling. Existing quad-remeshing techniques typically producโ€ฆ

Read Paper โ†’
Computer Science Preprint PDF DOI

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

Jun Yeon Won, Xin Jin, Shiqing Ma, Zhiqiang Lin ยท 2026

Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including computer security. In reverse engineering, LLMs are incโ€ฆ

Read Paper โ†’
โ† Prev Page 2 of 2858 Next โ†’