Expertini Research Research

Browse Research Papers

2,218+ open-access research outputs.

โœ• Clear
๐Ÿ” continuing ๐Ÿ“‚ Engineering
Showing 2218 results for "continuing" in Engineering
Engineering Preprint PDF DOI

Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos

Qixiu Li, Yu Deng, Yaobo Liang, Lin Luo, Lei Zhou, Chengtang Yao, Lingqi Zeng, Zhiyuan Feng, Huizhi Liang, Sicheng Xu, Yizhong Zhang, Xi Chen, Hao Chen, Lily Sun, Dong Chen, Jiaolong Yang, Baining Guo ยท 2025

This paper presents a novel approach for pretraining robotic manipulation Vision-Language-Action (VLA) models using a large corpus of unscripted real-life video recordings of human hand activities. Trโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Configuration-Dependent Robot Kinematics Model and Calibration

Chen-Lung Lu, Honglu He, Agung Julius, John T. Wen ยท 2025

Accurate robot kinematics is essential for precise tool placement in articulated robots, but non-geometric factors can introduce configuration-dependent model discrepancies. This paper presents a confโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping

Zhirui Dai, Qihao Qian, Tianxing Fan, Nikolay Atanasov ยท 2025

Reconstructing signed distance functions (SDFs) from point cloud data benefits many robot autonomy capabilities, including localization, mapping, motion planning, and control. Methods that support onlโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Audio dequantization using instantaneous frequency

Vojtech Kovanda, Pavel Rajmic ยท 2025

We present a dequantization method that employs a phase-aware regularizer, originally successfully applied in an audio inpainting problem. The method promotes a temporal continuity of sinusoidal compoโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Audio-Visual Speech Enhancement for Spatial Audio - Spatial-VisualVoice and the MAVE Database

Danielle Yaffe, Ferdinand Campe, Prachi Sharma, Dorothea Kolossa, Boaz Rafaely ยท 2025

Audio-visual speech enhancement (AVSE) has been found to be particularly useful at low signal-to-noise (SNR) ratios due to the immunity of the visual features to acoustic noise. However, a significantโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Manual2Skill++: Connector-Aware General Robotic Assembly from Instruction Manuals via Vision-Language Models

Chenrui Tie, Shengxiang Sun, Yudi Lin, Yanbo Wang, Zhongrui Li, Zhouhan Zhong, Jinxuan Zhu, Yiman Pang, Haonan Chen, Junting Chen, Ruihai Wu, Lin Shao ยท 2025

Assembly hinges on reliably forming connections between parts; yet most robotic approaches plan assembly sequences and part poses while treating connectors as an afterthought. Connections represent thโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Fast, Differentiable, GPU-Accelerated Ray Tracing for Multiple Diffraction and Reflection Paths

Jerome Eertmans, Sophie Lequeu, Benoit Legat, Laurent Jacques, Claude Oestges ยท 2025

We present a fast, differentiable, GPU-accelerated optimization method for ray path tracing in environments containing planar reflectors and straight diffraction edges. Based on Fermat's principle, ouโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

DexCanvas: Bridging Human Demonstrations and Robot Learning for Dexterous Manipulation

Xinyue Xu, Jieqiang Sun, Jing (Daisy) Dai, Siyuan Chen, Lanjie Ma, Ke Sun, Bin Zhao, Jianbo Yuan, Sheng Yi, Haohua Zhu, Yiwen Lu ยท 2025

We present DexCanvas, a large-scale hybrid real-synthetic human manipulation dataset containing 7,000 hours of dexterous hand-object interactions seeded from 70 hours of real human demonstrations, orgโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Real-Time Knee Angle Prediction Using EMG and Kinematic Data with an Attention-Based CNN-LSTM Network and Transfer Learning Across Multiple Datasets

Mojtaba Mollahossein, Gholamreza Vossoughi, Mohammad Hossein Rohban ยท 2025

Electromyography (EMG) signals are widely used for predicting body joint angles through machine learning (ML) and deep learning (DL) methods. However, these approaches often face challenges such as liโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Decoding non-invasive brain activity with novel deep-learning approaches

Richard Csaky ยท 2025

This thesis delves into the world of non-invasive electrophysiological brain signals like electroencephalography (EEG) and magnetoencephalography (MEG), focusing on modelling and decoding such data. Tโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Robust Recovery and Control of Cyber-physical Discrete Event Systems under Actuator Attacks

Samuel Oliveira, Mostafa Tavakkoli Anbarani, Gregory Beal, Ilya Kovalenko, Marcelo Teixeira, Andre B. Leal, Romulo Meira-Goes ยท 2025

Critical real-world applications strongly rely on Cyber-physical systems (CPS), but their dependence on communication networks introduces significant security risks, as attackers can exploit vulnerabiโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

A Dynamic Watermarking Technique for Matching Communication Addresses with Cars in a Visual Field

Woo-Hyun Ko, Jaewon Kim, Tzu-Hsiang Lin, Samin Moosavi, P. R. Kumar ยท 2025

We consider a problem faced by an intelligent roadside unit (RSU) monitoring a roadway by a video camera. Suppose the RSU notices that a particular car in its visual field needs to execute a specific โ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Coherent Load Profile Synthesis with Conditional Diffusion for LV Distribution Network Scenario Generation

Alistair Brash, Junyi Lu, Bruce Stephen, Blair Brown, Robert Atkinson, Craig Michie, Fraser MacIntyre, Christos Tachtatzis ยท 2025

Limited visibility of distribution network power flows at the low voltage level presents challenges to both distribution network operators from a planning perspective and distribution system operatorsโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

SAM2-3dMed: Empowering SAM2 for 3D Medical Image Segmentation

Yeqing Yang, Le Xu, Lixia Tian ยท 2025

Accurate segmentation of 3D medical images is critical for clinical applications like disease assessment and treatment planning. While the Segment Anything Model 2 (SAM2) has shown remarkable success โ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Attitude and Heading Estimation in Symmetrical Inertial Arrays

Yaakov Libero, Itzik Klein ยท 2025

Attitude and heading reference systems (AHRS) play a central role in autonomous navigation systems on land, air and maritime platforms. AHRS utilize inertial sensor measurements to estimate platform oโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Pseudo2Real: Task Arithmetic for Pseudo-Label Correction in Automatic Speech Recognition

Yi-Cheng Lin, Yu-Hsuan Li Liang, Hsuan Su, Tzu-Quan Lin, Shang-Tse Chen, Yun-Nung Chen, Hung-yi Lee ยท 2025

Robust ASR under domain shift is crucial because real-world systems encounter unseen accents and domains with limited labeled data. Although pseudo-labeling offers a practical workaround, it often intโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Real-Time Glass Detection and Reprojection using Sensor Fusion Onboard Aerial Robots

Malakhi Hopkins, Varun Murali, Vijay Kumar, Camillo J Taylor ยท 2025

Autonomous aerial robots are increasingly being deployed in real-world scenarios, where transparent obstacles present significant challenges to reliable navigation and mapping. These materials pose a โ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

A Trustworthy Industrial Fault Diagnosis Architecture Integrating Probabilistic Models and Large Language Models

Yue wu ยท 2025

There are limitations of traditional methods and deep learning methods in terms of interpretability, generalization, and quantification of uncertainty in industrial fault diagnosis, and there are coreโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

Scaling Multi-Talker ASR with Speaker-Agnostic Activity Streams

Xiluo He, Alexander Polok, Jesus Villalba, Thomas Thebaud, Matthew Maciejewski ยท 2025

An increasingly common training paradigm for multi-talker automatic speech recognition (ASR) is to use speaker activity signals to adapt single-speaker ASR models for overlapping speech. Although effeโ€ฆ

Read Paper โ†’
Engineering Preprint PDF DOI

CVSM: Contrastive Vocal Similarity Modeling

Christos Garoufis, Athanasia Zlatintsi, Petros Maragos ยท 2025

The availability of large, unlabeled datasets across various domains has contributed to the development of a plethora of methods that learn representations for multiple target (downstream) tasks throuโ€ฆ

Read Paper โ†’
โ† Prev Page 13 of 111 Next โ†’