Skip to main content
What types of page to search?

Alternatively use our A-Z index.

Dr Anh Nguyen

Reader in Artificial Intelligence and Robotics
Artificial Intelligence

Research outputs

What type of research output do you want to show?

2026

An Object Detection Benchmark of a Novel Byzantine Iconography Dataset

Theofilis, N., Nguyen, A., Payne, T. R., & Tamma, V. (2026). An Object Detection Benchmark of a Novel Byzantine Iconography Dataset. Journal on Computing and Cultural Heritage. doi:10.1145/3842383

DOI
10.1145/3842383
Journal article

Burn depth assessment by photoacoustic imaging: A review

Zhou, J., Liang, S., Park, J. H., Nguyen, A., & Seong, M. (2026). Burn depth assessment by photoacoustic imaging: A review. Methods, 252, 65-82. doi:10.1016/j.ymeth.2026.04.014

DOI
10.1016/j.ymeth.2026.04.014
Journal article

VeloRM: disentangling pre- and post-splicing RNA modification dynamics at single-cell resolution

Wang, H., Song, B., Wu, Z., Zhang, Y., Li, J., Zhang, Y., . . . Meng, J. (2026). VeloRM: disentangling pre- and post-splicing RNA modification dynamics at single-cell resolution. NUCLEIC ACIDS RESEARCH, 54(12), 18 pages. doi:10.1093/nar/gkag645

DOI
10.1093/nar/gkag645
Journal article

Stochastic semantic segmentation with Stochastic Mixture of Bottleneck Experts

Liew, Y. Z., Tan, A. H. P., Majeed, A. P. P. A., Nguyen, A., Paoletti, P., & Chen, W. (2026). Stochastic semantic segmentation with Stochastic Mixture of Bottleneck Experts. ARRAY, 30, 10 pages. doi:10.1016/j.array.2026.100752

DOI
10.1016/j.array.2026.100752
Journal article

TAVR-VLM: Risk-Conditioned Causal Grounding for Hallucination-Resistant Report Generation

DOI
10.48550/arxiv.2606.26874
Preprint

MA-KANet: Enhancing retinal vessel segmentation through multi-scale feature fusion and attention mechanisms

Li, G., Zhang, F., Ateeq, M., Nguyen, A., & Majeed, A. P. P. A. (2026). MA-KANet: Enhancing retinal vessel segmentation through multi-scale feature fusion and attention mechanisms. ICT Express, 12(3), 553-558. doi:10.1016/j.icte.2026.02.006

DOI
10.1016/j.icte.2026.02.006
Journal article

Design of a Magnetic Dual-Needle Biopsy Capsule Robot Resilient to Cross-Infection and Mixed Contamination

Ye, B., Chen, Z., Wang, S., Liu, S., Wang, B., Shu, Z., . . . Liu, S. (2026). Design of a Magnetic Dual-Needle Biopsy Capsule Robot Resilient to Cross-Infection and Mixed Contamination. IEEE TRANSACTIONS ON MEDICAL ROBOTICS AND BIONICS, 8(2), 894-905. doi:10.1109/TMRB.2026.3680996

DOI
10.1109/TMRB.2026.3680996
Journal article

SAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast Asia

Lu, Z., Zhang, C., Li, Y., Stefanidis, A., Nguyen, A., Razzak, I., . . . Jiang, Z. (2026). SAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast Asia. In Www 2026 Proceedings of the ACM Web Conference 2026 (pp. 9101-9112). doi:10.1145/3774904.3793041

DOI
10.1145/3774904.3793041
Conference Paper

SkinCLIP-VL: Consistency-Aware Vision-Language Learning for Multimodal Skin Cancer Diagnosis

DOI
10.48550/arxiv.2603.21010
Preprint

SAGE: Sustainable Agent-Guided Expert-tuning for Culturally Attuned Translation in Low-Resource Southeast Asia

DOI
10.48550/arxiv.2603.19931
Preprint

Enhanced Prediction of Post-Myocardial Infarction Complications: Dual-Modality Analysis with Optimized Flow Cytometry Preprocessing and Feature Visualization

Al-Dausari, N., Coenen, F., Nguyen, A., & Shantsila, E. (2026). Enhanced Prediction of Post-Myocardial Infarction Complications: Dual-Modality Analysis with Optimized Flow Cytometry Preprocessing and Feature Visualization. In Unknown Book (Vol. 2703 CCIS, pp. 93-105). doi:10.1007/978-3-032-06878-1_5

DOI
10.1007/978-3-032-06878-1_5
Chapter

MonoPredict-MI: Predicting Post-Myocardial Infarction Complications from Flow Cytometry–Derived Monocyte Subsets Using Neural Networks

Al-Dausari, N., Coenen, F., Nguyen, A., & Shantsila, E. (2026). MonoPredict-MI: Predicting Post-Myocardial Infarction Complications from Flow Cytometry–Derived Monocyte Subsets Using Neural Networks. In Proceedings of the 19th International Joint Conference on Biomedical Engineering Systems and Technologies (pp. 463-474). INSTICC. doi:10.5220/0014247900004070

DOI
10.5220/0014247900004070
Conference Paper

Out-of-Distribution Detection in Gastrointestinal Vision by Estimating Nearest Centroid Distance Deficit

Pokhrel, S., Bhandari, S., Ali, S., Lambrou, T., Anh, N., Shrestha, Y. R., . . . Bhattarai, B. (2026). Out-of-Distribution Detection in Gastrointestinal Vision by Estimating Nearest Centroid Distance Deficit. In Unknown Book (Vol. 15917, pp. 190-200). doi:10.1007/978-3-031-98691-8_14

DOI
10.1007/978-3-031-98691-8_14
Chapter

2025

Learning Human Motion with Temporally Conditional Mamba

Nguyen, Q., Le, T., Huang, B., Vu, M. N., Le, N., Vo, T., & Nguyen, A. (2025). Learning Human Motion with Temporally Conditional Mamba. In Proceedings SIGGRAPH Asia 2025 Conference Papers SA 2025. doi:10.1145/3757377.3763948

DOI
10.1145/3757377.3763948
Conference Paper

An explainable approach to learning feature representations of omics data using optimal transport to improve phenotype stratification

Gadhia, N. A., Ghosh, S., Krishna, R., Nguyen, A., & Smyrnakis, M. (2025). An explainable approach to learning feature representations of omics data using optimal transport to improve phenotype stratification. In Bcb 2025 Proceedings of the 16th ACM International Conference on Bioinformatics Computational Biology and Health Informatics. doi:10.1145/3765612.3767214

DOI
10.1145/3765612.3767214
Conference Paper

BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance

Le, H., Chung, N., Kieu, T., Nguyen, A., & Le, N. (2025). BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance. In Mm 2025 Proceedings of the 33rd ACM International Conference on Multimedia Co Located with mm 2025 (pp. 2831-2840). doi:10.1145/3746027.3754833

DOI
10.1145/3746027.3754833
Conference Paper

Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis

Pham, T. T., Nguyen, A., Deng, Z., Wu, C. C., Nguyen, H., & Le, N. (2025). Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis. In Mm 2025 Proceedings of the 33rd ACM International Conference on Multimedia Co Located with mm 2025 (pp. 257-266). doi:10.1145/3746027.3755039

DOI
10.1145/3746027.3755039
Conference Paper

GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning

Nguyen, H. H., Vuong, A., Nguyen, A., Reid, I., & Vu, M. N. (2025). GraspMamba: A Mamba-based Language-driven Grasp Detection Framework with Hierarchical Feature Learning. In 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) Vol. 00 (pp. 15808-15815). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/iros60139.2025.11245817

DOI
10.1109/iros60139.2025.11245817
Conference Paper

Modeling The States of Liquid Phase Change Pouch Actuators by Reservoir Computing

Caremel, C., Nguyen, K., Nguyen, A., Huber, M., Kawahara, Y., & Ta, T. D. (2025). Modeling The States of Liquid Phase Change Pouch Actuators by Reservoir Computing. In 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) Vol. 00 (pp. 3601-3608). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/iros60139.2025.11247242

DOI
10.1109/iros60139.2025.11247242
Conference Paper

More Reliable Pseudo-Labels, Better Performance: A Generalized Approach to Single Positive Multi-Label Learning

Tran, L., Vo, T., Nguyen, A., Dinh, S., & Nguyen, V. (2025). More Reliable Pseudo-Labels, Better Performance: A Generalized Approach to Single Positive Multi-Label Learning. In 2025 IEEE/CVF International Conference on Computer Vision (ICCV) Vol. 00 (pp. 1349-1358). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/iccv51701.2025.00133

DOI
10.1109/iccv51701.2025.00133
Conference Paper

SplineFormer: An Explainable Transformer Network for Autonomous Endovascular Navigation

Jianu, T., Doust, S., Li, M., Huang, B., Tuong, D., Hoan, N., . . . Anh, N. (2025). SplineFormer: An Explainable Transformer Network for Autonomous Endovascular Navigation. In 2025 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS) (pp. 7718-7725). doi:10.1109/IROS60139.2025.11246898

DOI
10.1109/IROS60139.2025.11246898
Conference Paper

Towards a Universal 3D Medical Multi-Modality Generalization via Learning Personalized Invariant Representation

Tan, Z., Yang, X., Pan, T., Liu, T., Jiang, C., Guo, X., . . . Cheng, Y. (2025). Towards a Universal 3D Medical Multi-Modality Generalization via Learning Personalized Invariant Representation. In 2025 IEEE/CVF International Conference on Computer Vision (ICCV) Vol. 00 (pp. 21895-21905). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/iccv51701.2025.02033

DOI
10.1109/iccv51701.2025.02033
Conference Paper

NUMINA: A Natural Understanding Benchmark for Multi-dimensional Intelligence and Numerical Reasoning Abilities

DOI
10.48550/arxiv.2509.16656
Preprint

Hybrid Gripper with Passive Pneumatic Soft Joints for Grasping Deformable Thin Objects

Tran, N. -D., Ly, H. -H., Nguyen, X. -T., Mac, T. -T., Nguyen, A., & Ta, T. D. (2025). Hybrid Gripper with Passive Pneumatic Soft Joints for Grasping Deformable Thin Objects. In 2025 IEEE International Conference on Robotics and Automation (ICRA) Vol. 00 (pp. 7858-7864). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/icra55743.2025.11127915

DOI
10.1109/icra55743.2025.11127915
Conference Paper

Online Trajectory Replanner for Dynamically Grasping Irregular Objects

Vu, M. N., Grander, F., Nguyen, A., & Unger, C. (2025). Online Trajectory Replanner for Dynamically Grasping Irregular Objects. In 2025 IEEE International Conference on Robotics and Automation (ICRA) Vol. 00 (pp. 7975-7981). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/icra55743.2025.11128477

DOI
10.1109/icra55743.2025.11128477
Conference Paper

A Novel Approach to Differential Expression Analysis of Co-Occurrence Networks for Small-Sampled Microbiome Data

Gadhia, N., Smyrnakis, M., Liu, P. -Y., Blake, D., Hay, M., Nguyen, A., . . . Krishna, R. (2025). A Novel Approach to Differential Expression Analysis of Co-Occurrence Networks for Small-Sampled Microbiome Data. IEEE TRANSACTIONS ON COMPUTATIONAL BIOLOGY AND BIOINFORMATICS, 22(4), 1311-1322. doi:10.1109/TCBBIO.2025.3554740

DOI
10.1109/TCBBIO.2025.3554740
Journal article

Adaptive Parametric Activation

Alexandridis, K. P., Deng, J., Nguyen, A., & Luo, S. (2025). Adaptive Parametric Activation. In Unknown Book (Vol. 15112 LNCS, pp. 455-476). doi:10.1007/978-3-031-72949-2_26

DOI
10.1007/978-3-031-72949-2_26
Chapter

Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?

Xie, J., Zhao, Y., Meng, Y., Zhao, H., Nguyen, A., & Zheng, Y. (2025). Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?. In 2025 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) (pp. 24309-24319). doi:10.1109/CVPR52734.2025.02264

DOI
10.1109/CVPR52734.2025.02264
Conference Paper

Attention-SSM Network for Predicting Springback Error in Single Point Incremental Forming

Chen, D., Oscoz, M. P., Martin Rebe, A., Hai, Y., Coenen, F., & Nguyen, A. (2025). Attention-SSM Network for Predicting Springback Error in Single Point Incremental Forming. In Proceedings of the International Joint Conference on Neural Networks. doi:10.1109/IJCNN64981.2025.11228939

DOI
10.1109/IJCNN64981.2025.11228939
Conference Paper

Autonomous Catheterization With Open-Source Simulator and Expert Trajectory

Jianu, T., Huang, B., Van Vo, T., Nhat Vu, M., Kang, J., Nguyen, H. C., . . . Nguyen, A. (2025). Autonomous Catheterization With Open-Source Simulator and Expert Trajectory. In Handbook of Robotic and Image Guided Surgery (pp. 131-148). doi:10.1016/B978-0-443-13912-3.00054-3

DOI
10.1016/B978-0-443-13912-3.00054-3
Chapter

CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling

Pham, T. T., Awasthi, A., Khan, S., Marti, E. D., Nguyen, T. P., Vo, K., . . . Le, N. (2025). CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling. In Proceedings of the IEEE International Conference on Computer Vision (pp. 21732-21743). doi:10.1109/ICCV51701.2025.02018

DOI
10.1109/ICCV51701.2025.02018
Conference Paper

EgoMusic-Driven Human Dance Motion Estimation with Skeleton Mamba

Nguyen, Q., Le, N., Huang, B., Vu, M. N., Tang, C., Nguyen, V., . . . Nguyen, A. (2025). EgoMusic-Driven Human Dance Motion Estimation with Skeleton Mamba. In Proceedings of the IEEE International Conference on Computer Vision (pp. 12023-12033). doi:10.1109/ICCV51701.2025.01118

DOI
10.1109/ICCV51701.2025.01118
Conference Paper

FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation

Pham, T. T., Ho, N. V., Bui, N. T., Phan, T., Brijesh, P., Adjeroh, D., . . . Le, N. (2025). FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation. In Unknown Book (Vol. 15477 LNCS, pp. 71-88). doi:10.1007/978-981-96-0960-4_5

DOI
10.1007/978-981-96-0960-4_5
Chapter

FedPGD: Federated Learning with Projected Gradient Descent for Catheter and Guidewire Segmentation

Kongtongvattana, C., Huang, B., Nguyen, H., Olajide, O., & Nguyen, A. (2025). FedPGD: Federated Learning with Projected Gradient Descent for Catheter and Guidewire Segmentation. In Unknown Book (Vol. 1419 LNNS, pp. 80-91). doi:10.1007/978-3-031-92011-0_7

DOI
10.1007/978-3-031-92011-0_7
Chapter

FlowMI-HybridNet: Integrating Data Handling and Neural Network for Enhanced Post-Myocardial Infarction Complication Prediction

Al-Dausari, N., Coenen, F., Nguyen, A., & Shantsila, E. (2025). FlowMI-HybridNet: Integrating Data Handling and Neural Network for Enhanced Post-Myocardial Infarction Complication Prediction. In Unknown Book (Vol. 1537 LNNS, pp. 327-341). doi:10.1007/978-981-96-9242-2_24

DOI
10.1007/978-981-96-9242-2_24
Chapter

Fractal Calibration for long-tailed object detection

Alexandridis, K. P., Elezi, I., Deng, J., Nguyen, A., & Luo, S. (2025). Fractal Calibration for long-tailed object detection. In Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition (pp. 15139-15150). doi:10.1109/CVPR52734.2025.01410

DOI
10.1109/CVPR52734.2025.01410
Conference Paper

GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System

Nguyen, Q., Le, T., Nguyen, H., Vo, T., Ta, T. D., Huang, B., . . . Nguyen, A. (2025). GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System. In 2025 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS) (pp. 14939-14946). doi:10.1109/IROS60139.2025.11246582

DOI
10.1109/IROS60139.2025.11246582
Conference Paper

Guide3D: A Bi-planar X-ray Dataset for 3D Shape Reconstruction

Jianu, T., Huang, B., Nguyen, H., Bhattarai, B., Do, T., Tjiputra, E., . . . Nguyen, A. (2025). Guide3D: A Bi-planar X-ray Dataset for 3D Shape Reconstruction. In Unknown Book (Vol. 15476, pp. 366-382). doi:10.1007/978-981-96-0917-8_21

DOI
10.1007/978-981-96-0917-8_21
Chapter

LP-Diff: Towards Improved Restoration of Real-World Degraded License Plate

Gong, H., Zhang, Z., Feng, Y., Anh, N., & Liu, H. (2025). LP-Diff: Towards Improved Restoration of Real-World Degraded License Plate. In 2025 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) (pp. 17831-17840). doi:10.1109/CVPR52734.2025.01661

DOI
10.1109/CVPR52734.2025.01661
Conference Paper

Lightweight Temporal Transformer Decomposition for Federated Autonomous Driving

Tuong, D., Nguyen, B. X., Tran, Q. D., Tjiputra, E., Chiu, T. -C., & Anh, N. (2025). Lightweight Temporal Transformer Decomposition for Federated Autonomous Driving. In 2025 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS) (pp. 12472-12479). doi:10.1109/IROS60139.2025.11246435

DOI
10.1109/IROS60139.2025.11246435
Conference Paper

NUMINA: A Natural Understanding Benchmark for Multi-dimensional Intelligence and Numerical Reasoning Abilities

Zeng, C., Wang, Y., Wang, Z., Wang, W., Yang, Z., Bao, M., . . . Yue, Y. (2025). NUMINA: A Natural Understanding Benchmark for Multi-dimensional Intelligence and Numerical Reasoning Abilities. In Emnlp 2025 2025 Conference on Empirical Methods in Natural Language Processing Findings of Emnlp 2025 (pp. 22575-22590). doi:10.18653/v1/2025.findings-emnlp.1229

DOI
10.18653/v1/2025.findings-emnlp.1229
Conference Paper

OVITA: Open-Vocabulary Interpretable Trajectory Adaptations

Maurya, A., Ghosh, T., Nguyen, A., & Prakash, R. (2025). OVITA: Open-Vocabulary Interpretable Trajectory Adaptations. IEEE ROBOTICS AND AUTOMATION LETTERS, 10(11), 11054-11061. doi:10.1109/LRA.2025.3606309

DOI
10.1109/LRA.2025.3606309
Journal article

Scalable Group Choreography via Variational Phase Manifold Learning

Le, N., Le, K., Bui, X., Do, T., Tjiputra, E., Tran, Q. D., & Nguyen, A. (2025). Scalable Group Choreography via Variational Phase Manifold Learning. In Unknown Book (Vol. 15076, pp. 293-311). doi:10.1007/978-3-031-72649-1_17

DOI
10.1007/978-3-031-72649-1_17
Chapter

Sheet Metal Forming Springback Prediction Using Image Geometrics (SPIG): A Novel Approach Using Heatmaps and Convolutional Neural Network

Chen, D., Oscoz, M. P., Hai, Y., Ander, M. R., Coenen, F., & Nguyen, A. (2025). Sheet Metal Forming Springback Prediction Using Image Geometrics (SPIG): A Novel Approach Using Heatmaps and Convolutional Neural Network. In International Joint Conference on Knowledge Discovery Knowledge Engineering and Knowledge Management Ic3k Proceedings Vol. 1 (pp. 28-38). doi:10.5220/0013673100004000

DOI
10.5220/0013673100004000
Conference Paper

Soft Joints Grippers for Grasping Deformable Thin Objects

Tran, N. -D., Ly, H. -H., Nguyen, X. -T., Mac, T. -T., Nguyen, A., & Ta, T. D. (2025). Soft Joints Grippers for Grasping Deformable Thin Objects. The Proceedings of JSME annual Conference on Robotics and Mechatronics (Robomec), 2025(0), 2p1-r01. doi:10.1299/jsmermd.2025.2p1-r01

DOI
10.1299/jsmermd.2025.2p1-r01
Journal article

Statistical modeling of immunoprecipitation efficiency of MeRIP-seq data enabled accurate detection and quantification of epitranscriptome

Wang, H., Chen, K., Wei, Z., Song, B., Zhu, M., Su, J., . . . Wang, Y. (2025). Statistical modeling of immunoprecipitation efficiency of MeRIP-seq data enabled accurate detection and quantification of epitranscriptome. COMPUTATIONAL AND STRUCTURAL BIOTECHNOLOGY JOURNAL, 27, 3742-3752. doi:10.1016/j.csbj.2025.08.030

DOI
10.1016/j.csbj.2025.08.030
Journal article

Translating Simulation Images to X-Ray Images via Multi-scale Semantic Matching

Kang, J., Jianu, T., Huang, B., Bhattarai, B., Ngan, L., Coenen, F., & Anh, N. (2025). Translating Simulation Images to X-Ray Images via Multi-scale Semantic Matching. In Unknown Book (Vol. 15265, pp. 95-104). doi:10.1007/978-3-031-73748-0_10

DOI
10.1007/978-3-031-73748-0_10
Chapter

Weakly-Supervised Learning via Multi-Lateral Decoder Branching for Tool Segmentation in Robot-Assisted Cardiovascular Catheterization

Omisore, O. M., Akinyemi, T., Anh, N., & Wang, L. (2025). Weakly-Supervised Learning via Multi-Lateral Decoder Branching for Tool Segmentation in Robot-Assisted Cardiovascular Catheterization. In 2025 IEEE INTERNATIONAL CONFERENCE ON ROBOTICS AND AUTOMATION (ICRA) (pp. 9325-9331). doi:10.1109/ICRA55743.2025.11128773

DOI
10.1109/ICRA55743.2025.11128773
Conference Paper

2024

A novel approach to differential expression analysis of co-occurrence networks for small-sampled microbiome data

DOI
10.48550/arxiv.2412.03744
Preprint

ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation

Tran, M., Bounsavy, W., Vo, K., Nguyen, A., Nguyen, T., & Le, N. (2024). ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation. In 2024 International Joint Conference on Neural Networks (IJCNN) Vol. 00 (pp. 1-8). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/ijcnn60899.2024.10650837

DOI
10.1109/ijcnn60899.2024.10650837
Conference Paper

CLIP-DR: Textual Knowledge-Guided Diabetic Retinopathy Grading with Ranking-aware Prompting

DOI
10.48550/arxiv.2407.04068
Preprint

Spring Back Prediction Using Point Series and Deep Learning

Bingqian, Y., Zeng, Y., Yang, H., Penalva Oscoz, M., Ortiz, M., Coenen, F., & Nguyen, A. (n.d.). Spring Back Prediction Using Point Series and Deep Learning. International Journal of Advanced Manufacturing Technology.

Journal article

Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation

Yamazaki, K., Hanyu, T., Vo, K., Pham, T., Tran, M., Doretto, G., . . . Le, N. (2024). Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation. In 2024 IEEE International Conference on Robotics and Automation (ICRA) Vol. 00 (pp. 9411-9417). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/icra57147.2024.10610193

DOI
10.1109/icra57147.2024.10610193
Conference Paper

WAVER: Writing-Style Agnostic Text-Video Retrieval Via Distilling Vision-Language Models Through Open-Vocabulary Knowledge

Le, H., Kieu, T., Nguyen, A., & Le, N. (2024). WAVER: Writing-Style Agnostic Text-Video Retrieval Via Distilling Vision-Language Models Through Open-Vocabulary Knowledge. In ICASSP 2024 - 2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Vol. 00 (pp. 3025-3029). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/icassp48485.2024.10446193

DOI
10.1109/icassp48485.2024.10446193
Conference Paper

DKE-Research at SemEval-2024 Task 2: Incorporating Data Augmentation with Generative Models and Biomedical Knowledge to Enhance Inference Robustness

DOI
10.48550/arxiv.2404.09206
Preprint

I-AI: A Controllable & Interpretable AI System for Decoding Radiologists’ Intense Focus for Accurate CXR Diagnoses

Pham, T. T., Brecheisen, J., Nguyen, A., Nguyen, H., & Le, N. (2024). I-AI: A Controllable & Interpretable AI System for Decoding Radiologists’ Intense Focus for Accurate CXR Diagnoses. In 2024 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) Vol. 00 (pp. 7835-7844). Institute of Electrical and Electronics Engineers (IEEE). doi:10.1109/wacv57701.2024.00767

DOI
10.1109/wacv57701.2024.00767
Conference Paper

3D Guidewire Shape Reconstruction from Monoplane Fluoroscopic Image

Jianu, T., Huang, B., Berthet-Rayne, P., Fichera, S., & Nguyen, A. (2024). 3D Guidewire Shape Reconstruction from Monoplane Fluoroscopic Image. In Unknown Book (Vol. 1132, pp. 84-94). doi:10.1007/978-3-031-70684-4_7

DOI
10.1007/978-3-031-70684-4_7
Chapter

CLIP-DR: Textual Knowledge-Guided Diabetic Retinopathy Grading with Ranking-Aware Prompting

Yu, Q., Xie, J., Anh, N., Zhao, H., Zhang, J., Fu, H., . . . Meng, Y. (2024). CLIP-DR: Textual Knowledge-Guided Diabetic Retinopathy Grading with Ranking-Aware Prompting. In Unknown Book (Vol. 15001, pp. 667-677). doi:10.1007/978-3-031-72378-0_62

DOI
10.1007/978-3-031-72378-0_62
Chapter

DKE-Research at SemEval-2024 Task 2: Incorporating Data Augmentation with Generative Models and Biomedical Knowledge to Enhance Inference Robustness

Wang, Y., Wang, Z., Wang, W., Chen, Q., Huang, K., Anh, N., & De, S. (2024). DKE-Research at SemEval-2024 Task 2: Incorporating Data Augmentation with Generative Models and Biomedical Knowledge to Enhance Inference Robustness. In PROCEEDINGS OF THE 18TH INTERNATIONAL WORKSHOP ON SEMANTIC EVALUATION, SEMEVAL-2024 (pp. 88-94). Retrieved from https://www.webofscience.com/

Conference Paper

Domain-specific Guided Summarization for Mental Health Posts

Qian, L., Wang, Y., Wang, Z., Zhang, H., Wang, W., Yu, T., & Nguyen, A. (2024). Domain-specific Guided Summarization for Mental Health Posts. In Paclic Pacific Asia Conference on Language Information and Computation.

Conference Paper

Fine-Grained Visual Classification using Self Assessment Classifier

Do, T., Trani, H., Tjiputra, E., Tran, Q. D., & Anh, N. (2024). Fine-Grained Visual Classification using Self Assessment Classifier. In 2024 IEEE CONFERENCE ON ARTIFICIAL INTELLIGENCE, CAI 2024 (pp. 597-602). doi:10.1109/CAI59869.2024.00117

DOI
10.1109/CAI59869.2024.00117
Conference Paper

Generating Valid and Natural Adversarial Examples with Large Language Models

Wang, Z., Wang, W., Chen, Q., Wang, Q., & Anh, N. (2024). Generating Valid and Natural Adversarial Examples with Large Language Models. In PROCEEDINGS OF THE 2024 27 TH INTERNATIONAL CONFERENCE ON COMPUTER SUPPORTED COOPERATIVE WORK IN DESIGN, CSCWD 2024 (pp. 1716-1721). doi:10.1109/CSCWD61410.2024.10580402

DOI
10.1109/CSCWD61410.2024.10580402
Conference Paper

HabiCrowd: A High Performance Simulator for Crowd-Aware Visual Navigation

An, V., Toan, N., Minh, N. V., Huang, B., Binh, H. T. T., Thieu, V., & Anh, N. (2024). HabiCrowd: A High Performance Simulator for Crowd-Aware Visual Navigation. In 2024 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS, IROS 2024 (pp. 5821-5827). doi:10.1109/IROS58592.2024.10801823

DOI
10.1109/IROS58592.2024.10801823
Conference Paper

Knowledge Distillation from Monolingual to Multilingual Models for Intelligent and Interpretable Multilingual Emotion Detection

Wang, Y., Wang, Z., Han, N., Wang, W., Chen, Q., Zhang, H., . . . Nguyen, A. (2024). Knowledge Distillation from Monolingual to Multilingual Models for Intelligent and Interpretable Multilingual Emotion Detection. In Wassa 2024 14th Workshop on Computational Approaches to Subjectivity Sentiment and Social Media Analysis Proceedings of the Workshop (pp. 470-475).

Conference Paper

Language-driven Grasp Detection

An, D. V., Minh, N. V., Baoru, H., Nghia, N., Hieu, L., Thieu, V., & Anh, N. (2024). Language-driven Grasp Detection. In 2024 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) (pp. 17902-17912). doi:10.1109/CVPR52733.2024.01695

DOI
10.1109/CVPR52733.2024.01695
Conference Paper

Language-driven Grasp Detection with Mask-guided Attention

Nan, V. V., Minh, N. V., Huang, B., An, V., Ngan, L., Thieu, V., & Anh, N. (2024). Language-driven Grasp Detection with Mask-guided Attention. In 2024 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS, IROS 2024 (pp. 7492-7498). doi:10.1109/IROS58592.2024.10802256

DOI
10.1109/IROS58592.2024.10802256
Conference Paper

Lightweight Language-driven Grasp Detection using Conditional Consistency Model

Nghia, N., Minh, N. V., Huang, B., Vuong, A., Ngan, L., Thieu, V., & Anh, N. (2024). Lightweight Language-driven Grasp Detection using Conditional Consistency Model. In 2024 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS 2024) (pp. 13719-13725). doi:10.1109/IROS58592.2024.10802007

DOI
10.1109/IROS58592.2024.10802007
Conference Paper

Multi-disease Detection in Retinal Images Guided by Disease Causal Estimation

Xie, J., Chen, X., Zhao, Y., Meng, Y., Zhao, H., Anh, N., . . . Zheng, Y. (2024). Multi-disease Detection in Retinal Images Guided by Disease Causal Estimation. In Unknown Book (Vol. 15001, pp. 743-753). doi:10.1007/978-3-031-72378-0_69

DOI
10.1007/978-3-031-72378-0_69
Chapter

Predicting Post Myocardial Infarction Complication: A Study Using Dual-Modality and Imbalanced Flow Cytometry Data

Aldausari, N., Coenen, F., Nguyen, A., & Shantsila, E. (2024). Predicting Post Myocardial Infarction Complication: A Study Using Dual-Modality and Imbalanced Flow Cytometry Data. In International Joint Conference on Knowledge Discovery Knowledge Engineering and Knowledge Management Ic3k Proceedings Vol. 1 (pp. 81-90). doi:10.5220/0012998300003838

DOI
10.5220/0012998300003838
Conference Paper

Preface

Yap, E. H., Majeed, A. P. P. A., Chen, W., Liu, P., Huang, X., Nguyen, A., & Kim, U. H. (2024). Preface. Lecture Notes in Networks and Systems, 1132 LNNS, v-vi.

Journal article

Preface

Yap, E. H., Abdul Majeed, A. P. P., Chen, W., Liu, P., Huang, X., Nguyen, A., & Kim, U. H. (2024). Preface. Lecture Notes in Networks and Systems, 1133 LNNS.

Journal article

Shape-Sensitive Loss for Catheter and Guidewire Segmentation

Kongtongvattana, C., Huang, B., Kang, J., Nguyen, H., Olufemi, O., & Nguyen, A. (2024). Shape-Sensitive Loss for Catheter and Guidewire Segmentation. In Unknown Book (Vol. 1132, pp. 95-107). doi:10.1007/978-3-031-70684-4_8

DOI
10.1007/978-3-031-70684-4_8
Chapter

Style Transfer for 2D Talking Head Generation

Pham, T. T., Do, T., Le, N., Le, N., Nguyen, H., Tjiputra, E., . . . Nguyen, A. (2024). Style Transfer for 2D Talking Head Generation. In 2024 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION WORKSHOPS, CVPRW (pp. 7500-7509). doi:10.1109/CVPRW63382.2024.00745

DOI
10.1109/CVPRW63382.2024.00745
Conference Paper

2023

Formulating a method to analyse the differential expression of co-occurrence networks for small-sampled microbiome data

Gadhia, N., Smyrnakis, M., Liu, P. -Y., Blake, D., Hay, M., Nguyen, A., . . . Krishna, R. (2023). Formulating a method to analyse the differential expression of co-occurrence networks for small-sampled microbiome data. In 14TH ACM CONFERENCE ON BIOINFORMATICS, COMPUTATIONAL BIOLOGY, AND HEALTH INFORMATICS, BCB 2023 (pp. 6 pages). doi:10.1145/3584371.3612969

DOI
10.1145/3584371.3612969
Conference Paper

SimPS-Net: Simultaneous Pose & Segmentation Network of Surgical Tools

Souipas, S., Nguyen, A., Laws, S., Davies, B., & Rodriguez Y Baena, F. (2023). SimPS-Net: Simultaneous Pose & Segmentation Network of Surgical Tools. In Proceedings of The 15th Hamlyn Symposium on Medical Robotics 2023 (pp. 71-72). The Hamlyn Centre for Robotic Surgery. doi:10.31256/hsmr2023.36

DOI
10.31256/hsmr2023.36
Conference Paper

A Client-Server Deep Federated Learning for Cross-Domain Surgical Image Segmentation

Subedi, R., Gaire, R. R., Ali, S., Nguyen, A., Stoyanov, D., & Bhattarai, B. (2023). A Client-Server Deep Federated Learning for Cross-Domain Surgical Image Segmentation. In Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics Vol. 14314 LNCS (pp. 21-33). doi:10.1007/978-3-031-44992-5_3

DOI
10.1007/978-3-031-44992-5_3
Conference Paper

Language-driven Scene Synthesis using Multi-conditional Diffusion Model

An, D. V., Minh, N. V., Toan, T. N., Huang, B., Dzung, N., Thieu, V., & Anh, N. (2023). Language-driven Scene Synthesis using Multi-conditional Diffusion Model. In ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 36 (NEURIPS 2023) (pp. 23 pages). Retrieved from https://www.webofscience.com/

Conference Paper

Learning to Terminate in Object Navigation

Song, Y., Nguyen, A., & Lee, C. Y. (2023). Learning to Terminate in Object Navigation. In Proceedings of Machine Learning Research Vol. 222 (pp. 1247-1262).

Conference Paper

Music-Driven Group Choreography

Le, N., Pham, T., Do, T., Tjiputra, E., Tran, Q. D., & Nguyen, A. (2023). Music-Driven Group Choreography. In 2023 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) (pp. 8673-8682). doi:10.1109/CVPR52729.2023.00838

DOI
10.1109/CVPR52729.2023.00838
Conference Paper

Preface

Bhattarai, B., Rau, A., Namburete, A., Caramalau, R., Stoyanov, D., Ali, S., & Nguyen, A. (2023). Preface (Vol. 14314 LNCS).

Book

Self-Supervised Learning for Point Clouds through Multi-crop Mutual Prediction

Zeng, C., Wang, W., Xiao, J., Nguyen, A., & Yue, Y. (2023). Self-Supervised Learning for Point Clouds through Multi-crop Mutual Prediction. In 2023 IEEE 4th International Conference on Pattern Recognition and Machine Learning Prml 2023 (pp. 160-167). doi:10.1109/PRML59573.2023.10348385

DOI
10.1109/PRML59573.2023.10348385
Conference Paper

2022

Deep Federated Learning for Autonomous Driving

Anh, N., Tuong, D., Minh, T., Nguyen, B. X., Chien, D., Tu, P., . . . Tran, Q. D. (2022). Deep Federated Learning for Autonomous Driving. In 2022 IEEE INTELLIGENT VEHICLES SYMPOSIUM (IV) (pp. 1824-1830). doi:10.1109/IV51971.2022.9827020

DOI
10.1109/IV51971.2022.9827020
Conference Paper

Generalised Zero-shot Learning for Entailment-based Text Classification with Externa Knowledge

Wang, Y., Wang, W., Chen, Q., Huang, K., Anh, N., & De, S. (2022). Generalised Zero-shot Learning for Entailment-based Text Classification with Externa Knowledge. In 2022 IEEE INTERNATIONAL CONFERENCE ON SMART COMPUTING (SMARTCOMP 2022) (pp. 19-25). doi:10.1109/SMARTCOMP55677.2022.00018

DOI
10.1109/SMARTCOMP55677.2022.00018
Conference Paper

Long-Tailed Instance Segmentation Using Gumbel Optimized Loss

Alexandridis, K. P., Deng, J., Nguyen, A., & Luo, S. (2022). Long-Tailed Instance Segmentation Using Gumbel Optimized Loss. In Unknown Book (Vol. 13670, pp. 353-369). doi:10.1007/978-3-031-20080-9_21

DOI
10.1007/978-3-031-20080-9_21
Chapter

2021

Graph-based Person Signature for Person Re-Identifications

Nguyen, B. X., Nguyen, B. D., Do, T., Tjiputra, E., Tran, Q. D., & Nguyen, A. (2021). Graph-based Person Signature for Person Re-Identifications. In 2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION WORKSHOPS, CVPRW 2021 (pp. 3487-3496). doi:10.1109/CVPRW53098.2021.00388

DOI
10.1109/CVPRW53098.2021.00388
Conference Paper

Autonomous Navigation with Mobile Robots Using Deep Learning and the Robot Operating System

Nguyen, A., & Tran, Q. D. (2021). Autonomous Navigation with Mobile Robots Using Deep Learning and the Robot Operating System. In Studies in Computational Intelligence (Vol. 962, pp. 177-195). doi:10.1007/978-3-030-75472-3_5

DOI
10.1007/978-3-030-75472-3_5
Chapter

Self-supervised generative adversarial network for depth estimation in laparoscopic images

Huang, B., Zheng, J. -Q., Nguyen, A., Tuch, D., Vyas, K., Giannarou, S., & Elson, D. S. (2021). Self-supervised generative adversarial network for depth estimation in laparoscopic images. In International Conference on Medical Image Computing and Computer-Assisted Intervention (pp. 227-237). Springer, Cham.

Conference Paper

2020

Autonomous Navigation in Complex Environments with Deep Multimodal Fusion Network

Nguyen, A., Nguyen, N., Tran, K., Tjiputra, E., & Tran, Q. D. (2020). Autonomous Navigation in Complex Environments with Deep Multimodal Fusion Network. In IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2020.

Conference Paper

BeetleBot: A Multi-Purpose AI-Driven Mobile Robot for Realistic Environments

Nguyen, A., Tjiputra, E., & Tran, Q. D. (2020). BeetleBot: A Multi-Purpose AI-Driven Mobile Robot for Realistic Environments. In Proceedings of UKRAS20 Conference: Robots into the real world.

Conference Paper

Collaborative Robot-Assisted Endovascular Catheterization with Generative Adversarial Imitation Learning

Chi, W., Dagnino, G., Kwok, T. M. Y., Nguyen, A., Kundrat, D., Abdelaziz, M. E. M. K., . . . Yang, G. Z. (2020). Collaborative Robot-Assisted Endovascular Catheterization with Generative Adversarial Imitation Learning. In 2020 International Conference on Robotics and Automation (ICRA).

Conference Paper

End-to-End Real-time Catheter Segmentation with Optical Flow-Guided Warping during Endovascular Intervention

Nguyen, A., Kundrat, D., Dagnino, G., Chi, W., Abdelaziz, M. E. M. K., Guo, Y., . . . Yang, G. -Z. (2020). End-to-End Real-time Catheter Segmentation with Optical Flow-Guided Warping during Endovascular Intervention. In 2020 IEEE International Conference on Robotics and Automation (ICRA).

Conference Paper

Nonlinearity Compensation in a Multi-DoF Shoulder Sensing Exosuit for Real-Time Teleoperation

Varghese, R. J., Nguyen, A., Burdet, E., Yang, G. -Z., & Lo, B. P. L. (2020). Nonlinearity Compensation in a Multi-DoF Shoulder Sensing Exosuit for Real-Time Teleoperation. In International Conference on Soft Robotics (RoboSoft), 2020..

Conference Paper

2019

Object captioning and retrieval with natural language

Nguyen, A., Tran, Q. D., Do, T. -T., Reid, I., Caldwell, D. G., & Tsagarakis, N. G. (2019). Object captioning and retrieval with natural language. In Proceedings of the IEEE/CVF International Conference on Computer Vision Workshops (pp. 0).

Conference Paper

Real-time 6DOF pose relocalization for event cameras with stacked spatial LSTM networks

Nguyen, A., Do, T. -T., Caldwell, D. G., & Tsagarakis, N. G. (2019). Real-time 6DOF pose relocalization for event cameras with stacked spatial LSTM networks. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops.

Conference Paper

Scene Understanding for Autonomous Manipulation with Deep Learning

Nguyen, A. (2019). Scene Understanding for Autonomous Manipulation with Deep Learning. arXiv preprint arXiv:1903.09761.

Journal article

V2CNet: A Deep Learning Framework to Translate Videos to Commands for Robotic Manipulation

Nguyen, A., Do, T. -T., Reid, I., Caldwell, D. G., & Tsagarakis, N. G. (2019). V2CNet: A Deep Learning Framework to Translate Videos to Commands for Robotic Manipulation. abs/1903.10869.

Journal article

2018

Affordancenet: An end-to-end deep learning approach for object affordance detection

Do*, T. -T., Nguyen*, A., & Reid, I. (2018). Affordancenet: An end-to-end deep learning approach for object affordance detection. In 2018 IEEE International Conference on Robotics and Automation (ICRA).

Conference Paper

Translating Videos to Commands for Robotic Manipulation with Deep Recurrent Neural Networks

Anh, N., Kanoulas, D., Muratore, L., Caldwell, D. G., & Tsagarakis, N. G. (2018). Translating Videos to Commands for Robotic Manipulation with Deep Recurrent Neural Networks. In 2018 IEEE INTERNATIONAL CONFERENCE ON ROBOTICS AND AUTOMATION (ICRA) (pp. 3782-3788). Retrieved from https://www.webofscience.com/

Conference Paper

2017

Object-based affordances detection with convolutional neural networks and dense conditional random fields

Nguyen, A., Kanoulas, D., Caldwell, D. G., & Tsagarakis, N. G. (2017). Object-based affordances detection with convolutional neural networks and dense conditional random fields. In 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (pp. 5908-5915). IEEE.

Conference Paper

Vision-based foothold contact reasoning using curved surface patches

Kanoulas, D., Zhou, C., Nguyen, A., Kanoulas, G., Caldwell, D. G., & Tsagarakis, N. G. (2017). Vision-based foothold contact reasoning using curved surface patches. In 2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids) (pp. 121-128). IEEE.

Conference Paper

2016

Detecting object affordances with convolutional neural networks

Nguyen, A., Kanoulas, D., Caldwell, D. G., & Tsagarakis, N. G. (2016). Detecting object affordances with convolutional neural networks. In 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (pp. 2765-2770). IEEE.

Conference Paper

Preparatory object reorientation for task-oriented grasping

Nguyen, A., Kanoulas, D., Caldwell, D. G., & Tsagarakis, N. G. (2016). Preparatory object reorientation for task-oriented grasping. In 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (pp. 893-899). IEEE.

Conference Paper

2014

Contextual labeling 3D point clouds with conditional random fields

Nguyen, A., & Le, B. (2014). Contextual labeling 3D point clouds with conditional random fields. In Asian Conference on Intelligent Information and Database Systems (pp. 581-590). Springer, Cham.

Conference Paper

2013

3D point cloud segmentation: A survey

Nguyen, A., & Le, B. (2013). 3D point cloud segmentation: A survey. In Robotics, Automation and Mechatronics (RAM), 2013 IEEE Conference on (pp. 225-230). IEEE.

Conference Paper