Dr Yipeng Qin
BSc (Hons), PhD, FHEA
- Ar gael fel goruchwyliwr ôl-raddedig
Timau a rolau for Yipeng Qin
Uwch Ddarlithydd; Arweinydd Effaith, Arloesi a Menter
Trosolwyg
Fi yw'r Arweinydd Effaith, Arloesi a Menter, Arweinydd y Grŵp Ymchwil Gweledigaeth Gyfrifiadurol (adran ymchwil Cyfrifiadura Gweledol), ac Uwch Ddarlithydd (Athro Cysylltiol) yn yr Ysgol Gwyddorau Cyfrifiadurol a Mathemategol. Rwyf hefyd yn aelod o Goleg Adolygu Cymheiriaid EPSRC. Derbyniais fy PhD mewn Cyfrifiadureg o'r Ganolfan Genedlaethol ar gyfer Animeiddio Cyfrifiadurol (NCCA), Prifysgol Bournemouth, y DU yn 2017, a fy BSc mewn Peirianneg Drydanol o Brifysgol Shanghai Jiao Tong, Tsieina yn 2013.
Mae fy niddordebau ymchwil yn gorwedd ar groesffordd dysgu peiriannau a'i gymwysiadau mewn gweledigaeth gyfrifiadurol, graffeg gyfrifiadurol, a rhyngweithio dynol-cyfrifiadur. Ar hyn o bryd, mae fy ngwaith yn canolbwyntio ar dair prif thema: (i) agor y "blwch du" dysgu dwfn, (ii) hyrwyddo deallusrwydd artiffisial creadigol (AI), a (iii) datblygu prototeipiau gwisgadwy ar gyfer monitro a dadansoddi symudiad. Y tu hwnt i'r meysydd hyn, rydw i hefyd yn ymwneud â phynciau cysylltiedig fel segmentu semantig, addasu parth, a dysgu lled-oruchwylio. Mae fy ymchwil wedi cael ei gydnabod gyda Gwobr Papur Gorau SIGGRAPH 2025 (Delwedd Clawr Newyddion) ac Ymgeisydd Gwobr Papur Gorau CVPR 2024.
Mae croeso i bob math o gydweithio!
Ar gyfer ymgeiswyr CSC: Mae Prifysgol Caerdydd yn elwa o bartneriaeth swyddogol gyda Chyngor Ysgoloriaethau Tsieina (CSC). Os oes gennych ddiddordeb mewn gwneud ymchwil gyda mi, cysylltwch â mi trwy e-bost cyn gynted â phosibl fel y gall terfynau cau / camau ychwanegol fod yn berthnasol.
Cyhoeddiad
2026
- Wei, X. et al., 2026. IntrinsicReal: Adapting IntrinsicAnything from synthetic to real objects. IEEE Transactions on Multimedia (10.1109/TMM.2026.3727723)
- Alshewaier, H. , Qin, Y. and Sun, X. 2026. Dual bounding box for medical image segmentation. Presented at: The 6th International Conference on Medical Imaging and Computer-Aided Diagnosis (MICAD 2025) London,UK 19-21 November 2025. Proceedings of 2025 International Conference on Medical Imaging and Computer-Aided Diagnosis (MICAD 2025). Vol. 1520. Springer Science. , pp.161-172. (10.1007/978-981-95-7425-4_15)
- Wang, Y. et al. 2026. Improved cinematic-guided camera language transfer in 3D scene. Presented at: International Conference on 3D Vision 2026 (3DV 2026) Vancouver, BC, Canada 20-23 March 2026. 2026 International Conference on 3D Vision (3DV). IEEE. , pp.1945-1955. (10.1109/3dv69130.2026.00183)
- Meng, Z. et al., 2026. Improving sparse IMU-based motion capture with motion label smoothing. Presented at: The 40th Annual AAAI Conference on Artificial Intelligence (AAAI) 2026 Singapore 20-27 January 2026. Vol. 40.Proceedings of the AAAI Conference on Artificial Intelligence Vol. 10. Washington DC, USA: AAAI Press. , pp.8034-8042. (10.1609/aaai.v40i10.37749)
- Kommers, C. et al., 2026. Computational hermeneutics: evaluating generative AI as a cultural technology. Frontiers in Artificial Intelligence 9 1753041. (10.3389/frai.2026.1753041)
- Chen, J. et al., 2026. 3DGS-HPC: Distractor-free 3D Gaussian Splatting with hybrid patch-wise classification. Presented at: The Forty-Third International Conference on Machine Learning (ICML) Seoul, South Korea 6-11 July 2026. Proceedings of the 43rd International Conference on Machine Learning.
- Gao, X. et al., 2026. CLOTHO: Canonicalizing IMUs from loose inertial garments for accurate human motion tracking. ACM Transactions on Graphics (10.1145/3842534)
- Hao, Y. et al., 2026. LoFA: learning to predict personalized priors for fast adaptation of visual generative models. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026 Colorado, USA 3-7 June 2026.
- Li, Z. et al., 2026. Direct language embedding enables Gaussian splatting for large scenes. Presented at: Findings of the Conference on Computer Vision and Pattern Recognition (CVPR Findings) 2026 Colorado, USA 3-7 June 2026.
- Meng, Z. et al., 2026. Distinguishing imitation error from intrinsic motion learning difficulty. Presented at: The Forty-Third International Conference on Machine Learning (ICML) Seoul, South Korea 6-11 July, 2026. Proceedings of the Forty-Third International Conference on Machine Learning (ICML).
- Ning, Y. et al., 2026. LookasideVLN: direction-aware aerial vision-and-language navigation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026 Colorado, USA 3-7 June 2026.
2025
- Jiang, Z. , Qin, Y. and Finnegan, D. 2025. ‘Chattable’ Avatars: Using LLMs to power visitor engagement with historical persons. Presented at: BCS 38th International Conference on Human Computer Interaction Cardiff, Wales 09 - 11 November. Proceedings of the 38th International BCS Human-Computer Interaction Conference. British Computer Society. , pp.91-102. (10.14236/ewic/BCSHCI2025.10)
- Zuo, C. et al., 2025. Transformer IMU calibrator: Dynamic on-body IMU calibration for inertial motion capture. Presented at: SIGGRAPH 2025 Vancouver, Canada 10-14 August 2025. Vol. 44.Vol. 4. New York, NY, USA: Association for Computing Machinery. , pp.45-45. (10.1145/3730937)
- He, Z. et al., 2025. VTON 360: High-fidelity virtual try-on from any viewing direction. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, USA 11-15 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.26388-26398. (10.1109/CVPR52734.2025.02457)
- Lai, P. et al., 2025. LLM-driven multimodal and multi-identity listening head generation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, USA 11 - 15 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.10656-10666. (10.1109/CVPR52734.2025.00996)
- Wu, Y. , Guo, S. and Qin, Y. 2025. MODA: Motion-drift augmentation for inertial human motion analysis. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, TN, USA 10-17 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.27771-27781. (10.1109/CVPR52734.2025.02586)
- Wu, Z. et al., 2025. Hierarchically controlled deformable 3D gaussians for talking head synthesis. Presented at: The 39th Annual AAAI Conference on Artificial Intelligence (AAAI) 2025 Pennsylvania, USA 25 February – 04 March 2025. Proceedings of the AAAI Conference on Artificial Intelligence. Vol. 39(8).Association for the Advancement of Artificial Intelligence. , pp.8532-8540. (10.1609/aaai.v39i8.32921)
- Ren, T. et al., 2025. Diverse motion in-betweening from sparse keyframes with dual posture stitching. IEEE Transactions on Visualization and Computer Graphics 31 (2), pp.1402-1413. (10.1109/TVCG.2024.3363457)
- Alwadee, E. J. et al. 2025. LATUP-Net: A lightweight 3D attention U-Net with parallel convolutions for brain tumor segmentation. Computers in Biology and Medicine 184 109353. (10.1016/j.compbiomed.2024.109353)
- Ying, E. et al., 2025. WristSketcher: Creating 2D dynamic sketches in AR with a sensing wristband. International Journal of Human-Computer Interaction 41 (1), pp.557-573. (10.1080/10447318.2024.2301857)
- Yao, Y. et al., 2025. ToF-IP: time-of-flight enhanced sparse inertial poser for real-time human motion capture. Presented at: The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025) San Diego, California, USA 2-7 December 2025. Advances in Neural Information Processing Systems 38. NeurIPS. , pp.88892-88911. (10.52202/085713-2677)
2024
- Zhao, G. et al., 2024. Exploration and exploitation of unlabeled data for open-set semi-supervised learning. International Journal of Computer Vision 132 , pp.5888-5904. (10.1007/s11263-024-02155-y)
- Zhan, L. et al., 2024. SATPose: Improving monocular 3D pose estimation with spatial-aware ground tactility. Presented at: ACM Multimedia 2024 Melbourne, Australia 28 October - 1 November 2024. MM '24: Proceedings of the 32nd ACM International Conference on Multimedia. ACM. , pp.6192-6201. (10.1145/3664647.3681654)
- Hou, B. et al., 2024. DCCTNet: Kidney tumors segmentation based on dual-level combination of CNN and transformer. Presented at: IEEE International Conference on Image Processing (ICIP 2024) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.3112-3116. (10.1109/ICIP51287.2024.10647912)
- Chen, J. et al., 2024. NeRF-HuGS: Improved neural radiance fields in non-static scenes using heuristics-guided segmentation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, WA, USA 17-21 June 2024. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.19436-19446. (10.1109/CVPR52733.2024.01838)
- Liang, Y. et al. 2024. Deep generative model based rate-distortion for image downscaling assessment. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, WA, USA 17-21 June 2024. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.19363-19372. (10.1109/CVPR52733.2024.01832)
- Ning, S. et al., 2024. PICTURE: PhotorealistIC virtual Try-on from UnconstRained dEsign. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, USA 16-22 June 2024. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). IEEE. , pp.6976-6985. (10.1109/CVPR52733.2024.00666)
- Zuo, C. et al., 2024. Loose inertial poser: Motion capture with IMU-attached loose-wear jacket. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, USA 17-21 June 2024. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). IEEE. , pp.2209-2219. (10.1109/CVPR52733.2024.00215)
- Liang, Y. et al. 2024. Efficient precision and recall metrics for assessing generative models using hubness-aware sampling. Presented at: The Forty-first International Conference on Machine Learning (ICML) Vienna, Austria 21-27 July 2024. Vol. 235., pp.29682-29699.
- Alwadee, E. et al. 2024. Assessing and enhancing the robustness of brain tumor segmentation using a probabilistic deep learning architecture [Abstract]. Proceedings of the 2024 ISMRM & ISMRT Annual Meeting (4526), pp.1-6. (10.58530/2024/4526)
- Alshewaier, H. , Qin, Y. and Sun, X. 2024. (ExMod) model for medical image segmentation using scribble annotations. Presented at: The 5th International Conference on Medical Imaging and Computer-Aided Diagnosis Manchester, UK 19-21 November 2024. Published in: Su, R. and Frangi, A. F. eds. Proceedings of 2024 International Conference on Medical Imaging and Computer-Aided Diagnosis. Vol. 1372.Lecture Notes in Electrical Engineering Singapore: Springer. , pp.133-143. (10.1007/978-981-96-3863-5_13)
- Chen, X. et al., 2024. Full-body human motion reconstruction with sparse joint tracking using flexible sensors. ACM Transactions on Multimedia Computing, Communications and Applications 20 (2) 44. (10.1145/3564700)
- Yan, Z. et al., 2024. Universal semi-supervised model adaptation via collaborative consistency training. Presented at: IEEE/CVF Winter Conference on Applications of Computer Vision (WACV 2024) Waikoloa, Hawaii, United States 4 - 8 January 2024. IEEE. , pp.861-871. (10.1109/WACV57701.2024.00092)
- Fang, J. et al., 2024. SuDA: Support-based domain adaptation for Sim2Real hinge joint tracking with flexible sensors. Presented at: The Forty-First International Conference on Machine Learning (ICML) Vienna, Austria 21 - 27 July 2024. Published in: Salakhutdinov, R. et al., Proceedings of the 41st International Conference on Machine Learning. Vol. 235.ML Research Press. , pp.22042-22061.
- Wu, Y. et al., 2024. Accurate and steady inertial pose estimation through sequence structure learning and modulation. Presented at: Thirty-Eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024) Vancouver, Canada 10-15 December 2024.
2023
- Zhan, L. et al., 2023. TouchEditor: Interaction design and evaluation of a flexible touchpad for text editing of head-mounted displays in speech-unfriendly environments. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies 7 (4), pp.1-29. 198. (10.1145/3631454)
- Wang, K. et al., 2023. Computational design of wiring layout on tight suits with minimal motion resistance. Presented at: The 16th ACM SIGGRAPH Conference and Exhibition on Computer Graphics and Interactive Techniques in Asia (SIGGRAPH ASIA 2023) Sydney, Australia 12 - 15 December 2023. Published in: Kim, J. , Lin, M. C. and Bickel, B. eds. SA '23: SIGGRAPH Asia 2023 Conference Papers. New York: Association for Computing Machinery. , pp.1-12. (10.1145/3610548.3618200)
- Song, S. et al. 2023. Feature proliferation — the "cancer" in StyleGAN and its treatments. Presented at: International Conference on Computer Vision (ICCV) 2023 Paris, France October 1 - 6, 2023. Proceedings of IEEE/CVF International Conference on Computer Vision. IEEE. , pp.2360-2370. (10.1109/ICCV51070.2023.00224)
- Huang, R. et al., 2023. Parametric implicit face representation for audio-driven facial reenactment. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 Vancouver, Canada 18 - 22 June 2023. Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.12759-12768. (10.1109/CVPR52729.2023.01227)
- Fang, F. et al., 2023. Handwriting velcro: Endowing AR glasses with personalized and posture-adaptive text input using flexible touch sensor. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies (IMWUT) 6 (4), pp.1-31. 163. (10.1145/3569461)
- Jones, O. , Poudevigne-Durance, T. and Qin, Y. 2023. Synthesis of time-series with missing observations using generative adversarial networks. Presented at: 34th Panhellenic Statistics Conference 19-22 May 2022. Greek Statistical Institute. , pp.154-166.
- Zhao, G. et al., 2023. Improved distribution matching for dataset condensation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 Vancouver, Canada 18 - 22 June 2023. Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.7856-7865. (10.1109/CVPR52729.2023.00759)
- Zhao, X. et al. 2023. CUDAS: Distortion-aware saliency benchmark. IEEE Access 11 , pp.58025-58036. (10.1109/ACCESS.2023.3283344)
- Zhou, W. et al. 2023. Reduced-reference quality assessment of point clouds via content-oriented saliency projection. IEEE Signal Processing Letters 30 , pp.354-358. (10.1109/LSP.2023.3264105)
- Zuo, C. et al., 2023. Self-adaptive motion tracking against on-body displacement of flexible sensors. Presented at: Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023) New Orleans, United States 10 - 16 December 2023. Published in: Oh, A. et al., Proceedings of the 37th Conference on Neural Information Processing Systems. Vol. 36.Neural information processing systems foundation. , pp.198465-198465.
2022
- Chen, C. et al., 2022. Real-world blind super-resolution via feature matching with implicit high-resolution priors. Presented at: the 30th ACM International Conference on Multimedia (ACMMM 2022) Lisbon, Portugal 10 - 14 October 2022. Proceedings of the 30th ACM International Conference on Multimedia (ACMMM 2022). ACM. , pp.1329-1338. (10.1145/3503161.3547833)
- Zhao, G. et al., 2022. Centrality and consistency: two-stage clean samples identification for learning with instance-dependent noisy labels. Presented at: European Conference on Computer Vision (ECCV 2022) Tel Aviv, Israel 23-27 October 2022. Published in: Avidan, S. et al., Proceedings of the Computer Vision – ECCV 2022. IEEE. , pp.21-37. (10.1007/978-3-031-19806-9_2)
- Yu, X. et al., 2022. PVSeRF: joint pixel-, voxel- and surface-aligned radiance field for single-image novel view synthesis. Presented at: 30th ACM International Conference on Multimedia (ACMMM 2022) Lisbon, Portugal 10 - 14 October 2022. Proceedings of the 30th ACM International Conference on Multimedia. New York: ACM. , pp.1572-1583. (10.1145/3503161.3547893)
- Yan, Z. et al., 2022. Multi-level consistency learning for semi-supervised domain adaptation. Presented at: 31st International Joint Conference on Artificial Intelligence (IJCAI-ECAI 2022) Vienna, Austria 23-29 July 2022. Published in: De Raedt, L. ed. Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence. International Joint Conferences on Artificial Intelligence Organization. , pp.1530-1536. (10.24963/ijcai.2022/213)
- Poudevigne-Durance, T. , Jones, O. D. and Qin, Y. 2022. MaWGAN: a generative adversarial network to create synthetic data from datasets with missing data. Electronics 11 (6) 837. (10.3390/electronics11060837)
- Liang, Y. et al. 2022. Exploring and exploiting hubness priors for high-quality GAN latent sampling. Presented at: The 39th International Conference on Machine Learning (ICML 2022) Baltimore, Maryland USA 17-23 July 2022. Vol. 162.ML Research Press. , pp.13271-13284.
2021
- Yan, Z. et al., 2021. Pixel-level intra-domain adaptation for semantic segmentation. Presented at: ACM Multimedia 2021 Chengdu, China 20-24 October 2021. MM '21: Proceedings of the 29th ACM International Conference on Multimedia. ACM. , pp.404-413. (10.1145/3474085.3475174)
- Zhu, Z. et al., 2021. Robust elbow angle prediction with aging soft sensors via output-level domain adaptation. IEEE Sensors Journal 21 (20), pp.22976-22984. (10.1109/JSEN.2021.3091004)
- Su, J. et al., 2021. Correcting corrupted labels using mode dropping of ACGAN. Presented at: 15th International Symposium on Medical Information and Communication Technology (ISMICT 2021) Xiamen, China 14-16 April 2021. 2021 15th International Symposium on Medical Information and Communication Technology (ISMICT). IEEE. , pp.98-103. (10.1109/ISMICT51748.2021.9434911)
- Chen, Z. et al., 2021. Human posture tracking with flexible sensors for motion recognition. Computer Animation and Virtual Worlds (10.1002/cav.1993)
2020
- Qin, Y. , Mitra, N. and Wonka, P. 2020. How does Lipschitz regularization influence GAN training?. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vevaldi, A. et al., Computer Vision – ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XVI. Lecture Notes in Computer Science Vol. 12361. Springer. , pp.310-326. (10.1007/978-3-030-58517-4_19)
- Abdal, R. , Qin, Y. and Wonka, P. 2020. Image2StyleGAN++: how to edit the embedded images?. Presented at: Conference on Computer Vision and Pattern Recognition (CVPR 2020) Seattle, Washington, USA 16-18 June 2020. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.8293-8302. (10.1109/CVPR42600.2020.00832)
- Zhu, P. et al., 2020. SEAN: image synthesis with semantic region-adaptive normalization. Presented at: Conference on Computer Vision and Pattern Recognition (CVPR 2020) Seattle, Washington, USA 14-19 June 2020. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.5103-5112. (10.1109/CVPR42600.2020.00515)
2019
- Abdal, R. , Qin, Y. and Wonka, P. 2019. Image2StyleGAN: How to embed images into the styleGAN latent space?. Presented at: International Conference on Computer Vision (ICCV) 2019 Seoul, South Korea 27 October 2019 - 3 November 2019. Proceedings of the International Conference on Computer Vision (ICCV) 2019. IEEE. , pp.4431-4440. (10.1109/ICCV.2019.00453)
2017
- Qin, Y. , Yu, H. and Zhang, J. 2017. Fast and memory-efficient Voronoi diagram construction on triangle meshes. Computer Graphics Forum 36 (5), pp.93-104. (10.1111/cgf.13248)
2016
- Qin, Y. et al. 2016. Fast and exact discrete geodesic computation based on triangle-oriented wavefront propagation. ACM Transactions on Graphics 35 (4) 125. (10.1145/2897824.2925930)
2015
- Yu, H. , Qin, Y. and Zhang, J. J. 2015. Eigenspace-based surface completeness. Journal of Electronic Imaging 24 (2) 023037. (10.1117/1.JEI.24.2.023037)
Cynadleddau
- Alshewaier, H. , Qin, Y. and Sun, X. 2026. Dual bounding box for medical image segmentation. Presented at: The 6th International Conference on Medical Imaging and Computer-Aided Diagnosis (MICAD 2025) London,UK 19-21 November 2025. Proceedings of 2025 International Conference on Medical Imaging and Computer-Aided Diagnosis (MICAD 2025). Vol. 1520. Springer Science. , pp.161-172. (10.1007/978-981-95-7425-4_15)
- Wang, Y. et al. 2026. Improved cinematic-guided camera language transfer in 3D scene. Presented at: International Conference on 3D Vision 2026 (3DV 2026) Vancouver, BC, Canada 20-23 March 2026. 2026 International Conference on 3D Vision (3DV). IEEE. , pp.1945-1955. (10.1109/3dv69130.2026.00183)
- Meng, Z. et al., 2026. Improving sparse IMU-based motion capture with motion label smoothing. Presented at: The 40th Annual AAAI Conference on Artificial Intelligence (AAAI) 2026 Singapore 20-27 January 2026. Vol. 40.Proceedings of the AAAI Conference on Artificial Intelligence Vol. 10. Washington DC, USA: AAAI Press. , pp.8034-8042. (10.1609/aaai.v40i10.37749)
- Chen, J. et al., 2026. 3DGS-HPC: Distractor-free 3D Gaussian Splatting with hybrid patch-wise classification. Presented at: The Forty-Third International Conference on Machine Learning (ICML) Seoul, South Korea 6-11 July 2026. Proceedings of the 43rd International Conference on Machine Learning.
- Hao, Y. et al., 2026. LoFA: learning to predict personalized priors for fast adaptation of visual generative models. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026 Colorado, USA 3-7 June 2026.
- Li, Z. et al., 2026. Direct language embedding enables Gaussian splatting for large scenes. Presented at: Findings of the Conference on Computer Vision and Pattern Recognition (CVPR Findings) 2026 Colorado, USA 3-7 June 2026.
- Meng, Z. et al., 2026. Distinguishing imitation error from intrinsic motion learning difficulty. Presented at: The Forty-Third International Conference on Machine Learning (ICML) Seoul, South Korea 6-11 July, 2026. Proceedings of the Forty-Third International Conference on Machine Learning (ICML).
- Ning, Y. et al., 2026. LookasideVLN: direction-aware aerial vision-and-language navigation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026 Colorado, USA 3-7 June 2026.
- Jiang, Z. , Qin, Y. and Finnegan, D. 2025. ‘Chattable’ Avatars: Using LLMs to power visitor engagement with historical persons. Presented at: BCS 38th International Conference on Human Computer Interaction Cardiff, Wales 09 - 11 November. Proceedings of the 38th International BCS Human-Computer Interaction Conference. British Computer Society. , pp.91-102. (10.14236/ewic/BCSHCI2025.10)
- Zuo, C. et al., 2025. Transformer IMU calibrator: Dynamic on-body IMU calibration for inertial motion capture. Presented at: SIGGRAPH 2025 Vancouver, Canada 10-14 August 2025. Vol. 44.Vol. 4. New York, NY, USA: Association for Computing Machinery. , pp.45-45. (10.1145/3730937)
- He, Z. et al., 2025. VTON 360: High-fidelity virtual try-on from any viewing direction. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, USA 11-15 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.26388-26398. (10.1109/CVPR52734.2025.02457)
- Lai, P. et al., 2025. LLM-driven multimodal and multi-identity listening head generation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, USA 11 - 15 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.10656-10666. (10.1109/CVPR52734.2025.00996)
- Wu, Y. , Guo, S. and Qin, Y. 2025. MODA: Motion-drift augmentation for inertial human motion analysis. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 Nashville, TN, USA 10-17 June 2025. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.27771-27781. (10.1109/CVPR52734.2025.02586)
- Wu, Z. et al., 2025. Hierarchically controlled deformable 3D gaussians for talking head synthesis. Presented at: The 39th Annual AAAI Conference on Artificial Intelligence (AAAI) 2025 Pennsylvania, USA 25 February – 04 March 2025. Proceedings of the AAAI Conference on Artificial Intelligence. Vol. 39(8).Association for the Advancement of Artificial Intelligence. , pp.8532-8540. (10.1609/aaai.v39i8.32921)
- Yao, Y. et al., 2025. ToF-IP: time-of-flight enhanced sparse inertial poser for real-time human motion capture. Presented at: The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025) San Diego, California, USA 2-7 December 2025. Advances in Neural Information Processing Systems 38. NeurIPS. , pp.88892-88911. (10.52202/085713-2677)
- Zhan, L. et al., 2024. SATPose: Improving monocular 3D pose estimation with spatial-aware ground tactility. Presented at: ACM Multimedia 2024 Melbourne, Australia 28 October - 1 November 2024. MM '24: Proceedings of the 32nd ACM International Conference on Multimedia. ACM. , pp.6192-6201. (10.1145/3664647.3681654)
- Hou, B. et al., 2024. DCCTNet: Kidney tumors segmentation based on dual-level combination of CNN and transformer. Presented at: IEEE International Conference on Image Processing (ICIP 2024) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.3112-3116. (10.1109/ICIP51287.2024.10647912)
- Chen, J. et al., 2024. NeRF-HuGS: Improved neural radiance fields in non-static scenes using heuristics-guided segmentation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, WA, USA 17-21 June 2024. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.19436-19446. (10.1109/CVPR52733.2024.01838)
- Liang, Y. et al. 2024. Deep generative model based rate-distortion for image downscaling assessment. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, WA, USA 17-21 June 2024. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.19363-19372. (10.1109/CVPR52733.2024.01832)
- Ning, S. et al., 2024. PICTURE: PhotorealistIC virtual Try-on from UnconstRained dEsign. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, USA 16-22 June 2024. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). IEEE. , pp.6976-6985. (10.1109/CVPR52733.2024.00666)
- Zuo, C. et al., 2024. Loose inertial poser: Motion capture with IMU-attached loose-wear jacket. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 Seattle, USA 17-21 June 2024. 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). IEEE. , pp.2209-2219. (10.1109/CVPR52733.2024.00215)
- Liang, Y. et al. 2024. Efficient precision and recall metrics for assessing generative models using hubness-aware sampling. Presented at: The Forty-first International Conference on Machine Learning (ICML) Vienna, Austria 21-27 July 2024. Vol. 235., pp.29682-29699.
- Alshewaier, H. , Qin, Y. and Sun, X. 2024. (ExMod) model for medical image segmentation using scribble annotations. Presented at: The 5th International Conference on Medical Imaging and Computer-Aided Diagnosis Manchester, UK 19-21 November 2024. Published in: Su, R. and Frangi, A. F. eds. Proceedings of 2024 International Conference on Medical Imaging and Computer-Aided Diagnosis. Vol. 1372.Lecture Notes in Electrical Engineering Singapore: Springer. , pp.133-143. (10.1007/978-981-96-3863-5_13)
- Yan, Z. et al., 2024. Universal semi-supervised model adaptation via collaborative consistency training. Presented at: IEEE/CVF Winter Conference on Applications of Computer Vision (WACV 2024) Waikoloa, Hawaii, United States 4 - 8 January 2024. IEEE. , pp.861-871. (10.1109/WACV57701.2024.00092)
- Fang, J. et al., 2024. SuDA: Support-based domain adaptation for Sim2Real hinge joint tracking with flexible sensors. Presented at: The Forty-First International Conference on Machine Learning (ICML) Vienna, Austria 21 - 27 July 2024. Published in: Salakhutdinov, R. et al., Proceedings of the 41st International Conference on Machine Learning. Vol. 235.ML Research Press. , pp.22042-22061.
- Wu, Y. et al., 2024. Accurate and steady inertial pose estimation through sequence structure learning and modulation. Presented at: Thirty-Eighth Annual Conference on Neural Information Processing Systems (NeurIPS 2024) Vancouver, Canada 10-15 December 2024.
- Wang, K. et al., 2023. Computational design of wiring layout on tight suits with minimal motion resistance. Presented at: The 16th ACM SIGGRAPH Conference and Exhibition on Computer Graphics and Interactive Techniques in Asia (SIGGRAPH ASIA 2023) Sydney, Australia 12 - 15 December 2023. Published in: Kim, J. , Lin, M. C. and Bickel, B. eds. SA '23: SIGGRAPH Asia 2023 Conference Papers. New York: Association for Computing Machinery. , pp.1-12. (10.1145/3610548.3618200)
- Song, S. et al. 2023. Feature proliferation — the "cancer" in StyleGAN and its treatments. Presented at: International Conference on Computer Vision (ICCV) 2023 Paris, France October 1 - 6, 2023. Proceedings of IEEE/CVF International Conference on Computer Vision. IEEE. , pp.2360-2370. (10.1109/ICCV51070.2023.00224)
- Huang, R. et al., 2023. Parametric implicit face representation for audio-driven facial reenactment. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 Vancouver, Canada 18 - 22 June 2023. Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.12759-12768. (10.1109/CVPR52729.2023.01227)
- Jones, O. , Poudevigne-Durance, T. and Qin, Y. 2023. Synthesis of time-series with missing observations using generative adversarial networks. Presented at: 34th Panhellenic Statistics Conference 19-22 May 2022. Greek Statistical Institute. , pp.154-166.
- Zhao, G. et al., 2023. Improved distribution matching for dataset condensation. Presented at: The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 Vancouver, Canada 18 - 22 June 2023. Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.7856-7865. (10.1109/CVPR52729.2023.00759)
- Zuo, C. et al., 2023. Self-adaptive motion tracking against on-body displacement of flexible sensors. Presented at: Thirty-seventh Conference on Neural Information Processing Systems (NeurIPS 2023) New Orleans, United States 10 - 16 December 2023. Published in: Oh, A. et al., Proceedings of the 37th Conference on Neural Information Processing Systems. Vol. 36.Neural information processing systems foundation. , pp.198465-198465.
- Chen, C. et al., 2022. Real-world blind super-resolution via feature matching with implicit high-resolution priors. Presented at: the 30th ACM International Conference on Multimedia (ACMMM 2022) Lisbon, Portugal 10 - 14 October 2022. Proceedings of the 30th ACM International Conference on Multimedia (ACMMM 2022). ACM. , pp.1329-1338. (10.1145/3503161.3547833)
- Zhao, G. et al., 2022. Centrality and consistency: two-stage clean samples identification for learning with instance-dependent noisy labels. Presented at: European Conference on Computer Vision (ECCV 2022) Tel Aviv, Israel 23-27 October 2022. Published in: Avidan, S. et al., Proceedings of the Computer Vision – ECCV 2022. IEEE. , pp.21-37. (10.1007/978-3-031-19806-9_2)
- Yu, X. et al., 2022. PVSeRF: joint pixel-, voxel- and surface-aligned radiance field for single-image novel view synthesis. Presented at: 30th ACM International Conference on Multimedia (ACMMM 2022) Lisbon, Portugal 10 - 14 October 2022. Proceedings of the 30th ACM International Conference on Multimedia. New York: ACM. , pp.1572-1583. (10.1145/3503161.3547893)
- Yan, Z. et al., 2022. Multi-level consistency learning for semi-supervised domain adaptation. Presented at: 31st International Joint Conference on Artificial Intelligence (IJCAI-ECAI 2022) Vienna, Austria 23-29 July 2022. Published in: De Raedt, L. ed. Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence. International Joint Conferences on Artificial Intelligence Organization. , pp.1530-1536. (10.24963/ijcai.2022/213)
- Liang, Y. et al. 2022. Exploring and exploiting hubness priors for high-quality GAN latent sampling. Presented at: The 39th International Conference on Machine Learning (ICML 2022) Baltimore, Maryland USA 17-23 July 2022. Vol. 162.ML Research Press. , pp.13271-13284.
- Yan, Z. et al., 2021. Pixel-level intra-domain adaptation for semantic segmentation. Presented at: ACM Multimedia 2021 Chengdu, China 20-24 October 2021. MM '21: Proceedings of the 29th ACM International Conference on Multimedia. ACM. , pp.404-413. (10.1145/3474085.3475174)
- Su, J. et al., 2021. Correcting corrupted labels using mode dropping of ACGAN. Presented at: 15th International Symposium on Medical Information and Communication Technology (ISMICT 2021) Xiamen, China 14-16 April 2021. 2021 15th International Symposium on Medical Information and Communication Technology (ISMICT). IEEE. , pp.98-103. (10.1109/ISMICT51748.2021.9434911)
- Qin, Y. , Mitra, N. and Wonka, P. 2020. How does Lipschitz regularization influence GAN training?. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vevaldi, A. et al., Computer Vision – ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XVI. Lecture Notes in Computer Science Vol. 12361. Springer. , pp.310-326. (10.1007/978-3-030-58517-4_19)
- Abdal, R. , Qin, Y. and Wonka, P. 2020. Image2StyleGAN++: how to edit the embedded images?. Presented at: Conference on Computer Vision and Pattern Recognition (CVPR 2020) Seattle, Washington, USA 16-18 June 2020. Proceedings of the Conference on Computer Vision and Pattern Recognition. IEEE. , pp.8293-8302. (10.1109/CVPR42600.2020.00832)
- Zhu, P. et al., 2020. SEAN: image synthesis with semantic region-adaptive normalization. Presented at: Conference on Computer Vision and Pattern Recognition (CVPR 2020) Seattle, Washington, USA 14-19 June 2020. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. IEEE. , pp.5103-5112. (10.1109/CVPR42600.2020.00515)
- Abdal, R. , Qin, Y. and Wonka, P. 2019. Image2StyleGAN: How to embed images into the styleGAN latent space?. Presented at: International Conference on Computer Vision (ICCV) 2019 Seoul, South Korea 27 October 2019 - 3 November 2019. Proceedings of the International Conference on Computer Vision (ICCV) 2019. IEEE. , pp.4431-4440. (10.1109/ICCV.2019.00453)
Erthyglau
- Wei, X. et al., 2026. IntrinsicReal: Adapting IntrinsicAnything from synthetic to real objects. IEEE Transactions on Multimedia (10.1109/TMM.2026.3727723)
- Kommers, C. et al., 2026. Computational hermeneutics: evaluating generative AI as a cultural technology. Frontiers in Artificial Intelligence 9 1753041. (10.3389/frai.2026.1753041)
- Gao, X. et al., 2026. CLOTHO: Canonicalizing IMUs from loose inertial garments for accurate human motion tracking. ACM Transactions on Graphics (10.1145/3842534)
- Ren, T. et al., 2025. Diverse motion in-betweening from sparse keyframes with dual posture stitching. IEEE Transactions on Visualization and Computer Graphics 31 (2), pp.1402-1413. (10.1109/TVCG.2024.3363457)
- Alwadee, E. J. et al. 2025. LATUP-Net: A lightweight 3D attention U-Net with parallel convolutions for brain tumor segmentation. Computers in Biology and Medicine 184 109353. (10.1016/j.compbiomed.2024.109353)
- Ying, E. et al., 2025. WristSketcher: Creating 2D dynamic sketches in AR with a sensing wristband. International Journal of Human-Computer Interaction 41 (1), pp.557-573. (10.1080/10447318.2024.2301857)
- Zhao, G. et al., 2024. Exploration and exploitation of unlabeled data for open-set semi-supervised learning. International Journal of Computer Vision 132 , pp.5888-5904. (10.1007/s11263-024-02155-y)
- Alwadee, E. et al. 2024. Assessing and enhancing the robustness of brain tumor segmentation using a probabilistic deep learning architecture [Abstract]. Proceedings of the 2024 ISMRM & ISMRT Annual Meeting (4526), pp.1-6. (10.58530/2024/4526)
- Chen, X. et al., 2024. Full-body human motion reconstruction with sparse joint tracking using flexible sensors. ACM Transactions on Multimedia Computing, Communications and Applications 20 (2) 44. (10.1145/3564700)
- Zhan, L. et al., 2023. TouchEditor: Interaction design and evaluation of a flexible touchpad for text editing of head-mounted displays in speech-unfriendly environments. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies 7 (4), pp.1-29. 198. (10.1145/3631454)
- Fang, F. et al., 2023. Handwriting velcro: Endowing AR glasses with personalized and posture-adaptive text input using flexible touch sensor. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies (IMWUT) 6 (4), pp.1-31. 163. (10.1145/3569461)
- Zhao, X. et al. 2023. CUDAS: Distortion-aware saliency benchmark. IEEE Access 11 , pp.58025-58036. (10.1109/ACCESS.2023.3283344)
- Zhou, W. et al. 2023. Reduced-reference quality assessment of point clouds via content-oriented saliency projection. IEEE Signal Processing Letters 30 , pp.354-358. (10.1109/LSP.2023.3264105)
- Poudevigne-Durance, T. , Jones, O. D. and Qin, Y. 2022. MaWGAN: a generative adversarial network to create synthetic data from datasets with missing data. Electronics 11 (6) 837. (10.3390/electronics11060837)
- Zhu, Z. et al., 2021. Robust elbow angle prediction with aging soft sensors via output-level domain adaptation. IEEE Sensors Journal 21 (20), pp.22976-22984. (10.1109/JSEN.2021.3091004)
- Chen, Z. et al., 2021. Human posture tracking with flexible sensors for motion recognition. Computer Animation and Virtual Worlds (10.1002/cav.1993)
- Qin, Y. , Yu, H. and Zhang, J. 2017. Fast and memory-efficient Voronoi diagram construction on triangle meshes. Computer Graphics Forum 36 (5), pp.93-104. (10.1111/cgf.13248)
- Qin, Y. et al. 2016. Fast and exact discrete geodesic computation based on triangle-oriented wavefront propagation. ACM Transactions on Graphics 35 (4) 125. (10.1145/2897824.2925930)
- Yu, H. , Qin, Y. and Zhang, J. J. 2015. Eigenspace-based surface completeness. Journal of Electronic Imaging 24 (2) 023037. (10.1117/1.JEI.24.2.023037)
Ymchwil
Mae fy niddordebau ymchwil yn canolbwyntio ar ddysgu peiriannau (ML) a'i gymwysiadau mewn gweledigaeth gyfrifiadurol (CV), graffeg gyfrifiadurol (CG), rhyngweithiadau dynol-cyfrifiadur (HCI), ac ati. Mae fy ymchwil bresennol yn troi o amgylch y tair thema ganlynol:
- Unboxing the deep learning black-box, which aims to understand the training dynamics of various deep neural networks, e.e. Generative Adversarial Networks (GANs).
- Deallusrwydd Artiffisial Creadigol (AI), gan gynnwys gwrthdroi GAN, datgysylltu delweddau rhanbarthol ar gyfer synthesis/trin delweddau amrywiol a rheoladwy, ac ati.
- Prototeipio dyfeisiau gwisgadwy ar gyfer monitro a dadansoddi symudiad, sy'n anelu at fodelu'r berthynas rhwng signalau synhwyrydd a symudiad dynol gan ddefnyddio dysgu peiriannau. Mae hon yn ymchwil ar y cyd gyda'r Athro Shihui Guo o Brifysgol Xiamen, Tsieina.
Mae gen i ddiddordeb hefyd mewn pynciau cysylltiedig eraill fel segmentu semantig, addasu parth, dysgu lled-oruchwylio, ac ati.
Prosiectau
- Tuag at Ddal Cynnig Graddadwy yn y Gwyllt, £2,810, PI, DTII Cyllid Ymateb Cyflym Seedcorn gan Brifysgol Caerdydd, Mai 2026 - Gorff 2026
- Gwella'r Diwylliant Ymchwil Ôl-raddedig mewn Ysgolion sy'n Uno, £1,000, PI, Cyllid yr Academi Ddoethurol gan Brifysgol Caerdydd, Mai 2026 - Rhagfyr 2026
- Cyllid Cydweithredu CU-XMU, £4,300, PI, Prifysgol Caerdydd, Mai 2026 - Rhagfyr 2026
- Lleihau Ymddygiad Gwrthgymdeithasol a Threisgar gan ddefnyddio Deallusrwydd Artiffisial, £29,352.5, Co-I, Cyfrif Cyflymu Effaith EPSRC (IAA), Mai 2026 - Rhagfyr 2026
- Cyllid Symudedd Ymchwil Taith, £925, PI, Llywodraeth Cymru, Hydref 2025 - Hydref 2025
- Sgyrsio Empathetic, £2,790, PI, Cyllid Interniaeth ar y campws Prifysgol Caerdydd, Meh 2025 - Awst 2025
- Rhyngwyneb Iaith Naturiol ar gyfer Cyfarwyddo Perfformiadau Cymeriadau Rhithwir, £59,981, PI, prosiect XR Network+ (EP/W020602/1) o EPSRC, Medi 2024 - Chwefror 2025
- Datgelu "Greddf" Modelau Cynhyrchiol Dwfn ar gyfer Creu Cynnwys Gweledol Teg a Diduedd, tua £72,922, Prif Oruchwylydd, EPSRC DTP, Rhif EP/T517951/1 (2599521), Hydref 2021 - Mawrth 2025
- Prototeipio Dillad Clyfar ar gyfer Monitro a Dadansoddi Cynnig ar lefel y Boblogaeth, £12,000, PI, Cyfnewidfeydd Rhyngwladol y Gymdeithas Frenhinol Cyfran Costau 2021 (NSFC), Rhif. IEC\NSFC\211022, Mawrth 2022 - Mawrth 2025
- Cyllid Cydweithredu CU-XMU, £1,880, Co-I, Prifysgol Caerdydd, Tachwedd 2024 - Rhagfyr 2024
- Cronfa Ymchwil Diwylliant Prifysgol Caerdydd 2024, £2,500, Co-I, CCAUC, Mai 2024 - Awst 2024
- Datblygu Map Ffordd – Creu Personau AI Chattable ar gyfer Amgueddfeydd ac Archifau, £4,000, PI, Cyllid Seedcorn y Sefydliad Arloesi Trawsnewid Digidol (DTII), Rhif. DTIIR3SC04, Ebrill 2024 - Awst 2024
- Chwyldroi Profiadau Ymwelwyr Treftadaeth Ddiwylliannol: Rhyddhau Pŵer Avatars Sgwrsio wedi'u Pweru gan AI mewn Arddangosfeydd Rhyngweithiol, £15,000, PI, Cyfrif Cyflymu Effaith Cysoni UKRI (IAA), Rhif 521632 (525436), Tachwedd 2023 - Mai 2024
- Charting New Frontiers: An Exploratory Expedition and Pilot Study on Chattable Virtual Avatars, Unveiling Ethical and Social Dimensions in Content Delivery, £5,000, PI, GW4 Crucible 2023 Seed Funding, No. Cru23_01, Medi 2023 - Mawrth 2024
- Optimeiddio Cynllun Synhwyrydd ar gyfer Ailadeiladu Siâp Corff 3D, £3,200, PI, Cyllid Cydweithredu Ymchwil Prifysgol Caerdydd-Xiamen, Gorff 2020 - Awst 2020
Gweithgareddau/Digwyddiadau
- Gwasanaethodd y 12fed Gynhadledd Ryngwladol ar Realiti Rhithwir (ICVR 2026) fel Cadeirydd y Pwyllgor Trefnu Lleol, Gorffennaf 28-30, 2026.
- Gweithdy Arweiniad Academaidd wedi'i gyfryngu gan AI, wedi'i drefnu ar y cyd â Dr. Jing Wu a'r Athro Paul Rosin o Brifysgol Caerdydd, Mehefin 10 a 17, 2026.
- Gweithdy Syniadau Tŷ Dr Jenner, wedi'i drefnu ar y cyd â Dr. Deborah Brewis o Brifysgol Caerfaddon, Hydref 27, 2025.
- Aelod o Bwyllgor Colocwiwm ECR LSW 2025, a drefnwyd gan Gymdeithas Ddysgedig Cymru (LSW), Mawrth - Mehefin 2025.
- Gŵyl y Gwyddorau Cymdeithasol, Deallusrwydd Artiffisial mewn diwylliant a threftadaeth 2024, a drefnir gan Dr. Jenny Kidd gyda chefnogaeth ein tîm, Tachwedd 7, 2024.
- Cydweithrediadauâ Hanes: dod â chymeriadau hanesyddol yn fyw trwy avatars wedi'u pweru gan AI, a drefnwyd gan yr Ymddiriedolaeth Genedlaethol gyda chefnogaeth ein tîm, Hydref 26-27, 2024.
- Gweithdai Datblygu Map Ffordd, a drefnwyd gan Dr. Yipeng Qin, Mai 31 a Gorffennaf 24, 2024.
- [Digwyddiad Ymylol AI UK 2024] Gorffennol yn Cwrdd â'r Dyfodol: A all AI Personas ddod â ffigurau hanesyddol yn fyw?, a drefnwyd gan Dr. Yipeng Qin, Mawrth 27, 2024
- Gweithdy Prifysgol Caerdydd a'r Ymddiriedolaeth Genedlaethol ar AI Avatars, wedi'i drefnu ar y cyd â Dr. Daniel Finnegan, Mawrth 13, 2024
- Gweithdy Avatars Rhithwir Chattable ar gyfer Amgueddfa ac Archifau, wedi'i drefnu ar y cyd â Dr. Barbara Caddick o Brifysgol Bryste, Mawrth 11, 2024.
Addysgu
Addysgu Modiwlaidd
- 2021/22 - nawr, CMT307 Dysgu Peiriant Cymhwysol
- 2021/22 - nawr, CMT316 Ceisiadau Dysgu Peiriannau: Prosesu Iaith Naturiol / Gweledigaeth Gyfrifiadurol
- 2019/20 - nawr, CM1205 Pensaernïaeth a Systemau Gweithredu, Arweinydd Modiwl
Dyfarniadau
- Hyrwyddwr dros Gydraddoldeb, Amrywiaeth a Chynhwysiant (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2025
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
- Cydweithrediad Dysgu ac Addysgu'r Flwyddyn (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2024
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
- Hyrwyddwr dros Gydraddoldeb, Amrywiaeth a Chynhwysiant (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2024
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
Arholwr Allanol
- Rhaglen BSc
- BSc Deallusrwydd Artiffisial (UI4AA), Prifysgol Derby, Mai 2025 - Medi 2030
- BSc AI a Gwyddor Data (UI4AD), Prifysgol Derby, Mai 2025 - Medi 2030
Bywgraffiad
Addysg a Chymwysterau
- 2017: PhD mewn Cyfrifiadureg, Y Ganolfan Genedlaethol ar gyfer Animeiddio Cyfrifiadurol, Prifysgol Bournemouth, y DU
- 2013: BEng mewn Peirianneg Drydanol, Prifysgol Shanghai Jiao Tong, Tsieina
Trosolwg gyrfa
- 2023 - presennol: Uwch Ddarlithydd, Prifysgol Caerdydd, DU
- 2019 - 2023: Darlithydd, Prifysgol Caerdydd, DU
- 2017 - 2019: Cymrawd Ymchwil Ôl-ddoethurol, Canolfan Cyfrifiadura Gweledol, Prifysgol Gwyddoniaeth a Thechnoleg King Abdullah, Saudi Arabia
Rôl weinyddol
Ysgol y Gwyddorau Cyfrifiadurol a Mathemategol (COMAT):
- 2026.08 - nawr, Arweinydd Effaith, Arloesi a Menter
Ysgol Cyfrifiadureg a Gwybodeg (COMSC):
- 2024.09 - 2026.07, Dirprwy Gyfarwyddwr Ymchwil
- 2024.02 - nawr, Arweinydd Grŵp Ymchwil Gweledigaeth Gyfrifiadurol
- 2024.04 - 2024.09, Tiwtor Derbyn PGT
- 2020.10 - 2024.09, Arweinydd Rhaglen PGT (COMSC) MSc Gwyddor Data a Dadansoddeg Data ac MSc Dadansoddeg Data ar gyfer y Llywodraeth
Anrhydeddau a dyfarniadau
(Ymchwil)
- Gwobr Papur Gorau SIGGRAPH 2025
- Delwedd Clawr Newyddion
- GAN: ACM SIGGRAPH Pwyllgorau
- Gwobr Adolygydd Rhagorol CVPR 2025
- 711 allan o 12,593 (5.6%)
- GAN: Y Sefydliad Gweledigaeth Gyfrifiadurol (CVF)
- Ymgeisydd Gwobr Papur Gorau CVPR 2024
- GAN: Y Sefydliad Gweledigaeth Gyfrifiadurol (CVF)
- Gwobr Arddangos Technoleg Orau CHCI 2024
- System Dal Cynnig Dynol yn seiliedig ar synwyryddion anadweithiol gwisgo rhydd
- GAN: Yr 20fed Gynhadledd Rhyngweithio Dynol-Cyfrifiadur Tsieineaidd (CHCI 2024)
(Addysgu)
- Hyrwyddwr dros Gydraddoldeb, Amrywiaeth a Chynhwysiant (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2025
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
- Cydweithrediad Dysgu ac Addysgu'r Flwyddyn (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2024
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
- Hyrwyddwr dros Gydraddoldeb, Amrywiaeth a Chynhwysiant (Enwebiad)
- Gwobrau Cyfoethogi Bywyd Myfyrwyr (ESLAs) 2024
- GAN: Prifysgol Caerdydd ac Undeb Myfyrwyr Caerdydd (CUSU)
Aelodaethau proffesiynol
- Coleg Adolygu Cymheiriaid EPSRC
- Aelod o ACM SIGGRAPH, Sefydliad Gweledigaeth Gyfrifiadurol (CVF), AsiaGraphics
- Labordy Ewropeaidd ar gyfer Dysgu a Systemau Deallus (ELLIS)
- Cymrodoriaeth HEA
- Cymdeithas Ddysgedig Cymru - Grŵp Cynghori ar gyfer Datblygu Ymchwilwyr
- Grŵp Diddordeb Arbennig y Dyniaethau a Gwyddor Data @ Sefydliad Alan Turing
- Rhwydwaith Ymchwilwyr Gyrfa Gynnar yr Academi Brydeinig
Safleoedd academaidd blaenorol
- 2023 - presennol: Uwch Ddarlithydd, Prifysgol Caerdydd, UK
- 2019 - 2023: Darlithydd, Prifysgol Caerdydd, DU
- 2017 - 2019: Cymrawd Ymchwil Ôl-ddoethurol, Canolfan Cyfrifiadura Gweledol, Brenin Abdullah Prifysgol Gwyddoniaeth a Thechnoleg, Saudi Arabia
Pwyllgorau ac adolygu
- Cadeirydd Ardal / Uwch Bwyllgor y Rhaglen: ICLR 2027, AAAI 2027, NeurIPS 2026, ICML 2026, ICML 2025, BMVC 2025, ICML 2024
- Adolygydd Grant:
- Turing AI Cymrodoriaethau Ymchwilwyr Sy'n Arwain y Byd
- Cymrodoriaethau Arweinwyr y Dyfodol UKRI (FLF)
- Coleg Adolygu Cymheiriaid EPSRC
- Swyddi
- Grant Ymchwil Safonol EPSRC
- Gwobr Ymchwilydd Newydd EPSRC (NIA)
- Rhaglen Ariannu Ychwanegol EPSRC ar gyfer Gwyddoniaeth Fathemategol
- Camau Ymchwil Cydgysylltiedig Gwlad Belg (CRAs)
- Bwrsariaethau Teithio a Threuliau ECR LSW
- Adolygydd Cyfnodolion:
- Trafodion IEEE ar Ddadansoddi Patrymau a Deallusrwydd Peiriant (TPAMI)
- Trafodion ACM ar Graffeg (TOG)
- Trafodion IEEE ar Brosesu Delweddau (TIP)
- Trafodion IEEE ar Ddelweddu a Graffeg Gyfrifiadurol (TVCG)
- Trafodion IEEE ar Gyfrifiadura Symudol (TMC)
- Trafodion IEEE ar Amlgyfrwng (TMM)
- Trafodion ar Ymchwil Dysgu Peiriannau (TMLR)
- Fforwm Graffeg Gyfrifiadurol (CGF)
- Adnabod Patrymau
- Niwrogyfrifiadura
- Y Cyfrifiadur Gweledol
- Animeiddio Cyfrifiadurol a Bydoedd Rhithwir (CAVW)
- Cyfrifiaduron a Graffeg
- Systemau Arbenigol gyda Chymwysiadau
- Prosesu Signalau, Delweddau a Fideo
- Prosesu a Rheoli Signal Biofeddygol
- Journal of Open Humanities Data
- Mynediad IEEE
- Gwybodeg Weledol
- Y We Fyd-eang
- Cyfrifiadura Delwedd a Gweledigaeth
- Ffiniau mewn Delweddu
- Aelod / Adolygydd Pwyllgor Rhaglen Cynhadledd (PC):
- SIGGRAFF ACM
- ACM SIGGRAPH Asia
- Cartref
- Cartref
- Cartref
- Mewngofnodi
- Trac setiau data a meincnodau NeurIPS
- ICLR
- Trac BlogPosts ICLR
- Cartref
- Cartref
- ACM MM
- Swyddi
- Ewrograffeg
- Graffeg y Môr Tawel
- Cartref
- Swyddi
- Cartref
- Cartref
- Cartref
- CARTREF
- ICMR
- IJCNN
- Cartref
- ICXR
- Cartref
- Adolygydd Llyfr: Gwasg CRC/Grŵp Taylor a Francis
Meysydd goruchwyliaeth
Mae gennyf ddiddordeb mewn goruchwylio myfyrwyr PhD yn y meysydd canlynol:
- AI ar gyfer Modelu Cynhyrchiol
- Synthesis a Thrin Delweddau
- ML / AI Dehongladwy ar gyfer Cynhyrchu Cynnwys Gweledol
- Rhagfarn a Thegwch mewn AI
- Prosesu Geometreg 3D
Ar gyfer ymgeiswyr CSC: Mae Prifysgol Caerdydd yn elwa o bartneriaeth swyddogol gyda Chyngor Ysgoloriaethau Tsieina (CSC). Os oes gennych ddiddordeb mewn gwneud ymchwil gyda mi, cysylltwch â mi trwy e-bost cyn gynted â phosibl fel y gall terfynau cau / camau ychwanegol fod yn berthnasol.
Goruchwyliaeth PhD Cyfredol
- Stephen Miles (2021/07 - nawr) - Adfer Delweddau gyda Dysgu Dwfn (Rhan-amser).
- Jinqi Wang (2022/10 - nawr) - Creu Cynnwys Anime sy'n seiliedig ar AI
- Zhuoling Jiang (2024/01 - nawr) - NPCs wedi'u gyrru gan AI ar gyfer Hapchwarae'r Genhedlaeth Nesaf (a ariennir gan yr ysgol)
- Yuan Wang (2024/10 - nawr) - Ffotograffiaeth wedi'i bweru gan AI
Cyd-oruchwylio
- Nada Saad M Alharbi (2023/10 - nawr) - Canfod Anhwylder Sbectrwm Awtistiaeth gan ddefnyddio Dysgu Atgyfnerthu Dwfn (wedi'i oruchwylio ar y cyd â Dr. Xianfang Sun).
- Yang Li (2023/10 - nawr) - Modelau Cynhyrchiol ar gyfer Dylunio Ffasâd Pensaernïol (wedi'i oruchwylio ar y cyd â Dr. Bailin Deng a'r Athro Wassim Jabi).
- Zhengwen Chen (2024/10 - nawr) - Mireinio Uwch o Fodelau Cynhyrchu Delweddau ar Raddfa Fawr wedi'u hyfforddi ymlaen llaw ar gyfer Perfformiad a Rheolaeth Well mewn Tasgau Arbenigol (cyd-oruchwylio gyda'r Athro Yukun Lai a Dr. Oktay Karakus)
Goruchwyliaeth gyfredol
Prosiectau'r gorffennol
- Hateef Alshewaier (2021/04 - 2026/05) - Dulliau Ensemble ar gyfer Dosbarthu Data Amlgyfrwng (wedi'i oruchwylio ar y cyd â Dr. Xianfang Sun).
- Ebtihal Alwadee (2021/10 - 2026/05) - Segmentu Tiwmor yr Ymennydd o MRI Aml-Foddol: Modelau Effeithlon ac Asesiad Cadernid Strwythuredig (wedi'i oruchwylio ar y cyd â Dr. Frank Langbein a Dr. Xianfang Sun).
- Shuang Song (2021/10 - 2026/02) - Deall a Rheoli Priodoleddau Wyneb gyda StyleGAN (CSC)
- Yuanbang Liang (2021/10 - 2025/09) - Samplu Ymwybyddiaeth Hubness ar gyfer Modelau Cynhyrchiol Dwfn mewn Cynhyrchu a Gwerthuso (EPSRC DTP).
- Xin Zhao (2019/10 - 2025/06) - Modelu Saliency ar gyfer Prosesu Delweddau Canfyddiadol (wedi'i oruchwylio ar y cyd â'r Athro Hantao Liu).
- Thomas Poudevigne-Durance (2019/10 - 2024/06) - Rhwydweithiau Gwrthwynebol Cynhyrchiol ar gyfer Cynyddu Digwyddiadau Prin (dan oruchwyliaeth ar y cyd â'r Athro Owen Jones, Ysgol Mathemateg).
Contact Details
+44 29208 75537
Abacws, Ystafell 2.20, Ffordd Senghennydd, Cathays, Caerdydd, CF24 4AG
Themâu ymchwil
Arbenigeddau
- Deallusrwydd artiffisial
- Graffeg cyfrifiadurol
- Golwg cyfrifiadurol
- Dysgu peirianyddol
- Rhyngweithio Cyfrifiadur Dynol