Dr Wei Zhou
(Gofynnwch i mi am fy rhagenwau)
BEng, PhD, SMIEEE, FHEA
- Ar gael fel goruchwyliwr ôl-raddedig
Timau a rolau for Wei Zhou
Uwch Ddarlithydd mewn Cyfrifiadura Gweledol Deallus
Trosolwyg
Ar hyn o bryd rwy'n Athro Cyswllt (Uwch Ddarlithydd) yn yr Ysgol Cyfrifiadureg a Gwybodeg, ac yn aelod o'r grŵp ymchwil Cyfrifiadura Gweledol . Rwyf hefyd yn aelod o bwyllgor Sefydliad Safonau Prydain a Choleg Adolygu Cymheiriaid EPSRC. Rwy'n gwasanaethu fel Cadeirydd Pwyllgor Etholiadau Pennod SPS IEEE UK ac Iwerddon. Roeddwn yn Gymrawd Ôl-ddoethurol ym Mhrifysgol Waterloo, Canada. Derbyniais y Ph.D. gradd o Brifysgol Gwyddoniaeth a Thechnoleg Tsieina yn 2021, ar y cyd â Phrifysgol Waterloo rhwng 2019 a 2021. Roeddwn i'n arfer bod yn athro gwadd yn y Gadair Gwybodeg Iechyd ym Mhrifysgol Dechnegol Munich, yr Ysgol Peirianneg Biofeddygol ym Mhrifysgol Shenzhen, yr Ysgol AI ym Mhrifysgol Technoleg Dalian; ysgolhaig gwadd yn y Sefydliad Cenedlaethol Gwybodeg, Japan; cynorthwyydd ymchwil gydag Intel; intern ymchwil gyda Microsoft Research ac Alibaba Group (y ddau gyda gwobrau rhagorol). Rwy'n Uwch Aelod o IEEE, gyda chysylltiadau â Chymdeithas Cylchedau a Systemau IEEE, Cymdeithas Prosesu Signalau IEEE, a Chymdeithas Cudd-wybodaeth Gyfrifiadurol IEEE. Cefais fy enwi yn rhifyn 2024 o restr Prifysgol Stanford o Wyddonwyr 2% Gorau'r Byd. Rwy'n Gymrawd yr Academi Addysg Uwch (FHEA).
Mae fy niddordebau ymchwil yn rhychwantu cyfrifiadura amlgyfrwng, prosesu delweddau canfyddiadol, a gweledigaeth gyfrifiadurol. Yn ystod y blynyddoedd diwethaf, rwyf wedi cyhoeddi dros 160 o bapurau cyfeirio yn y meysydd hyn, gan gynnwys y rhai a gyhoeddwyd mewn cyfnodolion mawreddog a chynadleddau rhyngwladol. Mae fy ymchwil wedi cael effaith yn y byd go iawn ar ddatblygiad technegau golwg lefel isel yn Alibaba Cloud. Cefais Wobr Sôn Anrhydeddus IEEE CASS VSPC Rising Star (y derbynnydd cyntaf ym maes ymchwil IQA ledled y byd) yn 2024, Gwobr Traethawd Doethurol Eithriadol ACM SIGMM China (1af ledled y wlad) yn 2022, a Gwobr Enillydd Grand Challenge (safle 1af) yn CVPR CLIC 2021.
Ewch i'm gwefan bersonol am fwy o fanylion.
Cyhoeddiadau cynrychioliadol
- Asesiad Ansawdd Delwedd Omnidirectional Stereosgopig Dall gan ddefnyddio Hierarchaeth Codio Rhagfynegol
Wei Zhou, A Kaup
Trafodion IEEE ar Amlgyfrwng, 2026 - Optimeiddio Ansawdd Canfyddiadol Uwch-Ddatrysiad Delwedd
Wei Zhou, Y Li, H Amirpour, X Hao, J Liu, P Wang, H Liu
Cynhadledd Ryngwladol IEEE ar Acwsteg, Lleferydd, a Phrosesu Signalau, 2026 - Ansawdd Profiad Amlgyfrwng yn Cwrdd â Deallusrwydd Peiriant
Wei Zhou, H Amirpour, T Hossfeld
Cofnodion ACM SIGMultimedia, 2026 - Asesiad Gweledol Canfyddiadol mewn Cyfathrebu Amlgyfrwng
Wei Zhou, H Amirpour
Cynhadledd Ryngwladol ACM ar Amlgyfrwng, 2025 - MCHM25: Cyfrifiadura Amlgyfrwng ar gyfer Iechyd a Meddygaeth
Wei Zhou, H Amirpour, L Yu, J Han, R Hong, PL Rosin
Cynhadledd Ryngwladol ACM ar Amlgyfrwng, 2025 - Asesiad Ansawdd Dyfnder Canfyddiadol o Ddelweddau Omnidirectional Stereosgopig
Wei Zhou, Z Wang
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo, 2024 | [Y mesur ansawdd dyfnder canfyddiadol cyntaf ar gyfer VR stereosgopig] - Asesiad Ansawdd Dall o Gymylau Pwynt 3D Trwchus gydag Ailsamplu dan Arweiniad Strwythur
Wei Zhou, Q Yang, W Chen, Q Jiang, G Zhai, W Lin
Trafodion ACM ar Gyfrifiadura Amlgyfrwng, Cyfathrebu a Chymwysiadau, 2024 | [Yr offeryn defnyddiol i werthuso ansawdd canfyddiadol cymylau pwynt heb gyfeirnod] - Gwerthusiad Ansawdd Delwedd Dehazed: O Anghysondeb Rhannol i Ganfyddiad Dall
Wei Zhou, R Zhang, L Li, G Yue, J Gong, H Chen, H Liu
Trafodion IEEE ar Gerbydau Deallus, 2024 | [Y metrig RR / NR cyntaf ar gyfer delweddau dehazed, profion ar ddata delweddau synthetig, byd go iawn a modurol, a gellir ei gymhwyso i ddad-hazing delweddau] - Asesiad ansawdd cyfeirio llai o gymylau pwynt trwy amcanestyniad saliency sy'n canolbwyntio ar gynnwys
Wei Zhou, G Yue, R Zhang, Y Qin, H Liu
Llythyrau Prosesu Signal IEEE, 2023 | [Y metrig RR cyntaf sy'n seiliedig ar ddelwedd ar gyfer cymylau pwynt 3D] - Asesiad Ansawdd Delwedd Omnidirectional Dall: Integreiddio Ystadegau Lleol a Semanteg Fyd-eang
Wei Zhou, Z Wang
Cynhadledd Ryngwladol IEEE ar Brosesu Delweddau, 2023 - Asesiad Ansawdd Uwch-Ddatrysiad Delwedd: Cydbwyso Ffyddlondeb Deterministic ac Ystadegol
Wei Zhou, Z Wang
Cynhadledd Ryngwladol ACM ar Amlgyfrwng, 2022 | [Mae'r gwaith cyntaf yn edrych ar ansawdd ailadeiladu uwch-ddatrysiad delwedd mewn gofod ffyddlondeb 2D] - Arolwg byr ar asesiad ansawdd ffrydio fideo addasol
Wei Zhou, X Min, H Li, Q Jiang
Cyfnodolyn Cyfathrebu Gweledol a Chynrychiolaeth Delwedd, 2022 - Asesiad Ansawdd Dim Cyfeirio ar gyfer Delweddau 360 Gradd trwy Ddadansoddi Gwybodaeth Amlamledd a Naturioldeb Lleol-Fyd-eang
Wei Zhou, J Xu, Q Jiang, Z Chen
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo, 2021 | [Yr algorithm ystadegau golygfeydd naturiol cyntaf ar gyfer asesu ansawdd VR] - Asesiad Ansawdd Uwch-Ddatrysiad Delwedd: Ffyddlondeb Strwythurol yn erbyn Naturioldeb Ystadegol
Wei Zhou, Z Wang, Z ChenCynhadledd Ryngwladol IEEE ar Ansawdd Profiad Amlgyfrwng, 2021 (Papur gwahoddedig)
- Dysgu Nodweddion Aml-Raddfa Dwfn ar gyfer Asesu Ansawdd Delwedd Ystumiedig
Wei Zhou, Z ChenSymposiwm Rhyngwladol IEEE ar Gylchedau a Systemau, 2021
- Asesiad Ansawdd Delwedd Maes Golau Dim Cyfeirio Tensor
Wei Zhou, L Shi, Z Chen, J Zhang
Trafodion IEEE ar Brosesu Delweddau, 2020 | [Mae'r gwaith cyntaf yn archwilio theori tensor ar gyfer gwerthuso ansawdd delwedd maes golau] - Asesiad Ansawdd Dall ar gyfer Uwchddatrysiad Delwedd gan ddefnyddio Rhwydweithiau Convolutional Dwy Ffrwd Dwfn
Wei Zhou, Q Jiang, Y Wang, Z Chen, W Li
Gwyddorau Gwybodaeth, 2020 - Agregu Nodweddion Gofodol Lleol a Byd-eang Dwfn ar gyfer Asesu Ansawdd Fideo Dall
Wei Zhou, Z Chen
Cynhadledd Ryngwladol IEEE ar Gyfathrebu Gweledol a Phrosesu Delweddau, 2020 - Rhwydweithiau Rhyngweithiol Deuol-Ffrwd ar gyfer Asesiad Ansawdd Delwedd Stereosgopig Heb Gyfeiriad
Wei Zhou, Z Chen, W Li
Trafodion IEEE ar Brosesu Delweddau, 2019 | [Mae'r gwaith rhagfynegi ansawdd delwedd stereosgopig cyntaf sy'n seiliedig ar ddysgu dwfn yn ystyried natur ryngweithiol hierarchaidd deuol-ffrwd y system weledol ddynol] - Rhagfynegiad ansawdd fideo stereosgopig yn seiliedig ar rwydweithiau niwral dwfn ffrwd ddeuol o'r diwedd i'r diwedd
Wei Zhou, Z Chen, W Li
Cynhadledd Pacific-Rim ar Amlgyfrwng, 2018
Arweinyddiaeth academaidd
Golygydd Cyswllt, Trafodion IEEE ar Rhwydweithiau Niwral a Systemau Dysgu, 2024-presennol
Golygydd Cyswllt, Trafodion ACM ar Gyfrifiadura, Cyfathrebu a Chymwysiadau Amlgyfrwng, 2025-presennol
Golygydd Cyswllt, Cydnabyddiaeth Patrymau, 2024-presennol
Os oes gennych ddiddordeb mewn gweithio gyda mi, anfonwch e-bost ataf os gwelwch yn dda. Sylwch mai dim ond myfyrwyr cyfatebol fydd yn cael eu hymateb.
Cyhoeddiad
2027
- Yang, Y. et al., 2027. BiSQAFusion: Multi-level fusion for personalized binaural speech quality assessment in hearing aids. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 139 (Part A) 104765. (10.1016/j.inffus.2026.104765)
- Liu, J. et al., 2027. SCMM: Calibrating cross-modal representations for text-based person search. Pattern Recognition 182 114775. (10.1016/j.patcog.2026.114775)
- Yan, W. et al., 2027. 3D object detection and knowledge distillation in autonomous driving: A survey. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 138 104706. (10.1016/j.inffus.2026.104706)
2026
- Shen, X. et al., 2026. BeatDance: Generating beat-consistent 3D dance with hierarchical spatial–temporal modeling. Pattern Recognition 180 114344. (10.1016/j.patcog.2026.114344)
- Wang, Z. et al., 2026. Robust low-light image enhancement in the wild via data synthesis and generative diffusion prior. Pattern Recognition 178 113336. (10.1016/j.patcog.2026.113336)
- Wang, H. et al., 2026. KSIQA: A knowledge-sharing model for no-reference image quality assessment. IEEE Transactions on Neural Networks and Learning Systems 37 (8), pp.3943-3955. (10.1109/tnnls.2026.3656757)
- Yang, D. et al., 2026. Toward empathetic care: an LLM-based multi-intention recognition framework for mental health and complex medical queries. IEEE Transactions on Affective Computing (10.1109/taffc.2026.3728254)
- Yu, L. et al., 2026. Blind image quality assessment via a hierarchical perceptual modulation network. ACM Transactions on Multimedia Computing, Communications and Applications (10.1145/3840393)
- Zhao, C. et al., 2026. Uncertainty-guided spatiotemporal consistency fusion network for infrared-visible video fusion under extremely low-light conditions. IEEE Transactions on Image Processing 35 , pp.8894-8909. (10.1109/tip.2026.3719477)
- Zou, M. et al., 2026. Pose-guided multi-cue explicit query construction for disambiguating human-object interactions. IEEE Transactions on Circuits and Systems for Video Technology 36 (7), pp.10794-10809. (10.1109/TCSVT.2026.3667102)
- Guo, Z. et al., 2026. GP-GS: Gaussian processes densification for 3D Gaussian Splatting. Presented at: 22nd International Conference on Intelligent Computing Toronto, ON, Canada 22-26 July 2026. Published in: Li, G. , Filipe, J. and Xu, Z. eds. Vol. 3037.[GP-GS: Gaussian Processes Densification for 3D Gaussian Splatting]. Singapore: Springer. , pp.139-150. (10.1007/978-981-92-3548-3_12)
- Yu, J. et al., 2026. One aligned LLM to serve them all: A transfer recipe for training VLMs without visual-language re-alignment. International Journal of Computer Vision 134 (7) 331. (10.1007/s11263-026-02930-z)
- Amirpour, H. et al., 2026. Quality-complexity trade-offs for sustainable media delivery. Presented at: 2026 18th International Conference on Quality of Multimedia Experience (QoMEX) Cardiff, UK 29 June 2026 - 3 July 2026. 2026 18th International Conference on Quality of Multimedia Experience (QoMEX). IEEE(10.1109/qomex69967.2026.11618319)
- Zhang, H. et al. 2026. Cross-modal interaction for multi-dimensional AI-generated image quality assessment. Presented at: 2026 18th International Conference on Quality of Multimedia Experience (QoMEX) Cardiff, UK 29 June 2026 - 03 July 2026. 2026 18th International Conference on Quality of Multimedia Experience (QoMEX). IEEE(10.1109/qomex69967.2026.11618318)
- Wu, J. et al., 2026. Digital human generation for games via tightness-aware multi-cue modeling. IEEE Transactions on Games (10.1109/tg.2026.3708023)
- Zhang, L. et al., 2026. Rethinking the effect of unimodal labels in multimodal sentiment analysis. ACM Transactions on Multimedia Computing, Communications, and Applications 22 (7) 204. (10.1145/3796718)
- Zhao, M. et al., 2026. A new semi-supervised video anomaly detection baseline in lack of anomalous samples. ACM Transactions on Multimedia Computing, Communications and Applications 22 (7) 189. (10.1145/3797034)
- Hao, X. et al., 2026. DADA++: Dual Alignment Domain Adaptation for unsupervised video-text retrieval. ACM Transactions on Multimedia Computing, Communications and Applications 22 (6) 175. (10.1145/3759252)
- Chen, B. et al., 2026. From global to granular: revealing IQA model performance via correlation surface. IEEE Transactions on Pattern Analysis and Machine Intelligence (10.1109/tpami.2026.3705184)
- Yin, Y. et al., 2026. Deep learning-based point cloud upsampling: A survey of methodologies, performance comparisons, and noise robustness analysis. Neurocomputing 681 133316. (10.1016/j.neucom.2026.133316)
- Hao, X. et al., 2026. Embodied spatial affordance: spatial-aware affordance learning for embodied navigation and manipulation. IEEE Transactions on Image Processing 35 , pp.6041-6054. (10.1109/tip.2026.3698366)
- Liu, J. et al. 2026. Saliency-guided action quality assessment: an AI-augmented framework for skill evaluation in physical education. Presented at: 2026 IEEE Conference on Artificial Intelligence (CAI) Granada, Spain 08 - 10 May 2026. 2026 IEEE Conference on Artificial Intelligence (CAI) Proceedings. IEEE. , pp.1997-2002. (10.1109/cai68641.2026.11536588)
- Yu, L. et al., 2026. DVLTA-VQA: Decoupled vision-language modeling with text-guided adaptation for blind video quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 36 (5), pp.6826-6837. (10.1109/tcsvt.2026.3657415)
- Zhou, W. et al. 2026. Quality of multimedia experience meets machine intelligence. ACM SIGMultimedia Records 18 (1) 3. (10.1145/3811013.3811016)
- You, L. et al., 2026. Integrating SAM supervision for 3D weakly supervised point cloud segmentation. IEEE Transactions on Image Processing 35 , pp.5212-5223. (10.1109/tip.2026.3691686)
- Kuang, Y. et al., 2026. A nature-inspired edge-cloud collaborative privacy framework for visual relocalization on consumer electronics. IEEE Transactions on Consumer Electronics (10.1109/TCE.2026.3692722)
- Amirpour, H. et al., 2026. BINR: Live video broadcasting quality assessment. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.22282-22286. (10.1109/icassp55912.2026.11460679)
- Wang, H. et al., 2026. Non-line-of-sight vehicle detection via audio-visual fusion. Presented at: ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.11807-11811. (10.1109/icassp55912.2026.11465095)
- Wang, Z. et al., 2026. VMambaMorph: A 3D multi-modality deformable image registration framework based on visual state space model with cross-scan module. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.8847-8851. (10.1109/icassp55912.2026.11464524)
- Zhou, W. et al. 2026. Perceptual quality optimization of image super-resolution. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-6 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.10537-10541. (10.1109/icassp55912.2026.11461596)
- Yang, Y. et al., 2026. Expressive human volumetric video generation with rich text. IEEE Transactions on Circuits and Systems for Video Technology 36 (4), pp.5424-5436. (10.1109/tcsvt.2025.3628996)
- Cen, J. et al., 2026. Depth-guided metric-aware temporal consistency for monocular video human mesh recovery. Presented at: International Conference on Acoustics, Speech, and Signal Processing Barcelona, Spain 3-8 May 2026. Proceedings of the 2026 International Conference on Acoustics, Speech, and Signal Processing. IEEE. , pp.10707-10711. (10.1109/icassp55912.2026.11463978)
- Chang, Y. et al., 2026. Perception-inspired network for stereo image quality assessment. IEEE Transactions on Image Processing (10.1109/tip.2026.3680564)
- Ni, Y. et al., 2026. Affine modulation-based audiogram fusion network for joint noise reduction and hearing loss compensation. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 127 (A) 103726. (10.1016/j.inffus.2025.103726)
- Yang, D. et al., 2026. MedAide: information fusion and anatomy of medical intents via LLM-based agent collaboration. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 127 (Part A) 103743. (10.1016/j.inffus.2025.103743)
- Li, Y. et al., 2026. Temporal inconsistency guidance for super-resolution video quality assessment. Presented at: The 40th Annual AAAI Conference on Artificial Intelligence Singapore 20-27 January 2026. Vol. 40 (8)., pp.6681-6689. (10.1609/aaai.v40i8.37599)
- Li, J. et al., 2026. SearchExpert: A GenAI-driven framework for reasoning-intensive multimedia information fusion through fine-tuning and reinforcement learning. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 126 (Part B) 103665. (10.1016/j.inffus.2025.103665)
- Zhu, J. et al., 2026. Is there a relationship between Mean Opinion Score (MOS) and Just Noticeable Difference (JND)?. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 01-04 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396801)
- Yuan, H. et al., 2026. DWCL: Dual-weighted contrastive learning for robust multi-view clustering. Engineering Applications of Artificial Intelligence 165 (PartB) 113532. (10.1016/j.engappai.2025.113532)
- Liu, W. et al., 2026. Exploiting independent query information for few-shot image segmentation. Displays 91 103179. (10.1016/j.displa.2025.103179)
- Yu, Q. et al., 2026. StealthMark: Harmless and stealthy ownership verification for medical segmentation via uncertainty-guided backdoors. IEEE Transactions on Image Processing 35 , pp.1290-1304. (10.1109/tip.2026.3655563)
- Liu, J. et al., 2026. Self-supervised unfolding network with shared reflectance learning for low-light image enhancement. IEEE Transactions on Image Processing 35 , pp.800-815. (10.1109/tip.2026.3652021)
- Li, H. et al., 2026. EHIN: Early-aware hierarchical interaction network for weakly-supervised referring image segmentation. Neurocomputing 659 131764. (10.1016/j.neucom.2025.131764)
- Xue, J. et al., 2026. Towards comprehensive interactive change understanding in remote sensing: A large-scale dataset and dual-granularity enhanced VLM. IEEE Transactions on Geoscience and Remote Sensing 64 4401516. (10.1109/tgrs.2025.3650151)
2025
- Ghanbari, M. et al., 2025. STACK: Spatial tower assembly using controlled kinetics. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396840)
- Hou, J. et al., 2025. Frequency-aware native resolution assessment of 8K omnidirectional images. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396856)
- Ju, Y. et al., 2025. Photometric regularization for 3D gaussian splatting in multi-view surface projection. IEEE Journal of Selected Topics in Signal Processing 19 (8), pp.1682-1693. (10.1109/jstsp.2025.3617861)
- Li, Y. et al., 2025. Unlocking implicit motion for evaluating image complexity. Displays 90 103131. (10.1016/j.displa.2025.103131)
- Li, Y. et al., 2025. Perception-oriented bidirectional attention network for image super-resolution quality assessment. IEEE Transactions on Image Processing 34 , pp.7728-7743. (10.1109/tip.2025.3633145)
- Liu, Y. et al., 2025. M 2 S 2 L: Mamba-based multi-scale spatial-temporal learning for video anomaly detection. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. IEEE(10.1109/vcip67698.2025.11396919)
- Wang, J. et al., 2025. CVBench: benchmarking and comparing video generation with large multimodal models. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396889)
- Wang, R. et al., 2025. Multi-view residual spatio-temporal topology adaptive graph convolutional network for urban road traffic accident prediction with multi-source risks. Array 28 100617. (10.1016/j.array.2025.100617)
- Zhang, Z. et al., 2025. Large multimodal models evaluation: a survey. SCIENCE CHINA Information Sciences 68 221301. (10.1007/s11432-025-4676-4)
- Liu, W. et al., 2025. Integrating large foundation models into multimodal named entity recognition with evidential fusion. Neurocomputing 652 131015. (10.1016/j.neucom.2025.131015)
- Li, X. et al., 2025. MSPoint-Gait: Multi-Scale Point cloud analysis for 3D gait recognition via cross-modal learning. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June - 4 July 2025. 2025 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/icme59968.2025.11209453)
- Ma, Y. et al., 2025. Analysing and predicting radiologists’ expertise using eye-tracking data: Insights for diagnostic decision-making. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June - 4 July 2025. IEEE. , pp.1-6. (10.1109/icme59968.2025.11209585)
- Yang, Y. et al., 2025. MCSMoG: Multi-Conditional Diffusion for stylized motion generation with parametric control. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June 2025 - 4 July 2025. 2025 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/icme59968.2025.11209286)
- Gao, C. et al., 2025. Compressed feature quality assessment: Dataset and baselines. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13450-13456. (10.1145/3746027.3758309)
- Gao, L. et al., 2025. EEmo-Bench: A benchmark for multi-modal large language models on image evoked emotion assessment. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.7064-7073. (10.1145/3746027.3755777)
- Ghanbari, M. et al., 2025. SDART: spatial dart AR simulation with hand-tracked input. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. New York, NY, United States: ACM. , pp.13543-13545. (10.1145/3746027.3754484)
- Meng, Y. et al., 2025. VideoForest: Person-anchored hierarchical reasoning for cross-video question answering. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.836-845. (10.1145/3746027.3754573)
- Miao, C. et al., 2025. MFFI: Multi-dimensional Face Forgery Image dataset for real-world scenarios. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13235-13242. (10.1145/3746027.3758280)
- Wang, P. et al., 2025. A spatial relationship aware dataset for robotics. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13332-13338. (10.1145/3746027.3758293)
- Zhou, W. and Amirpour, H. 2025. Perceptual visual quality assessment in multimedia communication. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.14340-14341. (10.1145/3746027.3759205)
- Gholami, A. et al., 2025. Perceptual quality assessment of spatial videos on Apple Vision Pro. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.20-28. (10.1145/3746269.3760422)
- Li, X. et al., 2025. DepthGait: Multi-scale cross-level feature fusion of RGB-derived depth and silhouette sequences for robust gait recognition. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.2333-2341. (10.1145/3746027.3755876)
- Wang, Z. et al., 2025. Evaluating perceptual color preferences in smartphone photography: dataset and challenges. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.12844-12850. (10.1145/3746027.3758227)
- Zou, M. et al., 2025. PhysLab: A benchmark dataset for multi-granularity visual parsing of physics experiments. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.12799-12806. (10.1145/3746027.3758221)
- Huang, F. et al., 2025. VQualA 2025 document image quality assessment challenge. Presented at: 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) Honolulu, HI, USA 19-20 October 2025. 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW). IEEE. , pp.3344-3353. (10.1109/ICCVW69036.2025.00351)
- Yao, X. et al., 2025. AMLPF-CLIP: Adaptive prompting and distilled learning for imbalanced histopathological image classification. IEEE Journal of Biomedical and Health Informatics (10.1109/jbhi.2025.3619343)
- Zeng, Y. et al., 2025. CLIP-DQA V2: Exploring CLIP for dehazed image quality assessment from a fragment-level perspective. IEEE Signal Processing Letters 32 , pp.3829-3833. (10.1109/lsp.2025.3615082)
- Amirpour, H. et al., 2025. VQM4HAS: A real-time quality metric for HEVC videos in HTTP Adaptive Streaming. IEEE Transactions on Multimedia 27 , pp.9619-9631. (10.1109/tmm.2025.3613110)
- Yu, Q. et al., 2025. Parameterized diffusion optimization enabled autoregressive ordinal regression for diabetic retinopathy grading. Presented at: MICCAI 2025 Daejeon, Republic of Korea 23-27 September 2025. Published in: Gee, J. C. et al., Proceedings Medical Image Computing and Computer Assisted Intervention. Lecture Notes in Computer Science. Vol. 15974.Switzerland: Springer Nature. , pp.450-460. (10.1007/978-3-032-05182-0_44)
- Liu, J. et al. 2025. Adaptive spatiotemporal graph transformer network for action quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 35 (7), pp.6628-6639. (10.1109/TCSVT.2025.3541456)
- Luo, Y. et al., 2025. Multi-attribute continual learning for blind image quality assessment. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) London 25-28 May 2025. 2025 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE(10.1109/iscas56072.2025.11043383)
- Zeng, Y. et al., 2025. CLIP-DQA: Blindly evaluating dehazed images from global and local perspectives using CLIP. Presented at: IEEE International Symposium on Circuits and Systems London, UK 25-28 May 2025. 2025 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE. (10.1109/iscas56072.2025.11043366)
- Chen, W. et al., 2025. No-reference point cloud quality assessment via graph convolutional network. IEEE Transactions on Multimedia 27 , pp.2489-2502. (10.1109/TMM.2024.3521845)
- Luo, Y. et al., 2025. Unsupervised low-light image enhancement with self-paced learning. IEEE Transactions on Multimedia 27 , pp.1808-1820. (10.1109/TMM.2024.3521752)
- Guanghui, Y. et al., 2025. Subjective and objective quality assessment of colonoscopy videos. IEEE Transactions on Medical Imaging 44 (2), pp.841-854. (10.1109/TMI.2024.3461737)
- Liu, W. et al., 2025. Physically-guided open vocabulary segmentation with weighted patched alignment loss. Neurocomputing 614 128788. (10.1016/j.neucom.2024.128788)
2024
- Zhou, W. and Wang, Z. 2024. Perceptual depth quality assessment of stereoscopic omnidirectional images. IEEE Transactions on Circuits and Systems for Video Technology 34 (12), pp.13452-13462. (10.1109/TCSVT.2024.3449696)
- Zhang, R. et al., 2024. Similarities between wheels and tracks: A "tire model" for tracked vehicles. IEEE Transactions on Vehicular Technology 73 (11), pp.16416-16431. (10.1109/TVT.2024.3421285)
- Zhang, X. et al., 2024. Neural image re-exposure. Computer Vision and Image Understanding 248 104094. (10.1016/j.cviu.2024.104094)
- Chen, W. et al., 2024. Dynamic hypergraph convolutional network for no-reference point cloud quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 34 (10), pp.10479-10493. (10.1109/TCSVT.2024.3410052)
- Wang, H. et al. 2024. Blind image quality assessment via adaptive graph attention. IEEE Transactions on Circuits and Systems for Video Technology 34 (10), pp.10299-10309. (10.1109/TCSVT.2024.3405789)
- Li, Y. et al., 2024. Deep bi-directional attention network for image super-resolution quality assessment. Presented at: IEEE International Conference on Multimedia and Expo (ICME) Niagra Falls, Canada 15-19 July 2024. 2024 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/ICME57554.2024.10687430)
- Zhou, T. et al., 2024. Adaptive mixed-scale feature fusion network for blind AI-generated image quality assessment. IEEE Transactions on Broadcasting 70 (3), pp.833-843. (10.1109/TBC.2024.3391060)
- Xu, J. et al., 2024. LISD: An efficient multi-task learning framework for LiDAR segmentation and detection. Presented at: IEEE International Conference on Image Processing (ICIP) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.3341-3347. (10.1109/ICIP51287.2024.10647535)
- Zhou, W. et al. 2024. Blind quality assessment of dense 3D point clouds with structure guided resampling. ACM Transactions on Multimedia Computing, Communications and Applications 20 (8)(10.1145/3664199)
- Shafiee Sarvestani, A. , Zhou, W. and Wang, Z. 2024. Perceptual crack detection for rendered 3D textured meshes. Presented at: IEEE International Conference on Quality of Multimedia Experience (QoMEX) Karlshamn, Sweden 18 - 20 June 2024. Proceedings of 16th International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.1-7. (10.1109/QoMEX61742.2024.10598253)
- Yan, W. et al., 2024. FVIFormer: flow-guided global-local aggregation transformer network for video inpainting. IEEE Journal of Emerging and Selected Topics in Circuits and Systems 14 (2), pp.235-244. (10.1109/JETCAS.2024.3392972)
- Yue, G. et al., 2024. Subjective and objective quality assessment of multi-attribute retouched face images. IEEE Transactions on Broadcasting 70 (2), pp.570-583. (10.1109/TBC.2024.3374043)
- Fu, J. et al., 2024. Vision-language consistency guided multi-modal prompt learning for blind AI generated image quality assessment. IEEE Signal Processing Letters 31 , pp.1820-1824. (10.1109/LSP.2024.3420083)
- Yue, G. et al., 2024. Dual-constraint coarse-to-fine network for camouflaged object detection. IEEE Transactions on Circuits and Systems for Video Technology 34 (5), pp.3286-3298. (10.1109/TCSVT.2023.3318672)
- Zhang, R. et al., 2024. A terramechanics-based dynamic model for motion control of unmanned tracked vehicles. IEEE Transactions on Intelligent Vehicles (10.1109/TIV.2024.3406582)
- Fu, J. , Zhou, W. and Chen, Z. 2024. Bayesian graph convolutional network for traffic prediction. Neurocomputing 582 127507. (10.1016/j.neucom.2024.127507)
- Yue, G. et al., 2024. Boundary refinement network for colorectal polyp segmentation in colonoscopy images. IEEE Signal Processing Letters 31 , pp.954-958. (10.1109/LSP.2024.3378106)
- Zhou, W. et al. 2024. Dehazed image quality evaluation: from partial discrepancy to blind perception. IEEE Transactions on Intelligent Vehicles 9 (2), pp.3843-3858. (10.1109/TIV.2024.3356055)
- Yi, X. , Jiang, Q. and Zhou, W. 2024. No-reference quality assessment of underwater image enhancement. Displays 81 102586. (10.1016/j.displa.2023.102586)
- Yue, G. et al., 2024. Subjective quality assessment of thermal infrared images. Presented at: IEEE International Conference on Image Processing (ICIP 2024) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.1212-1217. (10.1109/ICIP51287.2024.10648145)
2023
- Yue, G. et al., 2023. Subjective quality assessment of enhanced retinal images. Presented at: International Conference on Image Processing Kuala Lumpur, Malaysia 8–11 October 2023. 2023 IEEE International Conference on Image Processing Proceedings. IEEE. , pp.3005-3009. (10.1109/ICIP49359.2023.10222541)
- Zhou, W. and Wang, Z. 2023. Blind omnidirectional image quality assessment: integrating local statistics and global semantics. Presented at: International Conference on Image Processing Kuala Lumpur, Malaysia 8–11 October 2023. 2023 IEEE International Conference on Image Processing Proceedings. IEEE. , pp.1405-1409. (10.1109/ICIP49359.2023.10222049)
- Zhou, W. et al. 2023. A brief survey on adaptive video streaming quality assessment. Journal of Visual Communication and Image Representation 86 103526. (10.1016/j.jvcir.2022.103526)
- Sun, S. et al., 2023. GraphIQA: Learning Distortion Graph Representations for Blind Image Quality Assessment. IEEE Transactions on Multimedia 25 , pp.2912-2925. (10.1109/TMM.2022.3152942)
- Liu, J. et al., 2023. LIQA: lifelong blind image quality assessment. IEEE Transactions on Multimedia 25 , pp.5358-5373. (10.1109/TMM.2022.3190700)
- Zhou, W. et al. 2023. Reduced-reference quality assessment of point clouds via content-oriented saliency projection. IEEE Signal Processing Letters 30 , pp.354-358. (10.1109/LSP.2023.3264105)
2022
- Fu, J. et al., 2022. Adaptive hypergraph convolutional network for no-reference 360-degree image quality assessment. Presented at: MM '22: The 30th ACM International Conference on Multimedia Lisboa Portugal 10-14 October 2022. MM '22: Proceedings of the 30th ACM International Conference on Multimedia. Association for Computing Machinery. , pp.961-969. (10.1145/3503161.3548337)
- Zhou, W. and Wang, Z. 2022. Quality assessment of image super-resolution: balancing deterministic and statistical fidelity. Presented at: MM '22: The 30th ACM International Conference on Multimedia Lisboa Portugal 10-14 October 2022. MM '22: Proceedings of the 30th ACM International Conference on Multimedia. Association for Computing Machiner. , pp.934-942. (10.1145/3503161.3547899)
- Xu, J. et al., 2022. Quality assessment of multi-exposure image fusion by synthesizing local and global intermediate references. Displays 74 102188. (10.1016/j.displa.2022.102188)
- Lu, Y. et al., 2022. RTN: Reinforced transformer network for coronary CT angiography vessel-level image quality assessment. Presented at: MICCAI 2022 18-22 September 2022. Medical Image Computing and Computer Assisted Intervention – MICCAI 2022.. Vol. 13431.Lecture Notes in Computer Science Cham. Switzerland: Springer. (10.1007/978-3-031-16431-6_61)
- Zhou, W. et al. 2022. No-reference quality assessment for 360-degree images by analysis of multifrequency information and local-global naturalness. IEEE Transactions on Circuits and Systems for Video Technology 32 (4), pp.1778-1791. (10.1109/TCSVT.2021.3081182)
- Huang, S. et al., 2022. Perceptual evaluation of pre-processing for video transcoding. Presented at: 2021 International Conference on Visual Communications and Image Processing (VCIP) Munich, Germany 5-8 December 2021. 2021 International Conference on Visual Communications and Image Processing (VCIP). IEEE. , pp.1-5. (10.1109/VCIP53242.2021.9675438)
- Fang, Y. et al., 2022. A Bayesian deep image prior downscaling approach for high-resolution soil moisture estimation. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing 15 , pp.4571 - 4582. (10.1109/JSTARS.2022.3177081)
2021
- Mao, S. et al., 2021. Mechanism and application of capacitive-coupled memristive behavior based on a biomaterial developed memristive device. ACS Applied Electronic Materials 3 (12), pp.5537-5547. (10.1021/acsaelm.1c00951)
- Ling, S. et al., 2021. Re-visiting discriminator for blind free-viewpoint image quality assessment. IEEE Transactions on Multimedia 23 , pp.4245-4258. (10.1109/TMM.2020.3038305)
- Xu, J. et al., 2021. Perceptual quality assessment of internet videos. Presented at: MM '21: 29th ACM International Conference on Multimedia Virtual Event China 20 - 24 October 2021. MM '21: Proceedings of the 29th ACM International Conference on Multimedia. Association for Computing Machinery. , pp.1248-1257. (10.1145/3474085.3475486)
- Peng, Y. et al., 2021. Multi-metric fusion network for image quality assessment. Presented at: Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) Nashville, TN, USA 19-25 June 2021. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). IEEE. , pp.1857-1860. (10.1109/CVPRW53098.2021.00205)
- Zhou, W. , Wang, Z. and Chen, Z. 2021. Image super-resolution quality assessment: structural fidelity versus statistical naturalness. Presented at: 13th International Conference on Quality of Multimedia Experience (QoMEX) Montreal, QC, Canada 14-17 June 2021. 2021 13th International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.61-64. (10.1109/QoMEX51781.2021.9465479)
- Zhou, W. and Chen, Z. 2021. Deep multi-scale features learning for distorted image quality assessment. Presented at: International Symposium on Circuits and Systems (ISCAS) Daegu, Korea 22-28 May 2021. 2021 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE(10.1109/ISCAS51556.2021.9401285)
- Xu, J. , Zhou, W. and Chen, Z. 2021. Blind omnidirectional image quality assessment with viewport oriented graph convolutional networks. IEEE Transactions on Circuits and Systems for Video Technology 31 (5), pp.1724-1737. (10.1109/TCSVT.2020.3015186)
2020
- Zhou, W. and Chen, Z. 2020. Deep local and global spatiotemporal feature aggregation for blind video quality assessment. Presented at: 2020 IEEE International Conference on Visual Communications and Image Processing (VCIP) Macau 01-04 December 2020. 2020 IEEE International Conference on Visual Communications and Image Processing (VCIP). IEEE. (10.1109/VCIP49819.2020.9301764)
- Liu, J. et al., 2020. LIRA: Lifelong Image Restoration from Unknown Blended Distortions. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vedaldi, A. et al., Computer Vision – ECCV 2020. Lecture Notes in Computer Science Vol. 12363. Springer International Publishing. , pp.616-632. (10.1007/978-3-030-58523-5_36)
- Xu, J. et al., 2020. Binocular rivalry oriented predictive autoencoding network for blind stereoscopic image quality measurement. IEEE Transactions on Instrumentation and Measurement 70 5001413. (10.1109/TIM.2020.3026443)
- Jiang, Q. et al., 2020. A full-reference stereoscopic image quality measurement via hierarchical deep feature degradation fusion. IEEE Transactions on Instrumentation and Measurement 69 (12), pp.9784-9796. (10.1109/TIM.2020.3005111)
- Shi, L. et al., 2020. No-reference light field image quality assessment based on spatial-angular measurement. IEEE Transactions on Circuits and Systems for Video Technology 30 (11), pp.4114-4128. (10.1109/TCSVT.2019.2955011)
- Li, X. et al., 2020. Learning disentangled feature representation for hybrid-distorted image restoration. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vedaldi, A. et al., Computer Vision – ECCV 2020 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXIX. Lecture Notes in Computer Science Vol. 12374. Springer. , pp.313-329. (10.1007/978-3-030-58526-6_19)
- Zhou, W. et al. 2020. Blind quality assessment for image superresolution using deep two-stream convolutional networks. Information Sciences 528 , pp.205-218. (10.1016/j.ins.2020.04.030)
- Zhao, Y. et al., 2020. Infrared pedestrian detection with converted temperature map. Presented at: 2019 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) Lanzhou 18-21 November 2019. 2019 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC). IEEE. (10.1109/APSIPAASC47483.2019.9023228)
- Zhou, W. et al. 2020. Tensor oriented no-reference light field image quality assessment. IEEE Transactions on Image Processing 29 , pp.4070-4084. (10.1109/TIP.2020.2969777)
- Chen, Z. et al., 2020. Stereoscopic omnidirectional image quality assessment based on predictive coding theory. IEEE Journal of Selected Topics in Signal Processing 14 (1), pp.103-117. (10.1109/JSTSP.2020.2968182)
- Luo, Z. et al., 2020. No-reference light field image quality assessment based on micro-lens image. Presented at: Picture Coding Symposium, PCS Ningbo, China 12-15 November 2019. 2019 Picture Coding Symposium (PCS). IEEE. , pp.1-5. (10.1109/PCS48520.2019.8954551)
- Xu, J. et al., 2020. Quality assessment of stereoscopic 360-degree images from multi-viewports. Presented at: Picture Coding Symposium, PCS Ningbo, China 12-15 November 2019. 2019 Picture Coding Symposium (PCS). IEEE. , pp.1-5. (10.1109/PCS48520.2019.8954555)
2019
- Jin, X. et al., 2019. Unsupervised single image deraining with self-supervised constraints. Presented at: 2019 IEEE International Conference on Image Processing Taipei, Taiwan 22-25 September 2019. 2019 IEEE International Conference on Image Processing. IEEE. , pp.2761-2765. (10.1109/ICIP.2019.8803238)
- Jin, X. et al., 2019. AI-GAN: signal de-interference via asynchronous interactive generative adversarial network. Presented at: 2019 IEEE International Conference on Multimedia & Expo Workshops (ICMEW) Shanghai 08-12 July 2019. 2019 IEEE International Conference on Multimedia & Expo Workshops (ICMEW). IEEE. , pp.228-233. (10.1109/ICMEW.2019.00046)
- Zhou, W. , Chen, Z. and Li, W. 2019. Dual-stream interactive networks for no-reference stereoscopic image quality assessment. IEEE Transactions on Image Processing 28 (8), pp.3946-3958. (10.1109/TIP.2019.2902831)
- Zhao, S. et al., 2019. How do you perceive differently from an AI - A database for semantic distortion measurement. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) Sapporo, Japan 26-29 May 2019. 2019 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE. , pp.1-5. (10.1109/ISCAS.2019.8702774)
2018
- Jin, X. et al., 2018. Augmented coarse-to-fine video frame synthesis with semantic loss. Presented at: Pattern Recognition and Computer Vision, PRCV 2018 Guangzhou, China 23-26 November. Published in: Lai, J. et al., Pattern Recognition and Computer Vision, First Chinese Conference, PRCV 2018, Guangzhou, China, November 23-26, 2018, Proceedings, Part I. Lecture Notes in Computer Science Vol. 11256. Springer. , pp.439-452. (10.1007/978-3-030-03398-9_38)
- Xu, J. et al., 2018. Subjective quality assessment of stereoscopic omnidirectional image. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Lecture Notes in Computer Science Vol. 11164. Springer. , pp.589-599. (10.1007/978-3-030-00776-8_54)
- Zhang, W. et al., 2018. None ghosting artifacts stitching based on depth map for light field image. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Vol. 11164.Lecture Notes in Computer Science Vol. 11164. Springer. , pp.567-578. (10.1007/978-3-030-00776-8_52)
- Zhou, W. , Chen, Z. and Li, W. 2018. Stereoscopic video quality prediction based on end-to-end dual stream deep neural networks. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Vol. 11166.Lecture Notes in Computer Science Vol. 11166. Springer. , pp.482-492. (10.1007/978-3-030-00764-5_44)
- Shi, L. et al., 2018. Perceptual evaluation of light field image. Presented at: 2018 25th IEEE International Conference on Image Processing (ICIP) Athens, Greece 07-10 October 2018. 2018 25th IEEE International Conference on Image Processing (ICIP). IEEE. , pp.41-45. (10.1109/ICIP.2018.8451077)
- Jin, X. et al., 2018. A decomposed dual-cross generative adversarial network for image rain removal. Presented at: British Machine Vision Conference 2018 Newcastle, UK 3-6 September 2018.
- Hu, Y. et al., 2018. SDM: Semantic distortion measurement for video encryption. Presented at: 13th IEEE International Conference on Automatic Face & Gesture Recognition Xi'an, China 15-19 May 2018. 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2018). IEEE. , pp.764-768. (10.1109/FG.2018.00120)
- Zhou, Y. and Zhou, W. 2018. Visual comfort assessment for stereoscopic image retargeting. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) Florence, Italy 27-30 May 2018. 2018 IEEE International Symposium on Circuits and Systems (ISCAS). , pp.1-5. (10.1109/ISCAS.2018.8351198)
2017
- Chen, Z. , Zhou, W. and Li, W. 2017. Blind stereoscopic video quality assessment: from depth perception to overall experience. IEEE Transactions on Image Processing 27 (2), pp.721-734. (10.1109/TIP.2017.2766780)
2016
- Zhou, W. et al. 2016. 3D-HEVC visual quality assessment: database and bitstream model. Presented at: 8th International Conference on Quality of Multimedia Experience (QoMEX) Lisbon 6-8 June 2016. 2016 Eighth International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.1-6. (10.1109/QoMEX.2016.7498946)
Articles
- Yang, Y. et al., 2027. BiSQAFusion: Multi-level fusion for personalized binaural speech quality assessment in hearing aids. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 139 (Part A) 104765. (10.1016/j.inffus.2026.104765)
- Liu, J. et al., 2027. SCMM: Calibrating cross-modal representations for text-based person search. Pattern Recognition 182 114775. (10.1016/j.patcog.2026.114775)
- Yan, W. et al., 2027. 3D object detection and knowledge distillation in autonomous driving: A survey. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 138 104706. (10.1016/j.inffus.2026.104706)
- Shen, X. et al., 2026. BeatDance: Generating beat-consistent 3D dance with hierarchical spatial–temporal modeling. Pattern Recognition 180 114344. (10.1016/j.patcog.2026.114344)
- Wang, Z. et al., 2026. Robust low-light image enhancement in the wild via data synthesis and generative diffusion prior. Pattern Recognition 178 113336. (10.1016/j.patcog.2026.113336)
- Wang, H. et al., 2026. KSIQA: A knowledge-sharing model for no-reference image quality assessment. IEEE Transactions on Neural Networks and Learning Systems 37 (8), pp.3943-3955. (10.1109/tnnls.2026.3656757)
- Yang, D. et al., 2026. Toward empathetic care: an LLM-based multi-intention recognition framework for mental health and complex medical queries. IEEE Transactions on Affective Computing (10.1109/taffc.2026.3728254)
- Yu, L. et al., 2026. Blind image quality assessment via a hierarchical perceptual modulation network. ACM Transactions on Multimedia Computing, Communications and Applications (10.1145/3840393)
- Zhao, C. et al., 2026. Uncertainty-guided spatiotemporal consistency fusion network for infrared-visible video fusion under extremely low-light conditions. IEEE Transactions on Image Processing 35 , pp.8894-8909. (10.1109/tip.2026.3719477)
- Zou, M. et al., 2026. Pose-guided multi-cue explicit query construction for disambiguating human-object interactions. IEEE Transactions on Circuits and Systems for Video Technology 36 (7), pp.10794-10809. (10.1109/TCSVT.2026.3667102)
- Yu, J. et al., 2026. One aligned LLM to serve them all: A transfer recipe for training VLMs without visual-language re-alignment. International Journal of Computer Vision 134 (7) 331. (10.1007/s11263-026-02930-z)
- Wu, J. et al., 2026. Digital human generation for games via tightness-aware multi-cue modeling. IEEE Transactions on Games (10.1109/tg.2026.3708023)
- Zhang, L. et al., 2026. Rethinking the effect of unimodal labels in multimodal sentiment analysis. ACM Transactions on Multimedia Computing, Communications, and Applications 22 (7) 204. (10.1145/3796718)
- Zhao, M. et al., 2026. A new semi-supervised video anomaly detection baseline in lack of anomalous samples. ACM Transactions on Multimedia Computing, Communications and Applications 22 (7) 189. (10.1145/3797034)
- Hao, X. et al., 2026. DADA++: Dual Alignment Domain Adaptation for unsupervised video-text retrieval. ACM Transactions on Multimedia Computing, Communications and Applications 22 (6) 175. (10.1145/3759252)
- Chen, B. et al., 2026. From global to granular: revealing IQA model performance via correlation surface. IEEE Transactions on Pattern Analysis and Machine Intelligence (10.1109/tpami.2026.3705184)
- Yin, Y. et al., 2026. Deep learning-based point cloud upsampling: A survey of methodologies, performance comparisons, and noise robustness analysis. Neurocomputing 681 133316. (10.1016/j.neucom.2026.133316)
- Hao, X. et al., 2026. Embodied spatial affordance: spatial-aware affordance learning for embodied navigation and manipulation. IEEE Transactions on Image Processing 35 , pp.6041-6054. (10.1109/tip.2026.3698366)
- Yu, L. et al., 2026. DVLTA-VQA: Decoupled vision-language modeling with text-guided adaptation for blind video quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 36 (5), pp.6826-6837. (10.1109/tcsvt.2026.3657415)
- Zhou, W. et al. 2026. Quality of multimedia experience meets machine intelligence. ACM SIGMultimedia Records 18 (1) 3. (10.1145/3811013.3811016)
- You, L. et al., 2026. Integrating SAM supervision for 3D weakly supervised point cloud segmentation. IEEE Transactions on Image Processing 35 , pp.5212-5223. (10.1109/tip.2026.3691686)
- Kuang, Y. et al., 2026. A nature-inspired edge-cloud collaborative privacy framework for visual relocalization on consumer electronics. IEEE Transactions on Consumer Electronics (10.1109/TCE.2026.3692722)
- Yang, Y. et al., 2026. Expressive human volumetric video generation with rich text. IEEE Transactions on Circuits and Systems for Video Technology 36 (4), pp.5424-5436. (10.1109/tcsvt.2025.3628996)
- Chang, Y. et al., 2026. Perception-inspired network for stereo image quality assessment. IEEE Transactions on Image Processing (10.1109/tip.2026.3680564)
- Ni, Y. et al., 2026. Affine modulation-based audiogram fusion network for joint noise reduction and hearing loss compensation. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 127 (A) 103726. (10.1016/j.inffus.2025.103726)
- Yang, D. et al., 2026. MedAide: information fusion and anatomy of medical intents via LLM-based agent collaboration. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 127 (Part A) 103743. (10.1016/j.inffus.2025.103743)
- Li, J. et al., 2026. SearchExpert: A GenAI-driven framework for reasoning-intensive multimedia information fusion through fine-tuning and reinforcement learning. Information Fusion: An International Journal on Multi-Sensor, Multi-Source Information Fusion 126 (Part B) 103665. (10.1016/j.inffus.2025.103665)
- Yuan, H. et al., 2026. DWCL: Dual-weighted contrastive learning for robust multi-view clustering. Engineering Applications of Artificial Intelligence 165 (PartB) 113532. (10.1016/j.engappai.2025.113532)
- Liu, W. et al., 2026. Exploiting independent query information for few-shot image segmentation. Displays 91 103179. (10.1016/j.displa.2025.103179)
- Yu, Q. et al., 2026. StealthMark: Harmless and stealthy ownership verification for medical segmentation via uncertainty-guided backdoors. IEEE Transactions on Image Processing 35 , pp.1290-1304. (10.1109/tip.2026.3655563)
- Liu, J. et al., 2026. Self-supervised unfolding network with shared reflectance learning for low-light image enhancement. IEEE Transactions on Image Processing 35 , pp.800-815. (10.1109/tip.2026.3652021)
- Li, H. et al., 2026. EHIN: Early-aware hierarchical interaction network for weakly-supervised referring image segmentation. Neurocomputing 659 131764. (10.1016/j.neucom.2025.131764)
- Xue, J. et al., 2026. Towards comprehensive interactive change understanding in remote sensing: A large-scale dataset and dual-granularity enhanced VLM. IEEE Transactions on Geoscience and Remote Sensing 64 4401516. (10.1109/tgrs.2025.3650151)
- Ju, Y. et al., 2025. Photometric regularization for 3D gaussian splatting in multi-view surface projection. IEEE Journal of Selected Topics in Signal Processing 19 (8), pp.1682-1693. (10.1109/jstsp.2025.3617861)
- Li, Y. et al., 2025. Unlocking implicit motion for evaluating image complexity. Displays 90 103131. (10.1016/j.displa.2025.103131)
- Li, Y. et al., 2025. Perception-oriented bidirectional attention network for image super-resolution quality assessment. IEEE Transactions on Image Processing 34 , pp.7728-7743. (10.1109/tip.2025.3633145)
- Wang, R. et al., 2025. Multi-view residual spatio-temporal topology adaptive graph convolutional network for urban road traffic accident prediction with multi-source risks. Array 28 100617. (10.1016/j.array.2025.100617)
- Zhang, Z. et al., 2025. Large multimodal models evaluation: a survey. SCIENCE CHINA Information Sciences 68 221301. (10.1007/s11432-025-4676-4)
- Liu, W. et al., 2025. Integrating large foundation models into multimodal named entity recognition with evidential fusion. Neurocomputing 652 131015. (10.1016/j.neucom.2025.131015)
- Yao, X. et al., 2025. AMLPF-CLIP: Adaptive prompting and distilled learning for imbalanced histopathological image classification. IEEE Journal of Biomedical and Health Informatics (10.1109/jbhi.2025.3619343)
- Zeng, Y. et al., 2025. CLIP-DQA V2: Exploring CLIP for dehazed image quality assessment from a fragment-level perspective. IEEE Signal Processing Letters 32 , pp.3829-3833. (10.1109/lsp.2025.3615082)
- Amirpour, H. et al., 2025. VQM4HAS: A real-time quality metric for HEVC videos in HTTP Adaptive Streaming. IEEE Transactions on Multimedia 27 , pp.9619-9631. (10.1109/tmm.2025.3613110)
- Liu, J. et al. 2025. Adaptive spatiotemporal graph transformer network for action quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 35 (7), pp.6628-6639. (10.1109/TCSVT.2025.3541456)
- Chen, W. et al., 2025. No-reference point cloud quality assessment via graph convolutional network. IEEE Transactions on Multimedia 27 , pp.2489-2502. (10.1109/TMM.2024.3521845)
- Luo, Y. et al., 2025. Unsupervised low-light image enhancement with self-paced learning. IEEE Transactions on Multimedia 27 , pp.1808-1820. (10.1109/TMM.2024.3521752)
- Guanghui, Y. et al., 2025. Subjective and objective quality assessment of colonoscopy videos. IEEE Transactions on Medical Imaging 44 (2), pp.841-854. (10.1109/TMI.2024.3461737)
- Liu, W. et al., 2025. Physically-guided open vocabulary segmentation with weighted patched alignment loss. Neurocomputing 614 128788. (10.1016/j.neucom.2024.128788)
- Zhou, W. and Wang, Z. 2024. Perceptual depth quality assessment of stereoscopic omnidirectional images. IEEE Transactions on Circuits and Systems for Video Technology 34 (12), pp.13452-13462. (10.1109/TCSVT.2024.3449696)
- Zhang, R. et al., 2024. Similarities between wheels and tracks: A "tire model" for tracked vehicles. IEEE Transactions on Vehicular Technology 73 (11), pp.16416-16431. (10.1109/TVT.2024.3421285)
- Zhang, X. et al., 2024. Neural image re-exposure. Computer Vision and Image Understanding 248 104094. (10.1016/j.cviu.2024.104094)
- Chen, W. et al., 2024. Dynamic hypergraph convolutional network for no-reference point cloud quality assessment. IEEE Transactions on Circuits and Systems for Video Technology 34 (10), pp.10479-10493. (10.1109/TCSVT.2024.3410052)
- Wang, H. et al. 2024. Blind image quality assessment via adaptive graph attention. IEEE Transactions on Circuits and Systems for Video Technology 34 (10), pp.10299-10309. (10.1109/TCSVT.2024.3405789)
- Zhou, T. et al., 2024. Adaptive mixed-scale feature fusion network for blind AI-generated image quality assessment. IEEE Transactions on Broadcasting 70 (3), pp.833-843. (10.1109/TBC.2024.3391060)
- Zhou, W. et al. 2024. Blind quality assessment of dense 3D point clouds with structure guided resampling. ACM Transactions on Multimedia Computing, Communications and Applications 20 (8)(10.1145/3664199)
- Yan, W. et al., 2024. FVIFormer: flow-guided global-local aggregation transformer network for video inpainting. IEEE Journal of Emerging and Selected Topics in Circuits and Systems 14 (2), pp.235-244. (10.1109/JETCAS.2024.3392972)
- Yue, G. et al., 2024. Subjective and objective quality assessment of multi-attribute retouched face images. IEEE Transactions on Broadcasting 70 (2), pp.570-583. (10.1109/TBC.2024.3374043)
- Fu, J. et al., 2024. Vision-language consistency guided multi-modal prompt learning for blind AI generated image quality assessment. IEEE Signal Processing Letters 31 , pp.1820-1824. (10.1109/LSP.2024.3420083)
- Yue, G. et al., 2024. Dual-constraint coarse-to-fine network for camouflaged object detection. IEEE Transactions on Circuits and Systems for Video Technology 34 (5), pp.3286-3298. (10.1109/TCSVT.2023.3318672)
- Zhang, R. et al., 2024. A terramechanics-based dynamic model for motion control of unmanned tracked vehicles. IEEE Transactions on Intelligent Vehicles (10.1109/TIV.2024.3406582)
- Fu, J. , Zhou, W. and Chen, Z. 2024. Bayesian graph convolutional network for traffic prediction. Neurocomputing 582 127507. (10.1016/j.neucom.2024.127507)
- Yue, G. et al., 2024. Boundary refinement network for colorectal polyp segmentation in colonoscopy images. IEEE Signal Processing Letters 31 , pp.954-958. (10.1109/LSP.2024.3378106)
- Zhou, W. et al. 2024. Dehazed image quality evaluation: from partial discrepancy to blind perception. IEEE Transactions on Intelligent Vehicles 9 (2), pp.3843-3858. (10.1109/TIV.2024.3356055)
- Yi, X. , Jiang, Q. and Zhou, W. 2024. No-reference quality assessment of underwater image enhancement. Displays 81 102586. (10.1016/j.displa.2023.102586)
- Zhou, W. et al. 2023. A brief survey on adaptive video streaming quality assessment. Journal of Visual Communication and Image Representation 86 103526. (10.1016/j.jvcir.2022.103526)
- Sun, S. et al., 2023. GraphIQA: Learning Distortion Graph Representations for Blind Image Quality Assessment. IEEE Transactions on Multimedia 25 , pp.2912-2925. (10.1109/TMM.2022.3152942)
- Liu, J. et al., 2023. LIQA: lifelong blind image quality assessment. IEEE Transactions on Multimedia 25 , pp.5358-5373. (10.1109/TMM.2022.3190700)
- Zhou, W. et al. 2023. Reduced-reference quality assessment of point clouds via content-oriented saliency projection. IEEE Signal Processing Letters 30 , pp.354-358. (10.1109/LSP.2023.3264105)
- Xu, J. et al., 2022. Quality assessment of multi-exposure image fusion by synthesizing local and global intermediate references. Displays 74 102188. (10.1016/j.displa.2022.102188)
- Zhou, W. et al. 2022. No-reference quality assessment for 360-degree images by analysis of multifrequency information and local-global naturalness. IEEE Transactions on Circuits and Systems for Video Technology 32 (4), pp.1778-1791. (10.1109/TCSVT.2021.3081182)
- Fang, Y. et al., 2022. A Bayesian deep image prior downscaling approach for high-resolution soil moisture estimation. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing 15 , pp.4571 - 4582. (10.1109/JSTARS.2022.3177081)
- Mao, S. et al., 2021. Mechanism and application of capacitive-coupled memristive behavior based on a biomaterial developed memristive device. ACS Applied Electronic Materials 3 (12), pp.5537-5547. (10.1021/acsaelm.1c00951)
- Ling, S. et al., 2021. Re-visiting discriminator for blind free-viewpoint image quality assessment. IEEE Transactions on Multimedia 23 , pp.4245-4258. (10.1109/TMM.2020.3038305)
- Xu, J. , Zhou, W. and Chen, Z. 2021. Blind omnidirectional image quality assessment with viewport oriented graph convolutional networks. IEEE Transactions on Circuits and Systems for Video Technology 31 (5), pp.1724-1737. (10.1109/TCSVT.2020.3015186)
- Xu, J. et al., 2020. Binocular rivalry oriented predictive autoencoding network for blind stereoscopic image quality measurement. IEEE Transactions on Instrumentation and Measurement 70 5001413. (10.1109/TIM.2020.3026443)
- Jiang, Q. et al., 2020. A full-reference stereoscopic image quality measurement via hierarchical deep feature degradation fusion. IEEE Transactions on Instrumentation and Measurement 69 (12), pp.9784-9796. (10.1109/TIM.2020.3005111)
- Shi, L. et al., 2020. No-reference light field image quality assessment based on spatial-angular measurement. IEEE Transactions on Circuits and Systems for Video Technology 30 (11), pp.4114-4128. (10.1109/TCSVT.2019.2955011)
- Zhou, W. et al. 2020. Blind quality assessment for image superresolution using deep two-stream convolutional networks. Information Sciences 528 , pp.205-218. (10.1016/j.ins.2020.04.030)
- Zhou, W. et al. 2020. Tensor oriented no-reference light field image quality assessment. IEEE Transactions on Image Processing 29 , pp.4070-4084. (10.1109/TIP.2020.2969777)
- Chen, Z. et al., 2020. Stereoscopic omnidirectional image quality assessment based on predictive coding theory. IEEE Journal of Selected Topics in Signal Processing 14 (1), pp.103-117. (10.1109/JSTSP.2020.2968182)
- Zhou, W. , Chen, Z. and Li, W. 2019. Dual-stream interactive networks for no-reference stereoscopic image quality assessment. IEEE Transactions on Image Processing 28 (8), pp.3946-3958. (10.1109/TIP.2019.2902831)
- Chen, Z. , Zhou, W. and Li, W. 2017. Blind stereoscopic video quality assessment: from depth perception to overall experience. IEEE Transactions on Image Processing 27 (2), pp.721-734. (10.1109/TIP.2017.2766780)
Conferences
- Guo, Z. et al., 2026. GP-GS: Gaussian processes densification for 3D Gaussian Splatting. Presented at: 22nd International Conference on Intelligent Computing Toronto, ON, Canada 22-26 July 2026. Published in: Li, G. , Filipe, J. and Xu, Z. eds. Vol. 3037.[GP-GS: Gaussian Processes Densification for 3D Gaussian Splatting]. Singapore: Springer. , pp.139-150. (10.1007/978-981-92-3548-3_12)
- Amirpour, H. et al., 2026. Quality-complexity trade-offs for sustainable media delivery. Presented at: 2026 18th International Conference on Quality of Multimedia Experience (QoMEX) Cardiff, UK 29 June 2026 - 3 July 2026. 2026 18th International Conference on Quality of Multimedia Experience (QoMEX). IEEE(10.1109/qomex69967.2026.11618319)
- Zhang, H. et al. 2026. Cross-modal interaction for multi-dimensional AI-generated image quality assessment. Presented at: 2026 18th International Conference on Quality of Multimedia Experience (QoMEX) Cardiff, UK 29 June 2026 - 03 July 2026. 2026 18th International Conference on Quality of Multimedia Experience (QoMEX). IEEE(10.1109/qomex69967.2026.11618318)
- Liu, J. et al. 2026. Saliency-guided action quality assessment: an AI-augmented framework for skill evaluation in physical education. Presented at: 2026 IEEE Conference on Artificial Intelligence (CAI) Granada, Spain 08 - 10 May 2026. 2026 IEEE Conference on Artificial Intelligence (CAI) Proceedings. IEEE. , pp.1997-2002. (10.1109/cai68641.2026.11536588)
- Amirpour, H. et al., 2026. BINR: Live video broadcasting quality assessment. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.22282-22286. (10.1109/icassp55912.2026.11460679)
- Wang, H. et al., 2026. Non-line-of-sight vehicle detection via audio-visual fusion. Presented at: ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.11807-11811. (10.1109/icassp55912.2026.11465095)
- Wang, Z. et al., 2026. VMambaMorph: A 3D multi-modality deformable image registration framework based on visual state space model with cross-scan module. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-8 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.8847-8851. (10.1109/icassp55912.2026.11464524)
- Zhou, W. et al. 2026. Perceptual quality optimization of image super-resolution. Presented at: 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) Barcelona, Spain 3-6 May 2026. ICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE; 1999. , pp.10537-10541. (10.1109/icassp55912.2026.11461596)
- Cen, J. et al., 2026. Depth-guided metric-aware temporal consistency for monocular video human mesh recovery. Presented at: International Conference on Acoustics, Speech, and Signal Processing Barcelona, Spain 3-8 May 2026. Proceedings of the 2026 International Conference on Acoustics, Speech, and Signal Processing. IEEE. , pp.10707-10711. (10.1109/icassp55912.2026.11463978)
- Li, Y. et al., 2026. Temporal inconsistency guidance for super-resolution video quality assessment. Presented at: The 40th Annual AAAI Conference on Artificial Intelligence Singapore 20-27 January 2026. Vol. 40 (8)., pp.6681-6689. (10.1609/aaai.v40i8.37599)
- Zhu, J. et al., 2026. Is there a relationship between Mean Opinion Score (MOS) and Just Noticeable Difference (JND)?. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 01-04 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396801)
- Ghanbari, M. et al., 2025. STACK: Spatial tower assembly using controlled kinetics. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396840)
- Hou, J. et al., 2025. Frequency-aware native resolution assessment of 8K omnidirectional images. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396856)
- Liu, Y. et al., 2025. M 2 S 2 L: Mamba-based multi-scale spatial-temporal learning for video anomaly detection. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. IEEE(10.1109/vcip67698.2025.11396919)
- Wang, J. et al., 2025. CVBench: benchmarking and comparing video generation with large multimodal models. Presented at: 2025 International Conference on Visual Communications and Image Processing (VCIP) Klagenfurt, Austria 1-4 December 2025. 2025 International Conference on Visual Communications and Image Processing (VCIP). IEEE(10.1109/vcip67698.2025.11396889)
- Li, X. et al., 2025. MSPoint-Gait: Multi-Scale Point cloud analysis for 3D gait recognition via cross-modal learning. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June - 4 July 2025. 2025 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/icme59968.2025.11209453)
- Ma, Y. et al., 2025. Analysing and predicting radiologists’ expertise using eye-tracking data: Insights for diagnostic decision-making. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June - 4 July 2025. IEEE. , pp.1-6. (10.1109/icme59968.2025.11209585)
- Yang, Y. et al., 2025. MCSMoG: Multi-Conditional Diffusion for stylized motion generation with parametric control. Presented at: 2025 IEEE International Conference on Multimedia and Expo (ICME) Nantes, France 30 June 2025 - 4 July 2025. 2025 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/icme59968.2025.11209286)
- Gao, C. et al., 2025. Compressed feature quality assessment: Dataset and baselines. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13450-13456. (10.1145/3746027.3758309)
- Gao, L. et al., 2025. EEmo-Bench: A benchmark for multi-modal large language models on image evoked emotion assessment. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.7064-7073. (10.1145/3746027.3755777)
- Ghanbari, M. et al., 2025. SDART: spatial dart AR simulation with hand-tracked input. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. New York, NY, United States: ACM. , pp.13543-13545. (10.1145/3746027.3754484)
- Meng, Y. et al., 2025. VideoForest: Person-anchored hierarchical reasoning for cross-video question answering. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.836-845. (10.1145/3746027.3754573)
- Miao, C. et al., 2025. MFFI: Multi-dimensional Face Forgery Image dataset for real-world scenarios. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13235-13242. (10.1145/3746027.3758280)
- Wang, P. et al., 2025. A spatial relationship aware dataset for robotics. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.13332-13338. (10.1145/3746027.3758293)
- Zhou, W. and Amirpour, H. 2025. Perceptual visual quality assessment in multimedia communication. Presented at: MM '25: The 33rd ACM International Conference on Multimedia Dublin, Ireland 27-31 October 2025. MM '25: Proceedings of the 33rd ACM International Conference on Multimedia. ACM. , pp.14340-14341. (10.1145/3746027.3759205)
- Gholami, A. et al., 2025. Perceptual quality assessment of spatial videos on Apple Vision Pro. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.20-28. (10.1145/3746269.3760422)
- Li, X. et al., 2025. DepthGait: Multi-scale cross-level feature fusion of RGB-derived depth and silhouette sequences for robust gait recognition. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.2333-2341. (10.1145/3746027.3755876)
- Wang, Z. et al., 2025. Evaluating perceptual color preferences in smartphone photography: dataset and challenges. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.12844-12850. (10.1145/3746027.3758227)
- Zou, M. et al., 2025. PhysLab: A benchmark dataset for multi-granularity visual parsing of physics experiments. Presented at: MM '25:The 33rd ACM International Conference on Multimedia Dublin, Ireland 31 October 2025. IXR '25: Proceedings of the 3rd International Workshop on Interactive eXtended Reality. Dublin: ACM. , pp.12799-12806. (10.1145/3746027.3758221)
- Huang, F. et al., 2025. VQualA 2025 document image quality assessment challenge. Presented at: 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) Honolulu, HI, USA 19-20 October 2025. 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW). IEEE. , pp.3344-3353. (10.1109/ICCVW69036.2025.00351)
- Yu, Q. et al., 2025. Parameterized diffusion optimization enabled autoregressive ordinal regression for diabetic retinopathy grading. Presented at: MICCAI 2025 Daejeon, Republic of Korea 23-27 September 2025. Published in: Gee, J. C. et al., Proceedings Medical Image Computing and Computer Assisted Intervention. Lecture Notes in Computer Science. Vol. 15974.Switzerland: Springer Nature. , pp.450-460. (10.1007/978-3-032-05182-0_44)
- Luo, Y. et al., 2025. Multi-attribute continual learning for blind image quality assessment. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) London 25-28 May 2025. 2025 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE(10.1109/iscas56072.2025.11043383)
- Zeng, Y. et al., 2025. CLIP-DQA: Blindly evaluating dehazed images from global and local perspectives using CLIP. Presented at: IEEE International Symposium on Circuits and Systems London, UK 25-28 May 2025. 2025 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE. (10.1109/iscas56072.2025.11043366)
- Li, Y. et al., 2024. Deep bi-directional attention network for image super-resolution quality assessment. Presented at: IEEE International Conference on Multimedia and Expo (ICME) Niagra Falls, Canada 15-19 July 2024. 2024 IEEE International Conference on Multimedia and Expo (ICME). IEEE. , pp.1-6. (10.1109/ICME57554.2024.10687430)
- Xu, J. et al., 2024. LISD: An efficient multi-task learning framework for LiDAR segmentation and detection. Presented at: IEEE International Conference on Image Processing (ICIP) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.3341-3347. (10.1109/ICIP51287.2024.10647535)
- Shafiee Sarvestani, A. , Zhou, W. and Wang, Z. 2024. Perceptual crack detection for rendered 3D textured meshes. Presented at: IEEE International Conference on Quality of Multimedia Experience (QoMEX) Karlshamn, Sweden 18 - 20 June 2024. Proceedings of 16th International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.1-7. (10.1109/QoMEX61742.2024.10598253)
- Yue, G. et al., 2024. Subjective quality assessment of thermal infrared images. Presented at: IEEE International Conference on Image Processing (ICIP 2024) Abu Dhabi, United Arab Emirates 27-30 October 2024. Proceedings of International Conference on Image Processing. IEEE. , pp.1212-1217. (10.1109/ICIP51287.2024.10648145)
- Yue, G. et al., 2023. Subjective quality assessment of enhanced retinal images. Presented at: International Conference on Image Processing Kuala Lumpur, Malaysia 8–11 October 2023. 2023 IEEE International Conference on Image Processing Proceedings. IEEE. , pp.3005-3009. (10.1109/ICIP49359.2023.10222541)
- Zhou, W. and Wang, Z. 2023. Blind omnidirectional image quality assessment: integrating local statistics and global semantics. Presented at: International Conference on Image Processing Kuala Lumpur, Malaysia 8–11 October 2023. 2023 IEEE International Conference on Image Processing Proceedings. IEEE. , pp.1405-1409. (10.1109/ICIP49359.2023.10222049)
- Fu, J. et al., 2022. Adaptive hypergraph convolutional network for no-reference 360-degree image quality assessment. Presented at: MM '22: The 30th ACM International Conference on Multimedia Lisboa Portugal 10-14 October 2022. MM '22: Proceedings of the 30th ACM International Conference on Multimedia. Association for Computing Machinery. , pp.961-969. (10.1145/3503161.3548337)
- Zhou, W. and Wang, Z. 2022. Quality assessment of image super-resolution: balancing deterministic and statistical fidelity. Presented at: MM '22: The 30th ACM International Conference on Multimedia Lisboa Portugal 10-14 October 2022. MM '22: Proceedings of the 30th ACM International Conference on Multimedia. Association for Computing Machiner. , pp.934-942. (10.1145/3503161.3547899)
- Lu, Y. et al., 2022. RTN: Reinforced transformer network for coronary CT angiography vessel-level image quality assessment. Presented at: MICCAI 2022 18-22 September 2022. Medical Image Computing and Computer Assisted Intervention – MICCAI 2022.. Vol. 13431.Lecture Notes in Computer Science Cham. Switzerland: Springer. (10.1007/978-3-031-16431-6_61)
- Huang, S. et al., 2022. Perceptual evaluation of pre-processing for video transcoding. Presented at: 2021 International Conference on Visual Communications and Image Processing (VCIP) Munich, Germany 5-8 December 2021. 2021 International Conference on Visual Communications and Image Processing (VCIP). IEEE. , pp.1-5. (10.1109/VCIP53242.2021.9675438)
- Xu, J. et al., 2021. Perceptual quality assessment of internet videos. Presented at: MM '21: 29th ACM International Conference on Multimedia Virtual Event China 20 - 24 October 2021. MM '21: Proceedings of the 29th ACM International Conference on Multimedia. Association for Computing Machinery. , pp.1248-1257. (10.1145/3474085.3475486)
- Peng, Y. et al., 2021. Multi-metric fusion network for image quality assessment. Presented at: Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) Nashville, TN, USA 19-25 June 2021. 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). IEEE. , pp.1857-1860. (10.1109/CVPRW53098.2021.00205)
- Zhou, W. , Wang, Z. and Chen, Z. 2021. Image super-resolution quality assessment: structural fidelity versus statistical naturalness. Presented at: 13th International Conference on Quality of Multimedia Experience (QoMEX) Montreal, QC, Canada 14-17 June 2021. 2021 13th International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.61-64. (10.1109/QoMEX51781.2021.9465479)
- Zhou, W. and Chen, Z. 2021. Deep multi-scale features learning for distorted image quality assessment. Presented at: International Symposium on Circuits and Systems (ISCAS) Daegu, Korea 22-28 May 2021. 2021 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE(10.1109/ISCAS51556.2021.9401285)
- Zhou, W. and Chen, Z. 2020. Deep local and global spatiotemporal feature aggregation for blind video quality assessment. Presented at: 2020 IEEE International Conference on Visual Communications and Image Processing (VCIP) Macau 01-04 December 2020. 2020 IEEE International Conference on Visual Communications and Image Processing (VCIP). IEEE. (10.1109/VCIP49819.2020.9301764)
- Liu, J. et al., 2020. LIRA: Lifelong Image Restoration from Unknown Blended Distortions. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vedaldi, A. et al., Computer Vision – ECCV 2020. Lecture Notes in Computer Science Vol. 12363. Springer International Publishing. , pp.616-632. (10.1007/978-3-030-58523-5_36)
- Li, X. et al., 2020. Learning disentangled feature representation for hybrid-distorted image restoration. Presented at: 16th European Conference on Computer Vision (ECCV 2020) Glasgow, Scotland 23-28 August 2020. Published in: Vedaldi, A. et al., Computer Vision – ECCV 2020 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXIX. Lecture Notes in Computer Science Vol. 12374. Springer. , pp.313-329. (10.1007/978-3-030-58526-6_19)
- Zhao, Y. et al., 2020. Infrared pedestrian detection with converted temperature map. Presented at: 2019 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) Lanzhou 18-21 November 2019. 2019 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC). IEEE. (10.1109/APSIPAASC47483.2019.9023228)
- Luo, Z. et al., 2020. No-reference light field image quality assessment based on micro-lens image. Presented at: Picture Coding Symposium, PCS Ningbo, China 12-15 November 2019. 2019 Picture Coding Symposium (PCS). IEEE. , pp.1-5. (10.1109/PCS48520.2019.8954551)
- Xu, J. et al., 2020. Quality assessment of stereoscopic 360-degree images from multi-viewports. Presented at: Picture Coding Symposium, PCS Ningbo, China 12-15 November 2019. 2019 Picture Coding Symposium (PCS). IEEE. , pp.1-5. (10.1109/PCS48520.2019.8954555)
- Jin, X. et al., 2019. Unsupervised single image deraining with self-supervised constraints. Presented at: 2019 IEEE International Conference on Image Processing Taipei, Taiwan 22-25 September 2019. 2019 IEEE International Conference on Image Processing. IEEE. , pp.2761-2765. (10.1109/ICIP.2019.8803238)
- Jin, X. et al., 2019. AI-GAN: signal de-interference via asynchronous interactive generative adversarial network. Presented at: 2019 IEEE International Conference on Multimedia & Expo Workshops (ICMEW) Shanghai 08-12 July 2019. 2019 IEEE International Conference on Multimedia & Expo Workshops (ICMEW). IEEE. , pp.228-233. (10.1109/ICMEW.2019.00046)
- Zhao, S. et al., 2019. How do you perceive differently from an AI - A database for semantic distortion measurement. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) Sapporo, Japan 26-29 May 2019. 2019 IEEE International Symposium on Circuits and Systems (ISCAS). IEEE. , pp.1-5. (10.1109/ISCAS.2019.8702774)
- Jin, X. et al., 2018. Augmented coarse-to-fine video frame synthesis with semantic loss. Presented at: Pattern Recognition and Computer Vision, PRCV 2018 Guangzhou, China 23-26 November. Published in: Lai, J. et al., Pattern Recognition and Computer Vision, First Chinese Conference, PRCV 2018, Guangzhou, China, November 23-26, 2018, Proceedings, Part I. Lecture Notes in Computer Science Vol. 11256. Springer. , pp.439-452. (10.1007/978-3-030-03398-9_38)
- Xu, J. et al., 2018. Subjective quality assessment of stereoscopic omnidirectional image. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Lecture Notes in Computer Science Vol. 11164. Springer. , pp.589-599. (10.1007/978-3-030-00776-8_54)
- Zhang, W. et al., 2018. None ghosting artifacts stitching based on depth map for light field image. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Vol. 11164.Lecture Notes in Computer Science Vol. 11164. Springer. , pp.567-578. (10.1007/978-3-030-00776-8_52)
- Zhou, W. , Chen, Z. and Li, W. 2018. Stereoscopic video quality prediction based on end-to-end dual stream deep neural networks. Presented at: 19th Pacific-Rim Conference on Multimedia Hefei, China 21-22 Sept 2018. Published in: Hong, R. et al., Advances in Multimedia Information Processing – PCM 2018. Vol. 11166.Lecture Notes in Computer Science Vol. 11166. Springer. , pp.482-492. (10.1007/978-3-030-00764-5_44)
- Shi, L. et al., 2018. Perceptual evaluation of light field image. Presented at: 2018 25th IEEE International Conference on Image Processing (ICIP) Athens, Greece 07-10 October 2018. 2018 25th IEEE International Conference on Image Processing (ICIP). IEEE. , pp.41-45. (10.1109/ICIP.2018.8451077)
- Jin, X. et al., 2018. A decomposed dual-cross generative adversarial network for image rain removal. Presented at: British Machine Vision Conference 2018 Newcastle, UK 3-6 September 2018.
- Hu, Y. et al., 2018. SDM: Semantic distortion measurement for video encryption. Presented at: 13th IEEE International Conference on Automatic Face & Gesture Recognition Xi'an, China 15-19 May 2018. 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2018). IEEE. , pp.764-768. (10.1109/FG.2018.00120)
- Zhou, Y. and Zhou, W. 2018. Visual comfort assessment for stereoscopic image retargeting. Presented at: IEEE International Symposium on Circuits and Systems (ISCAS) Florence, Italy 27-30 May 2018. 2018 IEEE International Symposium on Circuits and Systems (ISCAS). , pp.1-5. (10.1109/ISCAS.2018.8351198)
- Zhou, W. et al. 2016. 3D-HEVC visual quality assessment: database and bitstream model. Presented at: 8th International Conference on Quality of Multimedia Experience (QoMEX) Lisbon 6-8 June 2016. 2016 Eighth International Conference on Quality of Multimedia Experience (QoMEX). IEEE. , pp.1-6. (10.1109/QoMEX.2016.7498946)
Ymchwil
Diddordebau ymchwil
Gweledigaeth Gyfrifiadurol a Ffotograffiaeth
Prosesu Signal Amlgyfrwng mewn Amgylcheddau Gweledol
Canfyddiad ac Optimeiddio Gweledol sy'n Canolbwyntio ar Bobl
Dehongli a Dadansoddi Delweddu Optegol
Arddangosfeydd a Chymwysiadau Deallus
Prosiectau a ariennir
Prosesu delweddau biofeddygol ar gyfer cymwysiadau gofal iechyd craff, Grant Ymchwil Taith, PI (2024-2025)
Archwilio canfod halltedd mewn gwerthuso ansawdd delwedd stereo VR, DUT Collaboration Funding, PI (2023-2024)
Asesiad ansawdd delwedd canfyddiadol aml-ffynhonnell ac ansawdd fideo, Cyllid Traethawd PhD Eithriadol, PI (2020-2022)
Cyhoeddiadau dethol (* Cyfraniad cyfartal, ^ Awdur cyfatebol, mwy o gyhoeddiadau ar Google Scholar):
- Asesiad ansawdd goddrychol a gwrthrychol o fideos colonosgopi
G Yue, L Zhang, J Du, T Zhou, Wei Zhou, W Lin
Trafodion IEEE ar Ddelweddu Meddygol (TMI), 2024 - Asesiad ansawdd cwmwl pwynt cyfeirio trwy rwydwaith cyfnewidiol graff
W Chen, Q Jiang, Wei Zhou, F Shao, G Zhai, W Lin
Trafodion IEEE ar Amlgyfrwng (TMM), 2024 - Gwella delwedd golau isel heb oruchwyliaeth gyda dysgu hunan-paced
Y Luo, X Chen, J Ling, Wei Zhou, G Yue
Trafodion IEEE ar Amlgyfrwng (TMM), 2024 - Asesiad ansawdd dyfnder canfyddiadol o ddelweddau omnidirectional stereosgopig
Wei Zhou, Z Wang
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2024 - Rhwydwaith cyfnewidiol hypergraff deinamig ar gyfer asesiad ansawdd cwmwl pwynt cyfeirio
W Chen, Q Jiang, Wei Zhou, L Xu, W Lin
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2024 - Asesiad ansawdd delwedd ddall trwy sylw graff addasol
H Wang, J Liu, H Tan, J Lou, X Liu, Wei Zhou, H Liu
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2024 - Asesiad ansawdd dall o gymylau pwynt 3D trwchus gyda strwythur ailsamplu dan arweiniad strwythur
Wei Zhou, Q Yang, W Chen, Q Jiang, G Zhai, W Lin
ACM Trafodion ar Gyfrifiadura Amlgyfrwng, Cyfathrebu a Chymwysiadau (TOMM), 2024 - Roedd cysondeb iaith gweledigaeth yn arwain dysgu prydlon aml-foddol ar gyfer asesiad ansawdd delwedd a gynhyrchir gan AI dall
J Fu, Wei Zhou^, Q Jiang, H Liu, G Zhai
IEEE Llythyrau Prosesu Signal (SPL), 2024 - Model deinamig sy'n seiliedig ar terramechanics ar gyfer rheoli symudiadau cerbydau trac di-griw
R Zhang, Wei Zhou, H Liu, J Gong, H Chen, Amir Khajepour
Trafodion IEEE ar Gerbydau Deallus (TIV), 2024 - Rhwydwaith ymasiad nodwedd ar raddfa gymysg addasol ar gyfer asesiad ansawdd delwedd dall a gynhyrchir gan AI
T Zhou, S Tan, Wei Zhou, Y Luo, Y Wang, G Yue
Trafodion IEEE ar Ddarlledu (TBC), 2024 - FVIFormer: rhwydwaith trawsnewidyddion agregu byd-eang dan arweiniad llif ar gyfer mewnbeintio fideo
W Yan, Y Sun, G Yue, Wei Zhou, H Liu
IEEE Journal on Pynciau sy'n Dod i'r Amlwg a Dethol mewn Cylchedau a Systemau (JETCAS), 2024 - Rhwydwaith mireinio ffiniau ar gyfer segmentu polyp colorectal mewn delweddau colonosgopi
G Yue, Y Li, W Jiang, Wei Zhou, T Zhou
IEEE Llythyrau Prosesu Signal (SPL), 2024 - Gwerthusiad ansawdd delwedd dehazed: o anghysondeb rhannol i ganfyddiad dall
Wei Zhou, R Zhang, L Li, G Yue, J Gong, H Chen, H Liu
Trafodion IEEE ar Gerbydau Deallus (TIV), 2024 - Asesiad ansawdd goddrychol a gwrthrychol o ddelweddau wyneb aml-briodoledd wedi'u hailgyffwrdd
G Yue, H Wu, W Yan, T Zhou, H Liu, Wei Zhou
Trafodion IEEE ar Ddarlledu (TBC), 2024 - Asesiad ansawdd llai o gymylau pwynt trwy amcanestyniad halltedd sy'n canolbwyntio ar gynnwys
Wei Zhou, G Yue, R Zhang, Y Qin, H Liu
Llythyrau Prosesu Signal IEEE (SPL), 2023 - Rhwydwaith bras-i-ddirwy ddeuol-gyfyngedig ar gyfer canfod gwrthrychau cuddliw
G Yue, H Xiao, H Xie, T Zhou, Wei Zhou, W Yan, B Zhao, T Wang, Q Jiang
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2023 - Asesiad ansawdd delwedd omnidirectional dall: integreiddio ystadegau lleol a semanteg fyd-eang
Wei Zhou, Z Wang
Cynhadledd Ryngwladol IEEE ar Brosesu Delweddau (ICIP), 2023 - Asesu ansawdd uwch-benderfyniad delwedd: cydbwyso ffyddlondeb penderfynol ac ystadegol
Wei Zhou, Z Wang
Cynhadledd Ryngwladol ACM ar Amlgyfrwng (ACM MM), 2022 - LIQA: Asesiad ansawdd delwedd ddall gydol oes
J Liu *, Wei Zhou*, X Li, J Xu, Z Chen
Trafodion IEEE ar Amlgyfrwng (TMM), 2022 - Rhwydwaith cyfnewidiol hypergraff addasol ar gyfer asesiad ansawdd delwedd 360 gradd dim-cyfeiriad
J Fu, C Hou, Wei Zhou^, J Xu, Z Chen
Cynhadledd Ryngwladol ACM ar Amlgyfrwng (ACM MM), 2022 - GraphIQA: Cynrychioliadau graff ystumiad dysgu ar gyfer asesiad ansawdd delwedd ddall
S Sun, T Yu, J Xu, Wei Zhou, Z Chen
Trafodion IEEE ar Amlgyfrwng (TMM), 2022 - Delwedd ddwfn Bayesaidd cyn israddio dull ar gyfer amcangyfrif lleithder pridd cydraniad uchelY Fan, L Xu, Y Chen, Wei Zhou, Alexander Wong, David A ClausiIEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing (JSTARS), 2022
- Arolwg byr ar asesu ansawdd ffrydio fideo addasol
Wei Zhou, X Min, H Li, Q Jiang
Journal of Visual Communication and Image Representation (JVCI), 2022 - Asesiad ansawdd dim cyfeiriad ar gyfer delweddau 360 gradd trwy ddadansoddi gwybodaeth amlamledd a naturioldeb lleol-fyd-eang
Wei Zhou, J Xu, Q Jiang, Z Chen
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2021 - Asesiad ansawdd delwedd omnidirectional ddall gyda rhwydweithiau cydgyfeiriant graff sy'n canolbwyntio ar y porthladd golwg
J Xu *, Wei Zhou*, Z Chen
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2021 - Asesiad ansawdd canfyddiadol o fideos rhyngrwydJ Xu, J Li, X Zhou, Wei Zhou, B Wang, Z ChenCynhadledd Ryngwladol ACM ar Amlgyfrwng (ACM MM), 2021
- Asesu ansawdd cydraniad uwch: ffyddlondeb strwythurol yn erbyn naturioldeb ystadegolWei Zhou, Z Wang, Z ChenCynhadledd Ryngwladol IEEE ar Ansawdd Profiad Amlgyfrwng (QoMEX), 2021 (Papur gwahoddedig, Oral)
- Mae nodweddion aml-raddfa dwfn yn dysgu ar gyfer asesu ansawdd delwedd ystumiedigWei Zhou, Z ChenSymposiwm Rhyngwladol IEEE ar Gylchedau a Systemau (ISCAS), 2021 (Oral)
- Asesiad ansawdd delwedd maes golau heb gyfeirio tenor oriented
Wei Zhou, L Shi, Z Chen, J Zhang
Trafodion IEEE ar Brosesu Delweddau (TIP), 2020 - Ail-ymweld â gwahaniaethydd ar gyfer asesiad ansawdd delwedd di-olwg ddall
S Ling, J Li, Z Che, Wei Zhou, J Wang, Patrick Le Callet
Trafodion IEEE ar Amlgyfrwng (TMM), 2020 - Rhwydwaith awto-amgodio rhagfynegol rhagfynegol sy'n canolbwyntio ar gystadleuaeth binocwlaidd ar gyfer mesur ansawdd delwedd stereosgopig dall
J Xu *, Wei Zhou*, Z Chen, S Ling, Patrick Le Callet
Trafodion IEEE ar Offeryniaeth a Mesur (TIM), 2020 - Mesur ansawdd delwedd stereosgopig cyfeiriad llawn trwy ymasiad diraddiad nodwedd ddwfn hierarchaidd
Q Jiang *, Wei Zhou *, X Chai, G Yue, F Shao, Z Chen
Trafodion IEEE ar Offeryniaeth a Mesur (TIM), 2020 - Asesiad ansawdd delwedd omnidirectional stereosgopig yn seiliedig ar theori codio rhagfynegol
Z Chen, J Xu, C Lin, Wei Zhou
IEEE Journal of Selected Topics in Signal Processing (JSTSP), 2020 - Asesiad ansawdd dall ar gyfer superresolution delwedd gan ddefnyddio rhwydweithiau cyfnewidiol dwy ffrwd dwfn
Wei Zhou, Q Jiang, Y Wang, Z Chen, W Li
Gwyddorau Gwybodaeth (INS), 2020 - Cydgrynhoi nodwedd gofodol lleol a byd-eang dwfn ar gyfer asesiad ansawdd fideo dall
Wei Zhou, Z Chen
Cynhadledd Ryngwladol IEEE ar Gyfathrebu Gweledol a Phrosesu Delweddau (VCIP), 2020 - Rhwydweithiau rhyngweithiol deuol ffrwd ar gyfer asesiad ansawdd delwedd stereosgopig heb gyfeiriad
Wei Zhou, Z Chen, W Li
Trafodion IEEE ar Brosesu Delweddau (TIP), 2019 - Asesiad ansawdd delwedd maes golau heb gyfeiriad yn seiliedig ar fesuriad gofodol-onglog
L Shi *, Wei Zhou*, Z Chen, J Zhang
Trafodion IEEE ar Gylchedau a Systemau ar gyfer Technoleg Fideo (TCSVT), 2019 - Rhagfynegiad ansawdd fideo stereosgopig yn seiliedig ar rwydweithiau niwral dwfn llif deuol o'r dechrau i'r diwedd
Wei Zhou, Z Chen, W Li
Cynhadledd Pacific-Rim ar Amlgyfrwng (PCM), 2018 - Asesiad ansawdd fideo stereosgopig dall: o ganfyddiad dyfnder i brofiad cyffredinol
Z Chen (cynghorydd), Wei Zhou, W Li
Trafodion IEEE ar Brosesu Delweddau (TIP), 2018
Addysgu
2024-2025, CM2101, Rhyngweithio Cyfrifiadur Dynol
2023-2024, CM6312, Mabwysiadu Technoleg (Academi Meddalwedd Genedlaethol), arweinydd modiwl
Goruchwyliwr prosiect ar gyfer myfyrwyr UG (blwyddyn 3)
Goruchwyliwr prosiect ar gyfer myfyrwyr MSc
Tiwtor personol ar gyfer myfyrwyr UG (blwyddyn 1 a blwyddyn 2)
Tiwtor personol ar gyfer myfyrwyr MSc
Bywgraffiad
Rwyf wedi bod yn Athro Cysylltiol (Uwch Ddarlithydd) yn yr Ysgol Cyfrifiadureg a Gwybodeg ym Mhrifysgol Caerdydd ers 2026. Rwy'n aelod o Grŵp Ymchwil Cyfrifiadura Gweledol. Rwyf hefyd yn aelod o bwyllgor Sefydliad Safonau Prydain. Rwy'n cymryd rhan yn y gweithgareddau safoni JPEG.
Cyn hynny, astudiais a gweithiais yn yr Adran Peirianneg Drydanol a Chyfrifiadurol ym Mhrifysgol Waterloo rhwng 2019 a 2023. Roeddwn i'n arfer bod yn athro gwadd ym Mhrifysgol Technoleg Dalian, ysgolhaig gwadd yn y Sefydliad Cenedlaethol Gwybodeg (Tokyo), cynorthwyydd ymchwil gydag Intel, intern ymchwil yn Microsoft Research ac Alibaba Cloud. Roedd gen i gefndir addysgol mewn Peirianneg Drydanol ym Mhrifysgol Gwyddoniaeth a Thechnoleg Tsieina.
Mae fy niddordebau ymchwil yn canolbwyntio'n bennaf ar brosesu delwedd / fideo, cyfrifiadura amlgyfrwng, canfyddiad cymhwysol, AI sy'n canolbwyntio ar ddynol, ffotograffiaeth gyfrifiadurol, delweddu, gweledigaeth, arddangosfeydd, a dysgu peiriannau. Rwy'n anelu at ddatblygu algorithmau prosesu gweledol effeithlon a deallus yn seiliedig ar ganfyddiad dynol, gan helpu i wella ansawdd profiad systemau gweledol dynol. Derbyniais Wobr Sôn Anrhydeddus IEEE CASS VSPC Rising Star (y derbynnydd cyntaf ym maes ymchwil IQA ledled y byd) yn 2024, Gwobr Traethawd Doethurol Eithriadol ACM SIGMM (Adran Tsieina, 1af ledled y wlad) yn 2022, a Gwobr Enillydd yr Her Fawr (safle 1af) yn CVPR CLIC 2021.
Anrhydeddau a dyfarniadau
2024 2% Gwyddonwyr Gorau Worldwide, Prifysgol Stanford
2024 Digileader gan Digital Futures, Sweden
2024 Pwll Tywod Ymchwil Rhyngddisgyblaethol o'r Academi Brydeinig, Grant Teithio
2024 Journal of Electronics & Information Technology Gwobr Adolygydd Eithriadol
Gwobr Sôn Anrhydeddus IEEE CASS VSPC Rising Star, y derbynnydd cyntaf ym maes ymchwil IQA ledled y byd
2023 Enillydd Gwobr Adolygydd Eithriadol Flynyddol y Synwyryddion
Gwobr Traethawd Hir Doethurol Eithriadol Tsieina 2022 ACM SIGMM, 1af ledled y wlad
2022 Cyrhaeddodd rownd derfynol Cymrodoriaeth Ôl-ddoethurol Arlywyddol o Brifysgol Dechnolegol Nanyang, Singapore
Cymrodoriaeth Ôl-ddoethurol MSCA 2022, Grant Teithio o Brifysgol Nantes, Ffrainc
Gwobr Enillydd Her Fawr 2021 ar gyfer IEEE CVPR CLIC Trac Metrig Canfyddiadol, 1af lle
Gwobr adolygydd gorau IEEE VCIP 2021
Gwobr Llywydd 2021, Academi Gwyddorau Tsieineaidd
2020 Traethawd Hir Doethurol Eithriadol Cyllid USTC
Ysgoloriaeth Genedlaethol 2020 gan USTC
2019 Outstanding Research Intern of Alibaba Group
2018 Gwobr Microsoft Research STAR of Tomorrow Excellence Award
2017 NII International Internship Funding Support, Japan
2016 Gwobr Traethawd Baglor Eithriadol Taleithiol
Meysydd goruchwyliaeth
Myfyrwyr presennol
Yixiao Li (Myfyriwr PhD Ymweld o BUAA, dan gefnogaeth Cyd-ariannu Rhyngwladol ar gyfer Myfyrwyr Doethurol)
- Papurau cydgysylltiedig: Adolygiad ail-rownd IEEE TNNLS (awdur cyntaf), IEEE TIP dan adolygiad (awdur cyntaf), CVPR dan adolygiad (awdur cyntaf), IEEE ICME 2024 (awdur cyntaf)
Huasheng Wang (myfyriwr PhD, CaerdyddU, goruchwyliwr arweiniol: Yr Athro Hantao Liu)
- Papurau a gydweithredir: IEEE TCSVT (awdur cyntaf)
Jiang Liu (myfyriwr PhD, CaerdyddU, goruchwyliwr arweiniol: Yr Athro Hantao Liu)
- Papurau cydgysylltiedig: Adolygiad ail-rownd IEEE TCSVT (awdur cyntaf)
Rwyf hefyd yn arholwyr Allanol (Prifysgol Shanghai Jiao Tong) o PhD Thesis.
Rhwng 2017 a 2021, roeddwn i'n arwain grŵp ymchwil yn USTC. Ers 2019, rwyf hefyd yn ffodus i weithio gyda llawer o fyfyrwyr talentog yn UWaterloo.
Armin Shafiee Sarvestani (myfyriwr PhD, UWaterloo)
Yipeng Du (Israddedig, UWaterloo)
Jinghan Zhou (myfyriwr PhD, UWaterloo)
Shiyu Huang (Prif fyfyriwr, USTC)
Yanding Peng (Prif fyfyriwr, USTC)
Yiting Lu (myfyriwr PhD, USTC)
Jun Fu (PhD, 2022 -> HW)
Jianzhao Liu (Meistr, 2022 -> Bytedance)
Ziyuan Luo (Meistr, 2021 -> Kwai)
Jiahua Xu (Meistr, 2021 -> Tencent, Ysgoloriaeth Genedlaethol, Gwobr Traethawd Traethawd Eithriadol Meistr
Ya Zhou (Meistr, 2020 -> Kwai)
Likun Shi (Meistr, 2019 -> SenseTime)
Xinyu Tang (Israddedig, 2019 -> Princeton PhD, Ysgoloriaeth Guo Moruo)
Chaoyi Lin (Meistr, 2018 -> Hikvision)
Contact Details
Themâu ymchwil
Arbenigeddau
- Prosesu delweddau
- Prosesu fideo
- Golwg cyfrifiadurol a chyfrifiant amlgyfrwng
- Deallusrwydd artiffisial