Dr Fernando Alva Manchego
(e/fe)
- Ar gael fel goruchwyliwr ôl-raddedig
Timau a rolau for Fernando Alva Manchego
Uwch Ddarlithydd
Trosolwyg
Rwy'n Uwch Ddarlithydd (~Athro Cysylltiol) yn Ysgol y Gwyddorau Cyfrifiadurol a Mathemategol ym Mhrifysgol Caerdydd. Mae fy ymchwil yn canolbwyntio ar dechnolegau sy'n cymhwyso Deallusrwydd Artiffisial ar gyfer hygyrchedd gwybodaeth. Yn benodol, mae fy ngwaith yn defnyddio dulliau Prosesu Iaith Naturiol i hwyluso darllen a deall. Mae gen i ddiddordeb arbennig mewn astudio galluoedd gwirioneddol systemau ar gyfer sawl tasg Generartion Iaith Naturiol, megis Cyfieithu Peiriannau, Crynhoi a Symleiddio Testun. Er mwyn gwneud hynny, mae fy nghydweithwyr a minnau'n creu adnoddau iaith, dylunio methodolegau gwerthuso neu fetrigau, ac yn gweithredu modelau gan ddefnyddio technegau dysgu peiriannau .
Mae fy niddordebau ymchwil yn cynnwys:
- Cynhyrchu Testun-i-Destun (e.e. Symleiddio Testun, Crynhoi, Peiriant Cyfieithu, ac ati)
- Gwerthusiad o Genheu Iaith Naturiol
- Cymorth Ysgrifennu
- Prosesu Iaith Naturiol ar gyfer Addysg
Cyhoeddiad
2026
- Alshatti, A. , Schockaert, S. and Alva Manchego, F. 2026. A meta-evaluation of automatic metrics for elaborative simplification. Presented at: LREC 2026 Workshop Mallorca, Spain 11-16 May 2026. Published in: Shardlow, M. et al., Proceedings of the Joint Workshop on Readability and Text Simplification (READIxTSAR) @ LREC 2026. LREC. , pp.193-209. (10.63317/3bhnb2uoif7o)
- Gutiérrez-Rolón, N. and Alva Manchego, F. 2026. Unsupervised labelling of mutation triggers in Welsh. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Mallorca, Spain 11-16 May 2026. Proceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026). ELRA. , pp.11631-11641. (10.63317/37oxwc9pnyfv)
- Gutiérrez-Rolón, N. et al., 2026. Proffiliadur: Welsh language text profiling toolkit. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Palma de Mallorca, Spain 13-15 May 2026. Proceedings of Learning Resources Evaluation Conference 2026 (LREC). ELRA. , pp.1129-1142. (10.63317/5c6yawn79s5h)
- Waqar, E. et al., 2026. CEFR-Cymraeg: A dataset and baseline models for language proficiency assessment in Welsh. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Palma de Mallorca, Spain 13-15 May 2026. Proceedings of Learning Resources Evaluation Conference 2026 (LREC). ELRA. , pp.3496-3505. (10.63317/2dvuy5ucr9g2)
2025
- Ayesh, M. , Gutiérrez Rolón, N. and Alva Manchego, F. 2025. CardiffNLP at CLEARS-2025: Prompting large language models for plain language and easy-to-read text rewriting. Presented at: IberLEF 2025 Zaragoza, Spain 23 September 2025. Published in: Jiménez-Zafra, S. M. et al., Proceedings of the Iberian Languages Evaluation Forum (IberLEF 2025). Vol. 4098.
- Imperial, J. M. et al., 2025. UniversalCEFR: Enabling open multilingual research on language proficiency assessment. Presented at: Empirical Methods in Natural Language Processing (EMNLP 2025) Suzhou, China 4-11 November 2025. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.9703-9755. (10.18653/v1/2025.emnlp-main.491)
- Khallaf, N. et al., 2025. FreeTxt: Analyse and visualise multilingual qualitative survey data for cultural heritage sites. Presented at: Recent Advances in Natural Language Processing (RANLP) 2025 Varna, Bulgaria 8 -10 September 2025. Proceedings of the 15th International Conference on Recent Advances in Natural Language Processing - Natural Language Processing in the Generative AI Era. , pp.541-545. (10.26615/978-954-452-098-4-063)
- Maddela, M. and Alva Manchego, F. 2025. Adapting sentence-level automatic metrics for document-level simplification evaluation. Presented at: NAACL 2025 Albuquerque, New Mexico 29 April - 04 May 2025. Published in: Chiruzzo, L. , Ritter, A. and Wang, L. eds. Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies. Vol. 1.Association for Computational Linguistics. , pp.6444–6459. (10.18653/v1/2025.naacl-long.327)
- Mei, P. et al., 2025. If ChatGPT can do it, where is my creativity? generative AI boosts performance but diminishes experience in creative writing. Computers in Human Behavior: Artificial Humans 4 100140. (10.1016/j.chbah.2025.100140)
- Chaudhary, A. et al. 2025. Exploring the safe integration of generative AI in cybersecurity education: Addressing challenges in transparency, accuracy, and security. Presented at: 4th Annual Advances in Teaching and Learning for Cyber Security Education Bristol, UK 2 July 2024. Published in: Legg, P. , Coull, N. and Clarke, C. eds. Advances in Teaching and Learning for Cyber Security Education. Vol. 1213.Lecture Notes in Networks and Systems Vol. 1. Springer Cham. , pp.1-21. (10.1007/978-3-031-77524-6_1)
2023
- Kew, T. et al., 2023. BLESS: Benchmarking Large Language Models on Sentence Simplification. Presented at: 2023 Conference on Empirical Methods in Natural Language Processing 6-10 December 2023. Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing. ACL. , pp.13291–13309. (10.18653/v1/2023.emnlp-main.821)
- Ushio, A. , Alva Manchego, F. and Camacho-Collados, J. 2023. A practical toolkit for multilingual question and answer generation. Presented at: 61st Annual Meeting of the Association for Computational Linguistics 9-14 July 2023. Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics: System Demonstrations. Vol. 3.Association for Computational Linguistics. , pp.86-94. (10.18653/v1/2023.acl-demo.8)
- Ushio, A. , Alva Manchego, F. and Camacho-Collados, J. 2023. An empirical comparison of LM-based question and answer generation methods. Presented at: The 61st Annual Meeting of the Association for Computational Linguistics 9-14 July 2023. Findings of the Association for Computational Linguistics: ACL 2023. Toronto, Canada: Association for Computational Linguistics. , pp.14262-14272. (10.18653/v1/2023.findings-acl.899)
2022
- Ushio, A. , Alva Manchego, F. and Camacho Collados, J. 2022. Generative language models for paragraph-level question generation. Presented at: Conference on Empirical Methods in Natural Language Processing Abu Dhabi, UAE 7-11 December 2022. Published in: Goldberg, Y. , Kozareva, Z. and Zhang, Y. eds. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.670-688. (10.18653/v1/2022.emnlp-main.42)
- Vasquez-Rodriguez, L. et al., 2022. A benchmark for neural readability assessment of texts in Spanish. Presented at: Workshop on Text Simplification, Accessibility, and Readability (TSAR-2022) Abu Dhabi, United Arab Emirates (Virtual) 8 December 2022. Proceedings of the Workshop on Text Simplification, Accessibility, and Readability (TSAR-2022). Stroudsburg, PA, USA: Association for Computational Linguistics. , pp.188-198.
- Miliani, M. et al., 2022. Neural readability pairwise ranking for sentences in Italian administrative language. Presented at: 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing Online only 20-23 November 2022. Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing. Vol. 1.Association for Computational Linguistics. , pp.849-866.
- Alva Manchego, F. and Shardlow, M. 2022. Towards readability-controlled machine translation of COVID-19 texts. Presented at: 23rd Annual Conference of the European Association for Machine Translation Ghent, Belgium 1-3 June 2022. Published in: Moniz, H. et al., Proceedings of the 23rd Annual Conference of the European Association for Machine Translation. European Association for Machine Translation. , pp.287–288.
- Bejarano, G. et al., 2022. PeruSIL: A framework to build a continuous Peruvian Sign Language interpretation dataset. Presented at: LREC2022: 10th Workshop on the Representation and Processing of Sign Languages: Multilingual Sign Language Resources Marseille, France 20-25 June 2022. Published in: Efthimiou, E. et al., Proceedings of the LREC2022 10th Workshop on the Representation and Processing of Sign Languages: Multilingual Sign Language Resources. European Language Resources Association. , pp.1-8.
- Murrugarra-Llerena, J. , Alva Manchego, F. and Murrugarra-LLerena, N. 2022. Improving embeddings representations for comparing higher education curricula: A use case in computing. Presented at: 2022 Conference on Empirical Methods in Natural Language Processing Abu Dhabi, United Arab Emirates 7-11 December 2022. Published in: Goldberg, Y. , Kozareva, Z. and Zhang, Y. eds. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.11299–11307. (10.18653/v1/2022.emnlp-main.776)
- Shardlow, M. and Alva Manchego, F. 2022. Simple TICO-19: A dataset for joint translation and simplification of COVID-19 texts. Presented at: LREC 2022: Thirteenth Language Resources and Evaluation Conference Marseille, France 20-25 June 2022. Published in: Calzolari, N. et al., Proceedings of the Thirteenth Language Resources and Evaluation Conference. European Language Resources Association. , pp.3093–3102.
2021
- Alva Manchego, F. , Scarton, C. and Specia, L. 2021. The (un)suitability of automatic evaluation metrics for text simplification. Computational Linguistics 47 (4), pp.861-889. (10.1162/coli_a_00418)
- Alva-Manchego, F. et al. 2021. deepQuest-py: large and distilled models for quality estimation. Presented at: 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP) Punta Cana, Dominican Republic 7-11 November 2021. Published in: Adel, H. and Shi, S. eds. Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics. , pp.382-389. (10.18653/v1/2021.emnlp-demo.42)
- Gajbhiye, A. et al., 2021. Knowledge distillation for quality estimation. Presented at: 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL-IJCNLP 2021) Bangkok, Thailand 1-6 August 2021. Published in: Zong, C. et al., Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021. Association for Computational Linguistics. , pp.5091-5099. (10.18653/v1/2021.findings-acl.452)
- Rivas Rojas, K. and Alva-Manchego, F. 2021. IAPUCP at SemEval-2021 task 1: Stacking fine-tuned transformers is almost all you need for lexical complexity prediction. Presented at: 15th International Workshop on Semantic Evaluation (SemEval 2021) Virtual 5-6 August 2021. Published in: Palmer, A. et al., Proceedings of the 15th International Workshop on Semantic Evaluation (SemEval-2021). Association for Computational Linguistics. , pp.144-149. (10.18653/v1/2021.semeval-1.14)
- Maddela, M. , Alva-Manchego, F. and Xu, W. 2021. Controllable text simplification with explicit paraphrasing. Presented at: 2021 Annual Conference of the North American Chapter of the Association for Computational Linguistics Virtual 06-11 June 2021. Published in: Toutanova, K. et al., Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies. ML Research Press. , pp.3536-3553. (10.18653/v1/2021.naacl-main.277)
2020
- Alva Manchego, F. , Scarton, C. and Specia, L. 2020. Data-Driven Sentence Simplification: Survey and benchmark. Computational Linguistics 46 (1), pp.135-187. (10.1162/coli_a_00370)
- Alva Manchego, F. et al. 2020. ASSET: A dataset for tuning and evaluation of sentence simplification models with multiple rewriting transformations. Presented at: ACL 2020: 58th Annual Meeting of the Association for Computational LinguisticsPublished in: Jurafsky, D. et al., Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. ACL. , pp.4688-4679. (10.18653/v1/2020.acl-main.424)
2019
- Alva Manchego, F. et al. 2019. EASSE: Easier Automatic Sentence Simplification Evaluation. Presented at: 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) Hong Kong, China 3-7 November 2019. Published in: Pado, S. and Huang, R. eds. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP): System Demonstrations. Association for Computational Linguistics. , pp.49-54. (10.18653/v1/D19-3009)
- Finnimore, P. et al., 2019. Strong baselines for complex word identification across multiple languages. Presented at: 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019) Minneapolis, NM, USA 2-7 June 2019. Published in: Burstein, J. , Doran, C. and Solorio, T. eds. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics. , pp.970-977. (10.18653/v1/N19-1102)
2016
- Vargas-Campos, I. and Alva Manchego, F. 2016. SciEsp: Structural analysis of abstracts written in Spanish. Computación y Sistemas 20 (3), pp.551-558. (10.13053/cys-20-3-2463)
2012
- Alva Manchego, F. E. and Rosa, J. L. G. 2012. Semantic role labeling for Brazilian Portuguese: A benchmark. Presented at: Advances in Artificial Intelligence – IBERAMIA 2012 13-16 November 2012. Published in: Pavon, J. , Duque-Mendez, N. D. and Fuentes-Fernandez, R. eds. Advances in Artificial Intelligence – IBERAMIA 2012: 13th Ibero-American Conference on AI, Cartagena de Indias, Colombia, November 13-16, 2012. Proceedings. Vol. 7637.Lecture Notes in Computer Science Springer. , pp.481-490. (10.1007/978-3-642-34654-5_49)
- Alva Manchego, F. E. and Rosa, J. L. G. 2012. Towards semi-supervised Brazilian Portuguese semantic role labeling: Building a benchmark. Presented at: PROPOR: International Conference on Computational Processing of the Portuguese Language 17-20 April 2012. Published in: Caseli, H. et al., Computational Processing of the Portuguese Language: 10th International Conference, PROPOR 2012, Coimbra, Portugal, April 17-20, 2012. Proceedings. Vol. 7243.Springer. , pp.210-217. (10.1007/978-3-642-28885-2_24)
Cynadleddau
- Alshatti, A. , Schockaert, S. and Alva Manchego, F. 2026. A meta-evaluation of automatic metrics for elaborative simplification. Presented at: LREC 2026 Workshop Mallorca, Spain 11-16 May 2026. Published in: Shardlow, M. et al., Proceedings of the Joint Workshop on Readability and Text Simplification (READIxTSAR) @ LREC 2026. LREC. , pp.193-209. (10.63317/3bhnb2uoif7o)
- Gutiérrez-Rolón, N. and Alva Manchego, F. 2026. Unsupervised labelling of mutation triggers in Welsh. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Mallorca, Spain 11-16 May 2026. Proceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026). ELRA. , pp.11631-11641. (10.63317/37oxwc9pnyfv)
- Gutiérrez-Rolón, N. et al., 2026. Proffiliadur: Welsh language text profiling toolkit. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Palma de Mallorca, Spain 13-15 May 2026. Proceedings of Learning Resources Evaluation Conference 2026 (LREC). ELRA. , pp.1129-1142. (10.63317/5c6yawn79s5h)
- Waqar, E. et al., 2026. CEFR-Cymraeg: A dataset and baseline models for language proficiency assessment in Welsh. Presented at: The Fifteenth Language Resources and Evaluation Conference (LREC 2026) Palma de Mallorca, Spain 13-15 May 2026. Proceedings of Learning Resources Evaluation Conference 2026 (LREC). ELRA. , pp.3496-3505. (10.63317/2dvuy5ucr9g2)
- Ayesh, M. , Gutiérrez Rolón, N. and Alva Manchego, F. 2025. CardiffNLP at CLEARS-2025: Prompting large language models for plain language and easy-to-read text rewriting. Presented at: IberLEF 2025 Zaragoza, Spain 23 September 2025. Published in: Jiménez-Zafra, S. M. et al., Proceedings of the Iberian Languages Evaluation Forum (IberLEF 2025). Vol. 4098.
- Imperial, J. M. et al., 2025. UniversalCEFR: Enabling open multilingual research on language proficiency assessment. Presented at: Empirical Methods in Natural Language Processing (EMNLP 2025) Suzhou, China 4-11 November 2025. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.9703-9755. (10.18653/v1/2025.emnlp-main.491)
- Khallaf, N. et al., 2025. FreeTxt: Analyse and visualise multilingual qualitative survey data for cultural heritage sites. Presented at: Recent Advances in Natural Language Processing (RANLP) 2025 Varna, Bulgaria 8 -10 September 2025. Proceedings of the 15th International Conference on Recent Advances in Natural Language Processing - Natural Language Processing in the Generative AI Era. , pp.541-545. (10.26615/978-954-452-098-4-063)
- Maddela, M. and Alva Manchego, F. 2025. Adapting sentence-level automatic metrics for document-level simplification evaluation. Presented at: NAACL 2025 Albuquerque, New Mexico 29 April - 04 May 2025. Published in: Chiruzzo, L. , Ritter, A. and Wang, L. eds. Proceedings of the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics: Human Language Technologies. Vol. 1.Association for Computational Linguistics. , pp.6444–6459. (10.18653/v1/2025.naacl-long.327)
- Chaudhary, A. et al. 2025. Exploring the safe integration of generative AI in cybersecurity education: Addressing challenges in transparency, accuracy, and security. Presented at: 4th Annual Advances in Teaching and Learning for Cyber Security Education Bristol, UK 2 July 2024. Published in: Legg, P. , Coull, N. and Clarke, C. eds. Advances in Teaching and Learning for Cyber Security Education. Vol. 1213.Lecture Notes in Networks and Systems Vol. 1. Springer Cham. , pp.1-21. (10.1007/978-3-031-77524-6_1)
- Kew, T. et al., 2023. BLESS: Benchmarking Large Language Models on Sentence Simplification. Presented at: 2023 Conference on Empirical Methods in Natural Language Processing 6-10 December 2023. Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing. ACL. , pp.13291–13309. (10.18653/v1/2023.emnlp-main.821)
- Ushio, A. , Alva Manchego, F. and Camacho-Collados, J. 2023. A practical toolkit for multilingual question and answer generation. Presented at: 61st Annual Meeting of the Association for Computational Linguistics 9-14 July 2023. Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics: System Demonstrations. Vol. 3.Association for Computational Linguistics. , pp.86-94. (10.18653/v1/2023.acl-demo.8)
- Ushio, A. , Alva Manchego, F. and Camacho-Collados, J. 2023. An empirical comparison of LM-based question and answer generation methods. Presented at: The 61st Annual Meeting of the Association for Computational Linguistics 9-14 July 2023. Findings of the Association for Computational Linguistics: ACL 2023. Toronto, Canada: Association for Computational Linguistics. , pp.14262-14272. (10.18653/v1/2023.findings-acl.899)
- Ushio, A. , Alva Manchego, F. and Camacho Collados, J. 2022. Generative language models for paragraph-level question generation. Presented at: Conference on Empirical Methods in Natural Language Processing Abu Dhabi, UAE 7-11 December 2022. Published in: Goldberg, Y. , Kozareva, Z. and Zhang, Y. eds. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.670-688. (10.18653/v1/2022.emnlp-main.42)
- Vasquez-Rodriguez, L. et al., 2022. A benchmark for neural readability assessment of texts in Spanish. Presented at: Workshop on Text Simplification, Accessibility, and Readability (TSAR-2022) Abu Dhabi, United Arab Emirates (Virtual) 8 December 2022. Proceedings of the Workshop on Text Simplification, Accessibility, and Readability (TSAR-2022). Stroudsburg, PA, USA: Association for Computational Linguistics. , pp.188-198.
- Miliani, M. et al., 2022. Neural readability pairwise ranking for sentences in Italian administrative language. Presented at: 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing Online only 20-23 November 2022. Proceedings of the 2nd Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 12th International Joint Conference on Natural Language Processing. Vol. 1.Association for Computational Linguistics. , pp.849-866.
- Alva Manchego, F. and Shardlow, M. 2022. Towards readability-controlled machine translation of COVID-19 texts. Presented at: 23rd Annual Conference of the European Association for Machine Translation Ghent, Belgium 1-3 June 2022. Published in: Moniz, H. et al., Proceedings of the 23rd Annual Conference of the European Association for Machine Translation. European Association for Machine Translation. , pp.287–288.
- Bejarano, G. et al., 2022. PeruSIL: A framework to build a continuous Peruvian Sign Language interpretation dataset. Presented at: LREC2022: 10th Workshop on the Representation and Processing of Sign Languages: Multilingual Sign Language Resources Marseille, France 20-25 June 2022. Published in: Efthimiou, E. et al., Proceedings of the LREC2022 10th Workshop on the Representation and Processing of Sign Languages: Multilingual Sign Language Resources. European Language Resources Association. , pp.1-8.
- Murrugarra-Llerena, J. , Alva Manchego, F. and Murrugarra-LLerena, N. 2022. Improving embeddings representations for comparing higher education curricula: A use case in computing. Presented at: 2022 Conference on Empirical Methods in Natural Language Processing Abu Dhabi, United Arab Emirates 7-11 December 2022. Published in: Goldberg, Y. , Kozareva, Z. and Zhang, Y. eds. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing. Association for Computational Linguistics. , pp.11299–11307. (10.18653/v1/2022.emnlp-main.776)
- Shardlow, M. and Alva Manchego, F. 2022. Simple TICO-19: A dataset for joint translation and simplification of COVID-19 texts. Presented at: LREC 2022: Thirteenth Language Resources and Evaluation Conference Marseille, France 20-25 June 2022. Published in: Calzolari, N. et al., Proceedings of the Thirteenth Language Resources and Evaluation Conference. European Language Resources Association. , pp.3093–3102.
- Alva-Manchego, F. et al. 2021. deepQuest-py: large and distilled models for quality estimation. Presented at: 2021 Conference on Empirical Methods in Natural Language Processing (EMNLP) Punta Cana, Dominican Republic 7-11 November 2021. Published in: Adel, H. and Shi, S. eds. Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. Association for Computational Linguistics. , pp.382-389. (10.18653/v1/2021.emnlp-demo.42)
- Gajbhiye, A. et al., 2021. Knowledge distillation for quality estimation. Presented at: 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL-IJCNLP 2021) Bangkok, Thailand 1-6 August 2021. Published in: Zong, C. et al., Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021. Association for Computational Linguistics. , pp.5091-5099. (10.18653/v1/2021.findings-acl.452)
- Rivas Rojas, K. and Alva-Manchego, F. 2021. IAPUCP at SemEval-2021 task 1: Stacking fine-tuned transformers is almost all you need for lexical complexity prediction. Presented at: 15th International Workshop on Semantic Evaluation (SemEval 2021) Virtual 5-6 August 2021. Published in: Palmer, A. et al., Proceedings of the 15th International Workshop on Semantic Evaluation (SemEval-2021). Association for Computational Linguistics. , pp.144-149. (10.18653/v1/2021.semeval-1.14)
- Maddela, M. , Alva-Manchego, F. and Xu, W. 2021. Controllable text simplification with explicit paraphrasing. Presented at: 2021 Annual Conference of the North American Chapter of the Association for Computational Linguistics Virtual 06-11 June 2021. Published in: Toutanova, K. et al., Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies. ML Research Press. , pp.3536-3553. (10.18653/v1/2021.naacl-main.277)
- Alva Manchego, F. et al. 2020. ASSET: A dataset for tuning and evaluation of sentence simplification models with multiple rewriting transformations. Presented at: ACL 2020: 58th Annual Meeting of the Association for Computational LinguisticsPublished in: Jurafsky, D. et al., Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. ACL. , pp.4688-4679. (10.18653/v1/2020.acl-main.424)
- Alva Manchego, F. et al. 2019. EASSE: Easier Automatic Sentence Simplification Evaluation. Presented at: 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) Hong Kong, China 3-7 November 2019. Published in: Pado, S. and Huang, R. eds. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP): System Demonstrations. Association for Computational Linguistics. , pp.49-54. (10.18653/v1/D19-3009)
- Finnimore, P. et al., 2019. Strong baselines for complex word identification across multiple languages. Presented at: 2019 Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT 2019) Minneapolis, NM, USA 2-7 June 2019. Published in: Burstein, J. , Doran, C. and Solorio, T. eds. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). Association for Computational Linguistics. , pp.970-977. (10.18653/v1/N19-1102)
- Alva Manchego, F. E. and Rosa, J. L. G. 2012. Semantic role labeling for Brazilian Portuguese: A benchmark. Presented at: Advances in Artificial Intelligence – IBERAMIA 2012 13-16 November 2012. Published in: Pavon, J. , Duque-Mendez, N. D. and Fuentes-Fernandez, R. eds. Advances in Artificial Intelligence – IBERAMIA 2012: 13th Ibero-American Conference on AI, Cartagena de Indias, Colombia, November 13-16, 2012. Proceedings. Vol. 7637.Lecture Notes in Computer Science Springer. , pp.481-490. (10.1007/978-3-642-34654-5_49)
- Alva Manchego, F. E. and Rosa, J. L. G. 2012. Towards semi-supervised Brazilian Portuguese semantic role labeling: Building a benchmark. Presented at: PROPOR: International Conference on Computational Processing of the Portuguese Language 17-20 April 2012. Published in: Caseli, H. et al., Computational Processing of the Portuguese Language: 10th International Conference, PROPOR 2012, Coimbra, Portugal, April 17-20, 2012. Proceedings. Vol. 7243.Springer. , pp.210-217. (10.1007/978-3-642-28885-2_24)
Erthyglau
- Mei, P. et al., 2025. If ChatGPT can do it, where is my creativity? generative AI boosts performance but diminishes experience in creative writing. Computers in Human Behavior: Artificial Humans 4 100140. (10.1016/j.chbah.2025.100140)
- Alva Manchego, F. , Scarton, C. and Specia, L. 2021. The (un)suitability of automatic evaluation metrics for text simplification. Computational Linguistics 47 (4), pp.861-889. (10.1162/coli_a_00418)
- Alva Manchego, F. , Scarton, C. and Specia, L. 2020. Data-Driven Sentence Simplification: Survey and benchmark. Computational Linguistics 46 (1), pp.135-187. (10.1162/coli_a_00370)
- Vargas-Campos, I. and Alva Manchego, F. 2016. SciEsp: Structural analysis of abstracts written in Spanish. Computación y Sistemas 20 (3), pp.551-558. (10.13053/cys-20-3-2463)
Bywgraffiad
I joined the School of Computer Science and Informatics at Cardiff University in January 2022.
Previously, I was a Postdoctoral Research Associate at the University of Sheffield and a member of the Natural Language Processing Group (2020-2021). I worked with Prof. Lucia Specia for the APE-QUEST (EU CEF Integration Project) and Bergamot (EU's Horizon 2020) projects on Quality Estimation for Machine Translation.
I hold a PhD in Computer Science from the University of Sheffield focused on Automatic Text Simplification. My thesis title was: "Automatic Sentence Simplification with Multiple Rewriting Transformations". I was supervised by Prof. Lucia Specia and Dr. Carolina Scarton.
Before that, I worked as Adjunct Professor at the Pontifical Catholic University of Peru (2013-2016), where I was a member of the Artificial Intelligence Group IA-PUCP. During my Masters, I was also a member of the Interinstitutional Center for Computational Linguistics at the University of São Paulo.
Meysydd goruchwyliaeth
I am interested in supervising PhD students in projects involving Natural Language Processing for Text Adaptation.
Please, check the relevant pages in FindAPhD for more information, depending on whether you are a self-funded student, or plan on applying to a School scholarhip. Do not hesitate to contact me if you have any questions!
Goruchwyliaeth gyfredol
Contact Details
+44 29225 14738
Abacws, Ystafell Room 4.64, Ffordd Senghennydd, Cathays, Caerdydd, CF24 4AG
Themâu ymchwil
Arbenigeddau
- .AI
- Deallusrwydd artiffisial
- Prosesu iaith naturiol
- Ieithyddiaeth gyfrifiadurol