Evaluación efectiva de usabilidad mediante técnicas de análisis y extracción de conocimiento
DOI:
https://doi.org/10.65234/interaccion.141Palabras clave:
Evaluación de usabilidad, Diseño centrado en el usuario, Técnicas de Inteligencia ArtificialResumen
Este trabajo propone la aplicación de técnicas para automatizar evaluaciones de usabilidad basadas en el protocolo Thinking Aloud en con el objetivo de superar las limitaciones inherentes a los enfoques manuales tradicionales. Para ello, se realiza una revisión sistemática de la literatura que permitirá identificar avances recientes y vacíos existentes en la aplicación de técnicas automatizadas. El análisis considera tecnologías emergentes como el reconocimiento automático de voz, el procesamiento de lenguaje natural y el análisis multimodal de audio y video, evaluando su potencial para capturar y procesar datos de interacción de manera eficiente y objetiva. Asimismo, se examinan los retos asociados a la integración de estas tecnologías, incluyendo aspectos relacionados con la fiabilidad, la reducción de sesgos y la escalabilidad del proceso. A partir de los hallazgos, se propone el diseño de una herramienta de soporte orientada a la combinación de métodos de aprendizaje y análisis multimodal para optimizar la detección de emociones y la extracción de conocimiento en tiempo real. Esta aproximación busca mejorar la calidad y la eficiencia de las evaluaciones de usabilidad, ofreciendo un marco metodológico que contribuya a la evolución hacia procesos más automatizados y menos dependientes de la intervención humana para aumentar la objetividad.
Referencias
Abdat, F., Maaoui, C., & Pruski, A. (2011). Human–computer interaction using emotion recognition from facial expression. UKSim 5th European Symposium on Computer Modeling and Simulation (pp. 196–201). https://doi.org/10.1109/EMS.2011.20 DOI: https://doi.org/10.1109/EMS.2011.20
Bakkialakshmi, V. S., & Sudalaimuthu, T. (2021). A survey on affective computing for psychological emotion recognition. 5th International Conference on Electrical, Electronics, Communication, Computer Technologies and Optimization Techniques (ICEECCOT) (pp. 480–486). https://doi.org/10.1109/ICEECCOT52851.2021.9707947 DOI: https://doi.org/10.1109/ICEECCOT52851.2021.9707947
Baltrušaitis, T., Robinson, P., & Morency, L. P. (2016). OpenFace: An open source facial behavior analysis toolkit. IEEE Winter Conference on Applications of Computer Vision (WACV) (pp. 1–10). https://doi.org/10.1109/WACV.2016.7477553 DOI: https://doi.org/10.1109/WACV.2016.7477553
Bartlett, M. S., Littlewort, G., Fasel, I., & Movellan, J. R. (2003). Real-time face detection and facial expression recognition: Development and applications to human–computer interaction. Proceedings of the 2003 Conference on Computer Vision and Pattern Recognition Workshop (p. 53). https://doi.org/10.1109/CVPRW.2003.10057 DOI: https://doi.org/10.1109/CVPRW.2003.10057
Bhavan, A., Sharma, M., Piplani, M., Chauhan, P., Hitkul, S., & Shah, R. R. (2020). Deep learning approaches for speech emotion recognition. In B. Agarwal, R. Nayak, N. Mittal, & S. Patnaik (Eds.), Deep learning-based approaches for sentiment analysis (Algorithms for Intelligent Systems). Springer. https://doi.org/10.1007/978-981-15-1216-2_10 DOI: https://doi.org/10.1007/978-981-15-1216-2_10
Bhardwaj, V., Joshi, A., Bajaj, G., Sharma, V., Rushiya, A., & Bharghavi, S. S. (2023). Emotion detection from facial expressions using augmented reality. 2023 5th International Conference on Inventive Research in Computing Applications (ICIRCA) (pp. 1–5). https://doi.org/10.1109/ICIRCA57980.2023.10220824 DOI: https://doi.org/10.1109/ICIRCA57980.2023.10220824
Boren, T., & Ramey, J. (2000). Thinking aloud: Reconciling theory and practice. IEEE Transactions on Professional Communication, 43(3), 261–278. https://doi.org/10.1109/47.867942 DOI: https://doi.org/10.1109/47.867942
Calvo, R. A., & D'Mello, S. (2010). Affect detection: An interdisciplinary review of models, methods, and their applications. IEEE Transactions on Affective Computing, 1(1), 18–37. https://doi.org/10.1109/T-AFFC.2010.7 DOI: https://doi.org/10.1109/T-AFFC.2010.1
El Ayadi, M., Kamel, M., & Karray, F. (2011). Survey on speech emotion recognition: Features, classification schemes, and databases. Pattern Recognition, 44(3), 572–587. https://doi.org/10.1016/j.patcog.2010.09.020 DOI: https://doi.org/10.1016/j.patcog.2010.09.020
Fernández, J., & Macías, J. A. (2021, September). Heuristic-based usability evaluation support: A systematic literature review and comparative study. In Proceedings of the XXI International Conference on Human-Computer Interaction (pp. 1-9). https://doi.org/10.1145/3471391.3471395 DOI: https://doi.org/10.1145/3471391.3471395
Hernando, R., & Macías, J. A. (2023). Development of usable applications featuring QR codes for enhancing interaction and acceptance: a case study. Behaviour & Information Technology, 42(4), 360-378. https://doi.org/10.1080/0144929X.2021.2022209 DOI: https://doi.org/10.1080/0144929X.2021.2022209
Hertzum, M., & Holmegaard, K. D. (2015). Thinking aloud influences perceived time. Human Factors, 57(1), 101–109. https://doi.org/10.1177/0018720814549709 DOI: https://doi.org/10.1177/0018720814540208
Hinton, G., Deng, L., Yu, D., Dahl, G. E., Mohamed, A.-R., Jaitly, N., Senior, A., Vanhoucke, V., Nguyen, P., Sainath, T. N., & Kingsbury, B. (2012). Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups. IEEE Signal Processing Magazine, 29(6), 82–97. https://doi.org/10.1109/MSP.2012.2205597 DOI: https://doi.org/10.1109/MSP.2012.2205597
Jahangir, R., Teh, Y., Wah, H., Faiqa, F., & Mujtaba, G. (2021). Deep learning approaches for speech emotion recognition: State of the art and research challenges. Multimedia Tools and Applications, 80(16), 23745–23812. https://doi.org/10.1007/s11042-020-09874-7 DOI: https://doi.org/10.1007/s11042-020-09874-7
Jiang, C., Qiu, Y., Gao, H., Fan, T., Li, K., & Wan, J. (2019). An edge computing platform for intelligent operational monitoring in Internet data centers. IEEE Access, 7, 133375–133387. https://doi.org/10.1109/ACCESS.2019.2939614 DOI: https://doi.org/10.1109/ACCESS.2019.2939614
Kitchenham, B., & Charters, S. (2007). Guidelines for performing systematic literature reviews in software engineering(EBSE 2007 Technical Report). EBSE. https://www.durham.ac.uk/media/durham-university/departments-/computer-science/research/technical-reports/Guidelines-for-Performing-Systematic-Literature-Reviews-in-Software-Engineering-2007.pdf
Kitchenham, B., Brereton, O. P., Budgen, D., Turner, M., Bailey, J., & Linkman, S. (2009). Systematic literature reviews in software engineering – A systematic literature review. Information and Software Technology, 51(1), 7–15. https://doi.org/10.1016/j.infsof.2008.09.009 DOI: https://doi.org/10.1016/j.infsof.2008.09.009
Khan, U. A., Xu, Q., Liu, Y., Lagstedt, A., Alamäki, A., & Kauttonen, J. (2024). Exploring contactless techniques in multimodal emotion recognition: Insights into diverse applications, challenges, solutions, and prospects. Multimedia Systems, 30(115). https://doi.org/10.1007/s00530-024-01302-2 DOI: https://doi.org/10.1007/s00530-024-01302-2
Kuhn, K., Kersken, V., Reuter, B., Egger, N., & Zimmermann, G. (2024). Measuring the accuracy of automatic speech recognition solutions. ACM Transactions on Accessible Computing, 16(3), Article 25, 23 pages. https://doi.org/10.1145/3636513 DOI: https://doi.org/10.1145/3636513
Li, S., & Deng, W. (2022). Deep facial expression recognition: A survey. IEEE Transactions on Affective Computing, 13(3), 1195–1215. https://doi.org/10.1109/TAFFC.2020.2981446 DOI: https://doi.org/10.1109/TAFFC.2020.2981446
Li, S., Huang, X., Wang, T., Zheng, J. & Lajoie, S. (2025). Using text mining and machine learning to predict reasoning activities from think-aloud transcripts in computer-assisted learning. Journal of Computing in Higher Education, 37, 477–496. https://doi.org/10.1007/s12528-024-09404-6 DOI: https://doi.org/10.1007/s12528-024-09404-6
Macías, J. A. (2008). Intelligent assistance in authoring dynamically generated web interfaces. World Wide Web, 11(2), 253–286. https://doi.org/10.1007/s11280-008-0043-3 DOI: https://doi.org/10.1007/s11280-008-0043-3
Macías, J. A. (2012). Enhancing interaction design on the semantic web: A case study. IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews), 42(6), 1365–1373. https://doi.org/10.1109/TSMCC.2012.2187052 DOI: https://doi.org/10.1109/TSMCC.2012.2187052
Macías, J. A., & Castells, P. (2001). A generic presentation modeling system for adaptive web-based instructional applications. In CHI'01 Extended Abstracts on Human Factors in Computing Systems (pp. 349-350). https://doi.org/10.1145/634067.63427 DOI: https://doi.org/10.1145/634067.634273
Macías, J. A., & Castells, P. (2002). Tailoring dynamic ontology-driven web documents by demonstration. In Proceedings Sixth International Conference on Information Visualisation (pp. 535-540). IEEE. https://doi.org/10.1109/IV.2002.1028826 DOI: https://doi.org/10.1109/IV.2002.1028826
Macías, J. A., & Culén, A. L. (2021). Enhancing decision-making in user-centered web development: a methodology for card-sorting analysis. World Wide Web, 24(6), 2099-2137. https://doi.org/10.1007/s11280-021-00950-y DOI: https://doi.org/10.1007/s11280-021-00950-y
McDonald, S., Cockton, G., & Irons, A. (2020). The impact of thinking-aloud on usability inspection. Proceedings of the ACM on Human-Computer Interaction, 4(CSCW1), 1–22. https://doi.org/10.1145/3397876 DOI: https://doi.org/10.1145/3397876
McDonald, S., Edwards, H. M., & Zhao, T. (2012). Exploring think-alouds in usability testing: An international survey. IEEE Transactions on Professional Communication, 55(2), 2–19. https://doi.org/10.1109/TPC.2011.2182569 DOI: https://doi.org/10.1109/TPC.2011.2182569
Mollahosseini, A., Hasani, B., & Mahoor, M. H. (2016). Going deeper in facial expression recognition using deep neural networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW 2016) (pp. 1–6). https://doi.org/10.1109/CVPRW.2016.7477553 DOI: https://doi.org/10.1109/WACV.2016.7477450
Padua, A. G. (2010). Propuesta de un proceso de revisión sistemática de experimentos en ingeniería del software. Proceedings of the 13th Ibero-American Conference on Software Engineering (CIbSE 2010) (pp. 313–318). Universidad Politécnica de Madrid.
Pang, B., & Lee, L. (2008). Opinion mining and sentiment analysis. Foundations and Trends in Information Retrieval, 2(1–2), 1–135. https://doi.org/10.1561/1500000011 DOI: https://doi.org/10.1561/1500000011
Poria, S., Cambria, E., Bajpai, R., & Hussain, A. (2017). A review of affect analysis: From unimodal analysis to multimodal fusion. Information Fusion, 37, 98–125. https://doi.org/10.1016/j.inffus.2017.02.003 DOI: https://doi.org/10.1016/j.inffus.2017.02.003
Quintal, C., & Macías, J. A. (2021). Measuring and improving the quality of development processes based on usability and accessibility. Universal Access in the Information Society, 20(2), 203–221. https://doi.org/10.1007/s10209-020-00726-7 DOI: https://doi.org/10.1007/s10209-020-00726-7
Ribeiro, M. T., Singh, S., & Guestrin, C. (2016). “Why should I trust you?”: Explaining the predictions of any classifier. Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining(pp. 1135–1144). https://doi.org/10.1145/2939672.2939778 DOI: https://doi.org/10.1145/2939672.2939778
Rojas, L. A., & Macías, J. A. (2015). An agile information-architecture-driven approach for the development of user-centered interactive software. Proceedings of the XVI International Conference on Human–Computer Interaction (Article No. 50, pp. 1–8). https://doi.org/10.1145/2829875.2829919 DOI: https://doi.org/10.1145/2829875.2829919
Rojas, L. A., & Macías, J. A. (2019). Toward collisions produced in requirements rankings: A qualitative approach and experimental study. Journal of Systems and Software, 158, 110417. https://doi.org/10.1016/j.jss.2019.110417 DOI: https://doi.org/10.1016/j.jss.2019.110417
Scherer, K. R. (1986). Vocal communication of emotion: A review of research paradigms. Speech Communication, 5(1–2), 1–49. https://doi.org/10.1016/0167-6393(86)90070-X
Scherer, K. R., & Ellgring, H. (2007). Multimodal expression of emotion: Affect programs or componential appraisal patterns? Emotion, 7(1), 158–171. https://doi.org/10.1037/1528-3542.7.1.158 DOI: https://doi.org/10.1037/1528-3542.7.1.158
Schuller, B., Batliner, A., Steidl, S., & Seppi, D. (2011). Recognising realistic emotions and affect in speech: State of the art and lessons learnt from the first challenge. Speech Communication, 53(9–10), 1062–1087. https://doi.org/10.1016/j.specom.2011.01.011 DOI: https://doi.org/10.1016/j.specom.2011.01.011
Selvaraju, R. R., Cogswell, M., Das, A., Vedantam, R., Parikh, D., & Batra, D. (2017). Grad-CAM: Visual explanations from deep networks via gradient-based localization. 2017 IEEE International Conference on Computer Vision (ICCV)(pp. 618–626). https://doi.org/10.1109/ICCV.2017.74 DOI: https://doi.org/10.1109/ICCV.2017.74
Soleymani, M., Lichtenauer, J., Pun, T., & Pantic, M. (2012). A multimodal database for affect recognition and implicit tagging. IEEE Transactions on Affective Computing, 3(1), 42–55. https://doi.org/10.1109/T-AFFC.2011.25 DOI: https://doi.org/10.1109/T-AFFC.2011.25
Souto, T., Silva, H., Leite, A., Baptista, A., Queirós, C., & Marques, A. (2019). Facial emotion recognition: Virtual reality program for facial emotion recognition—a trial program targeted at individuals with schizophrenia. Rehabilitation Counseling Bulletin, 63(2), 79–90. https://doi.org/10.1177/0034355219847284 DOI: https://doi.org/10.1177/0034355219847284
Sun, Y., Qiu, H., Zheng, Y., Wang, Z., & Zhang, C. (2020). SIFRank: A new baseline for unsupervised keyphrase extraction based on pre-trained language model. IEEE Access, 8, 10896–10906. https://doi.org/10.1109/ACCESS.2020.2965087 DOI: https://doi.org/10.1109/ACCESS.2020.2965087
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. Advances in Neural Information Processing Systems, 30, 5998–6008.
Zhang, J., Borchers, C., Aleven, V., & Baker, R. S. (2024). Using large language models to detect self-regulated learning in think-aloud protocols. Proceedings of the 17th International Conference on Educational Data Mining (pp. 157–168). International Educational Data Mining Society. https://doi.org/10.5281/zenodo.12729790 DOI: https://doi.org/10.35542/osf.io/hrtz6
Descargas
Publicado
Número
Sección
Licencia
Derechos de autor 2025 Shuoshuo Li, José Antonio Macías Iglesias

Esta obra está bajo una licencia internacional Creative Commons Atribución-NoComercial 4.0.
