Usability of the "Be Yourself" IA-powered, chatbot-based website for the prevention of adolescent pregnancy.
Objective. Analyze the usability, interaction patterns, and quality of conversational interaction of the AI-powered "Be Yourself" chatbot to foster self-determined motivation and prevent adolescent pregnancy. Post-intervention usability pilot study conducted with 74 Mexican school-going adolescents (aged 11-15), using the System Usability Scale (SUS) at two time points: immediately after the intervention and at two-month follow-up. Interaction metrics (number of messages) and confidence coefficients were analyzed within content related to self-determined motivation, pregnancy prevention, and sexual communication. Participants were 49.6% male and 50.4% female, with a mean age of 12.7 (SD=0.98). The chatbot recorded 74 chats and 1,236 messages. The SUS was assessed post-intervention (n=74) and at the two-month follow-up (n=70). Initially, usability was rated as 18.9% "Good" and 81.1% "Acceptable." At follow-up, the proportions across categories changed to 5.7% "Good," 92.9% "Acceptable," and 1.4% "Poor." The Wilcoxon signed-rank test showed a significant decrease in perceived usability (p<0.001; r=-0.56) between the two assessment points. The highest confidence coefficients were observed for communication with parents (93.2%) and pregnancy prevention (89.1%), whereas self-determined motivation was lower (79.1%). The chatbot demonstrated acceptable usability and high short-term acceptance as an informational tool for preventing adolescent pregnancy. However, the decrease in usability scores at the two-month follow-up indicates the need for improvements to foster sustained engagement and address behavioral aspects.
- Book Chapter
1
- 10.1007/978-3-319-96068-5_97
- Aug 8, 2018
The huge rate of road accidents in Iran demonstrates the necessity of implementing prevention interventions in this field. Traffic signs are communication tools for passing information to the road users. However, misunderstanding of their messages could be a major cause of road accidents. The aim of the present study was to measure the usability of traffic signs used in rural and urban areas of Iran. 356 Iranians licensed drivers (39.0 years ± 9.7) were requested to rate their estimation about the effectiveness of the 20 selected traffic signs on the Persian version of System Usability Score (SUS). The mean of usability score for all signs were 59.37 ± 15.3. Mean usability score of Indication signs was higher than two other groups (63.2 ± 29.3) followed by Mandatory signs (60.9 ± 13.5). Warning signs had the lowest level of usability (53.9 ± 15.2). Except two signs (“Unguarded railway crossing in 200 m”) with the SUS score of less than 50 and “Camping” with the SUS score of higher than 70, the effectiveness of all signs was assessed to be at moderate level (50 > SUS < 70). It could be concluded that the majority of Iran traffic signs are marginally usable for Iranian population. Taking into account ergonomic features in designing traffic signs, could result in increasing sign effectiveness and decreasing road accidents. Appropriate strategies should be implemented to design more user-friend traffic signs.
- Research Article
2
- 10.1080/09720510.2020.1736316
- Feb 17, 2020
- Journal of Statistics and Management Systems
This study has two main objectives; one is to present the empirical result that shows how the SUS scores vary when end-users perform usability assessment relying on experience and by executing the assigned list of specific tasks on academic websites. The results are obtained by using independent sample t-test that scrutinizes the slight but significant difference in mean SUS scores among two. Results also investigate that 85.29% of websites have shown higher mean SUS scores when usability assessment is conducted with the assigned list of specific tasks whereas only 14.70% of websites have shown higher mean SUS scores when usability assessment is performed based on the experience. Another aim is to determine the relationship which is found to be strong positive correlation between system effectiveness, an ISO metric, measured as task success rate and subjective usability scores, measured through System Usability Scale (SUS) for 50-academic websites. It concludes that successful completion of the tasks results in high usability scores.
- Research Article
74
- 10.1007/s10209-021-00820-4
- May 22, 2021
- Universal Access in the Information Society
This paper elaborates the empirical evidence of a usability evaluation of a VR and non-VR virtual tour application for a living museum. The System Usability Scale (SUS) was used in between participants experiments (Group 1: non-VR version and Group 2: VR version) with 40 participants. The results show that the mean scores of all components for the VR version are higher compared to the non-VR version, overall SUS score (72.10 vs 68.10), usability score (75.50 vs 71.70), and learnability (58.40 vs 57.00). Further analysis using a two-tailed independent t test showed no difference between the non-VR and VR versions. Additionally, no significant difference was observed between the groups in the context of gender, nationality, and prior experience (other VR tour applications) for overall SUS score, usability score, and learnability score. Α two-tailed independent t test indicated no significant difference in the usability score between participants with VR experience and no VR experience. However, a significant difference was found between participants with VR experience and no VR experience for both SUS score (t(38) = 2.17, p = 0.037) and learnability score (t(38) = 2.40, p = 0.021). The independent t test results indicated a significant difference between participant with and without previous visits to SCV for the usability score (t(38) = −2.31, p = 0.027), while there was no significant differences observed in other components. It can be concluded that both versions passed based on the SUS score. However, the sub-scale usability and learnability scores indicated some usability issue.
- Research Article
- 10.1007/s10508-025-03351-8
- Jan 7, 2026
- Archives of sexual behavior
The objective of the study was to test a structural equation model in which sexual communication, life goals, basic psychological needs, and self-determined motivation were incorporated in order to explain their effect on sexual behavior to prevent adolescent pregnancy. This research was conducted with a sample of 620 Mexican adolescents of both sexes. The results of the model showed satisfactory goodness-of-fit indexes and it was determined that sexual communication with their mother predicted their intrinsic life goals. Sexual communication with friends predicted life goals and the satisfaction of basic psychological needs, and self-determined motivation decreased. Sexual communication with the partner increased basic psychological needs satisfaction. The greater basic psychological needs satisfaction, the greater self-determined motivation. Self-determined motivation can increase sexual behaviors to prevent pregnancy.
- Research Article
21
- 10.1071/sh10025
- Jan 1, 2011
- Sexual Health
Early marriage is common in many developing countries, including India. Women who marry early have little power within their marriage, particularly in the sexual domain. Research is limited on women's ability to control their marital sexual experiences. We identified factors affecting sexual communication among married women aged 16-25, in Bangalore, India, and how factors associated with sexual communication differed from those influencing non-sexual agency. We ran ordered logit regression models for one outcome of sexual agency (sexual communication, n = 735) and two outcomes of non-sexual agency (fertility control, n = 735, and financial decision-making, n = 728). Sexual communication was more restricted (83 women (11.3%) with high sexual communication) than financial decision-making (183 women (25.1%) with high financial decision-making agency) and fertility control (238 women (32.4%) with high fertility control). Feeling prepared before the first sexual experience was significantly associated with sexual communication (odds ratio (OR) = 1.8; 95% confidence interval (CI) = 1.13-2.89). Longer marriage duration (OR 2.13; 95% CI = 1.42-3.20) and having worked pre-marriage (OR 1.38; 95% CI = 1.02-1.86) were also significant. Few other measures of women's resources increased their odds of sexual communication. Education, having children, pre-marital vocational training and marital intimacy were significant for non-sexual outcomes but not sexual communication. Policy-makers seeking to enhance young married women's sexual communication need to consider providing sex education to young women before they marry. More broadly, interventions designed to increase women's agency need to be tailored to the type of agency being examined.
- Research Article
48
- 10.2196/mhealth.6611
- Nov 10, 2016
- JMIR mHealth and uHealth
BackgroundAdolescents in the United States and globally represent a high-risk population for unintended pregnancy, which leads to high social, economic, and health costs. Access to smartphone apps is rapidly increasing among youth, but little is known about the strategies that apps employ to prevent pregnancy among adolescents and young adults. Further, there are no guidelines on best practices for adolescent and young adult pregnancy prevention through mobile apps.ObjectiveThis review developed a preliminary evaluation framework for the assessment of mobile apps for adolescent and young adult pregnancy prevention and used this framework to assess available apps in the Apple App Store and Google Play that targeted adolescents and young adults with family planning and pregnancy prevention support.MethodsWe developed an assessment rubric called Mobile Criteria for Adolescent Pregnancy Prevention (mCAPP) for data extraction using evidence-based and promising best practices from the literature. mCAPP comprises 4 domains: (1) app characteristics, (2) user interface features, (3) adolescent pregnancy prevention best practices, and (4) general sexual and reproductive health (SRH) features. For inclusion in the review, apps that advertised pregnancy prevention services and explicitly mentioned youth, were in English, and were free were systematically identified in the Apple App Store and Google Play in 2015. Screening, data extraction, and 4 interrater reliability checks were conducted by 2 reviewers. Each app was assessed for 92 facets of the mCAPP checklist.ResultsOur search returned 4043 app descriptions in the Apple App Store (462) and Google Play (3581). After screening for inclusion criteria, 22 unique apps were included in our analysis. Included apps targeted teens in primarily developed countries, and the most common user interface features were clinic and health service locators. While app strengths included provision of SRH education, description of modern contraceptives, and some use of evidence-based adolescent best practices, gaps remain in the implementation of the majority of adolescent best practices and user interface features. Of the 8 best practices for teen pregnancy prevention operationalized through mCAPP, the most commonly implemented best practice was the provision of information on how to use contraceptives to prevent pregnancy (15/22), followed by provision of accurate information on pregnancy risk of sexual behaviors (13/22); information on SRH communication, negotiation, or refusal skills (10/22); and the use of persuasive language around contraceptive use (9/22).ConclusionsThe quality and scope of apps for adolescent pregnancy prevention varies, indicating that developers and researchers may need a supportive framework. mCAPP can help researchers and developers consider mobile-relevant evidence-based best practices for adolescent SRH as they develop teen pregnancy prevention apps. Given the novelty of the mobile approach, further research is needed on the impact of mCAPP criteria via mobile channels on adolescent health knowledge, behaviors, and outcomes.
- Book Chapter
3
- 10.1007/978-981-15-1286-5_2
- Jan 1, 2020
The prime objective of this study is to empirically determine the effect of tasks difficulty on usability scores computed using the System Usability Scale (SUS). Usability dataset is created by involving twelve end-users that evaluate the usability of 15 academic websites in a laboratory. Each end-user performs three subsets of six tasks whose difficulty is varying from easy to impossible under six different categories. Results are obtained after applying two statistical techniques, one is ANOVA and other is the correlation with regression. Results show that the SUS scores vary from higher to lower values when end-users conduct usability assessment with a list of easy, moderate, and impossible tasks on academic websites. The results also indicate the effect of tasks difficulty on the correlation between the SUS scores and task success rate. Though, the strength of the correlation is strong with each subset of tasks but it varies that depends on the nature of the tasks.
- Research Article
- 10.3233/shti250086
- Apr 8, 2025
- Studies in health technology and informatics
Surgical video review improves performance, but video capture, editing, sharing, and platform usability challenges hinder implementation. Evaluating platform usability by end-users is vital for effective adoption. In this study, we utilized a standardized measurement scale, the System Usability Scale (SUS), to measure the usability of an existing local network video capture (LNVC) platform within a surgical department. The existing LNVC system consists of intraoperative camera capture with an attached hard drive for data retention. The SUS survey was administered to surgical residents and faculty at a surgical department. The SUS was distributed via an online survey tool, and the SUS was calculated for distinct user groups. The median usability score for the entire cohort was low (Score:43, SUS average score: 68, 10thpercentile score:40). The median score was worst among routine video users (Score: 36). For the first time, we demonstrate that LNVC has among the lowest usability scores in a single surgery department.
- Research Article
2
- 10.1177/1541931214581238
- Sep 1, 2014
- Proceedings of the Human Factors and Ergonomics Society Annual Meeting
This study examined whether the average usability score for a series of tasks was the same as the usability score for the product if usability was measured only after all the tasks had been completed. Fifty participants completed a set of tasks for five websites and fourteen mock voting ballots. Subjective usability assessment was made with the System Usability Scale (SUS). Participants completed the SUS either after each task (five or fourteen SUS administrations, respectively) or after completing the entire set of tasks (one SUS). The results show that the average SUS scores for the task-level assessments were significantly higher than the SUS scores for the test-level assessments. Results were similar for the ballot and website conditions. Task-level SUS scores on the Honda websites ( M = 65.5) were significantly higher than the test-level SUS scores ( M = 42.8), p < 0.0001. Similar results were observed in the ballot condition, where task-level usability assessments were higher ( M = 59.5) than test-level assessments ( M = 38.5), p < 0.0001. Practitioners and those interpreting SUS scores need to be aware of how these experimental differences can lead to different assessment metrics.
- Research Article
- 10.1038/s41598-025-21016-3
- Oct 24, 2025
- Scientific Reports
Plantar pressure measurements provide critical insights into the structural and functional attributes of the lower limbs and feet. While previous studies have evaluated the reliability of insole technology for assessing foot pressure distribution during linear walking, natural motion often involves a combination of straight walking and turning. This study aimed to validate the test-retest reliability of a wearable in-shoe plantar pressure monitoring system for both linear and curved walking trajectories and to determine the minimum distance required to achieve excellent reliability for each output variable. Thirty-one healthy participants (15 females and 16 males aged 19–25 years) were recruited. Each participant performed two testing sessions, 4–7 days apart, involving three walking conditions: linear walking (LIN), clockwise curved walking (CW), and counterclockwise curved walking (CCW). A wearable footwear system equipped with embedded pressure sensors was used to collect plantar pressure data. Five key parameters—the peak pressure (PP), pressure‒time integral (PTI), full width at half maximum (FWHM), maximum pressure gradient (MaxPG), and average pressure (AP)—were analyzed across eight foot regions. Reliability was assessed via intraclass correlation coefficients (ICCs), Bland‒Altman plots, and minimal detectable changes (MDCs). Additionally, the System Usability Scale (SUS) and Intrinsic Motivation Inventory (IMI) were administered to evaluate usability. The wearable system demonstrated good reliability, with ICC values of approximately 0.9 for most parameters across all walking conditions. For whole-foot analysis, all variables presented ICCs > 0.60, confirming high reliability. The minimum distances required to achieve an ICC ≥ 0.90 were 207 m for LIN, 255 m for CW, and 467 m for CCW. The usability scores (SUS and IMI) indicated high user satisfaction and system acceptability. The developed wearable plantar pressure system exhibited excellent reliability and usability for both linear and curved walking conditions. These findings support its potential for clinical and research applications, particularly in scenarios involving mixed walking trajectories. The study also highlights the importance of considering step count thresholds to ensure reliable assessments in diverse walking conditions.Supplementary InformationThe online version contains supplementary material available at 10.1038/s41598-025-21016-3.
- Research Article
- 10.47065/josh.v5i4.5487
- Jul 26, 2024
- Journal of Information System Research (JOSH)
PT Pegadaian Persero is one of the State-Owned Enterprises that helps the government to provide financial solutions to customers who are experiencing financial difficulties by providing macro-scale loans or providing legal credit to customers. PT Pegadaian creates software that makes it easier for employees and clients to process all types of business processes and transactions because the company has been impacted by advances in information technology. With the increase and development of information technology, PT Pegadaian realizes that the use of information technology is very helpful in the process of optimizing business growth. The software owned by PT Pegadaian is PASSION (Pegadaian Application Support System Integrated Online). The PASSION application is an information system used in productivity and daily operations at Pegadaian, so it is necessary to pay attention to the usability level of the application so that it can make it easier for employees to apply the application. The aim of this research is to measure employee satisfaction, effectiveness and usability scores with the PASSION application used. To analyze the effectiveness, level of satisfaction and usability of the PASSION application, the System Usability Scale (SUS) method is the appropriate solution and method for evaluating system usability based on industry standards. The System Usability Scale (SUS) method was chosen because respondents can complete questions quickly and easily by filling out a questionnaire consisting of ten statements and the results are a single score from 0 to 100, which makes it relatively easy for users to understand. Based on the calculation process that has been carried out using the System Usability Scale (SUS) method, a value of 69.125 is obtained, which is in the grade C or "Good" position and has an Acceptable range based on the Net Promoter Score (NPS). This value shows that the PASSION application has successfully calculated its feasibility or usability level using this method so that this application can be used by employees at PT Pegadaian.
- Research Article
- 10.60005/ijlens.v2i1.102
- Feb 28, 2025
- International Journal of Learning Media on Natural Science (IJLENS)
Educational technology platforms and tools generally exhibit good to high usability, as measured by the System Usability Scale (SUS). Several studies indicate that usability scores are not significantly influenced by user demographics such as age, gender, or device type. However, factors like the subject matter, personality traits, and educational stage can sometimes affect perceived usability. This overview will delve into these aspects, highlighting key findings from relevant studies. This study aims to evaluate the usability of the web-based learning resource center Kumatalibi.com using the System Usability Scale (SUS) method, which consists of 10 standardized questions to assess a system's usability quality. A total of 101 preservice biology teacher students participated in the study. The results showed that the average SUS score was 73. Based on this score, the website received a grade of C, placed in the 70th percentile, categorized as "Good" in adjective rating, "Marginal" in acceptability, and classified as "Passive" based on its Net Promoter Score (NPS). These findings indicate that the learning website is generally acceptable to users; however, improvements are needed to enhance user satisfaction. Enhancing usability may lead to better learning outcomes for preservice biology teacher students.
- Research Article
3
- 10.15408/jti.v18i1.40766
- Apr 30, 2025
- JURNAL TEKNIK INFORMATIKA
The e-Polvot system at the University of Science and Technology Indonesia (USTI) is a digital platform used for student elections, replacing traditional paper-based voting to enhance efficiency and minimize election fraud. This study evaluates the system using the System Usability Scale (SUS) to assess its usability, including efficiency, effectiveness, and user satisfaction. However, SUS alone does not determine failure points but provides a usability score that reflects user perception. A survey was conducted with 88 respondents from three different academic programs, which showed that while the system generally received a "Good" usability rating, certain areas require enhancement to improve user engagement and satisfaction. Based on the findings, this study recommends enhancing the user interface, providing targeted user training, and introducing additional features to broaden the system’s application across academic units. Additionally, the study highlights the potential for expanding the system's functionality beyond student elections, supporting activities such as departmental voting and organizational decision-making processes. These improvements aim to increase user satisfaction and usability, making the system a more effective tool for various academic and institutional contexts.
- Research Article
- 10.1067/j.cpradiol.2026.03.008
- Mar 1, 2026
- Current problems in diagnostic radiology
Electronic health record usability for management of actionable incidental imaging findings: A survey of primary care providers.
- Research Article
16
- 10.1080/10447318.2018.1437865
- Feb 13, 2018
- International Journal of Human–Computer Interaction
ABSTRACTMarketing researchers use geography to identify specific user groups for studies to more effectively describe their potential customer base. Since usability professionals often recruit users employing similar selection criteria as their marketing peers, the use of geographic information might also be relevant when selecting usability test participants. In total, 3,168 participants from across the United States rated the usability of different hardware, software, and web-based products using the System Usability Scale (SUS). SUS scores were compared across geographic divisions to determine if usability assessments differ by location. SUS scores were also compared across rural and urban areas to determine if usability assessment scores change with population density. There was a lack of evidence to support significant differences in usability scores across both US geographic areas and zones of population density. The findings suggest that people make similar system usability assessments regardless of the area of the United States in which they live.