Search bioRxiv⌕ Search

bioRxiv · 10.1101/2025.10.07.680870

Evaluating peer-to-peer bioinformatics education: a case study of student learning outcomes and community impact in an undergraduate multi-omic data analysis course

Abstract

Computational methodology has become ubiquitous in biomedical research with the rise of big data analysis and popularity of artificial intelligence and machine learning. However, undergraduate bioinformatics education has largely struggled to keep pace with the demand for bioinformatics skills, due to a combination of social and resource-based barriers. In this case study, we discuss the application and outcomes of Multi-Omic Data Analysis, a peer-to-peer learning-based undergraduate bioinformatics course offered by the Department of Quantitative and Computational Biology at the University of Southern California. Over eight semesters, a cohort of student instructors taught 2-3 weekly lectures to 107 undergraduate students in the Quantitative Biology Bachelor of Science degree program. Lectures covered a range of topics, including R and Python data analysis, scientific communication, and general research readiness as undergraduate students. We find that bioinformatics education courses structured around peer-to-peer learning have great potential to overcome many of the obstacles to comprehensive undergraduate bioinformatics education, and provide additional benefits related to student cohesion and community. We further discuss the longevity and feasibility of such courses, both specific to our program and in undergraduate universities at large. Author SummaryThere is a growing need for accessible and beginner-friendly bioinformatics education at the undergraduate level. Traditional coursework often fails to meet this demand due to disciplinary separation between biology and computer science, as well as limited institutional resources. We evaluate a unique approach involving a student-led undergraduate bioinformatics course at the University of Southern California to understand how peer-to-peer learning influences student development in computational biology. We examined over 100 students across eight academic semesters. Students varied in prior coding experience and research exposure. Through survey responses, we found that students left the course feeling more confident and prepared for research in faculty labs. Students reported improved competence in bioinformatics skills compared to before they took the course. We found that a major reason for this success was the peer-to-peer learning format. Students mentioned that learning from fellow students made the material more approachable and created a supportive community where they felt comfortable asking questions. Our work demonstrates that student-led initiatives can be a highly effective solution for filling critical gaps in university curricula, making bioinformatics education more accessible.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Boohar, W. R., Xu, K. Y., Black, N., Mogalipuvvu, M., Manley, K., Calabrese, P., Lee, J. S. H.. 2025-10-09. Evaluating peer-to-peer bioinformatics education: a case study of student learning outcomes and community impact in an undergraduate multi-omic data analysis course. https://doi.org/10.1101/2025.10.07.680870

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Perceived Risk and Barriers to Open and Responsible Research Across Fifteen UK Universities

Open research practices are increasingly promoted to improve research transparency, reproducibility, accessibility, and societal impact. Despite growing support from research funders, institutions, and policy initiatives, adoption remains uneven across disciplines and research communities. This study examined perceived risks and barriers associated with 14 FORRT Guideline areas using qualitative responses from the UK Reproducibility Network Open and Transparent Research Practices Survey (N = 2,567), conducted across 15 UK higher education institutions. Free-text responses describing risks and barriers were analysed using inductive thematic analysis. A total of 3,951 relevant comments generated 3,433 coded references to barriers and risks. Three interconnected clusters emerged. Individual concerns included lack of motivation, fear of losing intellectual credit, and concerns about exposing mistakes and criticism. Systemic and institutional barriers included lack of time and resources, inadequate infrastructure, insufficient training and support, lack of incentives and recognition, unclear guidance, disciplinary and methodological challenges, and tensions between open practices and intellectual property requirements. Ethical and quality-related concerns included risks to participant confidentiality and privacy, challenges associated with sensitive data, concerns about inappropriate application of open research practices across different research traditions, perceived impacts on research quality and innovation, and potential effects on public trust in research. Lack of time and resources was the most frequently reported barrier. Across responses, barriers were commonly described as interdependent, with shortcomings in funding, infrastructure, institutional support, incentives, and training reinforcing one another. Respondents largely supported the principles underpinning open research but highlighted substantial practical, professional, and ethical challenges to implementation. These findings suggest that increasing open research adoption requires more than policy mandates or awareness-raising activities. Sustainable uptake will depend on aligned incentives, adequate infrastructure and support, recognition of methodological diversity, and approaches that enable openness to be implemented responsibly across different research contexts.

scientific communication and education↗

Evaluating Large Language Models as Tools to Navigate Researchers in Rapidly Evolving Research Landscapes: A Case Study in Cancer Drug Response Prediction

Large Language Models (LLMs) have emerged as promising tools for assisting researchers in automating and accelerating the synthesis of literature reviews. However, their reliability is a significant concern due to issues like factual inaccuracies and hallucinations. The key question is whether LLMs can reliably provide comprehensive, up-to-date overviews and analyses. This study evaluates the performance of three leading LLMs (OpenAI's ChatGPT, Google's Gemini, and DeepSeek) on the complex task of generating a comprehensive survey paper on deep learning for cancer Drug Response Prediction (DRP). By testing both standard and Deep Research (DR) / Deep Think (DT) modes of LLMs with prompts of varying detail, this paper assesses key academic dimensions, including reference management, content quality, and analytical depth. Key findings reveal that while DR modes of LLMs significantly improve reliability by eliminating hallucinations, performance variations exist across models and prompts. A trade-off between reference quantity and integration quality was observed, and even the best-performing models lacked the analytical depth of human experts, often requiring extensive human supervision. The study concludes that LLMs currently serve as powerful assistive tools but still cannot replace the critical validation and synthesis provided by human researchers. Choosing the best LLM to use depends on the task in hand, while several strategies can be implemented to improve the produced output.

scientific communication and education↗

Undergraduate Biophysical Chemistry Series: Teaching through a Combination of a Purpose-built Textbook, Research-derived Biomolecular Samples and Computer Labs

Biophysics is a rapidly advancing field with an incredible breadth of topics. Thus, undergraduate biophysics instructors have to strategize and decide what topics they will cover in their courses. Educational institutions utilize a variety of biophysics textbooks. A common deficiency of each of the existing texts is that it serves well a given set of topics (theory, illustrations, practice problems) and leaves out other areas. A typical example includes good theory and problems for thermodynamics and kinetics while presenting molecular dynamics and various spectroscopic methods in a lacking or outdated way. The authors of this manuscript teach a capstone Biophysical Chemistry three-quarter series (Western Washington University/WWU, Bellingham, WA) which ideally should resonate with the general and major-specific courses the students take within their major at WWU. To achieve this goal and to enrich the traditional lecture-based delivery, the instructors have developed and brought together key pedagogical elements: purpose-built online textbook with a uniform structure of the academic content and practice problems, a study sample (oligopeptide) of biophysical significance with a growing set of experimental and computational data and student-centric in-class activities including computer labs. Our Biophysical series emphasizes concepts and methods of computational structural biology (Molecular Dynamics) and spectroscopic approaches (IR, UV and NMR). Here we describe the details of our integrative approach, summarize key outcomes and chart ways to advance the biophysical chemistry series further. Our textbook can be found through LibreText.

scientific communication and education↗