Search bioRxiv⌕ Search

bioRxiv · 10.1101/2022.10.21.513244

A cost-free CURE: Using bioinformatics to identify DNA-binding factors at a specific genomic locus

Abstract

Research experiences provide diverse benefits for undergraduates. Many academic institutions have adopted course-based undergraduate research experiences (CUREs) to improve student access to research opportunities. However, potential instructors of a CURE might still face financial or practical hurdles that prevent implementation. Bioinformatics research offers an alternative that is free, safe, compatible with remote learning, and may be more accessible for students with disabilities. Here, we describe a bioinformatics CURE that leverages publicly available datasets to discover novel proteins that target an instructor-determined genomic locus of interest. We use the free, user-friendly bioinformatics platform Galaxy to map ChIP-seq datasets to a genome, which removes the computing burden from students. Both faculty and students directly benefit from this CURE, as faculty can perform candidate screens and publish CURE results. Students gain not only basic bioinformatics knowledge, but also transferable skills, including scientific communication, database navigation, and primary literature experience. The CURE is flexible and can be expanded to analyze different types of high-throughput data or to investigate different genomic loci in any species.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Schmidt, C. A., Hodkinson, L. J., Comstra, H. S., Rieder, L. E.. 2022-10-24. A cost-free CURE: Using bioinformatics to identify DNA-binding factors at a specific genomic locus. https://doi.org/10.1101/2022.10.21.513244

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Evaluating Large Language Models as Tools to Navigate Researchers in Rapidly Evolving Research Landscapes: A Case Study in Cancer Drug Response Prediction

Large Language Models (LLMs) have emerged as promising tools for assisting researchers in automating and accelerating the synthesis of literature reviews. However, their reliability is a significant concern due to issues like factual inaccuracies and hallucinations. The key question is whether LLMs can reliably provide comprehensive, up-to-date overviews and analyses. This study evaluates the performance of three leading LLMs (OpenAI's ChatGPT, Google's Gemini, and DeepSeek) on the complex task of generating a comprehensive survey paper on deep learning for cancer Drug Response Prediction (DRP). By testing both standard and Deep Research (DR) / Deep Think (DT) modes of LLMs with prompts of varying detail, this paper assesses key academic dimensions, including reference management, content quality, and analytical depth. Key findings reveal that while DR modes of LLMs significantly improve reliability by eliminating hallucinations, performance variations exist across models and prompts. A trade-off between reference quantity and integration quality was observed, and even the best-performing models lacked the analytical depth of human experts, often requiring extensive human supervision. The study concludes that LLMs currently serve as powerful assistive tools but still cannot replace the critical validation and synthesis provided by human researchers. Choosing the best LLM to use depends on the task in hand, while several strategies can be implemented to improve the produced output.

scientific communication and education↗

SIMHYB 2: a software tool to explore and illustrate evolutionary forces in Population Genetics teaching and research. Application to Conservation Genetics

Practical approaches have become a standard in many scientific disciplines, including population genetics. By analyzing properly selected datasets, the students can calculate parameters and draw conclusions about genetic diversity, differentiation and evolution of populations with higher efficiency than if based exclusively on theoretical lessons. However, preparing the appropriate datasets is a hard task and a wrong selection can spoil a well-aimed practice. Here we present SO_SCPLOWIMC_SCPLOWHO_SCPLOWYBC_SCPLOW 2, a software tool specifically intended to ease the full understanding of evolutionary forces by the students and to help the teacher to prepare adequate datasets and examples for the practices. It simulates the course of a mixed population under user-defined reproductive and evolutionary conditions. Outputs can be easily adapted for downstream analysis with other popular tools as GO_SCPLOWENC_SCPLOWAO_SCPLOWLC_SCPLOWEO_SCPLOWXC_SCPLOW or SO_SCPLOWTRUCTUREC_SCPLOW. Thus, SO_SCPLOWIMC_SCPLOWHO_SCPLOWYBC_SCPLOW 2 is very suitable for project-based-learning approaches: students can produce their own datasets in different scenarios of genetic drift, migration, selective advantage, reproductive success... Additionally, SO_SCPLOWIMC_SCPLOWHO_SCPLOWYBC_SCPLOW 2 is the only simulation software available to date providing traceable pedigrees of individuals, being therefore very convenient for preparing datasets for parentage analysis, spatial genetic structure or conservation genetics study cases. Satisfactory results from its ongoing utilization in higher education and research are reported.

scientific communication and education↗

Postdoctoral Scholar Recruitment and Hiring Practices in STEM: A Pilot Study

Despite the importance of the postdoctoral position in the training of scientists for independent research careers, few studies have addressed recruiting and hiring of postdocs. We conducted a pilot study on postdoctoral hiring in the Division of Chemistry and Chemical Engineering at the California Institute of Technology to serve as a starting point to better understand postdoctoral recruiting and hiring processes. From this survey of both postdocs and faculty, together with the available literature, the picture emerges that the postdoc hiring process is more decentralized than either faculty hiring or graduate admissions. Postdoc positions are often filled through a passive process where the initial expression of interest from a prospective postdoc is through a "cold-call" contact to a prospective advisor. Individual faculty members are often responsible for developing and implementing their own outreach and recruitment plans and deciding who to hire into a postdoc position. The overall opacity of the processes and practices by which postdocs are identified, recruited, and hired make it difficult to pinpoint where interventions could be effective to ensure equitable hiring practices. Implementation of such practices is critical to training a diverse postdoc population and subsequently of the future STEM faculty recruited from this group.

scientific communication and education↗