-
BACKGROUND:
The Center for Data-Driven Discovery in Biomedicine (D3b – https://d3b.center) is seeking talented engineers to build and scale solutions for accelerating discovery and advancements in child health. We're looking for creative problem solvers who can leverage software and system engineering to perform large-scale genomic and phenotypic analysis. You will be involved in research projects that contain large cross-disease cohorts which are being characterized with next-generation sequencing. This role will have a strong focus on phenotype harmonization and association with genomic variation. Successful candidates will gain deep knowledge about important clinical and research phenotypes across a number of pediatric diseases. With this knowledge, this role will utilize controlled vocabularies, ontologies and associated methodologies to develop scalable systems that harmonize, store, and analyze these phenotypes in the context of genomics.
RESPONSIBILITIES:
Computational Environment Development and Optimization (60%):- Independently manage and evolve local large-scale bioinformatics pipelines.
- Independently manage and evolve large-scale bioinformatics high performance computing (HPC) capability primarily at a local level.
- Independently manage and evolve process for effectively using enterprise-provided large-scale bioinformatics storage frameworks.
- Work with IS and data center staff to ensure the appropriate installation, maintenance, and support of bioinformatics-dedicated hardware, software, and data storage.
- Work with IS staff to establish and maintain appropriate levels of availability, response time, and performance of bioinformatics software and systems.
- Establish and implement integration and testing procedures following industry best practices for production bioinformatics systems integration and deployment.
- Facilitate efficient transfer of bioinformatics data from data sources to data users with benchmarking and data quality checks.
- Operationally manage and evolve a robust heterogeneous UNIX, LINUX, OSX environment, including integration with enterprise resources.
- Contribute to structured benchmark-based evaluation of new technologies.
- Ensure the appropriate installation, maintenance, and support of bioinformatics-dedicated hardware, software, and data storage by collaborating with information systems and data coordination staff.
- Generally implement bioinformatics processing, storage, and manipulation of bioinformatics data in a primarily local environment.
- Engage with and participate in discussions related to vendor-purchased systems and services usually under supervision.
- Provide continuous assessment of commercial and open-source bioinformatics data processing solutions by applying structured benchmark evaluation.
- Identify and test application/pipeline defects and fixes.
- Troubleshoot data discrepancies.
- Serve as engineering resource on a variety of bioinformatics-focused projects.
- Serve as engineering facilitator by assessing all stakeholders, including bioinformatics management, bioinformatics scientists, information systems staff, and principal investigators.
- Mentor lower tier engineering individuals and groups as needed
- Advocate for developed solutions in discussions with external technology owners in order to ensure that enterprise systems allow for freedom of operation.
- Under supervision, contribute to the development of a formal bioinformatics engineering plan and development roadmap.
- Adopts and implements policies and standards for data quality, completeness, and reproducibility.
- Adopts and implements policies and standards for performance benchmarking and system stability.
- Maintain and audit all documentation required for transparency and reproducibility of operations and any relevant regulations (e.g., CAP, CLIA). Documentation my include configuration, processes, service records, asset inventories, topologies, admin manuals, job instructions, support contacts, and bug/issue tracking.
- Install, maintain, and provide technical support for all software installations and associated hardware.
Required Education: Bachelor's Degree in computational discipline or systems engineering.
Required Experience: At least three (3) years of experience in a production clinical or research bioinformatics data processing role required.
Additional Technical Requirements:- Extensive knowledge with high performance and parallel computing environments and data processing workflows essential.
- Extensive experience with data storage frameworks essential.
- Extensive knowledge of CPU- and IO-intensive bioinformatics data analysis applications.
- Demonstrated track record of optimizing systems to meet changing performance and load requirements.
- Extensive knowledge of HPC systems, job management applications, including methods for profiling performance, benchmarking, and optimizing multiple job types and scenarios in bioinformatics data processing.
- Ability to independently plan and execute pipelines and workflows of high complexity required.
- Ability to independently engineer systems relative to larger enterprise framework required.
- Strong UNIX/LINUX expertise required.
- Expertise in support mechanisms for applications written in common bioinformatics languages such as R, Python, Perl or similar required.
- Expertise in support mechanisms for common bioinformatics applications, data sources, and data formats required.
- Knowledge of common microarray, NGS, mass spectrometry, or other high-throughput data formats is required.
- Expertise with resources of genomic data sets and analysis tools, such as UCSC Genome Browser, Bioconductor, ENCODE, and NCBI databases is required.
- Demonstrated ability to develop and implement best practices for bioinformatics systems integration, testing, and deployment is required.
- Expert knowledge with cloud computing concepts and applications is required.
- Ability to lead discussions with various information systems and technology owners to achieve desired bioinformatics outcomes is required.
Preferred Education: Master's degree in computational discipline or systems engineering.
Preferred Experience: Four (4) years of experience or more in a production clinical or research bioinformatics data processing role.
LOCATION:
Philadelphia
HOW TO APPLY:
Please follow this link to apply: https://careers.chop.edu/job/Philadelphia-Bioinformatics-Engineer-II-PA-19146/782579000/
Discussion forums: Opportunity: Bioinformatics Engineer II @ The Children's Hospital of Philadelphia (CHOP) -- Philadelphia, PA (US)
Expanded view | Monitor forum | Save place
Start a new thread:
You have to be logged in to post a reply.