Data Science & Machine Learning Intern: Statistical Genetics

South San Francisco, CA

Applications have closed
The OpportunityGlobal drug development productivity is declining exponentially, with an overall failure to develop effective treatments for many increasingly prevalent complex diseases affecting millions of patients per year. We seek to solve this by combining cutting-edge machine learning techniques with recent advances in life sciences to drastically improve how drugs are discovered and developed.
This summer we are looking for highly motivated interns looking to work at the intersection of machine learning and life sciences. As a statistical genetics intern, you will conduct genetic association studies and other statistical genetics techniques that involve integration of genetic, genomic and clinical data to guide in vitro disease model development. You will also work closely with a cross-functional team of machine learning scientists, software engineers, life scientists, bioengineers, and medical scientists to integrate human level data with our high-throughput in-house in vitro genomic and phenotypic data to identify therapeutic targets and develop drugs that have high efficacy and low toxicity.
Insitro has a highly dynamic and collaborative culture in a custom-built open office in South San Francisco. You will work closely with machine learning engineers and scientists, biologists, chemists, microscopy experts, and automation engineers. You will be mentored by one of our senior researchers, who has significant experience in machine learning for genetics or statistical genetics in metabolism, neuroscience, or oncology. You will also attend our machine learning team meetings and will be exposed to a diverse set of novel technologies and machine learning concepts that tackle various biological questions. 
Join us, and help make a difference to patients!

About You- Working towards a BS, MS, or Ph.D. in an engineering, computer science, mathematics, statistics, life science, chemistry, physics, or a related discipline.- Proficiency in one or more general-purpose programming languages. We primarily use Python.- Demonstrated ability to use and develop cutting edge statistical and machine learning methods inspired by real problems.- Demonstrated ability to write high-quality, production-ready code (readable, well-tested, with well-designed APIs)- Experience with germline single nucleotide variants, structural variants, and/or somatic variants- Ability to communicate effectively and collaborate with people of diverse backgrounds and job functions.- Passion for making a difference in the world.

Nice to Have- Hands-on experience with genetic association testing (eQTL mapping, GWAS, EWAS, PheWAS, etc.) or experience mining modern, large-scale genetic databases (e.g. ExAC/gnomAD, UK Biobank, EBI GWAS Catalog, etc.)- Experience in Linux environment, database languages (e.g., SQL, No-SQL) and version control practices and tools such as Git or Mercurial.- Familiarity with the SciPy/PyData ecosystem (numpy, pandas, scipy, dask etc.).- Familiarity with cloud computing services (AWS or GCP).

Intern Benefits- Excellent medical, dental, and vision coverage; insitro pays 100% of premiums for employees- Excellent mental health and well-being support- Access to free onsite baristas and cafe with daily lunch and breakfast- Access to free onsite fitness center- Commuter benefits- Competitive pay- Flexible work schedule (on site and remote)
About insitroinsitro is a data-driven drug discovery and development company using machine learning and data at scale to transform the way that drugs are discovered and developed for patients. insitro is developing predictive machine learning models to discover underlying biologic state based on human cohort data and in-house generated cellular data at scale. These predictive models can be brought to bear on key bottlenecks in pharmaceutical R&D to advance novel targets and patient biomarkers, design therapeutics, and inform clinical strategy. insitro is advancing a wholly owned and partnered pipeline of biologic insights and molecules in neuroscience and metabolic diseases. Since formation in mid 2018, insitro has raised over $700 million from top tech, biotech, and crossover investors and from collaborations with pharmaceutical partners. For more information on insitro, please visit the company’s website at www.insitro.com.

Tags: APIs AWS Chemistry Computer Science Drug discovery Engineering GCP Git Linux Machine Learning Mathematics ML models NumPy Pandas Physics Python R R&D SciPy SQL Statistics Testing

Perks/benefits: Career development Competitive pay Flex hours Health care

Region: North America
Country: United States
Job stats:  13  2  0

More jobs like this

Explore more AI, ML, Data Science career opportunities

Find even more open roles in Artificial Intelligence (AI), Machine Learning (ML), Natural Language Processing (NLP), Computer Vision (CV), Data Engineering, Data Analytics, Big Data, and Data Science in general - ordered by popularity of job title or skills, toolset and products used - below.