Data Science & Machine Learning Intern: Research Engineering
South San Francisco, CA
This summer we are looking for highly motivated interns looking to work at the intersection of machine learning and life sciences. As a research engineering intern, you will work on projects developing and maintaining our core machine learning algorithms and infrastructure. Our team is focused on the craft of scientific software engineering, ensuring that our tools are scalable, correct, maintainable, and robust. We work on a variety of problem domains and datasets; we use multi-petabyte collections of biological data spanning high content imaging, functional genomics, and biomolecular structures, and you will help tackle the software challenges that arise from analyzing this data. While not required, some knowledge of biological or chemical data is especially valuable in understanding the unique requirements and applications of ML to biology and drug discovery.
Insitro has a highly dynamic and collaborative culture in a custom-built open office in South San Francisco. You will work closely with machine learning engineers and scientists, biologists, chemists, microscopy experts, and automation engineers. You will be mentored by one of our senior researchers, who has significant experience in developing scalable, high performance tooling for machine learning. You will also attend our machine learning team meetings and will be exposed to a diverse set of novel technologies and machine learning concepts that tackle various biological questions.
Join us, and help make a difference to patients!
About You- Working towards a BS, MS, or Ph.D. in an engineering, computer science, mathematics, statistics, life science, chemistry, physics, or a related discipline.- Proficiency in one or more general-purpose programming languages. We primarily use Python.- Demonstrated ability to use and develop cutting edge statistical and machine learning methods inspired by real problems.- Demonstrated ability to write high-quality, production-ready code (readable, well-tested, with well-designed APIs)- Experience with the full lifecycle of designing, implementing, deploying and maintaining robust scientific software- Ability to communicate effectively and collaborate with people of diverse backgrounds and job functions.- Passion for making a difference in the world.
Nice to Have- Experience building, shipping, and benchmarking ML systems, or working on tooling supporting those activities- Experience adapting, developing, and adding functionality to differentiable programming tools such as Pytorch, Jax or probabilistic programming tools such as Pyro/Numpyro- Experience in Linux environment, database languages (e.g., SQL, No-SQL) and version control practices and tools such as Git or Mercurial.- Familiarity with the SciPy/PyData ecosystem (numpy, pandas, scipy, dask etc.).- Familiarity with cloud computing services (AWS or GCP).
Intern Benefits- Excellent medical, dental, and vision coverage; insitro pays 100% of premiums for employees- Excellent mental health and well-being support- Access to free onsite baristas and cafe with daily lunch and breakfast- Access to free onsite fitness center- Commuter benefits- Competitive pay- Flexible work schedule (on site and remote)
About insitroinsitro is a data-driven drug discovery and development company using machine learning and data at scale to transform the way that drugs are discovered and developed for patients. insitro is developing predictive machine learning models to discover underlying biologic state based on human cohort data and in-house generated cellular data at scale. These predictive models can be brought to bear on key bottlenecks in pharmaceutical R&D to advance novel targets and patient biomarkers, design therapeutics, and inform clinical strategy. insitro is advancing a wholly owned and partnered pipeline of biologic insights and molecules in neuroscience and metabolic diseases. Since formation in mid 2018, insitro has raised over $700 million from top tech, biotech, and crossover investors and from collaborations with pharmaceutical partners. For more information on insitro, please visit the company’s website at www.insitro.com.
Tags: APIs AWS Biology Chemistry Computer Science Drug discovery Engineering GCP Git Linux Machine Learning Mathematics ML models NumPy Pandas Physics Python PyTorch R R&D Research SciPy SQL Statistics
Perks/benefits: Career development Competitive pay Flex hours Health care Team events
More jobs like this
Explore more AI, ML, Data Science career opportunities
Find even more open roles in Artificial Intelligence (AI), Machine Learning (ML), Natural Language Processing (NLP), Computer Vision (CV), Data Engineering, Data Analytics, Big Data, and Data Science in general - ordered by popularity of job title or skills, toolset and products used - below.
- Open Data Science Manager jobs
- Open Lead Data Analyst jobs
- Open MLOps Engineer jobs
- Open Senior Business Intelligence Analyst jobs
- Open Data Engineer II jobs
- Open Sr Data Engineer jobs
- Open Data Manager jobs
- Open Principal Data Engineer jobs
- Open Data Analytics Engineer jobs
- Open Power BI Developer jobs
- Open Product Data Analyst jobs
- Open Junior Data Scientist jobs
- Open Business Intelligence Developer jobs
- Open Data Scientist II jobs
- Open Senior Data Architect jobs
- Open Sr. Data Scientist jobs
- Open Manager, Data Engineering jobs
- Open Business Data Analyst jobs
- Open Big Data Engineer jobs
- Open Data Analyst Intern jobs
- Open Data Quality Analyst jobs
- Open Principal Data Scientist jobs
- Open Data Product Manager jobs
- Open Azure Data Engineer jobs
- Open Junior Data Engineer jobs
- Open Data quality-related jobs
- Open Business Intelligence-related jobs
- Open GCP-related jobs
- Open ML models-related jobs
- Open Data management-related jobs
- Open Privacy-related jobs
- Open Java-related jobs
- Open Finance-related jobs
- Open Data visualization-related jobs
- Open APIs-related jobs
- Open Deep Learning-related jobs
- Open PyTorch-related jobs
- Open Consulting-related jobs
- Open Snowflake-related jobs
- Open TensorFlow-related jobs
- Open PhD-related jobs
- Open CI/CD-related jobs
- Open NLP-related jobs
- Open Kubernetes-related jobs
- Open Data governance-related jobs
- Open Airflow-related jobs
- Open Hadoop-related jobs
- Open Databricks-related jobs
- Open LLMs-related jobs
- Open Data warehouse-related jobs