OPPORTUNITY
Mammoth Biosciences is seeking a Senior Software Engineer I, Research Informatics, to join our Computational team. This is an impactful group of bioinformaticians and software engineers that supports Mammoth's therapeutic CRISPR gene-editing programs from discovery through IND-enabling studies. As part of the Computational team, you’ll help build and run the analysis pipelines, model and manage the data those programs generate, and develop the internal tools scientists use to design experiments and interpret results.
An ideal candidate will work at the intersection of software engineering and biology. You’ll be responsible for managing and improving the systems the research organization depends on every day: the data warehouse and models that connect assay results to decisions, the internal web applications scientists use to launch analyses and access their data, and the cloud infrastructure everything runs on.
KEY RESPONSIBILITIES
Design and own relational data models and pipelines that drive scientific decisions, working in the context of existing production systems including Benchling
Define and refine how data is ingested, modeled, and made queryable
Collaborate with wet-lab scientists to build right-sized internal analysis tools
Design and improve the web applications scientists use to launch analyses and access data
Maintain the cloud infrastructure and automation tooling the research org runs on
Work directly with Computational team members and biologists to understand scientific needs, then troubleshoot and improve the relevant pipelines and systems
Enable responsible and reproducible use of AI across the company
Document systems and decisions to improve reproducibility and engineering standards
REQUIRED QUALIFICATIONS
Bachelor's, Master's, or PhD degree in computer sciences or biological field
7+ years of professional software engineering experience (5 years with Master's or 3 years with PhD) in the life sciences or related industry
Strong proficiency in Python
Hands-on experience with AWS services (e.g. S3, Batch, RDS, EC2, Lambda) and containerized deployment (Docker)
Experience with web development tools (ReactJS, HTML, CSS)
Experience in SQL, designing relational data models and working with production databases
Experience with AI coding tools
Experience working with scientists, preferably in genomics (with NGS data) or related field
Proficiency with git-based collaborative development, code review, automated testing, and CI/CD
Ability to work onsite 2x/per week at our Brisbane HQ
PREFERRED QUALIFICATIONS
Experience with workflow orchestration systems (Prefect, Nextflow, CWL, Airflow)
Experience with LIMS/ELN data modeling, preferably Benchling
Experience building and operating data pipelines or ETL systems, and comfort owning them end to end
Experience with analytics engineering tooling (DBT) and data warehouse design
Track record of building tools used by scientists or other non-engineer users, and of translating ambiguous requirements into working software
Demonstrated ability to work independently in a small team where priorities shift and everyone wears many hats
Background -- or real interest -- in the underlying biology, and the communication skills to work directly with bench scientists
Infrastructure-as-code experience (Terraform)
Thoughtful, productive use of AI coding tools, and a considered point of view on where they help and where they do not
BENEFITS
Company-paid health/vision/dental benefits
Unlimited vacation and generous sick time
Company-sponsored meals and snacks
Wellness, caregiver and ergonomics benefits
401(k) with company matching