Develop and run reproducible analysis pipelines for NGS and other omics data, perform QC, integrate and harmonize datasets, generate cohort summaries and visualizations, and communicate results to collaborators.
This is a remote position.
About the Organization
We are a research‑driven group working with large‑scale genomic and related biomedical datasets to support studies in areas such as rare disease, oncology, infectious disease, and neurology. Our work focuses on developing and applying computational methods that help collaborators interpret complex molecular data and generate results that can inform research and, where applicable, clinical decision‑making.
The team includes scientists, analysts, and software professionals who collaborate closely with partners in academic, clinical, and industry settings.
Position Overview
The Bioinformatics Analyst will be responsible for developing and running data analysis workflows, performing quality control, and summarizing results from next‑generation sequencing and other omics data. The role combines hands‑on data analysis with the design and maintenance of reproducible computational pipelines.
This position is suited to someone who enjoys working directly with data, building robust workflows, and communicating findings to a variety of stakeholders.
Key Responsibilities
- Process and analyze genomic and other omics datasets (for example, whole‑genome, whole‑exome, RNA‑seq, or similar assays), including alignment, quality assessment, variant detection, and annotation.
- Develop, document, and maintain reproducible analysis pipelines using modern workflow or pipeline tools and scripting languages.
- Implement best practices for data quality control, including monitoring run performance, detecting technical issues, and proposing corrective actions.
- Integrate data from multiple sources, harmonize formats and metadata, and prepare analysis‑ready datasets.
- Work with common bioinformatics tools and file formats (e.g., FASTQ, BAM/CRAM, VCF, BED, GFF/GTF) in a Unix/Linux environment.
- Develop and execute exploratory analyses, including cohort‑level summaries, visualization of key metrics, and interpretation of variant and gene‑level results.
-
- Prepare clear, well‑structured reports, figures, and presentation materials describing methods, assumptions, and findings for collaborators with diverse backgrounds.
- Contribute to the evaluation and adoption of new algorithms, tools, and workflows in bioinformatics and data analysis.
- Collaborate with other team members on study design, analysis plans, timelines, and prioritization of tasks.
Requirements
Required Qualifications
- Graduate degree or equivalent experience in Bioinformatics, Computational Biology, Genomics, Computer Science, Statistics, or a related field.
- Practical experience with next‑generation sequencing data analysis (such as WGS, WES, or RNA‑seq), including quality control, alignment, and variant or expression analysis.
- Proficiency in at least one scripting language commonly used in bioinformatics (e.g., Python or R), and familiarity with relevant scientific or data‑analysis libraries.
- Experience working in Unix/Linux environments, including shell scripting and command‑line tools.
- Familiarity with standard genomics file formats and commonly used open‑source tools for sequence data processing and variant analysis.
- Exposure to workflow or pipeline management tools (such as Nextflow, Snakemake, CWL, WDL, or comparable systems), and an understanding of reproducible analysis practices.
- Strong organizational skills, attention to detail, and the ability to manage multiple analysis tasks in parallel while meeting agreed timelines.
- Clear written and verbal communication skills, including the ability to describe analytical approaches and results to non‑specialists.
Preferred Qualifications
- Experience with clinical or population‑based genomic datasets in any disease area.
- Familiarity with high‑performance or distributed computing environments used for computational biology workloads.
- Experience building and maintaining ETL (extract–transform–load) workflows and working with relational or NoSQL databases.
- Background in statistics, statistical genetics, or related quantitative disciplines.
- Contributions to shared code bases or open‑source projects in bioinformatics or data analysis.
- Experience generating visualizations or dashboards for scientific data using tools such as R Shiny, Plotly, Dash, or similar frameworks.
Working Style
- Comfortable working in a collaborative environment with researchers, analysts, and software professionals.
- Able to estimate effort, communicate progress, and flag risks or issues early.
- Curious and willing to learn new analytical methods, tools, and technologies.
Similar Jobs
Information Technology • Mobile • Social Impact • Software
Manage and support software implementation projects for clients, including planning, requirements gathering, stakeholder communication, integrations, documentation, training, risk tracking, quality management, and deployment. Collaborate with technical leads, analysts, engineers, partners, and vendors in an agile environment. Use AI-assisted tools to improve project documentation and analysis while mentoring teams and delivering technology solutions that support global health and development.
Top Skills:
Ai-Assisted ToolsCommcareDatabasesEmrsExcelLarge Language ModelsSaaSSystems Integration
Fintech
Design, develop, test, deploy, troubleshoot, and maintain scalable web applications and APIs using React, TypeScript, Node.js, GraphQL, and cloud technologies. Collaborate with product, design, architecture, and engineering teams; contribute to software standards, automation, documentation, security, performance, and reliability. Mentor engineers, participate in code reviews, work within Agile processes, and complete a technical assessment during hiring.
Top Skills:
Apollo GraphqlCi/Cd PipelinesDockerGitGCPGraphQLJavaScriptJestNode.jsPostgresReactTypescript
Fintech
Delivers accurate, timely reporting and actionable analysis for internal and external stakeholders. Handles ad hoc requests, validates and reconciles data, performs quality checks, documents report logic, manages request queues and SLAs, and identifies automation opportunities. Uses SQL and Python for data extraction, transformation, validation, analysis, and recurring process automation. Maintains report templates and collaborates with data, visualization, product, operations, compliance, and account management teams while protecting sensitive data.
Top Skills:
BigQueryClaude CodeGitGithub CopilotLookerPower BIPythonSnowflakeSQLTableau
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

.png)
