Sr. DevOps Engineer

Sr. DevOps Engineer at TaxGPT — San Francisco, CA, US

  • Company: TaxGPT
  • Location: San Francisco, CA, US
  • Employment type: FULL_TIME
  • Salary: USD 140000–160000 / year
  • Posted: 2026-07-17

About this role

About Us

TaxGPT is revolutionizing the tax and accounting space with AI-driven solutions tailored for accountants, tax professionals, and SMBs. We're building an AI co-pilot to transform tax workflows, drive efficiency, and simplify compliance. Recently named one of Business Insider’s 30 Early-Stage Startups Most Likely to Become Tech’s Next Unicorns, we'd love for you to join our growing team!

Location: US Remote

Pay Range: $140k to $160k + Equity + Variable Comp up to 20% of base salary

Benefits offered: Medical, dental, vision, 401k + 3% match, life insurance

About the Role

We are looking for a Senior DevOps Engineer with 7+ years of experience to help build, scale, and secure our infrastructure, deployment systems, and developer operations. This person will play a critical role in improving reliability, performance, security, and engineering velocity across the company.

You will work closely with software engineers, product teams, and technical leadership to design and maintain systems that support fast development, stable production environments, and long-term scalability.

This is a senior individual contributor role for someone who can operate with high autonomy, own critical infrastructure, make strong architectural decisions, and raise the technical bar across the engineering organization.

What You’ll Do

Infrastructure and Platform Ownership

Own the design, implementation, and long-term health of cloud infrastructure and internal platform systems

Build and maintain scalable, secure, and reliable infrastructure across development, staging, and production environments

Improve infrastructure automation, environment consistency, and operational resilience

Manage networking, compute, storage, observability, and access controls across core systems

CI/CD and Developer Productivity

Design, improve, and maintain CI/CD pipelines for fast, safe, and repeatable deployments

Build tooling and workflows that improve developer experience and reduce operational friction

Standardize release processes, deployment practices, and rollback strategies

Identify bottlenecks in development and deployment workflows and drive improvements

Reliability, Monitoring, and Incident Response

Improve system reliability, availability, and performance through strong operational practices

Build and maintain monitoring, alerting, logging, and incident response systems

Lead root cause analysis and drive permanent fixes for recurring operational issues

Establish and improve standards around uptime, recovery, and production readiness

Security and Compliance

Implement infrastructure security best practices across environments and workflows

Strengthen access controls, secrets management, auditability, and system hardening

Partner with engineering leadership to reduce operational and security risk

Support compliance, backup, disaster recovery, and resilience initiatives where needed

Technical Leadership

Lead architectural decisions related to infrastructure, deployment systems, and platform reliability

Partner with engineering leaders to shape long-term infrastructure strategy

Mentor engineers on infrastructure, deployment, observability, and operational best practices

Raise the team’s standards through documentation, design reviews, code reviews, and process improvements.

What We’re Looking For

Required Qualifications

7+ years of experience in DevOps, Site Reliability Engineering, Platform Engineering, or Infrastructure Engineering

Strong experience designing and managing production infrastructure in cloud environments such as AWS, GCP, or Azure

Deep experience with CI/CD systems, infrastructure automation, and deployment pipelines

Strong knowledge of containers and orchestration, including tools such as Docker and Kubernetes

Experience with Infrastructure as Code tools such as Terraform, Pulumi, or CloudFormation

Strong experience with monitoring, logging, and observability tools

Experience improving reliability, security, and scalability in production systems

Strong scripting or coding ability in languages such as Python, Bash, or Go

Strong understanding of networking, system design, access control, and cloud security fundamentals

Preferred Qualifications

Experience supporting fast-moving startup engineering teams

Experience building internal developer platforms or self-service infrastructure tooling

Familiarity with modern security and compliance practices

Experience with incident management, postmortems, and operational maturity improvements

Experience working closely with backend and application engineering teams to improve system design and delivery quality

How We Define Success in This Role

A strong Senior DevOps Engineer in this role:

operates with high autonomy and needs little day-to-day direction

owns critical infrastructure and improves its long-term health

makes strong architectural and operational decisions

improves developer velocity without compromising reliability or security

mentors others and raises the engineering bar across the team

brings structure and clarity to complex infrastructure and operational challenges

Key Responsibilities:

Autonomy

Works independently on complex infrastructure and reliability problems

Identifies risks and improvements before they become urgent issues

Translates broad engineering goals into clear technical plans

System Ownership

Owns critical infrastructure, deployment systems, and platform reliability

Takes responsibility for scalability, resilience, maintainability, and operational health

Drives long-term fixes, not just short-term patches

Mentoring and Collaboration

Guides engineers on infrastructure and operational best practices

Improves team effectiveness through documentation, reviews, and technical support

Collaborates closely with engineering, product, and leadership teams

Architectural Decision-Making

Makes sound decisions on cloud architecture, deployment strategy, observability, and security

Evaluates tradeoffs carefully across cost, speed, reliability, and complexity

Builds systems that scale with company needs

Why Join Us

You’ll have the opportunity to shape the foundation of our engineering platform and help define how infrastructure, reliability, and developer operations scale as the company grows. This is a high-impact role for someone who enjoys ownership, technical depth, and building systems that make the entire engineering organization stronger.

Apply for this position