
Job Description
Senior DevOps Engineer
📍 NYC (Hybrid) or US Remote | Competitive + Equity
About Avantos
Avantos is building the AI-native operating system for financial services — transforming fragmented data into a single, intelligent system that powers workflows, automation, and decision-making.
We’re a product-led, fast-moving team at the intersection of AI, fintech, and modern infrastructure.
The Role
We're seeking a Senior DevOps Engineer / Site Reliability Engineer to own and evolve our infrastructure, reliability, and deployment practices. You'll be responsible for building the foundational platform that enables our engineering teams to ship quickly and reliably while maintaining the security and compliance standards required in financial services.
What You’ll Do
Design, implement, and maintain our AWS cloud infrastructure using infrastructure-as-code principles with Terraform
Build and optimize CI/CD pipelines to enable rapid, safe deployments across multiple environments
Own observability strategy—implement comprehensive monitoring, logging, and alerting systems using Datadog and other tooling
Architect and manage containerized workloads on ECS Fargate and evaluate migration paths to Kubernetes
Establish and enforce security best practices, working closely with compliance teams on financial services requirements
Design and implement disaster recovery, backup, and business continuity strategies
Optimize system performance, cost efficiency, and resource utilization across AWS services
Collaborate with engineering teams to improve service reliability, reduce toil, and establish SLOs/SLIs
Participate in incident response and conduct thorough post-mortems to drive continuous improvement
Mentor engineers on DevOps practices, cloud architecture patterns, and operational excellence
What We’re Looking For
8+ years of experience in DevOps, SRE, or infrastructure engineering roles
Expert-level proficiency with AWS services including ECS Fargate, ALB, Cognito, S3, SQS, and related services
Deep hands-on experience with Terraform for managing complex, multi-account AWS environments
Strong scripting and automation skills in Python and/or Bash
Proven experience designing and implementing CI/CD pipelines (GitHub Actions, ArgoCD, or similar)
Solid understanding of containerization technologies (Docker) and orchestration platforms (Kubernetes/ECS)
Experience with observability and monitoring tools (Datadog, CloudWatch, or equivalent)
Deep knowledge of networking, security, and AWS best practices
Strong problem-solving abilities and experience troubleshooting complex distributed systems
Excellent communication skills and ability to work cross-functionally with engineering teams
Bonus
Have startup or early-stage experience
Experience with PostgreSQL performance tuning and RDS management
Why Join
Build the foundation layer of an AI-native product
High ownership + direct impact
The Fit
You’re a systems thinker with a passion for reliability and automation. You love tackling complex infrastructure challenges, thrive in fast-paced environments, and are motivated by the responsibility of building secure, scalable platforms from the ground up. You collaborate deeply, care about operational excellence, and are energized by turning ambiguity into robust solutions that help teams move faster and safer.
Optimize Your Resume for This Job
Get a match score and see exactly which keywords you're missing
Job Details
- Category
- Software
- Employment Type
- Full Time
- Location
- Remote (Hybrid)
- Posted
About Avantos.ai
AvantOS is a revolutionary AI platform that redefines client onboarding and servicing, making it a competitive advantage for financial institutions, delivering unmatched agility, scalability, and transparency.
More Roles at Avantos.ai





Similar Software Roles



Found this role interesting?