Senior Central Cloud Infrastructure Engineer
metropolis · Seattle, Washington, United States
Experience: 5+ years
## Who we are
The real world is the next frontier, and at Metropolis, we are creating the artificial intelligence to make it responsive. We are pioneering the Recognition Economy — a future where mundane repetition disappears and being known unlocks access, comfort and belonging everywhere you go. From transforming parking into a seamless drive-in, drive-out experience for millions of Members to expanding our intelligence layer across retail and hospitality, we are building a world that feels instinctive and magical. The future isn’t coming; it’s here, and we need builders, innovators and problem solvers to help us create it.
## Who you are
Metropolis is seeking a Senior Central Cloud Infrastructure Engineer to serve as a principal contributor to our central cloud platform infrastructure. In this role, you will architect, scale, and maintain highly resilient environments across multiple products while driving infrastructure modernization and legacy system integrations. The ideal candidate combines deep technical mastery of modern cloud-native systems with a strong collaborative mindset to interface across engineering teams, troubleshoot complex environments, and provide technical guidance to scale the team's overall competency.
## What you'll do
• Architect, implement, and maintain scalable, reliable, and secure cloud infrastructure environments across multiple company products
• Design and manage infrastructure modernization initiatives (CI/CD, IaC, K8s, AI, etc) and lead the complex integration of legacy or established systems into modern, cloud-native environments
• Craft comprehensive disaster recovery strategies and engineer production-grade technical solutions to safeguard system uptime
• Serve as a technical focal point for troubleshooting, system deep-dives, and resolving infrastructure bottlenecks across both Linux and Windows operating systems
• Ensure global cloud environments continuously meet and exceed stringent security guidelines and technical compliance frameworks, specifically PCI and SOC2
• Lead incident response and root-cause analysis, and champion blameless postmortems and error-budget discipline across the team
• Provide technical leadership and guidance to junior and mid-level engineers, helping to scale the infrastructure team's overall technical competency
## What we're looking for
• Hold a BS degree in Computer Science (or a relevant technical/engineering discipline) and 5+ years of hands-on experience, or equivalent professional experience (8+ years) working in high-growth cloud environments
• Demonstrate deep production-level technical experience using AWS, Terraform, Atmos, and Python (or equivalent automation language - Bash/Go)
• Demonstrate hands-on expertise with a modern observability stack (Datadog or equivalent: APM/distributed tracing, logs, metrics, RUM), including defining SLOs/SLIs and managing monitors-as-code
• Leverage production experience with service mesh and L7 proxies (Envoy/Istio) and Kubernetes ingress in high-throughput request paths
• Utilize advanced operational knowledge of Kubernetes/EKS, Helm, and overall container orchestration
• Apply strong secrets-management and rotation practices (Infisical, Vault, or AWS Secrets Manager) in a PCI/SOC2 context
• Showcase hands-on experience working with managed streaming data infrastructure via AWS MSK (Kafka) alongside various SQL and NoSQL AWS managed database solutions, including Amazon Aurora
• Maintain a strong understanding of cloud networking architecture, security protocols, and firewalls
• Exhibit expert competency managing, configuring, and troubleshooting Linux (and Windows where applicable) server environments
• Use excellent written and verbal communication skills to translate complex infrastructure concepts into clear, actionable technical documentation
## While not required, these are a plus:
• Possess previous experience working inside innovative, high-growth tech start-ups or fast-paced corporate engineering environments
4 Days in Office: Metropolis values in-person collaboration to drive innovation, strengthen culture, and enhance the Member experience. Our corporate team members hold to our office-first model, which requires employees to be on-site at least four days a week, fostering organic interactions that spark creativity and connection
When you join Metropolis, you'll join a team of world-class product leaders and engineers, building an ecosystem of technologies at the intersection of parking, mobility, and real estate. Our goal is to build an inclusive culture where everyone has a voice and the best idea wins. You will play a key role in building and maintaining this culture as our organization grows. The anticipated base salary for this position is
The real world is the next frontier, and at Metropolis, we are creating the artificial intelligence to make it responsive. We are pioneering the Recognition Economy — a future where mundane repetition disappears and being known unlocks access, comfort and belonging everywhere you go. From transforming parking into a seamless drive-in, drive-out experience for millions of Members to expanding our intelligence layer across retail and hospitality, we are building a world that feels instinctive and magical. The future isn’t coming; it’s here, and we need builders, innovators and problem solvers to help us create it.
## Who you are
Metropolis is seeking a Senior Central Cloud Infrastructure Engineer to serve as a principal contributor to our central cloud platform infrastructure. In this role, you will architect, scale, and maintain highly resilient environments across multiple products while driving infrastructure modernization and legacy system integrations. The ideal candidate combines deep technical mastery of modern cloud-native systems with a strong collaborative mindset to interface across engineering teams, troubleshoot complex environments, and provide technical guidance to scale the team's overall competency.
## What you'll do
• Architect, implement, and maintain scalable, reliable, and secure cloud infrastructure environments across multiple company products
• Design and manage infrastructure modernization initiatives (CI/CD, IaC, K8s, AI, etc) and lead the complex integration of legacy or established systems into modern, cloud-native environments
• Craft comprehensive disaster recovery strategies and engineer production-grade technical solutions to safeguard system uptime
• Serve as a technical focal point for troubleshooting, system deep-dives, and resolving infrastructure bottlenecks across both Linux and Windows operating systems
• Ensure global cloud environments continuously meet and exceed stringent security guidelines and technical compliance frameworks, specifically PCI and SOC2
• Lead incident response and root-cause analysis, and champion blameless postmortems and error-budget discipline across the team
• Provide technical leadership and guidance to junior and mid-level engineers, helping to scale the infrastructure team's overall technical competency
## What we're looking for
• Hold a BS degree in Computer Science (or a relevant technical/engineering discipline) and 5+ years of hands-on experience, or equivalent professional experience (8+ years) working in high-growth cloud environments
• Demonstrate deep production-level technical experience using AWS, Terraform, Atmos, and Python (or equivalent automation language - Bash/Go)
• Demonstrate hands-on expertise with a modern observability stack (Datadog or equivalent: APM/distributed tracing, logs, metrics, RUM), including defining SLOs/SLIs and managing monitors-as-code
• Leverage production experience with service mesh and L7 proxies (Envoy/Istio) and Kubernetes ingress in high-throughput request paths
• Utilize advanced operational knowledge of Kubernetes/EKS, Helm, and overall container orchestration
• Apply strong secrets-management and rotation practices (Infisical, Vault, or AWS Secrets Manager) in a PCI/SOC2 context
• Showcase hands-on experience working with managed streaming data infrastructure via AWS MSK (Kafka) alongside various SQL and NoSQL AWS managed database solutions, including Amazon Aurora
• Maintain a strong understanding of cloud networking architecture, security protocols, and firewalls
• Exhibit expert competency managing, configuring, and troubleshooting Linux (and Windows where applicable) server environments
• Use excellent written and verbal communication skills to translate complex infrastructure concepts into clear, actionable technical documentation
## While not required, these are a plus:
• Possess previous experience working inside innovative, high-growth tech start-ups or fast-paced corporate engineering environments
4 Days in Office: Metropolis values in-person collaboration to drive innovation, strengthen culture, and enhance the Member experience. Our corporate team members hold to our office-first model, which requires employees to be on-site at least four days a week, fostering organic interactions that spark creativity and connection
When you join Metropolis, you'll join a team of world-class product leaders and engineers, building an ecosystem of technologies at the intersection of parking, mobility, and real estate. Our goal is to build an inclusive culture where everyone has a voice and the best idea wins. You will play a key role in building and maintaining this culture as our organization grows. The anticipated base salary for this position is
70,000.00 USD to
25,000.00 USD annually. The actual base salary offered is determined by a number of variables, including, as appropriate, the applicant's qualifications for the position, years of relevant experience, distinctive skills, level of education attained, certifications or other professional licenses held, and the location of residence and/or place of employment. Base salary is one component of Metropolis’s total compensation package, which may also include access to or eligibility for healthcare benefits, a 401(k) plan, short-term and long-term disability coverage, basic life insurance, a lucrative stock option plan, bonus plans and more. #LI-AR1 #LI-Onsite
Metropolis may utilize an automated employment decision tool (AEDT) to assess or evaluate your candidacy for employment or promotion. AEDTs are used to assist in assessing a candidate’s application relative to the required job qualifications and responsibilities listed in the job posting.
As part of this process, Metropolis retains data relevant to your candidacy, including personal information, for a period that is reasonably necessary for the use of the tool. If you are hired for the position, your data may become part of your employee records.
Metropolis Technologies is an equal opportunity employer. We make all hiring decisions based on merit, qualifications, and business needs, without regard to race, color, religion, sex (including gender identity, sexual orientation, or pregnancy), national origin, disability, veteran status, or any other protected characteristic under federal, state, or local law.
Careeroza — One-stop Zone for Aspirants
Study material, Careeroza mentorship, tech jobs, and career guidance on careeroza.com.
Public study materials
- Express.js — Web APIs & middleware (expressjs)
- Application setup · basic
- Routing deep dive · basic
- 1. MVC / Layered Architecture · medium
- Validation · medium
- File uploads · medium
- Sessions & auth (stateful) · medium
- Passport & strategies · medium
- Templating & SSR · medium
- WebSockets & SSE · medium
- Security middleware · advance
- Reverse proxies & trust · advance
- Performance · advance
- API design & versioning · advance
- Testing with Supertest · advance
- GraphQL & tRPC (overview) · advance
- Deployment checklist · advance
- Middlewares · basic
- Request & Response · basic
- AWS Crash Course (AWS)
- What is Cloud ? · basic
- What is AWS ? · basic
- If not cloud ? · basic
- Cloud Computing · basic
- AWS Pricing · basic
- AWS Shared Responsibility Model · basic
- AWS Management Console · basic
- AWS SDKs · basic
- AWS IAM · medium
- Users, Groups, Roles · medium
- Policies · medium
- AWS Organizations · medium
- AWS Cognito · medium
- AWS Directory Service · medium
- AWS KMS (Key Management Service) · medium
- AWS Secrets Manager · medium
- AWS Shield · medium
- AWS WAF · medium
- AWS Inspector · medium
- AWS GuardDuty · medium
- EC2 · advance
- Launching EC2 Instances · advance
- EBS Volumes · advance
- Security Groups · advance
- Key Pairs · advance
- Elastic IP · advance
- User Data Scripts · advance
- Auto Scaling · advance
- Load Balancers · advance
- ALB · advance
- NLB · advance
- Serverless Compute ,AWS Lambda, Lambda Layers · advance
- Event-Driven Architecture · advance
- ECS · advance
- EKS · advance
- Django (Django)
- What is Django · basic
- Installing Django · basic
- Features of Django · basic
- MVT Architecture · basic
- Django vs Flask · basic
- Creating Project & Creating App · basic
- Django Project Structure · basic
- URL Routing · basic
- Views · basic
- Templates · basic
- Static & Media Files · medium
- Models · medium
- ORM (Object Relational Mapping) · medium
- Model Relationships · medium
- Migrations · medium
- Django Admin · medium
- Forms · medium
- Authentication · medium
- Authorization · medium
- Middleware · medium
- Signals · medium
- Class Based Views Deep Dive · advance
- Generic Views · advance
- File Handling · advance
- Django REST Framework (DRF) · advance
- Advanced ORM · advance
- Caching · advance
- Asynchronous Django · advance
- Background Tasks · advance
- Interview Questions · interview-questions
- System Designing (System Designing)
- Day-1 : What is system Designing ? · basic
- Day-2 : Vertical vs. Horizontal Scaling · basic
- Day-3:How to do vertical scaling ? · basic
- Day4:How to do horizaontal scaling ? · basic
- Day:5TCP vs UDP · basic
- Day6:IP & DNS · basic
- Day7:Client-Server Model · basic
- Day8:HTTP & HTTPS · basic
- Databases (SQL vs NoSQL) · medium
- Caching · medium
- Day9:Latency & Throughput · basic
- Load Balancing · medium
- Indexes & Query Optimization · medium
- CDN · medium
- Proxies · medium
- Message Queues · medium
- Horizontal vs Vertical Scaling · medium
- Database Replication · advance
- Database Sharding · advance
- Consistent Hashing · advance
- CAP Theorem · advance
- Rate Limiting · advance
- Service Discovery · advance
- Event-Driven Architecture · advance
- API Gateway · advance
- Distributed Consensus · expert
- Microservices · expert
- Observability · expert
- Idempotency · expert
- PACELC Theorem · expert
- Two-Phase Commit · expert
- Back-of-Envelope Estimation · expert
- Designing for Failure · expert
- Node.js — Server-side JavaScript (nodejs)
- Getting started · basic
- JavaScript on the server · basic
- CommonJS modules · basic
- ES modules (ESM) · basic
- npm & package management · basic
- Asynchronous JavaScript in Node · basic
- The event loop · basic
- Essential core utilities · basic
- process & configuration · basic
- File system basics · basic
- HTTP & HTTPS servers · medium
- Streams · medium
- Events & EventEmitter · medium
- Advanced filesystem · medium
- crypto · medium
- Compression & encoding · medium
- Child processes · medium
- net, dgram & DNS · medium
- readline, timers & scheduling · medium
- Testing & diagnostics (intro) · medium
- Worker threads · advance
- cluster & multi-process scaling · advance
- Performance & tuning · advance
- Debugging & observability · advance
- Security hardening · advance
- Native addons & N-API · advance
- Architecture patterns · advance
- Graceful shutdown · advance
- 100 Questions · interview questions
- Python (Python)
- Python Fundamentals · basic
- Control Flow · basic
- Strings · basic
- Collections / Data Structures · basic
- Functions · basic
- Modules and Packages · basic
- File Handling · basic
- Exception Handling · basic
- Object-Oriented Programming (OOP · medium
- Advanced Python Concepts · advance
- Functional Programming · advance
- Multithreading & Multiprocessing · advance
- Async Programming · advance
- JavaScript (JavaScript)
- JS Introduction · basic
- Variables & Data Types · basic
- Operators · basic
- Control Flow · basic
- Functions · basic
- Scope & Execution · basic
- Closures · basic
- Objects · basic
- Arrays · basic
- Strings · basic
- DOM Manipulation · basic
- Browser APIs · medium
- Asynchronous JavaScript · medium
- Fetch & APIs · medium
- ES6+ Features · medium
- OOP in JavaScript · medium
- Prototype & Inheritance · advance
- Advanced Functions · advance
- Memory Management · advance
- Error Handling · advance
- Modules · advance
- Advanced Async Concepts · advance
- Functional Programming · advance
- JavaScript Internals · advance
- Performance Optimization · advance
- MongoDB — Documents & data modeling (MongoDB)
- Introduction · basic
- Shell, Compass & tools · basic
- Databases & collections · basic
- CRUD operations · basic
- Indexes deep dive · medium
- Explain plans & performance · medium
- Aggregation framework · medium
- Schema design patterns · medium
- Mongoose basics · medium
- Mongoose advanced · medium
- Drivers & connection · medium
- Operators for updates & arrays · medium
- Replication & read preferences · advance
- Write concern & read concern · advance
- Multi-document transactions · advance
- Change streams · advance
- Sharding (overview) · advance
- Atlas Search & full-text · advance
- GridFS & large files · advance
- Backup, restore & ops · advance
- SQL (SQL)
- SQL Fundamentals · basic
- Database Operations · basic
- Table Operations · basic
- CRUD Operations · basic
- Filtering & Operators · basic
- SQL Functions · basic
- GROUPING Data · basic
- Joins · medium
- Constraints · basic
- Subqueries · medium
- Set Operators · medium
- Views · medium
- Indexes · medium
- Normalization · advance
- Transactions · advance
- Stored Procedures & Functions · advance
- Triggers · advance
- Advanced SQL · advance
- Query Optimization · advance
- Database Design · advance
- SQL Security · advance
- Backup & Recovery · advance
- Questions · interview questions
- System Architecture (System Architecture)
- Fundamentals of System Architecture · basic
- Distributed System Basics · basic
- System Reliability Concepts · basic
- Scaling Concepts · basic
- Networking Basics · basic
- Web Communication · basic
- API Communication · basic
- Proxy & Delivery Systems · basic
- Web Architecture Basics · basic
- Rendering Architectures · basic
- Frontend Advanced Concepts · basic
- Message Queue Basics · medium
- What is Load Balancer · medium
- Load Balancing Algorithms · medium
- API Design Basics · medium
- API Protection · medium
- Authentication Basics · medium
- Security Tokens · medium
- Security Threats · medium
- Encryption & Security · medium
- SQL Database Basics · medium
- SQL Scaling Concepts · medium
- NoSQL Databases · medium
- Database Optimization · medium
- Replication Strategies · medium
- Caching Basics · medium
- Cache Storage Systems · medium
- Cache Strategies · medium
- Event-Driven Systems · advance
- Queue Reliability · advance
- Microservices Basics · advance
- Microservice Communication · advance
- Distributed Transactions · advance
- DevOps Basics · advance
- Automation Tools · advance
- Deployment Strategies · advance
- Monitoring Basics · advance
- Monitoring Tools · advance
- Distributed System Concepts · advance
- Distributed Algorithms · advance
- Angular (Angular)
- Angular Fundamentals · basic
- Project Structure · basic
- Components & Templates · basic
- Data Binding · basic
- Directives · basic
- Pipes · basic
- Component Communication · basic
- Lifecycle Hooks · basic
- Routing Basics · basic
- Routing · basic
- API Calls · basic
- Forms · medium
- Routing · medium
- Services & Dependency Injection · medium
- RxJS & Observables · medium
- Authentication & Security · medium
- Component Interaction · medium
- State Management Basics · medium
- Error Handiling · medium
- Perfomance Basic · medium
- Real World Features · medium
- Advanced Angular Architecture · advance
- Change Detection · advance
- Advanced RxJS · advance
- State Management · advance
- Dynamic Rendering · advance
- Perfomance Optimization · advance
- Modern Angular · advance
- STAR (Situation, Task, Action, and Result) (Situation Based Questions)
Metropolis may utilize an automated employment decision tool (AEDT) to assess or evaluate your candidacy for employment or promotion. AEDTs are used to assist in assessing a candidate’s application relative to the required job qualifications and responsibilities listed in the job posting.
As part of this process, Metropolis retains data relevant to your candidacy, including personal information, for a period that is reasonably necessary for the use of the tool. If you are hired for the position, your data may become part of your employee records.
Metropolis Technologies is an equal opportunity employer. We make all hiring decisions based on merit, qualifications, and business needs, without regard to race, color, religion, sex (including gender identity, sexual orientation, or pregnancy), national origin, disability, veteran status, or any other protected characteristic under federal, state, or local law.