Senior DevOps / Platform Engineer
New
T
TruelogicFintech Wealth Management
United StatesFull-TimeSenior
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 6+ years
- Required Skills
- AWSPostgreSQLPythonKubernetesCI/CDTerraform
Requirements
- 6+ years of hands-on experience in DevOps, Site Reliability Engineering (SRE), Platform Engineering, or Infrastructure Engineering.
- Deep, practical experience managing and operating AWS environments.
- Strong production-level experience designing and scaling Kubernetes clusters.
- Proven track record of scaling high-availability production systems.
- Strong understanding of PostgreSQL administration, performance tuning, and scaling.
- Extensive infrastructure-as-code experience utilizing Terraform or similar technologies.
- Robust experience building and managing CI/CD pipelines and modern observability tooling.
- Deep understanding of cloud, network, and application security, with a proven ability to manage sensitive or regulated data.
- Ability to operate independently and take ownership in a fast-moving, startup environment.
- Background in fintech, wealth management, banking, lending, payments, or other transaction-oriented platforms is highly desirable.
- Prior experience designing infrastructure specifically for financial-data platforms or working within SOC 2 compliant environments is preferred.
- Proficiency or hands-on experience with Python is a plus.
Responsibilities
- Own and evolve cloud infrastructure running on AWS.
- Design, operate, and scale Kubernetes-based production environments.
- Improve system reliability, availability, and performance as transaction volume and financial data workloads rapidly grow.
- Build and maintain infrastructure-as-code, CI/CD pipelines, observability, and deployment tooling.
- Partner closely with backend and data engineers on PostgreSQL performance, database reliability, backups, and scaling.
- Strengthen infrastructure security and access controls to safeguard highly sensitive financial data.
- Identify architectural bottlenecks and proactively improve system scalability and resilience.
- Help define infrastructure and operational standards as the engineering organization expands.
- Participate in incident response, root-cause analysis, and continuous reliability improvements.
View Full Description & ApplyYou'll be redirected to the employer's site