Platform Engineer
Key Responsibilities
• Lead the maintenance and optimization of our AWS data infrastructure, focusing on proactive improvements and regular maintenance
• Manage database platforms including Aurora PostgreSQL and MySQL, performing upgrades, rehydrations, and performance tuning
• Implement and maintain infrastructure as code using Terraform for consistent and repeatable deployments
• Configure and manage AWS data services including DMS and Glue
• Design and implement monitoring and alerting systems to ensure platform health and performance
• Develop automation scripts and tools to improve operational efficiency
• Spearhead cloud cost optimization initiatives and implement FinOps best practices for our data platform
• Collaborate with internal teams including data engineers, analysts, and scientists to understand their platform needs
• Participate in on-call rotation to support critical platform issues when they arise
• Document platform architecture, operations procedures, and best practices
• Provide technical guidance and mentorship to junior team members
Required Qualifications
• 5+ years of experience managing AWS infrastructure with a focus on data services
• Strong hands-on experience with relational databases, particularly PostgreSQL and MySQL
• Experience with infrastructure as code, specifically Terraform
• Knowledge of AWS data services including DMS, Glue, and RDS/Aurora
• Solid understanding of database administration, including performance tuning and troubleshooting
• Experience with cloud cost management and FinOps practices
• Proficiency with Linux/Unix systems and shell scripting
• Experience with monitoring tools and implementing observability solutions
• Understanding of network security principles and best practices in cloud environments
Preferred Qualifications
• AWS certifications (Data Engineer, Solutions Architect)
• Experience with database replication and high availability configurations
• Knowledge of data modeling and schema design
• Experience with containerization technologies (Docker, Kubernetes)
• Understanding of disaster recovery and business continuity planning
• Familiarity with CI/CD pipelines and DevOps practices
• Knowledge of Python or other programming languages for automation