OverviewThis is a hybrid role requiring 3 days a week You must be work authorized in the United States without the need for employer sponsorship. Must have Ad Tech / MarTech industry experience, specifically in e-commerce, travel, and finance.As a Software Engineer 2 on the Engineering Experience (EngExp) platform team, you help run and improve the platform that powers CJ's production systems across multiple AWS regions. "Platform" here is broad - it is the Kubernetes clusters, but also the observability stack every squad depends on, the CI/CD and artifact infrastructure their builds run through, the AWS networking that connects them, and the secrets and access systems that gate them. EngExp owns all of it. This is a hands-on, operationally-focused role: you'll take on well-scoped work across these systems, ship it end to end, and build real depth in them with senior engineers alongside you. We are looking for someone genuinely curious about how these systems work under the hood and eager to grow into owning them.
ResponsibilitiesThe Systems You Work On:EngExp owns the systems below. You'll contribute to operating them and build depth in them over time, starting with guidance from senior engineers:
- Observability & monitoring - Prometheus, Alertmanager, Grafana, and OpenTelemetry across production regions. You'll add and tune metrics, dashboards, and alert rules, and learn how Prometheus and Alertmanager actually behave at scale - cardinality, recording rules, and alert routing. This is real depth to grow into, not just dashboard-building.
- Kubernetes & cloud infrastructure - multi-region EKS clusters: routine maintenance, node group changes, controller upgrades, and add-on configuration, with senior engineers alongside.
- AWS networking - VPCs, subnets and CIDR management, security groups, and Route53. You'll help keep prod networking healthy and learn the multi-region topology.
- CI/CD & artifact management - GitLab CI/CD pipelines, GitOps delivery through ArgoCD, and the Nexus artifact repository.
- Access & identity - IAM roles and service accounts for apps in the clusters, Vault-managed secrets, and fulfilling access requests - then helping turn the repetitive ones into self-service.
- Cost observability - OpenCost and cleanup work (e.g. orphaned EBS volumes) so cost is attributable rather than shared overhead.
What You'll Do:- Pick up well-scoped platform work across the systems above and ship it end to end
- Build and maintain GitLab CI/CD pipelines and GitOps delivery through ArgoCD
- Write and review infrastructure-as-code with Terraform for AWS resources
- Add and tune observability - Prometheus metrics, Grafana dashboards, Alertmanager rules
- Help fulfill and then automate recurring requests (ingress, DNS, service accounts, IAM roles)
- Help investigate platform incidents and document what you learn
- Learn to recognize common failure modes (IP exhaustion, resource limits, reconciliation lag) and escalate them early
Technologies We Use:- Kubernetes / EKS (multi-cluster, multi-region), Karpenter, cert-manager, external-dns
- Prometheus, Alertmanager, Grafana, OpenTelemetry (and long-term storage / sharding for Prometheus)
- AWS networking (VPC, VPC peering, Transit Gateway, Route53, NAT Gateway, security groups, subnet/CIDR design across accounts and regions)
- Terraform, AWS (IAM, EKS, S3, EBS)
- ArgoCD, GitLab CI/CD, Nexus (artifact registry), Docker, container image build pipelines
- Vault, OpenCost
Engineering Practices We Employ:- Agile software development
- Infrastructure as Code (IaC)
- Pair programming
- Test-Driven Development (TDD)
- Continuous Delivery
QualificationsWhat We Look For:- 1-2 years of experience in software or infrastructure engineering
- Bachelor's degree or equivalent experience
- Comfortable in the terminal and reading YAML, Terraform, or similar declarative config
- Some exposure to cloud (AWS or equivalent) and containers - production Kubernetes experience is a plus, not a requirement
- Genuine interest in operating real systems - especially observability - and growing deep in them
- Willing to learn how systems fail and to ask questions when something looks off
- Effective communication; enjoys pair programming and code review
Nice to Have:- Hands-on exposure to Kubernetes, Terraform, or GitLab/GitHub CI/CD
- Any experience with Prometheus/Grafana or another metrics stack
- Scripting in Python, Bash, or Go
What Success Looks Like:- You reliably deliver well-scoped platform work end to end with decreasing oversight
- You're building genuine depth in at least one system we own, not just familiarity with the tools
- Other engineers find the workflows you touch easier and more predictable to use
Additional informationThis is a hybrid role requiring 3 days a week in office.
Compensation Range: USD $87,210.00 - USD $131,230.00/Annually. This is the pay range the Company believes it will pay for this position at the time of this posting. Consistent with applicable law, compensation will be determined based on the skills, qualifications, and experience of the applicant along with the requirements of the position, and the Company reserves the right to modify this pay range at any time. Temporary roles may be eligible to participate in our freelancer/temporary employee medical plan through a third-party benefits administration system once certain criteria have been met. Temporary roles may also qualify for participation in our 401(k) plan after eligibility criteria have been met. For regular roles, the Company will offer medical coverage, dental, vision, disability, 401k, and paid time off. The Company anticipates the application deadline for this job posting will be 8/21/2026.