Senior Platform Operations and Network EngineerPosted today

The opportunity

Upwork Inc. ’s (Nasdaq: UPWK) family of companies connects businesses with global, AI-enabled talent across every contingent work type including freelance, fractional, and payrolled.

What you'll do

  • Operate and improve shared Cloud infrastructure across multiple providers,: accounts, and environments, with a focus on availability, security, scalability, and cost efficiency.

  • Manage the infrastructure lifecycle, including provisioning, configuration,: patching, upgrades, capacity planning, backup validation, recovery testing, and decommissioning.

  • Build and maintain reusable infrastructure using Terraform Enterprise or: Terraform Cloud. Improve module quality, testing, versioning, policy enforcement, and drift management.

  • Use configuration management tools such as Chef or Ansible to maintain: consistent systems and reduce reliance on manual changes.

  • Improve infrastructure reliability through actionable monitoring, alerting,: service health indicators, operational runbooks, and tested recovery procedures.

  • Participate in the production on-call rotation, lead complex infrastructure: incident response, and convert incident findings into preventive controls and automation.

What they're looking for

  • Own and improve network infrastructure across cloud and corporate: environments, including VPCs, routing, DNS, load balancing, private connectivity, segmentation, firewalls, VPNs, switching, office networks, and wireless access.
  • Ensure reliable and secure connectivity between employees, offices, cloud: environments, platform workloads, and approved external services.
  • Establish consistent network architecture, observability, change management,: documentation, and recovery practices in partnership with Corporate Technology and Security.
  • Operate EKS infrastructure in depth, including cluster and node lifecycle,: autoscaling, networking, ingress, DNS, storage integration, upgrades, capacity, security, observability, and recovery. Help converge traditional cloud infrastructure and Kubernetes runtime infrastructure through common automation, lifecycle controls, reliability standards, and ownership.