Responsibilities
- Develop, design, create, modify, and/or test software services to ensure optimal performance and capacity for growth.
- Own services, databases, and frameworks that enable use of public cloud platforms by Meta teams, and ensure services run without incident.
- Write and review code, develop documentation and capacity plans, and debug the problems in real time in highly complex software systems.
- Serve as an escalation contact for service incidents.
- Work on problems of moderate scope where analysis of situations or data requires a review of a variety of factors.
- Exercise judgment within defined procedures and practices to determine appropriate action.
Minimum Qualifications
- Bachelor's degree (or foreign degree equivalent) in Computer Science, Engineering, Information Systems, Analytics, Mathematics, Physics, Applied Sciences, or a related field and three years of work experience in job offered or computer-related occupation. Requires three years of experience in the following:
- UNIX or Linux operating system fundamentals
- TCP/IP network fundamentals
- Coding in at least one of the following higher-level programming languages: PHP, Python, C++, or Java
- Software frameworks and APIs
- Performing 'guerilla capacity planning' for internet service architectures
- Internet service architectures (such as load balancing, LAMP, or CDN’s)
- Configuring and maintaining applications using at least one of the following: web servers, load balancers, relational databases, storage systems, or messaging systems
- Relational Databases including MySQL
- Network protocols including at least one of the following: NFS, DHCP, NTP, SSH, DNS, or SNMP
- Maintaining web-based applications using at least one of the following: Apache, Memcached, or Squid
- Storage Systems including NFS
- Network Management tools like DHCP, NTP, SSH, DNS, or SNMP
- Diagnosing and troubleshooting issues ranging from low-level hardware issues to large scale failures within datacenter clusters
- Experience utilizing high performance query engines (Presto or Spark) for big data
- Cloud-based services such as AWS S3, EC2
- Container orchestration systems such as Kubernetes and Docker
- Infrastructure as code platforms including Helm and Terraform
- Continuous integration and deployment systems and
- Observability frameworks (e.g., Prometheus, fluentbit, Grafana)
$174,978/year to $209,000/year + bonus + equity + benefits
Learn more about this Employer on their Career Site
