Skip to main content
A

Sr. Engineer, AWS Platform

Ayar Labs

Location

San Jose, CA

Salary

Not specified

Type

fulltime

Posted

Today

via linkedin

Job Description

Senior Engineer, AWS Platform

Location: San Jose, CA (Headquarters) - Onsite Required

Ayar Labs is shattering AI data bottlenecks by moving data at the speed of light. As pioneers of co-packaged optics (CPO), we are using light instead of electricity to move data faster, further, and with a fraction of the energy needed to fuel the explosive growth of AI models.

Backed by industry giants like NVIDIA, AMD and Intel and manufactured in partnership with the world’s leading semiconductor ecosystem, Ayar Labs’ co-packaged optics solution is key to unleashing next-generation AI scale-up architectures.

Ayar Labs is looking for an experienced AWS engineer to build and run the computing environment our engineers use to design chips. Most of that work will run on AWS, with some systems remaining on local hardware. You will help move workloads to the cloud and make sure the platform can support the demands of chip development.

You will take projects from an initial problem through implementation and into production. Working with engineering and IT, you will make architecture decisions, resolve problems, and improve how the platform operates. The role includes production support and on-call responsibilities, with support from our existing systems administrators.

Key Responsibilites

  • Build and operate an AWS platform that gives engineers dependable access to the compute and data they need. Make practical decisions about performance, security, cost, and how cloud and local systems work together.
  • Lead workload migrations from planning through production. Understand what each workload needs, test the approach, and manage the transition without disrupting engineering work.
  • Investigate problems that slow down or interrupt design work. Follow issues across the cloud platform and connected systems, identify the cause, and put lasting fixes in place.
  • Automate provisioning and routine operations so the environment is repeatable, changes can be tested, and the platform is straightforward to maintain.
  • Work with engineers to anticipate demand ahead of major design milestones. Ensure capacity, data access, and software licensing can support planned runs, and explain constraints early.
  • Protect design data and keep the platform recoverable. Establish appropriate access, monitoring, and recovery practices, and take responsibility for resolving production incidents.
  • Measure whether changes improve engineering turnaround time, reliability, and cost. Document decisions and operating procedures so others can support and extend the platform.

Basic Requirements

  • 7\+ years building and operating production infrastructure, including at least 3 years of hands-on responsibility for AWS environments.
  • A track record of taking a compute-intensive workload from design through deployment or migration on AWS and supporting it in production. You should be able to explain what you implemented, the tradeoffs you made, and the results.
  • The ability to diagnose difficult problems in Linux-based environments and across cloud, network, and storage boundaries. You have resolved production incidents and made changes that prevented them from recurring.
  • Experience implementing and maintaining production infrastructure through code, using Terraform or OpenTofu. You can build useful automation and make changes safely in an environment other people depend on.
  • Sound judgment about how to build and operate a secure AWS environment. You can balance performance, reliability, and cost, and recognize when a simple approach is sufficient.
  • The ability to turn an engineering need into a workable solution, carry it through implementation, and explain decisions clearly. You leave documentation that helps others operate what you build.

Preferred Qualifications

  • Supporting chip-design or other engineering workloads, particularly where large compute jobs, shared data, and software licenses need to work together.
  • Building an AWS environment that integrates with local infrastructure, including secure remote access for engineers.

Salary Range: $140,000 - $160,000

NOTE TO RECRUITERS:

Principals only. We are not accepting resumes from recruiters for this position. Remuneration for recruiting activities is only applicable subject to a signed and executed agreement between the parties. Please don’t send candidates to Ayar Labs, and do not contact our managers.

Ayar Labs is an Equal Opportunity Employer and is strongly committed to all policies which will afford equal opportunity employment to all qualified persons without regard to age, sex, national origin, race, color, ethnicity, creed, religion, gender identity, sexual orientation, disability, veteran status, or any other characteristic protected by law. It is the policy of Ayar Labs to provide reasonable accommodation when requested by a qualified applicant or employee with a disability, unless such accommodation would cause an undue hardship. Veterans are more than welcome and encouraged to apply.

Looking for more opportunities?

Browse thousands of graduate jobs and entry-level positions.

Browse All Jobs