Job Description
Description
AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon's Simple Storage Service (S3) and Amazon Elastic Compute Cloud (EC2), to consistently released new product innovations that continue to set AWS's services and features apart in the industry. As a member of the UC organization, you'll support the development and management of Compute, Database, Storage, Internet of Things (Iot), Platform, and Productivity Apps services in AWS, including support for customers who require specialized security solutions for their cloud services.
Looking for a highly-skilled Senior Machine Learning Engineer, to lead the development and delivery of technologies to push the boundaries of efficient inference for open source Generative Artificial Intelligence (GenAI) models. This role offers the exciting chance to work in a highly technical domain at the boundary between fundamental AI research and production engineering such as Quantization, Speculative Decoding, and Long Context for inference efficiency.
Key job responsibilities
* Create solutions that facilitate the usage and building of artificial intelligence workflows and optimize them for cost and latency.
* Collaborate with cross-functional teams of engineers and scientists to identify and solve complex problems in GenAI
- Design, prototype, and evaluate new inference engines and optimization techniques
- Participate in deep-dive analysis and profiling of production code. Optimize inference performance across various platforms
- Hold a high bar for technical excellence within the team and across the organization
About the team
Diverse Experiences
AWS values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying.
Why AWS?
Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating - that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.
Inclusive Team Culture
AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger, more collaborative teams. Our continual innovation is fueled by the bold ideas, fresh perspectives, and passionate voices our teams bring to everything we do.
Mentorship & Career Growth
We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.
Work/Life Balance
*** Please continue to use the below tagline in all job postings as the statement has been approved by all stakeholders and aligns with Amazon's working culture.
We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there's nothing we can't achieve.
EEO/Accommodations
Amazon is committed to a diverse and inclusive workplace. Amazon is an equal opportunity employer and does not discriminate on the basis of race, national origin, gender, gender identity, sexual orientation, protected veteran status, disability, age, or other legally protected status. For individuals with disabilities who would like to request an accommodation, please let us know and we will connect you to our accommodation team. You may also reach them directly by visiting
Basic Qualifications
- 5+ years of non-internship professional software development experience
- 5+ years of programming with at least one software programming language experience
- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- Experience as a mentor, tech lead or leading an engineering team
Preferred Qualifications
- 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
- Bachelor's degree in computer science or equivalent
- - Experience with inference frameworks such as vllm, PyTorch, TensorFlow, TensorRT etc.
- - Proficiency in performance optimization on GPU or Trainiums
- - Proficiency in kernel programming for accelerated hardware using programming models such as (but not limited to) CUDA
- - Experience with latency-sensitive optimizations and real-time inference
- - Knowledge of model optimization techniques
- - Strong communication skills and ability to work in a collaborative environment
- - Passion for solving complex problems and driving innovation in AI technology
- - (MS/Phd) in Mathematics or Electrical engineering, with sufficient coding expertise CUDA/kernals preferred
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company's reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.
Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit for more information. If the country/region you're applying in isn't listed, please contact your Recruiting Partner.
Our compensation reflects the cost of labor across several US geographic markets. The base pay for this position ranges from $151,300/year in our lowest geographic market up to $261,500/year in our highest geographic market. Pay is based on a number of factors including market location and may vary depending on job-related knowledge, skills, and experience. Amazon is a total compensation company. Dependent on the position offered, equity, sign-on payments, and other forms of compensation may be provided as part of a total compensation package, in addition to a full range of medical, financial, and/or other benefits. For more information, please visit . This position will remain posted until filled. Applicants should apply via our internal or external career site.
Job Tags
Internship, Local area,
Similar Jobs
The Loose Tooth Pediatric Dentistry
...About the Job: Our amazing state of the art Pediatric dental office is expanding! We are looking to add another enthusiastic, energetic, reliable and self motivated dental assistant to our already amazing team! We are a non-corporate, privately owned practice. We serve...
SynergisticIT
...excellent reputation with the clients.Currently, We are looking for entry-level software programmers, Java Full stack developers, Python/Java developers, Data analysts/ Data Scientists, Machine Learning engineers.Who Should Apply Recent Computer science/Engineering /...
Sandpiper Productions
...About us Join our team of professionals and apply for our elite brand ambassador job in Massachusetts and be part of something great! Starting pay $30.00/hour. Female-owned and known for our professionalism and progressive approach, we specialize in consumer...
Huntington Ingalls Industries
...meeting you.To learn more about Mission Technologies, click here for a short video: Who We AreWe are seeking experienced Fiber Optic Technicians for fleet modernization in San Diego, CA area. The selected candidates will be supporting our Systems Integrations...
Joseph Toyota of Cincinnati
Joseph Buick GMC has an opening for a Qualified General Motors Technician. This opening is for an experienced Technician. General Motors Experienced preferred but can fast track training for the right individual. Responsibilities Include but Not Limited To:*...