Join us if you're excited about pushing the boundaries of what's possible in physical AI working with world-class scientists and engineers and seeing your innovations deployed at unprecedented scale.
We are seeking an exceptional Principal Applied Scientist to drive technical innovation in visual reasoning foundation models. You will be a technical leader who sets the research direction, architects novel solutions, and delivers breakthrough results that advance the state of the art while solving real-world business problems.
You will lead the efforts of building a next-generation visual reasoning engine powered by frontier Large Video Models (LVMs). Your mission is to build a system that rivals human understanding of the physical world-moving far beyond the static perception of detection and tracking into the realm of deep spatial-temporal reasoning. This is not a passive computer vision tool; it is an agentic collaborator capable of interpreting natural language instructions navigating unstructured environments and executing complex tasks.
You will sit at the high-stakes intersection of LVMs, LLMs, and Agentic AI engineering systems that don't just 'see' but reason and act within the physical world. You will own end-to-end technical solutions from research to production deployment driving innovation through hands-on research prototyping and deployment while delivering production impact.
Key Responsibilities
• Direct the technical vision for next-gen visual reasoning pioneering the use of LVMs to solve high-dimensional spatial-temporal problems
• Design and implement novel deep learning architectures combining a multitude of modalities including image video and geospatial data
• Solve computational challenges to train foundation models at scale taking advantage of latest developments in hardware and deep learning libraries
• Architect scalable solutions that deliver real-time insights across diverse physical environments
• Build agentic AI systems that autonomously execute end-to-end workflows transforming visual data into actionable business intelligence
• Collaborate with multiple science and engineering teams to build adaptations that power use cases across diverse domains
• Publish research at top-tier conferences (CVPR NeurIPS ICML) and establish technical thought leadership in visual reasoning and multi-modal AI
• Mentor scientists and engineers while maintaining significant hands-on contribution to technical solutions
• Influence product roadmaps through deep technical expertise and business acumen
Key job responsibilities
- Direct the technical vision for next-gen visual reasoning, pioneering the use of LVMs to solve high-dimensional spatial-temporal problems
- Design and implement novel deep learning architectures combining a multitude of modalities, including image, video, and geospatial data
- Solve computational challenges to train foundation models at scale, taking advantage of latest developments in hardware and deep learning libraries
- Architect scalable solutions that deliver real-time insights across diverse physical environments
- Build agentic AI systems that autonomously execute end-to-end workflows, transforming visual data into actionable business intelligence
- Collaborate with multiple science and engineering teams to build adaptations that power use cases across diverse domains
- Publish research at top-tier conferences (CVPR, NeurIPS, ICML) and establish technical thought leadership in visual reasoning and multi-modal AI
- Mentor scientists and engineers while maintaining significant hands-on contribution to technical solutions
- Influence product roadmaps through deep technical expertise and business acumen
A day in the life
• Develop and implement novel foundation model architectures working hands-on with data and our extensive training and evaluation infrastructure
• Guide and support fellow scientists in solving complex technical challenges from spatiotemporal reasoning to efficient multi-task learning
• Guide and support fellow engineers in building scalable and reusable infrastructure to support model training evaluation and inference
• Lead focused technical initiatives from conception through deployment ensuring successful integration with production systems
• Drive technical discussions with the team and key stakeholders
• Conduct experiments and prototype new ideas
• Mentor team members while maintaining significant hands-on contribution to technical solutions
About the team
Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying.
Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.
We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud.
Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity and AmazeCon conferences, inspire us to never stop embracing our uniqueness.
We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.