AWS Jobs
366 open roles mentioning AWS
Generative AI Inference Engineer
Stability AI
We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.
Solutions Architect - Public Sector
Cohere
Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.
Cloud Security Engineer
Applied Intuition
Applied Intuition is seeking a highly focused Cloud Security Engineer to play a crucial role in securing our infrastructure across diverse multi-cloud environments (AWS, Azure, GCP, OCI), with a heavy emphasis on Kubernetes cluster hardening. You will establish robust guardrails, enforce Identity and Access Management policies, and maintain our Cloud Security Posture Management (CSPM) to prevent insecure deployments and ensure continuous compliance. This role involves working alongside our Corporate Security & Infrastructure team to secure our cloud footprint.
$60k - $300k
Member of Technical Staff - Platform Foundations
Reflection ai
Reflection is building a company-wide foundations platform to accelerate every engineering team by providing reliable, scalable developer infrastructure, SRE capabilities, and high-throughput data ingestion tooling. This team builds and operates the core platform layer that all engineering teams depend on, defining opinionated golden paths for cloud infrastructure, networking, and access patterns. The goal is to ensure engineers can ship quickly and safely without sacrificing reliability, security, or cost predictability, enabling the development of open foundational models.
Software Engineer, Site Reliability
Hebbia
Hebbia is seeking a Site Reliability Engineer who approaches the role with a software engineering mindset. You will be responsible for the entire lifecycle of critical production systems, focusing on design, development, and enhancement rather than just operation. This involves writing production-quality code to ensure platform reliability at scale, collaborating with product engineering teams to integrate reliability into architectural decisions from the outset, and building essential internal tooling for all engineers. The role emphasizes coding, including instrumenting services, optimizing performance, developing deployment platforms, and translating incident learnings into permanent architectural improvements.
$160k - $350k
Forward Deployed Engineer (India)
cartesia
Cartesia is seeking a Forward Deployed Engineer to join their mission of building real-time multimodal intelligence. This role involves embedding directly with enterprise customers to deliver agentic voice AI solutions into their production environments. The Forward Deployed Engineer will be responsible for maximizing customer success and revenue by transforming Cartesia's core product into deployed, high-impact solutions. This is an opportunity to work on cutting-edge AI models and experiences, with a focus on innovation and rapid execution.
Solutions Architect
Modal
Modal is seeking a high-impact Solutions Architect to lead technical strategy for its most strategic enterprise accounts. In this role, you will act as the executive technical counterpart to Enterprise Account Executives, guiding complex evaluations, shaping infrastructure modernization roadmaps, and promoting multi-product adoption for AI/ML workloads. This is a strategic, consultative position that requires deep architectural knowledge, executive presence, and the ability to influence significant infrastructure decisions. You will collaborate directly with CTOs, VPs of Engineering, and ML platform leaders to reimagine how AI infrastructure is built and operated. If you excel in fast-paced technical sales environments and are passionate about shaping the infrastructure powering modern AI companies, this role is for you.
Security and Compliance Manager
sierra.ai
Sierra is a leading platform for customer-facing AI agents, partnering with major global brands to transform customer service and business growth. We are primarily an in-person company based in San Francisco, with expanding offices internationally. Our culture is built on core values of Trust, Customer Obsession, Craftsmanship, Intensity, and a commitment to balancing Family. The company was co-founded by Bret Taylor, former co-CEO of Salesforce and CTO of Facebook, and Clay Bavor, who spent 18 years at Google leading initiatives like Google Labs, AR/VR, and Google Lens.
$53k - $800k
Software Engineer, Data
Heyge
HeyGen is seeking a Software Engineer with data engineering responsibilities to bridge the gap between core application development and large-scale data infrastructure. You will help build the data foundational layers for our next-generation features, enabling AI models to function in real-time, building robust pipelines for multimedia, and powering engaging user experiences. This role is crucial for developing cutting-edge features like PPT-to-video converters and interactive, conversational video capabilities.
$180k - $220k
Software Engineer, Growth Infrastructure
Replit
Replit is seeking an experienced Growth Infrastructure Engineer to build and maintain the technical foundation for scalable growth experiments, high-performance data pipelines, and automated systems. This role is at the intersection of growth, product, and infrastructure, requiring deep technical engineering skills combined with an understanding of experimentation and data-driven optimization. You will collaborate with product, data science, and backend teams to ensure growth initiatives run smoothly and scale efficiently across systems. This is a unique opportunity to be an early member of a new Growth team, with significant ownership and influence over technical direction and product outcomes, working on a product that has experienced massive user growth.
Senior Software Engineer, Core Infrastructure
Harvey
As a Software Engineer on the Core Infrastructure team at Harvey, you will play a critical role in designing and building new infrastructure systems while equally scaling and strengthening our existing infrastructure. This foundation powers every user interaction with Harvey, processing billions of prompt tokens and millions of daily requests across our global legal AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring Harvey remains resilient and efficient as it scales products, regions, customers, and usage. Your contributions will directly impact the reliability, scalability, and security of our platform as we serve the world's leading law firms and professional service providers.
$200k - $250k
MTS, Security
fireworks ai
Fireworks is seeking a Security Engineer to play a key role in designing, implementing, and operating security controls across AI infrastructure, AI platforms, and internal systems. This role is crucial for strengthening our security posture and supporting rapid growth, ensuring the confidentiality, integrity, and availability of data, models, and infrastructure as organizations increasingly rely on large language models and cloud-native AI services. You will be instrumental in building trust by embedding security across all layers of our technology stack.
Member of Technical Staff - Platform Engineering
Modal
AI needs a new infrastructure layer, and Modal is building it. We provide instant GPU access, sub-second container starts, and native storage, enabling customers to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. As a rapidly growing cloud infrastructure company, we are seeking to dramatically improve our platform's reliability while scaling our team and customer base. This role is ideal for individuals with deep systems thinking, a passion for reliability, and a drive to enable others to move faster at scale.
Deployed Engineer (UK)
Langchain
LangChain is seeking a Deployed Engineer to join their team, focusing on building and running AI agents in production. This role involves working on complex applied AI systems that real teams depend on, with a fast feedback loop and visible impact. You will collaborate with customer engineering teams to co-architect and co-build production AI agents, own the technical win in pre-sales by designing POCs and answering technical questions, and help customers deploy and operate agent-based applications. You will also advise customers on architecture and best practices, run technical demos and trainings, and contribute reusable patterns and code. This position offers the opportunity to directly shape how AI agents are built and adopted in the real world.
Site Reliability Engineer, Inference Infrastructure
Cohere
Cohere is seeking a Site Reliability Engineer to join the Model Serving team. This role is crucial for developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and create customized deployments to meet their specific needs, contributing to the widespread adoption of AI.
Staff Software Engineer, Inference Infrastructure
Cohere
Cohere is seeking Members of Technical Staff to join the Model Serving team. This role focuses on developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and build customized deployments to meet their specific needs.
Forward Deployed Engineer - Systems
Modal
Modal is seeking an experienced Forward Deployed Engineer (FDE) to partner with our sales team and drive technical sales success. As an FDE, you will be the technical voice in our sales process, working directly with Account Executives to help enterprise customers understand how Modal can transform their AI/ML infrastructure. You will partner with Account Executives to identify, qualify, and close strategic enterprise opportunities, lead technical discovery sessions with prospective customers to understand their current infrastructure, pain points, and requirements, and design and present compelling technical solutions that demonstrate how Modal addresses customer needs. You will also architect migration paths from existing cloud infrastructure (AWS, GCP, Azure) to Modal's serverless platform, conduct technical demos, experiments, and proof-of-concepts that showcase Modal's capabilities, and navigate complex technical evaluations and address security, compliance, and integration concerns. Additionally, you will build trusted advisor relationships with technical decision-makers, collaborate with product and engineering teams to communicate customer feedback and influence product roadmap, and support contract negotiations by providing technical expertise.
Senior Backend Software Engineer, AI Observability & Evals Platform (LangSmith)
Langchain
LangChain is seeking a Senior Backend Engineer to join the LangSmith team, which focuses on building the core platform for observability, evaluation, and production reliability of AI systems. In this role, you will be responsible for developing the backend systems that power LangChain's observability and evals platform, enabling developers to monitor and evaluate their AI applications at scale. While the primary focus is on backend development, experience with full-stack or frontend engineering, performance tuning, and debugging production issues will be highly beneficial.
$175k - $240k
Security Engineer, Cloud
Rogo.ai
Rogo is seeking a Staff Security Engineer to lead the design and implementation of cloud security architecture across AWS and GCP. This is a hands-on role for an engineer experienced in building and operating secure cloud platforms at scale, using code, systems design, and automation to solve security challenges. You will own the technical direction of cloud security, focusing on secure primitives, large-scale Terraform authoring, identity and network architecture, and embedding security into the core platform. The role requires a senior technical leader who is also tactical, writing production code, reviewing infrastructure changes, and providing pragmatic security solutions.
Forward Deployed Engineer, Infrastructure Specialist (Europe)
Cohere
Cohere is a leading security-first enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers to join our team and contribute to the widespread adoption of AI. This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications, acting as a bridge between our core North product and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors like finance, healthcare, and telecommunications, working with esteemed clients and focusing on Agentic AI.
$20k - $40k