Azure Jobs
230 open roles mentioning Azure
Staff Site Reliability Engineer
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This is a rare opportunity to help build a generational company at an inflection point, with strong product-market fit and world-class investor support. The team moves fast, takes ownership, and is deeply committed to the mission, operating with intensity and pushing for excellence. As a Staff Software Engineer on the Site Reliability team, you will ensure the reliability, scalability, and performance of our legal AI platform, owning the systems that keep our platform fast, secure, and always on. Your work will be critical in scaling across 50+ regions and automating mission-critical operations to ensure Harvey remains resilient as we grow. If you are passionate about building robust systems and reducing complexity through automation, we encourage you to apply.
Staff Software Engineer, Core Infrastructure
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This is a rare opportunity to help build a generational company at an inflection point, with strong product-market fit and significant investor backing. The team moves fast, takes ownership, and is deeply committed to the mission, operating with intensity and pushing for excellence. As a Staff Software Engineer on the Core Infrastructure team, you will play a critical role in designing, building, scaling, and strengthening the infrastructure that powers every user interaction with Harvey's AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring the platform remains resilient and efficient as it scales.
Staff Software Engineer, Databases
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI with an enterprise-grade platform. We are a fast-scaling company with strong product-market fit, seeking individuals who want to do the best work of their careers alongside driven colleagues. Our values are Decisiveness, Simplicity, and Job's Not Finished, emphasizing quick action, scalable solutions, and continuous improvement. As a Staff Database Engineer, you will define and evolve our global PostgreSQL database infrastructure. This leadership role involves setting technical strategy for our fleet, establishing standards for migration governance, driving multi-region scaling architecture, and ensuring data layer reliability. You will partner with senior engineers on roadmap decisions while also being hands-on in diagnosing and resolving critical issues.
$191k - $325k
Generative AI Inference Engineer
Stability AI
We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.
Solutions Architect - Public Sector
Cohere
Cohere is seeking a Solutions Architect to play a significant role in growing the company's Defence and National Security business. This dynamic role requires a blend of strategic thinking and hands-on execution, focusing on building customer demos and proof-of-concepts that highlight the business value of Cohere's AI platform. As the technical relationship owner, you will collaborate with stakeholders to understand their objectives and translate them into technical solutions. You will also serve as the voice of the customer, liaising between clients and the product team, providing guidance on best practices, identifying platform improvements, and cultivating technical champions within customer organizations to drive adoption and gather feedback.
Staff / Principal Platform Engineer - USA
inworld
Inworld is seeking a Staff / Principal Platform Engineer to take end-to-end ownership of building, securing, and scaling our AI products. This role involves being the driving force behind our cloud infrastructure, partnering with engineers to deploy and evolve services across major cloud providers using tools like Terraform and ArgoCD. You will identify needs and drive initiatives forward, directly shaping how the company operates and innovates.
$280k - $350k
Cloud Security Engineer
Applied Intuition
Applied Intuition is seeking a highly focused Cloud Security Engineer to play a crucial role in securing our infrastructure across diverse multi-cloud environments (AWS, Azure, GCP, OCI), with a heavy emphasis on Kubernetes cluster hardening. You will establish robust guardrails, enforce Identity and Access Management policies, and maintain our Cloud Security Posture Management (CSPM) to prevent insecure deployments and ensure continuous compliance. This role involves working alongside our Corporate Security & Infrastructure team to secure our cloud footprint.
$60k - $300k
Solutions Architect
Modal
Modal is seeking a high-impact Solutions Architect to lead technical strategy for its most strategic enterprise accounts. In this role, you will act as the executive technical counterpart to Enterprise Account Executives, guiding complex evaluations, shaping infrastructure modernization roadmaps, and promoting multi-product adoption for AI/ML workloads. This is a strategic, consultative position that requires deep architectural knowledge, executive presence, and the ability to influence significant infrastructure decisions. You will collaborate directly with CTOs, VPs of Engineering, and ML platform leaders to reimagine how AI infrastructure is built and operated. If you excel in fast-paced technical sales environments and are passionate about shaping the infrastructure powering modern AI companies, this role is for you.
Security and Compliance Manager
sierra.ai
Sierra is a leading platform for customer-facing AI agents, partnering with major global brands to transform customer service and business growth. We are primarily an in-person company based in San Francisco, with expanding offices internationally. Our culture is built on core values of Trust, Customer Obsession, Craftsmanship, Intensity, and a commitment to balancing Family. The company was co-founded by Bret Taylor, former co-CEO of Salesforce and CTO of Facebook, and Clay Bavor, who spent 18 years at Google leading initiatives like Google Labs, AR/VR, and Google Lens.
$53k - $800k
Software Engineer, Growth Infrastructure
Replit
Replit is seeking an experienced Growth Infrastructure Engineer to build and maintain the technical foundation for scalable growth experiments, high-performance data pipelines, and automated systems. This role is at the intersection of growth, product, and infrastructure, requiring deep technical engineering skills combined with an understanding of experimentation and data-driven optimization. You will collaborate with product, data science, and backend teams to ensure growth initiatives run smoothly and scale efficiently across systems. This is a unique opportunity to be an early member of a new Growth team, with significant ownership and influence over technical direction and product outcomes, working on a product that has experienced massive user growth.
Senior Software Engineer, Core Infrastructure
Harvey
As a Software Engineer on the Core Infrastructure team at Harvey, you will play a critical role in designing and building new infrastructure systems while equally scaling and strengthening our existing infrastructure. This foundation powers every user interaction with Harvey, processing billions of prompt tokens and millions of daily requests across our global legal AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring Harvey remains resilient and efficient as it scales products, regions, customers, and usage. Your contributions will directly impact the reliability, scalability, and security of our platform as we serve the world's leading law firms and professional service providers.
$200k - $250k
MTS, Security
fireworks ai
Fireworks is seeking a Security Engineer to play a key role in designing, implementing, and operating security controls across AI infrastructure, AI platforms, and internal systems. This role is crucial for strengthening our security posture and supporting rapid growth, ensuring the confidentiality, integrity, and availability of data, models, and infrastructure as organizations increasingly rely on large language models and cloud-native AI services. You will be instrumental in building trust by embedding security across all layers of our technology stack.
Site Reliability Engineer, Inference Infrastructure
Cohere
Cohere is seeking a Site Reliability Engineer to join the Model Serving team. This role is crucial for developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and create customized deployments to meet their specific needs, contributing to the widespread adoption of AI.
Staff Software Engineer, Inference Infrastructure
Cohere
Cohere is seeking Members of Technical Staff to join the Model Serving team. This role focuses on developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and build customized deployments to meet their specific needs.
Forward Deployed Engineer - Systems
Modal
Modal is seeking an experienced Forward Deployed Engineer (FDE) to partner with our sales team and drive technical sales success. As an FDE, you will be the technical voice in our sales process, working directly with Account Executives to help enterprise customers understand how Modal can transform their AI/ML infrastructure. You will partner with Account Executives to identify, qualify, and close strategic enterprise opportunities, lead technical discovery sessions with prospective customers to understand their current infrastructure, pain points, and requirements, and design and present compelling technical solutions that demonstrate how Modal addresses customer needs. You will also architect migration paths from existing cloud infrastructure (AWS, GCP, Azure) to Modal's serverless platform, conduct technical demos, experiments, and proof-of-concepts that showcase Modal's capabilities, and navigate complex technical evaluations and address security, compliance, and integration concerns. Additionally, you will build trusted advisor relationships with technical decision-makers, collaborate with product and engineering teams to communicate customer feedback and influence product roadmap, and support contract negotiations by providing technical expertise.
Senior Backend Software Engineer, AI Observability & Evals Platform (LangSmith)
Langchain
LangChain is seeking a Senior Backend Engineer to join the LangSmith team, which focuses on building the core platform for observability, evaluation, and production reliability of AI systems. In this role, you will be responsible for developing the backend systems that power LangChain's observability and evals platform, enabling developers to monitor and evaluate their AI applications at scale. While the primary focus is on backend development, experience with full-stack or frontend engineering, performance tuning, and debugging production issues will be highly beneficial.
$175k - $240k
Forward Deployed Engineer, Infrastructure Specialist (Europe)
Cohere
Cohere is a leading security-first enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers to join our team and contribute to the widespread adoption of AI. This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications, acting as a bridge between our core North product and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors like finance, healthcare, and telecommunications, working with esteemed clients and focusing on Agentic AI.
$20k - $40k
Senior Software Engineer, Site Reliability Engineer
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are a fast-scaling company with strong product-market fit, seeking ambitious individuals to help build a generational company. Our team operates with intensity, ownership, and a commitment to our mission, valuing decisiveness, simplicity, and continuous improvement. As a Software Engineer on the Site Reliability team, you will ensure the reliability, scalability, and performance of our legal AI platform, owning the systems that keep our platform fast, secure, and always on. Your work will be crucial in maintaining platform resilience as we grow across 50+ regions.
$200k - $260k
Staff Software Engineer, Site Reliability Engineer
Harvey
Why Harvey At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come. This is a rare chance to help build a generational company at a true inflection point. With 1500+ customers in 60+ countries, strong product-market fit, and world-class investor support, we’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched. Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you. At Harvey, the future of professional services is being written today — and we’re just getting started. Role Overview As a Staff Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow. If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you. This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees. What You’ll Do - Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions - Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements - Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention - Establish best practices for security, compliance, and reliability and collaborate across teams to drive these principles throughout the software lifecycle - Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality - Provide technical mentorship and leadership, promoting best practices and fostering team growth What You Have - 10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams - Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.) - Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.) - Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.) - Strong programming skills (Python, Bash, Go, or similar languages) - Proven track record of diagnosing complex system problems and implementing durable solutions - Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles - Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence Compensation Range $238,000 - $290,000 USD Depending on your location, an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices [here]. #LI-AN2 Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing [email protected]
$238k - $290k
Senior Backend Engineer
Legora
Legora is redefining how legal work gets done with an AI-native workspace that helps legal professionals move faster, think more clearly, and operate with sharper precision. We analyze thousands of documents in minutes and power end-to-end workflows, cutting through complexity so teams can focus on judgment, strategy, and outcomes. As a Senior Backend Engineer, you will build, review, and ship amazing products that assist legal professionals globally. You will have the opportunity to work across the entire tech stack with clear ownership and autonomy.