Senior Systems Engineer - Enterprise AI Platforms

$206k - $275k • Remote • San Francisco Office (Fremont St) • FullTime

Posted 23h ago

Job Location

San Francisco Office (Fremont St)

Tech Stack

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

AI Infrastructure Engineer

About the job

Lambda, The Superintelligence Cloud, is seeking a Senior Systems Engineer to join their Enterprise AI Platforms team. This role is crucial for building and scaling the internal systems that power Lambda's business operations, partnering with various departments to implement tools, automate workflows, and ensure data integrity. The ideal candidate will design, write, and deliver software and services to enhance the availability, scalability, reliability, and efficiency of Lambda’s internal IT systems and platforms, while also automating responses to non-exceptional events and influencing new designs for large-scale distributed systems.

Responsibilities

  • Design, write, and deliver software and services to improve the availability, scalability, reliability, and efficiency of Lambda’s internal IT systems and platforms.
  • Solve problems relating to mission-critical services and build automation to prevent problem recurrence.
  • Collaborate with Lambda Engineering and internal teams to influence and create new designs, architectures, standards, and methods for large-scale distributed systems.
  • Engage in service capacity planning and demand forecasting, software performance analysis, and system tuning.
  • Produce documentation and related artifacts for the systems you are responsible for.
  • Design and deploy low-code/no-code and custom internal platforms integrated with LLM workflows.
  • Deploy, monitor, and optimize internal AI proxy gateways to route model traffic, manage API rate limits, optimize latency, and control inference costs.
  • Integrate AI agents into internal IT workflows such as Slack/Discord bots, incident triage, log parsing, user provisioning, and ticket routing.
  • Manage deployments, serverless functions, and infra pipelines to ensure high uptime for all internal systems.
  • Implement guardrails, data-privacy controls, and secret management to ensure sensitive corporate data isn't exposed to third-party LLM providers.

Requirements

  • 6+ years in software engineering, site reliability engineering, or IT systems engineering where you built and ran things in code.
  • Experience with tools like Appsmith, custom Vercel-hosted frontends.
  • Experience with AI agents and LLM workflows.
  • Experience managing AI infrastructure and gateways.
  • Experience with CI/CD pipelines (Vercel, AWS/GCP, etc.).
  • Excellent communication skills, with a desire to record and document issues and solutions.
  • An enthusiastic, go-for-it attitude and a drive to fix what is broken.
  • An urge for delivering quickly and effectively, and iterating fast.
  • Experience and interest in ML/AI workloads and compute (Nice to Have).
  • Practical experience implementing and managing paging, alerting, and on-call scheduling flows (Nice to Have).

Benefits

  • Generous cash & equity compensation
  • Health, dental, and vision coverage for you and your dependents
  • Wellness and commuter stipends for select roles
  • 401k Plan with 2% company match (USA employees)
  • Flexible paid time off plan

About Lambda

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.