Solutions Architect (Inference)
London
Posted 1mo ago
Job Location
London
Tech Stack
Remote Work Policy
On-site
Categories
Applied AI Engineer
About the job
As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.
Responsibilities
- Act as a technical advisor to strategic customers, supporting the ideation and development of applications using OSS models on Together AI.
- Run complex demonstrations and POCs of Together’s entire stack, including hardware and software solutions.
- Collaborate with sales to qualify new prospects and support existing customers in building Generative AI solutions.
- Build and maintain strong relationships with customer leadership and stakeholders.
- Deliver feedback to Product, Engineering, and Research teams to evolve the platform.
- Build educational content and tooling for internal and external use around Together’s solutions.
Requirements
- 7+ years of experience in a customer-facing technical role, with at least 2 years in a pre-sales function.
- Demonstrable background in building and delivering technical solutions in a Solutions Consulting / Pre-Sales environment, leveraging open-source models, ideally within inference-driven use cases.
- Excellent communication and interpersonal skills, with the ability to explain complex technical concepts to non-technical stakeholders.
- Ability to consult with customers to map business needs to technical solutions.
- Strong technical background with knowledge of AI, ML, GPU technologies, and their integration into HPC environments.
- Strong understanding of training, fine-tuning, and inference in the context of open source LLMs.
- Proficiency in Python and JavaScript, with experience building and delivering prototypes on API platforms.
- Familiarity with infrastructure services (e.g., Kubernetes, SLURM), infrastructure as code solutions (e.g., Ansible), container infrastructure (Docker), and scripting/programming languages (Python, Javascript).
- Strong sense of ownership and willingness to learn new skills.
- Ability to operate in dynamic environments, manage multiple projects, and prioritize effectively.