SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Together AI is seeking a Technical Support Engineer to serve as the first line of defense for customers building training, fine-tuning, and inference solutions on the platform. You'll engage directly with customers to resolve complex technical challenges involving GPU clusters, Kubernetes-based inference endpoints, and generative AI services.
Key responsibilities include:
- Troubleshoot and resolve customer issues with inference endpoints, GPU infrastructure, and model deployment
- Act as a customer-facing SRE, ensuring endpoint health, stability, and performance
- Become a product expert across Together AI's Gen AI solutions, serving as the last line of technical defense before escalation to Engineering
- Monitor dashboards and detect anomalies; escalate issues with data-backed analysis
- Manage customer communications during incidents, translating technical findings into clear updates
- Execute infrastructure changes via pull requests (infrastructure-as-code) for endpoint configuration, model deployment, and capacity scaling
- Flag engine-level bugs with logs and reproduction steps for the engineering team
- Collaborate across Engineering, Research, and Product teams to address customer concerns
- Identify patterns in support cases and work with teams to influence product roadmap decisions
- Maintain documentation of system configurations, troubleshooting guides, and FAQs
Schedule: Full-time position working US daytime hours. Initially Monday–Friday for onboarding and ramp-up. After full ramp, transitions to a 4-day shift (Saturday, Sunday, plus two weekdays), 10 hours per day, with 2 additional hours of on-call coverage on weekends.
This role is ideal for a deeply technical professional passionate about AI infrastructure and customer success, comfortable with on-call responsibilities and weekend work.
About Together AI
AI / Data / Infrastructure — cloud platform for open-source and generative AI model training and inference.