SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Binance, the world's largest cryptocurrency exchange by trading volume, is seeking a Senior DevOps Engineer to manage and optimize cloud infrastructure supporting a global blockchain ecosystem serving 300+ million users across 100+ countries.
You will take ownership of production infrastructure stability, handling incidents and post-mortem analysis to drive continuous improvements. Key responsibilities include designing, deploying, monitoring, and troubleshooting Kafka and Redis clusters in production environments, ensuring optimal performance and reliability at scale. You'll work closely with development teams to enable seamless application deployments and manage cloud infrastructure across AWS and Alicloud for performance, cost, and reliability optimization.
A significant part of this role involves developing DevOps platforms—such as online load testing and change management systems—and pioneering the integration of AI and LLMs (OpenAI, Dify, Agno, LangChain) into infrastructure operations. You'll leverage AI-driven insights to enhance automation, implement intelligent alert triage, conduct root cause analysis, and build chat-based operations (ChatOps) capabilities. This is an opportunity to shape the future of AIOps at a leading fintech organization.
Binance offers a flat organizational structure, autonomy in fast-paced projects, a results-driven culture with career growth opportunities, competitive compensation, and flexible work-from-home arrangements (varying by team).
REQUIREMENTS:
- 5+ years of hands-on experience in Kafka and Redis operations in large-scale production environments; ability to collaborate with developers on code optimization
- Proficiency in at least one of: Python, Go, or Java; SQL programming required
- Hands-on experience with containerization and orchestration (Docker, Kubernetes)
- Strong experience with CI/CD tools (GitHub Actions, Ansible, Terraform, etc.)
- Minimum 3 years of AWS cloud platform experience; GCP, Azure, or Alicloud experience is a plus
- Excellent problem-solving and troubleshooting skills
- Strong team collaboration and partnership-building abilities
- Practical experience building or operating AIOps systems (anomaly detection, alert correlation, automated healing, RCA)
- Familiarity with LLM-based DevOps automation (e.g., chat-based ops assistants, AI-driven observability workflows)
- Experience using or integrating tools like Dify, Agno, or LangChain into operational workflows