What you will do:
- Work at the intersection of development and site reliability. Creating SRE tools and systems, as well as supporting existing infrastructure and platforms.
- Ensure the reliability, availability, and performance of Zilliz’s distributed database systems.
- Develop and implement strategies for monitoring, incident management, and disaster recovery.
- Automate system operations and maintenance tasks to improve efficiency and reduce manual intervention.
- Design and build tools to manage and monitor infrastructure, ensuring scalability and robustness.
- Collaborate with software engineers to enhance system reliability, scalability, and performance.
- Maintain and improve the CI/CD pipeline to ensure smooth and rapid deployment of changes.
- Actively contribute to the Milvus Vector Database open-source community, focusing on improving reliability and operational efficiency.
What we are looking for:
- 4+ years of experience in site reliability engineering or similar roles with a focus on cloud-native systems.
- Proficiency in scripting languages such as Python, Go, or Java.
- Strong knowledge of container orchestration technologies like Kubernetes and Docker.
- Expertise with cloud platforms such as AWS, GCP, or Azure, and their respective monitoring and management tools.
- Experience with infrastructure as code tools such as Terraform or Ansible.
- Familiarity with CI/CD tools such as Jenkins, GitLab CI, or Argo.
- Proven ability to troubleshoot complex distributed systems and resolve issues promptly.
- Bachelor’s degree or above in computer science, software engineering, or other relevant disciplines.
- Ability to thrive in a fast-paced, startup environment and handle multiple projects simultaneously.
- Experience with Open Source Milvus Vector Database is nice to have
Top Skills
What We Do
Zilliz is a leading vector database company for production-ready AI. Built by the engineers who created Milvus, the world's most popular open-source vector database, Zilliz is on a mission to unleash data insights with AI. The company builds next-generation database technologies to help organizations rapidly create AI/ML applications, and unlock the potential of unstructured data. By taking the burden of complex data infrastructure management off of its users, Zilliz is committed to bringing the power of AI to every corporation, every organization, and every individual.
Headquartered in San Francisco, Zilliz is backed by a number of prestigious investors, including Aramco's Prosperity7 Ventures, Temasek's Pavilion Capital, Hillhouse Capital, 5Y Capital, Yunqi Partners, Trustbridge Partners and others. Zilliz's technologies and products help over 1000 organizations worldwide easily create AI applications in various scenarios, including computer vision, image retrieval, video analysis, NLP, recommendation engines, targeted ads, customized search, smart chatbots, fraud detection, network security, new drug discovery, and much more. Learn more at zilliz.com or follow @zilliz_universe.