Наши требования
Bachelor’s degree in Computer Science, Engineering, or a related field
Proven experience in any cloud (AWS/GCP/Azure)
Experience with implementing SRE practices such as SLO/SLI, Error budgets, Postmortems, Reducing Toil, capacity planning, and Incident Management
Knowledge of Python or other scripting/programming language
Strong background in monitoring tools
Proficiency in CI/CD tools, infrastructure as code, and configuration management
Solid knowledge of container orchestration technologies (Kubernetes, Docker)
Expertise in deployment and management of LLMs, including technologies like RAG
Certification in Kubernetes, AWS/GCP/Azure, or similar technologies
Proven experience in DevOps
Knowledge of managing and optimizing AI/ML models in production environments, including basic deployment, monitoring, and maintenance