Чем предстоит заниматься
Build evaluation and regression detection infrastructure
Define AI infrastructure readiness with GPU driver validation and capacity planning
Deploy and tune vLLM inference servers across customer environments
Implement observability for AI stack with tracing and dashboards
Maintain LiteLLM gateway and Helm charts
Manage model lifecycle with canary rollouts and rollback plans
Optimize LLM inference performance