SRE资深(AI运维建设/开发)-安全与风控 - 安全与风控
Core
Design, build, and maintain scalable systems and tools for AI-driven operations (AIOps) to ensure the reliability and stability of large-scale security infrastructure.
Role type
Senior Site Reliability Engineer (AI Operations & Development)
Builds
AI capability platforms for security operations, automated stability solutions, and resource lifecycle management tools.
Domain
Cybersecurity / AI Engineering / Cloud Infrastructure
Deliverable
production ML models
Required skills
Go, Python, Shell, Linux, Kubernetes, Load Balancer, Nginx, Ansible, Argo CD, Prometheus, Grafana, TCP/IP, HTTP, DNS, NAT, Network Routing, Network Switching
Preferred skills
AI Agent frameworks, Security & Risk Control business scenarios, System operations platform development, English fluency
Responsibilities
Ensure reliable, stable, and efficient operation of ultra-large-scale security infrastructure; Build AI capability platforms to empower business operations; Explore and apply advanced AI technologies to solve business challenges; Provide end-to-end AI operations solutions covering stability, deployment, iteration, and resource management.
