Bachelor's degree or above in Computer Science or a related field
3–8 years in software or data engineering (we'll calibrate level to your experience)
Strong production skills in Java and/or Python
Solid SQL and hands-on experience with at least one of MySQL/PostgreSQL, ClickHouse, or a comparable columnar/time-series store (schema design, query performance)
Experience building and operating data pipelines (batch and/or streaming) in production
Experience integrating with external APIs (REST/WebSocket) and handling real-world constraints — auth, rate limits, partial outages, schema drift
An ownership and operational mindset: you care about data quality, monitoring, reproducibility, incident response, and documentation
Proactive in using AI tools (AI coding assistants, LLMs) to work faster and smarter, with curiosity about applying AI/ML to data problems
Eagerness to learn — open to picking up Rust and new storage technologies as the platform evolves
Strong communication skills; comfortable working across business and technical teams
Fluency in Chinese and English
Rust, or a strong interest in learning it for performance-critical components
Hands-on experience with Apache Iceberg / lakehouse, ClickHouse, or Kafka (Spark a plus for Iceberg integration)
Exposure to Kubernetes, GitLab CI/CD, and cloud infrastructure (Alibaba Cloud, AWS, or GCP)
Workflow orchestration tooling (Airflow, Celery, or similar)
Exposure to AI/ML tooling (MLflow, model serving) or building/serving data for ML and AI workloads
Experience with market data, trading, or financial/crypto data at scale