주요업무
• Design, build, and maintain high-throughput data pipelines processing petabyte-scale market data, execution logs, alternative data, and reference data
• Operate and evolve our analytical data platform — columnar stores, workflow orchestration, and custom tooling
• Ingest and normalize data from exchanges, vendors, and internal systems into consistent, queryable formats
• Implement data quality checks, lineage tracking, and monitoring to guarantee correctness and freshness
• Optimize storage layouts, query performance, and pipeline latency for both batch and near-real-time workloads
• Build and maintain data infrastructure supporting AI/ML workflows — training data preparation, feature pipelines, and model serving
• Collaborate directly with quantitative researchers, risk managers, and infrastructure teams to translate data requirements into reliable code and infrastructure
• Own the full lifecycle: capacity planning, schema evolution, incident response, and documentation