Sr. Associate Director, Data and Analytics (AI Data Engineer)

HSBC
Shenzhen
Workplace: OnsiteFull timeFunction: Data Analytics & Business IntelligenceSkills: ["Attention to data governance","Monitoring","Data quality","Privacy-by-design","Stakeholder collaboration"]

Build and operate structured and unstructured data pipelines that power KYC, credit, and portfolio AI solutions. Develop embedding/vectorisation pipelines with governed vector indices, implement data quality/observability/lineage, and apply privacy-by-design for PII handling. Establish dataset versioning and reproducible evaluations, partner with Agent Engineers to define requirements and feedback loops, and support production operations to continuously improve data services.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
HSBC
HSBC
2 days ago

Sr. Associate Director, Data and Analytics (AI Data Engineer)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Build and operate structured and unstructured data pipelines that power KYC, credit, and portfolio AI solutions. Develop embedding/vectorisation pipelines with governed vector indices, implement data quality/observability/lineage, and apply privacy-by-design for PII handling. Establish dataset versioning and reproducible evaluations, partner with Agent Engineers to define requirements and feedback loops, and support production operations to continuously improve data services.
Location: Shenzhen
Workplace: Onsite
Employment Type: Full time
Job Function: Data Analytics & Business Intelligence
Seniority: Director level

Key Responsibilities

  • •Build and operate structured domain data pipelines for onboarding, customer/entity data, credit facilities/exposures, limits, risk grades, portfolio hierarchies, and performance/arreas.
  • •Build unstructured data pipelines for domain documents with parsing, metadata enrichment, deduplication, and retention handling.
  • •Develop embedding/vectorisation pipelines and manage vector indices with refresh/deletion strategies aligned to data governance.
  • •Implement data quality, observability, and lineage with automated testing, SLAs, anomaly detection, monitoring, and runbooks.
  • •Apply privacy-by-design (PII handling, masking/tokenisation, access controls, audit trails) and support dataset versioning/reproducibility for AI evaluations; partner with Agent Engineers to define evaluation datasets and feedback loops.

Key Requirements

  • •Deliver reliable, fresh, well-governed datasets and indices for KYC, credit, and portfolio AI solutions.
  • •Improve retrieval relevance and grounding by enhancing metadata, input chunking, and data quality.
  • •Reduce operational incidents via strong testing, monitoring, and disciplined change management.
  • •Make AI datasets reproducible and auditable to support reviews and control expectations.
  • •Build and operate both structured domain data pipelines and unstructured domain document pipelines (parsing, metadata enrichment, deduplication, retention handling).
Experience:AIData engineeringData governanceKYCCreditPortfolio
Skills:Attention to data governanceMonitoringData qualityPrivacy-by-designStakeholder collaboration

Company Brief

HSBC
Global banking and financial services organisation offering retail, commercial, corporate and investment banking, wealth management, and global markets services across Europe, Asia, the Americas and the Middle East.
Industry: Banking
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: London, United Kingdom
Founded: 1865
Glassdoor
Glassdoor: 3.6
WebsiteLinkedIn