Sr Data Specialist - R01564133

Brillio
India
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 5-10 yearsEducation: bachelorsSkills: ["Data quality","Governance","Reliability","Cross-functional collaboration","Performance optimization"]

Design, build, and optimize scalable data pipelines and platforms using Spark/PySpark on Databricks. Own ETL/ELT workflows for large-scale processing across AWS (EMR, S3, data lake/warehouse) and Hadoop/Hive ecosystems. Develop data ingestion and consumption APIs, implement orchestration with Airflow/Autosys, and improve performance through partitioning, caching, and query tuning. Ensure data quality, governance, and reliability while collaborating with Analytics, Product, and Engineering teams.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Brillio
Brillio
5 months ago

Sr Data Specialist - R01564133

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 12 hours agoStatus: Live

Job Summary

Design, build, and optimize scalable data pipelines and platforms using Spark/PySpark on Databricks. Own ETL/ELT workflows for large-scale processing across AWS (EMR, S3, data lake/warehouse) and Hadoop/Hive ecosystems. Develop data ingestion and consumption APIs, implement orchestration with Airflow/Autosys, and improve performance through partitioning, caching, and query tuning. Ensure data quality, governance, and reliability while collaborating with Analytics, Product, and Engineering teams.
Location: India
Workplace: Hybrid
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Mid level

Key Responsibilities

  • •Design and build scalable data pipelines using Spark/PySpark on Databricks.
  • •Develop and optimize ETL/ELT workflows for large-scale data processing.
  • •Work with AWS EMR, S3, and the Hadoop ecosystem for distributed processing.
  • •Build and maintain data lake and data warehouse solutions, including performance tuning.
  • •Develop APIs for data ingestion/consumption, implement orchestration (Airflow/Autosys), and collaborate with cross-functional teams.

Key Requirements

  • •Minimum 5+ years of data engineering experience and strong hands-on work on Databricks and the Big Data stack.
  • •Strong expertise in Apache Spark/PySpark and SQL (PostgreSQL or similar).
  • •Hands-on experience with AWS EMR and S3, plus Hadoop/Hive ecosystem knowledge.
  • •Must-have Python experience and strong knowledge of ETL pipelines and data engineering fundamentals.
  • •Experience integrating data ingestion/consumption APIs and ensuring data quality, governance, and reliability.
Experience:5-10 yearsBig dataCloud data engineeringDistributed systems
Education:Bachelor's in Computer Science, Engineering, or related field
Skills:Data qualityGovernanceReliabilityCross-functional collaborationPerformance optimization
Tech Stack:DatabricksApache SparkPySparkSQLPostgreSQLAWS EMRS3HadoopHivePythonScalaUNIXShell scriptingETLELTAWS Step FunctionsAWS GlueAWS LambdaAWS EC2Amazon Athena

Company Brief

Brillio
Provides digital transformation and IT consulting services, including cloud migration, analytics, digital engineering, and managed services to accelerate business outcomes for enterprises across industries.
Industry: Consulting
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Funding: Private Equity Backed
Headquarters: Santa Clara, United States
Founded: 2014
WebsiteLinkedIn