Senior Data Engineer

Cloudera
Budapest
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 5+ yearsEducation: bachelorsSkills: ["Python","SQL","Spark","Hadoop","Kafka","Hive","Impala","Cursor","GitHub Copilot","Gemini","RAG","LLM","Airflow","NiFi","Parquet","Avro","HDFS","Kudu","HBase"]

Senior Data Engineer at Cloudera responsible for building scalable data pipelines and GenAI-enabled data platforms. You’ll lead AI-first development with Vibe Coding, design real-time and batch data flows, and deploy self-service GenAI tools on Cloudera’s native platform, enabling business users to resolve data needs quickly while ensuring data integrity.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cloudera
Cloudera
3 months ago

Senior Data Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 15 days agoStatus: Live

Job Summary

Senior Data Engineer at Cloudera responsible for building scalable data pipelines and GenAI-enabled data platforms. You’ll lead AI-first development with Vibe Coding, design real-time and batch data flows, and deploy self-service GenAI tools on Cloudera’s native platform, enabling business users to resolve data needs quickly while ensuring data integrity.
Location: Budapest
Workplace: Hybrid
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Manager level

Key Responsibilities

  • •Collaborate with Data Architects, Operational Architects, and Data Analysts to understand data and operational requirements across business units.
  • •Partner with data owners to ensure reliable data ingestion for traditional analytics and GenAI applications.
  • •Master Vibe Coding and AI-orchestrated development to accelerate data pipelines and GenAI apps, reducing end-to-end timelines.
  • •Develop and implement data transformations following specifications and AI-first workflows.
  • •Design robust architectures for real-time, near real-time, and batch data processing to meet complex business needs.
  • •Design and deploy GenAI-powered Self-Service tools, including automated documentation generators and natural language interfaces.
  • •Implement monitoring and CI/CD automation to ensure data quality and reliability of AI-supported data services.
  • •Standardize AI-first engineering workflows to ensure high-quality, well-documented code delivery.

Pay and Benefits

Perks:Paid LeaveRemote WorkLearning BudgetWellness StipendVolunteer Time

Key Requirements

  • •5+ years of experience as a Data Engineer.
  • •Proven experience with AI-first approaches and Vibe Coding with production-ready data pipelines.
  • •Proficiency with AI-assisted coding tools such as Cursor, GitHub Copilot, or Gemini.
  • •Strong system design skills for batch and real-time/streaming data processing.
  • •Proficiency in Python and SQL with ETL experience
Experience:5+ yearsData engineeringBig dataAIGenAIETL
Education:Bachelor's
Skills:PythonSQLSparkHadoopKafkaHiveImpalaCursorGitHub CopilotGeminiRAGLLMAirflowNiFiParquetAvroHDFSKuduHBase
Tech Stack:PythonSQLSparkHadoopKafkaHiveImpalaCursorGitHub CopilotGeminiAirflowNiFiParquetAvroHDFSKuduHBaseRAGLLMSpark Streaming

Company Brief

Cloudera
Provides a hybrid data platform for managing, analyzing, and securing enterprise data across cloud and on-premises environments, enabling machine learning, analytics, and data engineering at scale.
Industry: Data Infrastructure
Company Size: Enterprise (1,001+ employees)
Revenue: USD 500M to 1B
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 2008
Glassdoor
Glassdoor: 3.8
WebsiteLinkedIn