This page was automatically translated and may contain errors. View in English.
C

Databricks Engineer

CGI

Hyderabad, Telangana, India · 全职

抢先申请

经验
3年以上
薪水
INR 900,000 – INR 1,700,000 / year
职位空缺
1
发布
2小时前
工作模式
在办公室
学历
任何毕业生
合格
Any graduate degree holder can apply for this position.
恢复
需要申请

你的工作地点

职位描述

Role Overview

As a Databricks Engineer, you will be responsible for designing, building, and maintaining large-scale ETL/ELT data pipelines utilizing Databricks and PySpark technologies. This role involves transforming and processing significant datasets through Python and Spark to support scalable and efficient data workflows. You will engage with complex SQL query development for data extraction, transformation, validation, and reporting purposes. Collaborating with various stakeholders, including data analysts and business teams, to gather requirements is a key aspect of the position.

Key Responsibilities

  • Develop and sustain scalable ETL/ELT pipelines using Databricks and PySpark.
  • Implement data processing solutions using Python and Apache Spark for large-scale transformations.
  • Create advanced SQL queries for data extraction, validation, and reporting.
  • Enhance data workflows focusing on performance, scalability, and reliability.
  • Handle structured and semi-structured data from diverse sources.
  • Collaborate with analysts and stakeholders to understand and fulfill data needs.
  • Conduct data quality assessments, troubleshoot issues, and perform root cause analyses.
  • Optimize Spark jobs and SQL queries to improve execution times.
  • Adhere to coding standards and best practices including documentation and version control.
  • Engage in code reviews and assist with production environment deployments.

Required Skills and Experience

  • At least 3 years of experience in data engineering roles.
  • Proficient hands-on experience with the Databricks platform.
  • Strong working knowledge of PySpark and Apache Spark frameworks.
  • Solid programming capability in Python.
  • Advanced SQL proficiency, including complex joins, window functions, and query performance optimization.
  • Experience in constructing ETL/ELT pipelines.
  • Understanding of data warehousing principles and data modeling techniques.
  • Familiarity with Git or comparable version control systems.
  • Excellent analytical skills along with problem-solving and debugging expertise.

Eligibility Criteria

Candidates holding any graduate degree are eligible to apply.

如果您希望收到回复,请留下您的信息——我们不会将您的信息用于其他用途。

点击浏览拖放,或 粘贴 截图

PNG、JPG、GIF、MP4、WebM、MOV 格式 · 每个文件最大 20MB · 最多 5 个文件

🤖
在线·即时人工智能帮助