ShyftLabs Logo

ShyftLabs

Lead Data Engineer (Databricks)

Posted 8 Days Ago
Be an Early Applicant
Hybrid
Coimbatore, Tamil Nadu
Expert/Leader
Hybrid
Coimbatore, Tamil Nadu
Expert/Leader
Lead the design, development, and optimization of scalable ETL/ELT pipelines on Databricks using Python, PySpark, and SQL. Integrate databases, files, Amazon S3, and REST APIs; implement data transformations, dimensional models, Delta Lake Medallion Architecture, and Unity Catalog. Manage Databricks Jobs and Workflows, ensure data quality, tune Spark performance, and collaborate cross-functionally. Guide engineers and drive technical decisions in a lead capacity.
The summary above was generated by AI
Position Overview

We are looking for a Data Engineer with hands-on experience in building scalable data pipelines and data engineering solutions on the Databricks Lakehouse Platform. The ideal candidate should have strong expertise in Python, PySpark, SQL, Databricks, AWS, and REST API integrations for data ingestion, managing large volumes of data, and data export
ShyftLabs is a growing data product company that was founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.
 

Job Responsibilities:

    Design, develop, and maintain scalable ETL/ELT pipelines using Databricks,
    PySpark, and SQL.
    ● Integrate data from multiple sources, including databases, Amazon S3, files, and REST APIs.
    ● Build data pipelines with Databricks Unity Catalog.
    ● Implement business logic, data transformations, and dimensional data models.
    ● Create, schedule, monitor, and optimize Databricks Jobs and Workflows.
    ● Design and manage Delta Lake tables using Medallion Architecture (Bronze, Silver,Gold).
    ● Ensure data quality through validations, error handling, logging, and monitoring.
    ● Optimize Spark workloads for performance, scalability, and reliability.
    ● Collaborate with cross-functional teams to deliver production-ready data solutions.

Basic Qualification:

    Strong expertise in Python, PySpark, and Advanced SQL.
    ● Hands-on experience with the Databricks Lakehouse Platform.
    ● Good understanding of Unity Catalog, Delta Lake, Databricks Workflows/Jobs,
    Clusters, Notebooks, Repos, and Medallion Architecture.
    ● Experience integrating with REST APIs for data ingestion and data export.
    ● Strong knowledge of ETL/ELT development, batch processing, incremental loading,
    and data transformation.
    ● Experience with data modeling (Star Schema, Snowflake Schema, Fact & Dimension
    tables, SCD concepts).
    ● Understanding of data warehousing concepts and best practices.
    ● Experience working with structured and semi-structured data (CSV, JSON, Parquet,
    Delta).
    ● Knowledge of partitioning, file optimization, Spark performance tuning, and query
    optimization.
    ● Experience with Git and CI/CD best practices

Preferred Qualifications:

  • 9+ years of experience in Data Engineering, including 3+ years of hands-on experience with Databricks.
  • Prior experience in a Lead Data Engineer / Technical Lead role, with experience guiding engineers and driving technical decisions.
  • Strong hands-on experience with Databricks, Apache Spark, and SQL.
  • Experience designing, developing, and optimizing ETL/ELT data pipelines.
  •  Experience with Auto Loader, Spark Declarative pipelines, Kafka, Airflow, or dbt is a plus. 
  • Databricks certification is an added advantage. Give me Jd for lead role 

We are proud to offer a competitive salary alongside a strong insurance package. We pride ourselves on the growth of our employees, offering extensive learning and development resources.

Similar Jobs

Yesterday
Easy Apply
Remote or Hybrid
India
Easy Apply
Senior level
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Manage enterprise customer accounts as a trusted technical advisor, driving platform adoption, measurable business value, and technical ROI. Develop success plans, troubleshoot complex hardware, software, and API issues, manage escalations, conduct root-cause analysis, and lead technical reviews. Collaborate with Sales, Support, Product, and Engineering while communicating effectively with technical and executive stakeholders. Use AI tools to improve customer insights and outcomes, and contribute to knowledge sharing and team development.
Top Skills: AIAPIsGainsightGongInternet Of Things (Iot)JIRAPaasPythonSaaSSalesforceTableauZendesk
Yesterday
Remote or Hybrid
India
Entry level
Entry level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Tests liquidity risk and regulatory reporting releases across PRA, EBA, HKMA, and US Federal Reserve requirements. Responsibilities include SQL-based data validation, test strategy and execution, defect management, UAT dashboards, requirements gathering, data-gap analysis, issue resolution, and stakeholder communication. The role also supports finance data quality initiatives, regulatory compliance, operating-model transitions, change implementation, and coordination across Finance, Risk, Business, and IT teams.
Top Skills: ExcelMicrosoft PowerpointMicrosoft ProjectMicrosoft VisioMicrosoft WordQuality CentreSQL
Yesterday
In-Office
Chennai, Tamil Nadu, IND
Expert/Leader
Expert/Leader
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads global logistics data governance, accuracy, stewardship, and data quality initiatives. Manages Data Operations analysts, develops data strategy and quality metrics, validates data models, and supports planning and analytics. Partners with digital, supply chain, manufacturing, finance, logistics, and external freight-forwarding teams in a global matrix. Oversees recruitment, coaching, change adoption, project alignment, IMEX standards, and continuous improvement across transactional, planning, and master data.
Top Skills: Business Intelligence ToolsDatabase ManagementDataikuETLPower BIPythonRelational DatabasesSQLTableau

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account