Designs, builds, and maintains scalable data pipelines, ETL processes, data architectures, and integrations. Ensures data quality, governance, security, compliance, and accessibility while monitoring performance and troubleshooting issues. Collaborates with data scientists, analysts, and engineers to define requirements and deliver reliable solutions. Participates in code reviews, documentation, data modeling, schema design, and CI/CD implementation for data workflows.
We
are seeking a detail-oriented and highly motivated Data Engineer to join our
growing Data & Analytics team.
In this role, you will be responsible for designing, building, and maintaining
robust data pipelines and infrastructure that power insights across the
organization. You’ll work closely with data scientists, analysts, and engineers
to ensure the integrity, accessibility, and scalability of our data systems.
Key
Responsibilities:
- Design, develop, and maintain scalable
data pipelines and ETL processes.
- Build and optimize data architecture to
ensure data quality and consistency.
- Integrate data from diverse internal
and external sources.
- Collaborate with cross-functional teams
to define data requirements and deliver solutions.
- Implement best practices for data
governance, security, and compliance.
- Monitor pipeline performance and
perform real-time troubleshooting of data issues.
- Participate in code reviews and
contribute to documentation and standards.
Required
Qualifications & Skills:
- Bachelor’s degree in computer science,
Engineering, or a related field (or equivalent practical experience).
- 5+ years of professional experience in
data engineering.
- Solid understanding of SQL and
proficiency in at least one programming language, such as Python, Java, or
Scala.
- Practical experience building and
maintaining data pipelines using tools like Apache Airflow or DBT.
- Hands-on experience with cloud
platforms (AWS, GCP, Azure) and data warehousing solutions (Redshift, BigQuery,
Snowflake).
- Familiarity with big data technologies
and frameworks, including Spark, Kafka, and Hadoop.
- Demonstrated ability to solve complex
problems with a strong focus on detail.
- Experience implementing CI/CD practices
for data workflows.
- Working knowledge of data modelling
principles and schema design.
- Exposure to machine learning pipelines
or real-time analytics systems is a plus
Similar Jobs
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Design and build data services and pipelines supporting machine learning products. Develop automation tools for model deployment, maintain production data systems, improve infrastructure, participate in code reviews, and collaborate across engineering, data science, and product teams. The role requires expertise in distributed systems, large-scale data processing, CI/CD, container orchestration, and AI-enabled workflow improvements.
Top Skills:
AirflowAWSAws BatchCi/CdDockerEmrGlueGoKafkaKubernetesKv StoresLinuxPythonRelational DatabasesSagemakerSparkSpinnaker
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Provide L1-L3 production support for enterprise MongoDB Atlas environments, including monitoring, scaling, backups, upgrades, performance tuning, security, networking, disaster recovery, and incident resolution. Build Terraform-based infrastructure automation, support multi-cloud deployments across Azure, AWS, and GCP, partner with application and engineering teams, coordinate vendor escalations, and participate in 24/7 on-call coverage.
Top Skills:
AWSAzureGCPGitGithub ActionsGoJavaKubernetesMongodb AtlasPowershellPythonShellSQL ServerTerraform
Automotive
Designs and maintains scalable ETL pipelines, data models, databases, and cloud-based analytics workflows. Manages PostgreSQL, MySQL, MongoDB, and Google Cloud SQL environments, performs database optimization, develops AutoML solutions, and deploys data engineering pipelines to production. Works within Agile teams using Scrum or Kanban, participating in sprint planning, backlog grooming, and daily stand-ups.
Top Skills:
AutomlData ModelingETLGoogle BigqueryGoogle Cloud DataflowGoogle Cloud DataprocGoogle Cloud PlatformGoogle Cloud SqlGoogle Cloud StorageKanbanMongoDBMySQLPostgresScrum
What you need to know about the Chennai Tech Scene
To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.



