Build and maintain Snowflake and SQL Server data pipelines for an operational data store, including ingestion, transformation, schema modeling, migrations, CI/CD, data quality testing, and legacy batch-to-API modernization. Collaborate with domain teams and architects to support customer, loan, payment, and interaction data. Ensure data accuracy, lineage, observability, and performance while applying generative AI tools to accelerate development and testing.
-
Key Responsibilities
-
- Assess existing SQL Server (Atlas) and Snowflake environments to identify pipeline gaps, data quality issues, and structural inefficiencies
-
- Build and maintain data pipelines supporting the Operational Data Store (ODS), including ingestion, transformation, and loading of key business entities
-
- Implement, version-control, and enforce data dictionary standards, logical/physical entity models, and naming conventions established by the Data Architect
-
- Design, implement, and maintain CI/CD pipelines (e.g., GitHub Actions, GitLab CI) to automate the deployment of Snowflake schemas, database structures, and pipeline code
-
- Manage database schema migrations as code using tools such as dbt, Schemachange, Flyway, or Terraform
-
- Execute the batch-to-API migration roadmap — re-engineering Informatica/EDW batch flows to consume FDR APIs
-
- Implement automated data quality frameworks, testing validations (e.g., dbt tests, Great Expectations), and data parity reconciliation routines between legacy and new ODS environments
-
- Collaborate with domain teams to engineer pipelines for customer, loan, payment, and interaction data entities
-
- Apply Snowflake best practices in structuring schemas, managing compute, and separating analytical from operational workloads
-
- Partner with the internal data architect and downstream consumers to ensure data accuracy, lineage, and observability across the platform
-
- Leverage generative AI tools (e.g., coding assistants, LLMs) to accelerate pipeline development, optimize query performance, write documentation, and generate automated tests
-
Must-Have Skills
-
- 7+ years in data engineering roles with a strong focus on pipeline development, CI/CD setup, and database modeling
-
- Strong programming skills in SQL (Sql Server and Postgres) and Python for data engineering, automation, and pipeline development
-
- Proficiency in Snowflake — including schema design, performance tuning, and data loading patterns
-
- Hands-on experience building DevOps/DataOps pipelines (Git, CI/CD tools, automated deployment strategies)
-
- Hands-on experience with database migration tools / schema management systems (e.g., dbt, Schemachange, Flyway)
-
- Hands-on experience with ETL/ELT tools and batch processing (Informatica experience strongly preferred)
-
- Solid understanding of ODS concepts, dimensional modeling, and managing data dictionaries
-
- Experience integrating with REST/API-based data sources as part of modernization or migration efforts
-
- Demonstrated experience using Generative AI coding assistants (e.g., Copilot, ChatGPT, Gemini) to enhance coding productivity, write tests, and troubleshoot complex SQL/Python scripts
-
- Background in financial services or lending domain
-
Nice to Have
-
- Experience with modern orchestrators (e.g., Apache Airflow, Prefect, Dagster) or Snowflake-native orchestration (Tasks, Dynamic Tables)
-
- Familiarity with deploying semantic layers or metadata-driven catalog systems (Collibra, Alation) to support downstream AI/ML and Natural Language query agents
-
- Experience working with FDR or similar loan servicing platforms
-
- Familiarity with data mesh principles and domain-oriented data ownership
-
- Exposure to AI/ML data pipeline requirements and feature engineering
-
- Experience with data governance and cataloguing tools (Collibra, Alation, etc.)
-
- Background in student lending or consumer finance
Photon Chennai, Tamil Nadu, IND Office
DLF IT Park 1/124 Mount Poonamallee Road Sivaji Gardens Manapakkam , Chennai, India, 600089
Similar Jobs
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads complex, cross-functional scientific learning projects for Medical Affairs. Develops curricula, e-learning, and blended learning programs; partners with medical, scientific, agency, and learning systems teams; manages multiple projects, timelines, budgets, and priorities; evaluates emerging e-learning technologies; and updates existing learning resources across therapeutic areas.
Top Skills:
Articulate StorylineDigital Learning TechnologyE-Learning Authoring ToolsLearning Management Systems
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Lead technicians to install, maintain, calibrate, and troubleshoot semiconductor lab and manufacturing equipment. Ensure ISO/IEC 17025 calibration compliance, manage equipment tracking, drive RCA/RCCA and preventive actions, coordinate vendors and cross-functional teams, support audits, and provide training while maintaining safety, ESD, and cleanroom standards.
Top Skills:
3D X-RayArtificial IntelligenceBend TesterCleanroomCsamEquipment Tracking SystemEsdFibFtirHast ChamberIso/Iec 17025KaizenLean Six SigmaLinuxReflow OvenSemShock TesterSoak ChamberTemp CycleTemperature ChamberTesterThbWindowsX-Section
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Lead operation, maintenance, troubleshooting, and continuous improvement of facility gas generation, storage, and distribution systems (LN2, N2, CO2, O2, argon, helium, hydrogen, CDA). Ensure gas purity, availability, safety compliance (EHS, LOTO, PTW), perform preventive and corrective maintenance, support commissioning/expansion, maintain documentation, and use PLC/BMS/SCADA and CMMS/SAP tools. Proficiency with GenAI/Copilot tools is expected.
Top Skills:
Agent BotsBmsCmmsCo2 Supply SystemsCopilotCopilot LibraryCryogenic Tank & Vaporizer SystemsGas Cabinets & VmbsGas Detection SystemsGen Ai ToolsNitrogen Generation SystemsOxygen Distribution SystemsP&IdPlcSAPScadaSpecialty Gas Panels & Manifolds
What you need to know about the Chennai Tech Scene
To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.


.jpeg)