Clario Logo

Clario

Lead Data Engineer (GenAI / LLM Applications)

Posted 3 Days Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Bangalore, Bengaluru Urban, Karnataka
Senior level
In-Office or Remote
Hiring Remotely in Bangalore, Bengaluru Urban, Karnataka
Senior level
Designs and maintains scalable data platforms, architectures, ETL pipelines, and complex SQL solutions across multiple databases. Develops Python applications and LLM-powered solutions using LangChain, RAG, and AWS Bedrock. Implements Airflow orchestration, data quality and governance practices, AWS deployments, CI/CD, and production support. Collaborates with data scientists, analysts, product managers, and engineering teams to translate requirements into reliable data-driven solutions and reporting.
The summary above was generated by AI

We are looking for a skilled and motivated Lead Engineer to join our Data Science and Delivery group at Clario, a part of Thermo Fisher Scientific. This role combines software development, data engineering, and analytical problem‑solving to design, build, and maintain scalable data platforms that support clinical trial operations and business intelligence. You will work across the full software development lifecycle (SDLC)—from requirements gathering through production support—collaborating closely with data scientists, analysts, product managers, and engineering teams to deliver high‑quality, data‑driven solutions.

What We Offer

  • Competitive compensation aligned with local market practices

  • Comprehensive health and wellness benefits

  • Paid time off and company holidays

  • Opportunities for professional development, learning, and career growth

  • The flexibility of working from Bangalore or remotely within India, while collaborating with global teams

What You’ll Be Doing

  • Design, develop, and maintain scalable software architectures and data pipelines that integrate with analytical and operational systems.

  • Write clean, reusable, and well‑tested Python code using frameworks such as Flask and related libraries.

  • Leverage AI‑assisted development tools, including GitHub Copilot and LangChain, to design, build, and integrate LLM‑powered solutions such as retrieval‑augmented generation (RAG) pipelines, intelligent agents, and automated workflows using AWS Bedrock or similar services.

  • Develop and optimize complex SQL across Oracle, MS SQL Server, PostgreSQL, and Snowflake, including procedures, functions, views, analytical functions, and dynamic SQL.

  • Design and implement ETL pipelines using Snowflake and related data processing technologies.

  • Implement scheduling and orchestration using Apache Airflow or similar workflow orchestration frameworks.

  • Establish and maintain data quality frameworks, versioning, and governance practices to ensure data reliability, integrity, and compliance.

  • Develop and maintain data architectures and models for both structured and unstructured data sources.

  • Troubleshoot production issues and drive continuous improvement in software quality, performance, and reliability.

  • Deploy, manage, and support solutions on AWS, including storage, compute, and pipeline services.

  • Create source‑to‑target mappings and support data and code migration initiatives.

  • Partner with stakeholders to gather requirements, translate business needs into technical solutions, and produce clear, well‑structured documentation.

  • Collaborate with product managers, analysts, and cross‑functional teams to deliver data‑driven insights and reporting using tools such as Plotly and Power BI.

What We Look For

  • Bachelor’s or higher degree in Computer Science, Information Technology, or a related technical field.

  • 5+ years of professional experience in software engineering, data engineering, or data‑focused development roles.

  • Strong proficiency in Python, including frameworks and libraries such as Django or Flask, pandas, NumPy, Plotly, and ag‑Grid.

  • Strong SQL expertise with Oracle, MS SQL Server, PostgreSQL, and/or Snowflake.

  • Proven experience writing complex SQL, including analytical and window functions, subqueries, all join types, DML/DDL/TCL statements, CASE expressions, and performance tuning.

  • Working knowledge of cloud platforms, with a preference for AWS (S3, EC2, Secrets Manager, Bedrock, Lambda).

  • Experience using AI‑assisted development tools and frameworks such as GitHub Copilot and LangChain for building LLM‑powered applications and workflows.

  • Experience with Git‑based version control systems and CI/CD pipelines.

  • Familiarity with data modeling concepts for both structured and unstructured data.

  • Strong analytical thinking, problem‑solving abilities, and communication skills.

  • Willingness to work across all phases of the SDLC, including requirements gathering, design, development, deployment, and production support.

  • Preferred experience includes exposure to the clinical trial lifecycle or clinical data management, data visualization tools (Plotly, Power BI), front‑end technologies (HTML5, CSS3, JavaScript), collaboration tools (Jira, Confluence, Microsoft Teams), and hands‑on data analysis or data cleansing using programming languages, SQL, and Excel.

At Clario, our purpose is to transform lives by unlocking better evidence. It’s a cause that unites and inspires us. It’s why we come to work—and how we empower our people to make a positive impact every day. Whether you’re starting your clinical data career or building long‑term expertise, your work helps bring life‑changing therapies to patients faster.

Similar Jobs

Yesterday
Easy Apply
Remote or Hybrid
India
Easy Apply
Senior level
Senior level
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Manage enterprise customer accounts as a trusted technical advisor, driving platform adoption, measurable business value, and technical ROI. Develop success plans, troubleshoot complex hardware, software, and API issues, manage escalations, conduct root-cause analysis, and lead technical reviews. Collaborate with Sales, Support, Product, and Engineering while communicating effectively with technical and executive stakeholders. Use AI tools to improve customer insights and outcomes, and contribute to knowledge sharing and team development.
Top Skills: AIAPIsGainsightGongInternet Of Things (Iot)JIRAPaasPythonSaaSSalesforceTableauZendesk
Yesterday
Remote or Hybrid
India
Entry level
Entry level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Tests liquidity risk and regulatory reporting releases across PRA, EBA, HKMA, and US Federal Reserve requirements. Responsibilities include SQL-based data validation, test strategy and execution, defect management, UAT dashboards, requirements gathering, data-gap analysis, issue resolution, and stakeholder communication. The role also supports finance data quality initiatives, regulatory compliance, operating-model transitions, change implementation, and coordination across Finance, Risk, Business, and IT teams.
Top Skills: ExcelMicrosoft PowerpointMicrosoft ProjectMicrosoft VisioMicrosoft WordQuality CentreSQL
Yesterday
Remote or Hybrid
India
Expert/Leader
Expert/Leader
Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Leads commercial finance partnering for East India, Nepal, and Bangladesh. Responsibilities include budgeting, forecasting, performance management, revenue and pricing analysis, trade investment and profitability management, financial governance, reporting, team leadership, stakeholder influence, and finance process transformation through standardization and automation.

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account