YipitData Logo

YipitData

Data Engineer - Global Team (India)

Posted 13 Days Ago
Remote
Hiring Remotely in India
Senior level
Remote
Hiring Remotely in India
Senior level
Build and maintain end-to-end data pipelines using PySpark and SQL on Databricks/Delta. Establish data modeling and pipeline best practices, create AI-ready analytical datasets, troubleshoot pipeline issues, and collaborate with stakeholders to embed business logic and optimize reliability, efficiency, and performance.
The summary above was generated by AI

About Us:

YipitData is the leading market research and analytics firm for the disruptive economy and most recently raised $475M from The Carlyle Group at a valuation of over $1B. Every day, our proprietary technology analyzes billions of alternative data points to uncover actionable insights across sectors like software, AI, cloud, e-commerce, ridesharing, and payments.

Our data and research teams transform raw data into strategic intelligence, delivering accurate, timely, and deeply contextualized analysis that our customers—ranging from the world’s top investment funds to Fortune 500 companies—depend on to drive high-stakes decisions. From sourcing and licensing novel datasets to rigorous analysis and expert narrative framing, our teams ensure clients get not just data, but clarity and confidence.

We operate globally with offices in the US (NYC, Austin, Miami, Mountain View), APAC (Hong Kong, Shanghai, Beijing, Guangzhou, Singapore), and India. Our award-winning, people-centric culture—recognized by Inc. as a Best Workplace for three consecutive years—emphasizes transparency, ownership, and continuous mastery.

What It’s Like to Work at YipitData:

YipitData isn’t a place for coasting—it’s a launchpad for ambitious, impact-driven professionals. From day one, you’ll take the lead on meaningful work, accelerate your growth, and gain exposure that shapes careers.

Why Top Talent Chooses YipitData:

  • Ownership That Matters: You’ll lead high-impact projects with real business outcomes
  • Rapid Growth: We compress years of learning into months
  • Merit Over Titles: Trust and responsibility are earned through execution, not tenure
  • Velocity with Purpose: We move fast, support each other, and aim high—always with purpose and intention

If your ambition is matched by your work ethic—and you're hungry for a place where growth, impact, and ownership are the norm—YipitData might be the opportunity you’ve been waiting for.

About The Role:

We are seeking a highly skilled Data Engineer to join our dynamic Data Engineering team. The ideal candidate possesses 4-7 years of data engineering experience. An excellent candidate should have a solid understanding of Spark.Pyspark, and SQL, and have data pipeline experience. Hired individuals will play a crucial role in helping to build out our data engineering team to support one of our strategic pipelines and optimize for reliability, efficiency, and performance.

Additionally, Data Engineering serves as the gold standard for all other YipitData analyst teams, building and maintaining the core pipelines and tooling that power our products. This high-impact, high-visibility team is instrumental to the success of our rapidly growing business. This is a unique opportunity to be the first hire in this team, with the potential to build and lead the team as their responsibilities expand.

This is a remote opportunity based in India. 

During training and onboarding, we will expect several hours of overlap with US working hours.Afterward, standard IST working hours are permitted with the exception of 1-2 days per week, when you will join meetings with the US team.

As Our Data Engineer You Will:

  • Report directly to a Manager of Data Engineering, who will provide significant, hands-on training on cutting-edge data tools and techniques.
  • Build and maintain end-to-end data pipelines.
  • Help with setting best practices for our data modeling and pipeline builds.
  • Create AI-ready analytical datasets designed with the structure, metadata, documentation, and business context needed for effective use by AI agents and insight-driven applications.
  • Become an expert at solving complex data pipeline issues using PySpark and SQL.
  • Collaborate with stakeholders to incorporate business logic into our central pipelines.
  • Deeply learn Databricks, Spark, AI, and other ETL toolings developed internally.

You Are Likely To Succeed If:

  • You hold a Bachelor’s or Master’s degree in Computer Science, STEM, or a related technical discipline.
  • You have 4+ years of experience as a Data Engineer or in other technical functions.
  • You are excited about solving data challenges and learning new skills.
  • You have a great understanding of working with data or building data pipelines.
  • You are comfortable working with large-scale datasets using PySpark, Delta, and Databricks.
  • You understand business needs and the rationale behind data transformations to ensure alignment with organizational goals and data strategy.
  • You are eager to constantly learn new technologies.
  • You are a self-starter who enjoys working collaboratively with stakeholders.
  • You have exceptional verbal and written communication skills.
  • Nice to have: Experience with Airflow, dbt, Snowflake, or equivalent.

What We Offer:

Our compensation package includes comprehensive benefits, perks, and a competitive salary: 

  • We care about your personal life and we mean it. We offer vacation time, parental leave, team events, learning reimbursement, and more!
  • Your growth at YipitData is determined by the impact that you are making, not by tenure, unnecessary facetime, or office politics. Everyone at YipitData is empowered to learn, self-improve, and master their skills in an environment focused on ownership, respect, and trust.

We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, marital status, disability, gender, gender identity or expression, or veteran status. We are proud to be an equal-opportunity employer.

Job Applicant Privacy Notice



Similar Jobs

54 Minutes Ago
Remote
India
Senior level
Senior level
Cloud • Information Technology • Productivity • Software • Automation
The role involves monitoring and resolving customer integration issues on the Boomi platform, troubleshooting errors, and maintaining customer relationships while providing technical support.
Top Skills: As2BoomiEdiEdifactFlat FilesGraphQLGroovyJavaJavaScriptJSONMs SqlMySQLOftpPostgresPythonRestSftpSoapX12XML
An Hour Ago
Remote or Hybrid
India
Senior level
Senior level
Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Serve as the ledger Subject Matter Expert driving GL and sub-ledger integration, accounting treatments, ledger process mapping, reconciliations, controls and period-end close. Translate finance requirements into functional specifications, partner with technology on target ledger architecture, support audits and governance, mentor finance team members, and approve functional readiness for releases.
An Hour Ago
In-Office or Remote
Senior level
Senior level
Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Lead technical pre-sales engagements for large enterprise customers in APAC (India), build trusted advisor relationships, position Atlassian solutions to meet customer goals, collaborate with sales and partners, and drive revenue outcomes while staying current on product and technical knowledge.
Top Skills: Atlassian ProductsDevOpsItsmScaled Agile

What you need to know about the Chennai Tech Scene

To locals, it's no secret that South India is leading the charge in big data infrastructure. While the environmental impact of data centers has long been a concern, emerging hubs like Chennai are favored by companies seeking ready access to renewable energy resources, which provide more sustainable and cost-effective solutions. As a result, Chennai, along with neighboring Bengaluru and Hyderabad, is poised for significant growth, with a projected 65 percent increase in data center capacity over the next decade.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account