Senior Data Engineer - Databricks

Other Jobs To Apply

No other job posts for this day.

About SugarAI

SugarAI is redefining CRM for the age of AI.

We’re delivering on the original promise of CRM—turning fragmented customer and revenue signals into clear, prioritized action. Instead of more dashboards or surface-level insights, we help teams focus on what matters most and know exactly what to do next.

More than two decades after our founding, we’re entering a new chapter with clarity and momentum—building intelligent, intuitive solutions that work within the flow of how teams actually sell and serve. We’re focused on solving complex, real-world challenges where relationships, context, and precision make all the difference.

Our global team is united by a shared commitment to impact, ownership, and continuous growth. We create an environment where thoughtful ideas move quickly, where people are trusted to lead, and where flexibility supports how great work gets done.

If you’re excited to help shape what’s next in AI-driven CRM—and build technology that drives real outcomes—we’d love to meet you.

Where You Fit In:

The Sugar Predict platform powers revenue intelligence for mid-market enterprises by fusing ERP and CRM data into actionable insights. As a Senior Data Engineer, you will own the Databricks pipelines that make this possible, driving production reliability, cost efficiency, and platform growth through customer onboarding and legacy modernization. You will work closely with ML engineers, product teams, and the Enterprise Architecture team to ensure the data backbone behind Sugar Predict is always fast, clean, and ready to deliver at a global scale.

Impact You Will Make in the Role:

  • Own Databricks production support for the Sugar Predict data platform, including monitoring, alerting, and incident response across all production data flows
  • Maintain and report on SLA performance metrics for data pipeline delivery, ensuring visibility into platform health and accountability across internal and external stakeholders
  • Identify and implement pipeline optimizations that reduce Databricks compute costs, improve throughput, and reduce processing windows while tracking impacts through measurable KPIs
  • Migrate legacy ETL/ELT pipelines to Databricks, building automation tooling to reduce manual intervention and ensure uninterrupted data delivery during transitions
  • Support new customers onboarding by provisioning, validating, and hardening tenant data pipelines that deliver reliable, isolated data from day one
  • Design and build high-performance Databricks pipelines that ingest, transform, and serve ERP and CRM data at scale across both Azure and AWS environments
  • Own the Delta Lake architecture including schema design, partitioning strategies, data quality enforcement, and incremental processing patterns
  • Enforce data security best practices across Databricks environments, including role-based access control, secrets management, and compliance requirements for enterprise CRM and ERP data
  • Implement data quality monitoring and observability across pipeline health and ML model inputs, ensuring data integrity that directly supports Sugar Predict prediction accuracy
  • Apply and enforce multi-tenant data isolation patterns ensuring reliable, secure data delivery across Sugar Predict enterprise customers
  • Partner with the Enterprise Architecture team to ensure Sugar Predict data pipelines integrate seamlessly with the broader SugarAI product ecosystem
  • Support a globally distributed operation through on-call rotation and after-hours incident response, meeting SLAs across multiple time zones
  • Maintain technical documentation, runbooks, and architectural decision records, contributing to team knowledge sharing and operational readiness across on-call and incident response scenarios
  • Apply CI/CD best practices to data pipeline development, including version control, automated testing, and deployment tooling to ensure reliable and repeatable pipeline delivery

What You Will Bring:

  • 4+ years of data engineering experience
  • At least 2 years on Databricks or the Apache Spark ecosystem across Azure and/or AWS
  • Proficiency in PySpark, SQL, and Python with a strong track record building and operating production-grade pipelines under SLA constraints
  • Hands-on experience with Delta Lake including schema evolution, ACID transactions, optimize/vacuum lifecycle, and both incremental and streaming processing patterns
  • Hands-on experience with pipeline performance tuning and compute optimization in production Databricks environments
  • Solid working knowledge of PostgreSQL including query optimization, schema design, and use as a source or sink in production data pipelines
  • Experience supporting and maintaining legacy ETL tooling (SSIS, Informatica, custom Python/SQL pipelines, or similar) in production
  • Experience supporting large-scale multi-tenant architectures with a focus on tenant isolation, per-tenant performance, and data privacy, including navigating tools and platforms that default to single-tenant assumptions
  • Proven ability to work collaboratively across data science, product, and infrastructure teams, owning end-to-end delivery in a cross-functional environment
  • Strong understanding of data governance, security, and compliance principles, including access control, data privacy, and protection of sensitive enterprise data across multi-tenant environments

Preferred Qualifications/Experience:

  • Experience operating Databricks workspaces across both Azure and AWS, including cost governance, cluster management, and cross-cloud data access
  • Experience optimizing Databricks workloads in a Serverless environment, including compute cost governance and performance tuning for serverless compute
  • Experience with Microsoft SQL Server in a data engineering or ETL context
  • Exposure to ML feature engineering or feature stores (Databricks Feature Store, Feast, or similar) supporting predictive analytics
  • Experience with customer onboarding automation or IaC patterns for provisioning tenant data pipelines at scale
  • Databricks Certified Data Engineer Associate or Professional certification
Benefits and Perks:

Beyond a stellar work environment, friendly people, and inspiring work, we have some sweet benefits and perks:

  • Excellent healthcare package for you and your family
  • Savings and Investment – 401(k) match
  • Unlimited Paid Time Off
  • Paid Parental Leave
  • Online Legal Services (Rocket Lawyer)
  • Financial Planning Services (Origin)
  • Discounted Pet Insurance (Embrace Pet Insurance)
  • Corporate Benefit Program (Working Advantage). This benefit offers you exclusive travel and entertainment offers and special discounts that are not available to the general public
  • Health and Wellness Reimbursement Program
  • Travel Discounts
  • Educational Resources - Career & Personal Development Program
  • Employee Referral Bonus Program
  • We are a merit-based company - many opportunities to learn, excel and grow your career!
Back to blog

Common Interview Questions And Answers

1. HOW DO YOU PLAN YOUR DAY?

This is what this question poses: When do you focus and start working seriously? What are the hours you work optimally? Are you a night owl? A morning bird? Remote teams can be made up of people working on different shifts and around the world, so you won't necessarily be stuck in the 9-5 schedule if it's not for you...

2. HOW DO YOU USE THE DIFFERENT COMMUNICATION TOOLS IN DIFFERENT SITUATIONS?

When you're working on a remote team, there's no way to chat in the hallway between meetings or catch up on the latest project during an office carpool. Therefore, virtual communication will be absolutely essential to get your work done...

3. WHAT IS "WORKING REMOTE" REALLY FOR YOU?

Many people want to work remotely because of the flexibility it allows. You can work anywhere and at any time of the day...

4. WHAT DO YOU NEED IN YOUR PHYSICAL WORKSPACE TO SUCCEED IN YOUR WORK?

With this question, companies are looking to see what equipment they may need to provide you with and to verify how aware you are of what remote working could mean for you physically and logistically...

5. HOW DO YOU PROCESS INFORMATION?

Several years ago, I was working in a team to plan a big event. My supervisor made us all work as a team before the big day. One of our activities has been to find out how each of us processes information...

6. HOW DO YOU MANAGE THE CALENDAR AND THE PROGRAM? WHICH APPLICATIONS / SYSTEM DO YOU USE?

Or you may receive even more specific questions, such as: What's on your calendar? Do you plan blocks of time to do certain types of work? Do you have an open calendar that everyone can see?...

7. HOW DO YOU ORGANIZE FILES, LINKS, AND TABS ON YOUR COMPUTER?

Just like your schedule, how you track files and other information is very important. After all, everything is digital!...

8. HOW TO PRIORITIZE WORK?

The day I watched Marie Forleo's film separating the important from the urgent, my life changed. Not all remote jobs start fast, but most of them are...

9. HOW DO YOU PREPARE FOR A MEETING AND PREPARE A MEETING? WHAT DO YOU SEE HAPPENING DURING THE MEETING?

Just as communication is essential when working remotely, so is organization. Because you won't have those opportunities in the elevator or a casual conversation in the lunchroom, you should take advantage of the little time you have in a video or phone conference...

10. HOW DO YOU USE TECHNOLOGY ON A DAILY BASIS, IN YOUR WORK AND FOR YOUR PLEASURE?

This is a great question because it shows your comfort level with technology, which is very important for a remote worker because you will be working with technology over time...