Role Overview
Sublime Data is a boutique data engineering and AI consulting firm, registered Databricks partner, building specialized expertise in data platform modernization and agentic AI implementation for clients.
Responsibilities
- Design, build, and maintain data pipelines using Databricks (PySpark, Delta Lake, Delta Live Tables, Workflows)
- Work directly with client stakeholders to understand requirements, provide updates, and troubleshoot issues — this is a client-facing role, not purely back-office delivery
- Participate in technical discovery and client interviews as part of our presales and onboarding process
- Collaborate with our architecture team on solution design for new client engagements
- Support data migration and modernization projects across cloud platforms (Azure primary, AWS/GCP as needed)
- Travel to client locations as required for project kickoffs, workshops, or on-site delivery phases
Requirements
- 2-5 years of total data engineering experience, with at least 2 years of hands-on Databricks experience (PySpark, SQL, Delta Lake)
- Databricks certification (Data Engineer Associate/Professional) is a strong plus, not mandatory
- Strong SQL and Python fundamentals
- Exposure to at least one other data platform — Snowflake, Microsoft Fabric, or native AWS/GCP data services — is a genuine advantage
- Comfortable communicating directly with clients — confident in explaining technical work to non-technical or semi-technical stakeholders
- Ability to join within 15 days or immediately
- Based in or willing to relocate to Ahmedabad — this is an in-office role, not remote
- Willingness to travel for client engagements as needed
Skills
- Databricks
- PySpark
- Python
- SQL
- Delta Lake
Nice to Have
- Experience with Unity Catalog, MLflow, or Databricks Workflows specifically
- Prior client-facing consulting or services company experience
- Exposure to any GenAI/LLM tooling (LangChain, RAG pipelines)
- Experience in BFSI, healthcare/pharma, or energy sector data projects