Design, build, and optimize Databricks-based big data pipelines and ETL workflows using Spark, SQL, and Python/Scala. Maintain data lakehouse architectures, enforce governance and security, automate CI/CD and IaC, monitor cluster performance, and collaborate with data teams to enable AI/ML workflows.
Position Overview:
ShyftLabs is seeking a skilled Databricks Engineer to support in designing, developing, and optimizing big data solutions using the Databricks Unified Analytics Platform. This role requires strong expertise in Apache Spark, SQL, Python, and cloud platforms (AWS/Azure/GCP). The ideal candidate will collaborate with cross-functional teams to drive data-driven insights and ensure scalable, high-performance data architectures.
ShyftLabs is a growing data product company that was founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.
Job Responsiblities
- Design, implement, and optimize big data pipelines in Databricks.
- Develop scalable ETL workflows to process large datasets.
- Leverage Apache Spark for distributed data processing and real-time analytics.
- Implement data governance, security policies, and compliance standards.
- Optimize data lakehouse architectures for performance and cost-efficiency.
- Collaborate with data scientists, analysts, and engineers to enable advanced AI/ML workflows.
- Monitor and troubleshoot Databricks clusters, jobs, and performance bottlenecks.
- Automate workflows using CI/CD pipelines and infrastructure-as-code practices.
- Ensure data integrity, quality, and reliability in all pipelines.
Basic Qualifications
- Bachelor’s or Master’s degree in Computer Science, Data Engineering, or a related field.
- 3+ years of hands-on experience with Databricks and Apache Spark.
- Proficiency in SQL, Python, or Scala for data processing and analysis.
- Experience with cloud platforms (AWS, Azure, or GCP) for data engineering.
- Strong knowledge of ETL frameworks, data lakes, and Delta Lake architecture.
- Experience with CI/CD tools and DevOps best practices.
- Familiarity with data security, compliance, and governance best practices.
- Strong problem-solving and analytical skills with an ability to work in a fast-paced environment.
Preferred Qualifications
- Databricks certifications (e.g., Databricks Certified Data Engineer, Spark Developer).
- Hands-on experience with MLflow, Feature Store, or Databricks SQL.
- Exposure to Kubernetes, Docker, and Terraform.
- Experience with streaming data architectures (Kafka, Kinesis, etc.).
- Strong understanding of business intelligence and reporting tools (Power BI, Tableau, Looker).
- Prior experience working with retail, e-commerce, or ad-tech data platforms.
We are proud to offer a competitive salary alongside a strong insurance package. We pride ourselves on the growth of our employees, offering extensive learning and development resources.
ShyftLabs Hyderabad, Telangana, IND Office
Hyderabad, Telangana, India
ShyftLabs Hyderabad, Telangana, IND Office
Hyderabad, India
Similar Jobs
Financial Services
Leads forward-deployed engineering engagements delivering enterprise data solutions on Databricks and Snowflake. Owns delivery strategy, operating models, architecture, governance, privacy, security, compliance, budgets, vendors, metrics, and executive communications. Partners with product and technology leaders to develop reusable platform capabilities, while hiring, coaching, and managing senior engineering teams. The role requires enterprise-scale cloud data platform expertise across multi-cloud and hybrid environments, including AWS, data governance, metadata, lineage, retention, and regulatory controls.
Top Skills:
Amazon S3AWSCi/CdDatabricksSnowflake
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design, build, and operate scalable cloud-native ETL/ELT pipelines and data platforms (Databricks, Snowflake) ensuring data quality, governance, performance, and enablement of analytics and AI/ML workloads. Partner with stakeholders, apply data modeling and transformations, modernize legacy ETL, follow CI/CD and secure engineering practices, and mentor junior engineers.
Top Skills:
AzureAzure DevopsCi/CdDatabricksGitGithub CopilotPysparkPythonSnowflakeSQLStructured Streaming
Financial Services
Build and maintain a secure, scalable KYC and risk-assessment data platform. Develop production-quality data-intensive code, create reusable frameworks, advise cross-functional teams, apply cloud-native and big-data tooling, and use SDLC and enterprise AI-assisted development to automate and improve delivery.
Top Skills:
Ai-Assisted Development ToolsAirflowApi DesignAWSAzureData Lake ArchitecturesDatabricksDynatraceGCPGrafanaJavaKafkaMemcachedMicroservicesNoSQLObservability ToolsPysparkPythonRedisSparkSplunkSql (Relational Databases)TemporalVector Stores
What you need to know about the Hyderabad Tech Scene
Because of its proximity to leading research institutions and a government committed to the city's growth, Hyderabad's tech scene is booming. With plans to establish India's first "AI city," the city is on track to become one of the world's most anticipated tech hubs, with companies like TransUnion, Schrödinger and Freshworks, among others, already calling the city home.


