Design, build, and optimize Databricks-based big data pipelines and ETL workflows using Spark, SQL, and Python/Scala. Maintain data lakehouse architectures, enforce governance and security, automate CI/CD and IaC, monitor cluster performance, and collaborate with data teams to enable AI/ML workflows.
Position Overview:
ShyftLabs is seeking a skilled Databricks Engineer to support in designing, developing, and optimizing big data solutions using the Databricks Unified Analytics Platform. This role requires strong expertise in Apache Spark, SQL, Python, and cloud platforms (AWS/Azure/GCP). The ideal candidate will collaborate with cross-functional teams to drive data-driven insights and ensure scalable, high-performance data architectures.
ShyftLabs is a growing data product company that was founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.
Job Responsiblities
- Design, implement, and optimize big data pipelines in Databricks.
- Develop scalable ETL workflows to process large datasets.
- Leverage Apache Spark for distributed data processing and real-time analytics.
- Implement data governance, security policies, and compliance standards.
- Optimize data lakehouse architectures for performance and cost-efficiency.
- Collaborate with data scientists, analysts, and engineers to enable advanced AI/ML workflows.
- Monitor and troubleshoot Databricks clusters, jobs, and performance bottlenecks.
- Automate workflows using CI/CD pipelines and infrastructure-as-code practices.
- Ensure data integrity, quality, and reliability in all pipelines.
Basic Qualifications
- Bachelor’s or Master’s degree in Computer Science, Data Engineering, or a related field.
- 3+ years of hands-on experience with Databricks and Apache Spark.
- Proficiency in SQL, Python, or Scala for data processing and analysis.
- Experience with cloud platforms (AWS, Azure, or GCP) for data engineering.
- Strong knowledge of ETL frameworks, data lakes, and Delta Lake architecture.
- Experience with CI/CD tools and DevOps best practices.
- Familiarity with data security, compliance, and governance best practices.
- Strong problem-solving and analytical skills with an ability to work in a fast-paced environment.
Preferred Qualifications
- Databricks certifications (e.g., Databricks Certified Data Engineer, Spark Developer).
- Hands-on experience with MLflow, Feature Store, or Databricks SQL.
- Exposure to Kubernetes, Docker, and Terraform.
- Experience with streaming data architectures (Kafka, Kinesis, etc.).
- Strong understanding of business intelligence and reporting tools (Power BI, Tableau, Looker).
- Prior experience working with retail, e-commerce, or ad-tech data platforms.
We are proud to offer a competitive salary alongside a strong insurance package. We pride ourselves on the growth of our employees, offering extensive learning and development resources.
ShyftLabs Hyderabad, Telangana, IND Office
Hyderabad, Telangana, India
ShyftLabs Hyderabad, Telangana, IND Office
Hyderabad, India
Similar Jobs
Financial Services
Design, develop, and maintain secure, scalable full‑stack applications and data platform solutions. Build production code, troubleshoot complex issues, create architecture/design artifacts, leverage cloud and Databricks, apply CI/CD and secure coding practices, and use enterprise AI-assisted development tools responsibly to improve risk reporting systems.
Top Skills:
Ai-Assisted Development ToolsAmazon Web Services (Aws)AngularC#Ci/CdDatabricksJavaPivotal Cloud Foundry (Pcf)ReactSQLTableau
Healthtech • Biotech • Pharmaceutical
Lead design, implementation, and optimization of Azure data solutions. Administer Databricks, build scalable Data Lake and database architectures, implement IaC (Terraform/ARM/Bicep), create ADF pipelines, drive cloud migrations, enforce security/compliance, mentor teams, and optimize performance and CI/CD automation.
Top Skills:
Arm TemplatesAzureAzure BicepAzure Cosmos DbAzure Data FactoryAzure Data LakeAzure DatabricksAzure MigrateAzure Sql DatabaseAzure Sql Managed InstanceAzure SynapseCi/Cd PipelinesData Migration AssistantDatabricks ClustersDatabricks NotebooksTerraform
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and maintain data pipelines using the Azure ecosystem. Optimize streaming and batch processing with Databricks and manage Medallion Architecture. Collaborate with stakeholders to ensure data quality and support analytics use cases. Implement testing automation and best practices for governance in cloud-based systems.
Top Skills:
SparkAzure Data FactoryAzure Data LakeAzure DatabricksAzure SynapsePower BIPythonScalaSQL
What you need to know about the Hyderabad Tech Scene
Because of its proximity to leading research institutions and a government committed to the city's growth, Hyderabad's tech scene is booming. With plans to establish India's first "AI city," the city is on track to become one of the world's most anticipated tech hubs, with companies like TransUnion, Schrödinger and Freshworks, among others, already calling the city home.



