Search Jobs

Search by job, company or skills

Databricks Data Engineer

Databricks Data Engineer

RIDIK Pte Ltd
7-8 Years
SGD 9,500 - 10,500 per month
Quick Apply
  • Posted a month ago
  • Over 50 applicants have applied

Job Description

Essential Technical Skills

  • Data Engineering: Strong foundation in data engineering principles, ETL/ELT processes, and data pipeline design patterns
  • PySpark: Proven hands-on experience developing data pipelines using PySpark, including DataFrames API, Spark SQL, and performance optimization
  • Databricks Platform: Practical experience with Databricks workspace, cluster management, notebooks, and job orchestration
  • Workspace AI Agent: Knowledge of Databricks Workspace AI Agent capabilities and integration
  • Data Modelling: Experience implementing data models including dimensional modeling, data vault, or lakehouse architectures
  • Delta Lake: Understanding of Delta Lake features including ACID transactions, schema evolution, and optimization techniques
  • Python: Strong Python programming skills for data processing and automation

 

Additional Technical Skills

  • SQL proficiency for data querying and transformation
  • Experience with cloud platforms (Azure, AWS, or GCP)
  • Understanding of data governance and security best practices
  • Knowledge of streaming data processing (Structured Streaming)
  • Familiarity with DevOps practices and CI/CD pipelines
  • Experience with version control systems (Git)
  • Understanding of data quality frameworks and testing methodologies

 

Professional Experience

  • Minimum 8 years in data engineering or related roles
  • At least 2-3 years of hands-on experience with Databricks platform
  • Proven track record of refactoring legacy code to modern frameworks
  • Experience building and maintaining production data pipelines at scale
  • Background working across multiple data sources and formats
  • Experience in agile development environments

 

Required Certifications

  • Databricks Certified Data Engineer Associate OR Databricks Certified Data Engineer Professional

Data Pipeline Development & Operations

  • Design, build, and operate scalable and reliable data pipelines on the Databricks platform
  • Develop end-to-end data workflows from ingestion through transformation to consumption
  • Implement robust error handling, monitoring, and alerting mechanisms
  • Ensure data pipeline reliability, performance, and maintainability
  • Optimize pipeline performance through efficient Spark job design and cluster configuration
  • Manage and orchestrate complex data workflows using Databricks Jobs and workflows

 

Legacy Code Modernization

  • Refactor legacy code and data pipelines to PySpark for improved performance and scalability
  • Migrate traditional ETL processes to modern ELT patterns on Databricks
  • Assess existing codebases and identify opportunities for optimization and modernization
  • Ensure backward compatibility and data integrity during migration processes
  • Document refactoring approaches and create migration playbooks
  • Collaborate with stakeholders to minimize disruption during code transitions

 

RESPONSIBILITIES:

Data Engineering Excellence

  • Implement data quality checks and validation frameworks
  • Design and maintain Delta Lake tables with appropriate optimization strategies
  • Develop reusable code libraries and frameworks for common data engineering tasks
  • Follow software engineering best practices including version control, testing, and CI/CD
  • Participate in code reviews and provide constructive feedback to team members
  • Troubleshoot and resolve data pipeline issues in production environments

 

Collaboration & Knowledge Sharing

  • Work closely with data architects, analysts, and business stakeholders
  • Collaborate with Infrastructure (Infra), Applications (Apps), and Cyber teams
  • Share knowledge and best practices with Team NCS
  • Mentor junior data engineers on PySpark and Databricks technologies
  • Document technical solutions and maintain comprehensive documentation

More Info

Job Type:
Function:
Employment Type:

Key Skills

About Company

RIDIK, a subsidiary of CLPS Inc, is part of a global leading information technology consulting and solutions service provider focusing on the banking, insurance, and financial service sectors.

As a wholly-owned subsidiary of CLPS Incorporation (Nasdaq: CLPS), we leverage global resources to deliver innovative, tailored solutions across Asia Pacific, North America, and the Middle East.

We have more than 3000 employees working across 8 countries and 8 development centres. Our development centres have been certified with ISO 9001, 27001, and CMMi L5. For more information: please visit: https://www.clpsglobal.com/.

Similar Jobs

6-10 yrs
Singapore
Skills:
ScalaPrometheusDevopsCloudwatchGcpMLopsTerraformDatabricksSplunkAzurePythonAWSDLTUnity CatalogDelta LakeSQL Warehouses
6-10 yrs
SGD 6,000 - 12,000 per month
Singapore
Skills:
ScalaPrometheusDevopsCloudwatchGcpMLopsTerraformDatabricksSplunkAzurePythonAWSDLTUnity CatalogDelta LakeSQL Warehouses
5-7 yrs
Singapore
Skills:
PysparkGitPower BiSparkTableauDatabricksSqlPython
5-8 yrs
SGD 6,000 - 8,500 per month
Singapore, Kallang
Skills:
data engineering Apache HadoopMachine LearningPysparkApache SparkData IntegrationSqlDockerApache KafkaData LakeDatabricksData WarehousingKubernetesPythonEtlBig Data FrameworkCloud Computing ApplicationCI/CD
8-12 yrs
SGD 7,000 - 12,000 per month
Singapore
Skills:
snowflake cml PysparkKafkaTerraformDockerXGBoostOpenshiftPythonAWSGoogle Cloud PlatformScalaApache SparkImpalaSqlJenkinsGitHiveSpark StreamingClouderaDatabricksAzureKubernetesAirflowApache IcebergDremioscikit-learnMLflowFlinkSpark MLlibTrinoApache HudiDelta Lake