Search by job, company or skills

Databricks Engineer

3-5 Years
SGD 8,000 - 8,800 per month
Early Applicant
  • Posted a month ago
  • Be among the first 10 applicants

Job Description

Job Role Name: Databricks Engineer

Qualifications

  • 3 or more years of experience in data engineering with scalable pipelines
  • Strong experience designing data solutions including data modelling and distributed computing architectures
  • Hands-on experience with data processing jobs using PySpark, Spark SQL, and Databricks notebooks/jobs
  • Experience orchestrating data pipelines with ADF, Airflow, or similar tools
  • Experience with both real-time and batch data processing
  • Experience building pipelines on Azure, with AWS experience beneficial
  • Proficiency in SQL including window functions and performance optimization
  • Understanding of DevOps tools, Git workflows, and CI/CD pipelines
  • Familiarity with Scrum methodology and experience working in Scrum teams
  • Ability to apply Scrum practices in a practical project context
  • Strong problem-solving and collaborative mindset
  • Experience with streaming technologies such as Apache Kafka, Apache Flink, or AWS Kinesis
  • Ability to design and implement real-time data processing pipelines

Certification:
Databricks Certified Data Engineer Associate and Databricks Certified Data Engineer Professional are preferred.

Job Description

The Data Engineer will be responsible for designing, developing, and maintaining scalable and reliable data pipelines on Databricks and cloud platforms. The role requires integrating diverse data sources, ensuring high-quality data processing, and supporting analytics, reporting, and machine learning workloads. The role involves collaborating closely with analytics, product, and infrastructure teams to enhance the company's data platform while adhering to best practices for governance, monitoring, and reliability.

What will you do

  • Develop and maintain ETL pipelines for centralized data storage systems (e.g. Delta Lake).
  • Integrate data from databases, APIs, log files, streaming platforms, and external providers
  • Develop data transformation routines to clean, normalize, and aggregate data
  • Apply data processing techniques to handle complex or inconsistent datasets
  • Contribute to frameworks and best practices for code development and deployment
  • Implement data governance in alignment with company standards
  • Partner with analytics and product leaders to design and operationalize pipelines
  • Collaborate with infrastructure leaders to advance cloud-based data platforms
  • Explore new tools and techniques leveraging Azure, Databricks, or related platforms
  • Monitor data pipelines to detect and resolve issues promptly
  • Develop monitoring tools, alerts, and automated error-handling mechanisms
  • Analyze business requirements and identify data extraction requirements
  • Attend and refinement sessions with users
  • Develop and maintain ETL pipelines for ingestion, transformation, validation, and loading
  • Optimize performance and batch scheduling
  • Develop dashboards, reports, scorecards, and data visualizations
  • Perform SIT, data profiling and confirm data accuracy
  • Validate completeness and consistency of ETL Loads
  • Support UAT and production implementation

More Info

Job Type:
Industry:
Employment Type:

Job ID: 151947965

Similar Jobs

Singapore

Skills:

PysparkScalaApache SparkSqlApache AirflowGitGcpKinesisApache KafkaDatabricksAzurePythonAWSDatabricks Lakeflow JobsMLflowvector databasesMedallion ArchitectureUnity CatalogdbtCI CD pipelinesDelta LakeStructured StreamingDatabricks Workflows

Singapore, Raffles Link / Raffles Place

Skills:

dbt (data build tool)PysparkApache AirflowKinesisPythonAWSSpark SQLScalaApache SparkUnit TestingSqlGitGcpApache KafkaDatabricksAzureMLflowStructured StreamingDatabricks Lakehouse platformDatabricks WorkflowsDatabricks Lakeflow JobsVector databasesUnity CatalogCI/CD pipelinesDelta Lake

Singapore

Skills:

data engineering Data EngineerAWSDatabricksPysparkLambdaAmazon Web ServicesApache SparkSpark SQLSqlData ModelingEtlELTBig DataAws S3Amazon S3AWS GlueAws LambdaAmazon RedshiftAWS IAMData WarehousingData LakeData IntegrationData TransformationData MigrationCloud MigrationGitDevopsPythonAws CloudData ArchitectureData AnalyticsData GovernanceData QualityAgileSdlcAWS Data EngineerData PipelinesScalable Data PipelinesDelta LakeLakehouse ArchitectureAmazon AthenaCloud Data EngineeringData LakehouseData ProcessingWorkflow OrchestrationBatch ProcessingReal-Time Data ProcessingPerformance OptimizationCI/CDAWS Certified Data EngineerAWS Certified Solutions ArchitectDatabricks Certified Data EngineerData PlatformData OptimizationStakeholder Management

Singapore

Skills:

Spark SQLGitApache FlinkPysparkApache KafkaDatabricksAzureSqlAirflowAWS Kinesis

Singapore

Skills:

PysparkApache SparkDatabricksSqlDelta Lake

Beware of Scammers

We don’t charge money for job offers