
Search by job, company or skills
Intro to the Role and Team
We are expanding our Singapore Infrastructure & Data Systems Hub to power Luma's next-generation multimodal foundation models and Physical AI pipelines. As a Data Engineer, you will join an agile engineering group that builds and operates the core data engines feeding our AI models, working with our global research and infrastructure teams on large-scale data acquisition, distributed storage, and multimodal dataset processing.
What You'll Own
What You'll Bring
Basic Requirements:
Nice-to-Haves:
About Luma: Luma's mission is to build unified general intelligence that can generate, understand, and operate in the physical world. We believe multimodality is critical for intelligence - the next step beyond language models comes from vision. Luma is an equal opportunity employer.
Job ID: 153353587
Skills:
Data Modelling, Pyspark, Data Warehousing, Azure Databricks, Azure Sql, Sql, ELT, Git, Azure Data Factory, Azure Data Lake, Python, Azure DevOps, Etl, data pipelines, lakehouse architecture, CI CD deployment practices, Microsoft Fabric, Azure data services
Skills:
Java, Ranger, Hadoop, Scala, Prometheus, Bitbucket, Grafana, OpenShift Container Platform, Jenkins, Hive, Docker, Ansible, Spark, Splunk, Python, Kubernetes, HDFS, Quantexa
Skills:
Spark SQL, S3, Power Bi, Cloudformation, AWS Glue, Tableau, Emr, Redshift, Sql, Nosql, Lambda, Kinesis, Terraform, Databricks, Python, Informatica Data Management Cloud, Databricks Delta Lake, MLflow, IDMC, R, Athena
Skills:
Java, Machine Learning, Hadoop Ecosystem, Scala, Big Data Technologies, Data Modeling, Dashboards, Nlp, Spark, Python, Data Processing, data warehouse design, Flink, data visualization tools, reporting systems
Skills:
data wrangling , Hadoop, Power Bi, Big Data Technologies, Informatica, Sql, Hive, Docker, Spark, Data Visualization, Dbms, Talend, Kubernetes, Python, DevSecOps Methodology, DI ETL technology, MS Access, Microservices Architecture, Podman