
Search by job, company or skills
. Bachelor's degree in Computer Science, IT, Engineering, or a related field with a demonstrated continuous learning ethos.
. Must have a minimum of 8+ years IT experience with at least 5+ years hands-on data engineering or datapipeline development
. Expert-level SQL proficiency with strong expertise in SQL Server, including queryoptimization, indexing, and performance tuning
. Advanced Python programming skills for data processing, automation, and production-grade pipeline development
. Kubernetes expertise - Design, deploy, andmanage containerized data pipelines in on-premise environments
. Strong data modellingexpertise- Both relational and non-relational concepts
. Proven experience with flexible lakehouse/data lake architecture - Multi-layer datalakes, partitioning strategies, and metadata management, Iceberg tables, and optimization
. CI/CD and DevOps practices- Setting up CI/CD pipelines, Git, automated testing, andinfrastructure-as-code tools
. ETL/ELT orchestration experience-Apache Airflow or similar tools for scheduling and monitoring batch and real-time jobs
. Hands-on experience with at least one NoSQL database (MongoDB, Cassandra, etc.)
. Hands-on experience with Apache Spark and PySpark for distributed data processing andperformance optimization
. Data security andgovernance- Role-based access control, data masking, and compliance frameworks
. Proven ability to work autonomously on complex projects while maintaining high codequality standards
. Excellent problem-solving, communication, and cross-functional collaboration skills
Job ID: 152522653
Skills:
Cassandra, Pyspark, SQL Server, Apache Spark, Data Modeling, ELT, Apache Airflow, Docker, MongoDB, Kubernetes, Python, Etl, Data security and governance, Lakehouse architecture
Skills:
Aws Lambda, S3, Aws Services, Pyspark, AWS Glue, Sql, Jenkins, Cloudwatch, Terraform, Iam, Gitlab, Data Modelling, AWS, Airflow, AWS Step Functions, GitHub Actions, data lakes, dbt, Lakehouse data lake architecture
Skills:
Spark, Tableau, Sql, Apache Parquet, Airflow, Metabase, ClickHouse, dbt, Delta Lake, Milvus
Skills:
snowflake , S3, Hadoop, Prometheus, Emr, Grafana, Redshift, Sql, Lambda, Cloudwatch, Spark, Databricks, Python, Airflow, MLflow, SageMaker, Lake Formation, dbt, Glue, Vertex AI, Athena
Skills:
snowflake , Pyspark, Amazon S3, Amazon Kinesis, AWS Glue, Sql, Apache Airflow, Spark, Amazon Rds, Python, AWS, Airflow, Amazon Step Functions, Amazon Lambda