Search Jobs

Search by job, company or skills

Python-Spark Scala

Python-Spark Scala

Infosys Limited
Early Applicant
  • Posted 22 days ago
  • Be among the first 10 applicants

Job Description

Key Responsibilities:

  • 5 9 years of experience in Data Engineering Big Data development
  • Strong hands on experience with Python and Apache Spark Scala PySpark
  • Deep understanding of Spark architecture RDD DataFrames Spark SQL
  • Experience in distributed computing and big data processing
  • Strong knowledge of SQL and data modeling concepts
  • Experience with data pipeline development ETL ELT
  • Familiarity with Linux Unix environments
  • Experience with version control tools Git

Technical Requirements:

  • Primary skills Technology Big Data Data Processing Spark Technology Functional Programming Scala Technology Machine Learning Python

Additional Responsibilities:

  • Experience with cloud platforms AWS Azure or GCP
  • Hands on with Databricks EMR Spark clusters
  • Knowledge of streaming technologies Kafka Spark Streaming Structured Streaming
  • Experience with workflow orchestration tools Airflow Oozie
  • Familiarity with Delta Lake Lakehouse architecture
  • Exposure to NoSQL databases MongoDB Cassandra
  • Knowledge of CI CD and DevOps practices

Preferred Skills:

Technology->AI-Data science->PYTHON,Technology->Big Data - Data Processing->Spark->SparkSQL,Technology->Functional Programming->Scala

More Info

Job Type:
Industry:
Employment Type:

Key Skills

About Company

Similar Jobs

5-9 yrs
Bengaluru, India
Skills:
Devops, Sql, Databricks, Emr, Pyspark, Kafka, Linux, Spark Streaming, Etl, Cassandra, Gcp, Git, Scala, Oozie, ELT, Unix, Data Modeling, Apache Spark, AWS, Spark SQL, Python, Azure, MongoDB, DataFrames, Airflow, Delta Lake, Structured Streaming, Spark architecture, RDD
5-8 yrs
Bengaluru, India
Skills:
containerization , Hadoop, Kafka, Redis, Nosql, Hive, RDBMS, data pipelines, orchestration tools, unit integration and end-to-end tests, observability tools, data technologies, large-scale data processing, CI CD pipelines, secure coding practices, ML services
4-8 yrs
Bengaluru, India
Skills:
Java, Spark SQL, Hadoop, Cloudformation, Scala, Apache Spark, Big Data, Cloud Architecture, Git, Hive, Spark Streaming, Docker, Terraform, Linux, Presto, Ansible, Shell scripting, Postgres, Gitlab, Oracle, AWS, Airflow, Service-oriented architecture