We have an opening of one Data Engineer profile to fill in. JD is given below, please share compatible profiles.
4+ years of overall IT experience, including hands-on experience with Big Data technologies.
Strong hands-on experience in Python and PySpark.
o Python should be used extensively for application development, ETL (Extract, Transform, Load) processes, and Data Lake curation.
Experience in building PySpark applications using Spark DataFrames in Python with development tools such as Jupyter Notebook and PyCharm (IDE).
Proven experience in optimizing Spark jobs that process large-scale datasets and high-volume workloads.
Hands-on experience with version control systems, particularly Git.
Experience working with AWS Analytics services, including:
o Amazon EMR
o Amazon Athena
o AWS Glue
Experience with AWS Compute and Storage services, including:
o AWS Lambda
o Amazon EC2
o Amazon S3
o Other AWS services such as Amazon SNS
Working knowledge of Bash/Shell scripting is an added advantage.