Role description
Job Role : Python Pyspark Developer:
Location : Irving,Texas
At least 5 years of development work experience in Hadoop programming HDFS using pysparkHive based Data warehouse projects and with good Shell Scripting experience
Experience in Big Data technologies including Pyspark Hive and Hadoop
Experience in Pyspark programming languages
You have a good understanding of organizational strategy architecture patterns Microservices Event Driven and technology choices and coaching the team in execution in alignment to these guidelines
You can apply organizational technology patterns effectively in projects and make recommendations on alternate options
You have handson experience working with large volumes of data including different patterns of data ingestion processing batch realtime movement storage and access for both internal and external to BU and ability to make independent decisions within scope of project
You have a good understanding of data structures and algorithms
You can test debug and fix issues within established SLAs
You can design software that is easily testable and observable
You understand how teams goals fit a business need
You can identify business problems at the project level and provide solutions
You understand data access patterns streaming technology data validation data performance cost optimization
Strong SQL skills