Lead Big Data Engineer Remote -Atlanta, GA at Atlanta, Georgia, USA |
Email: [email protected] |
From: UTTAM BARMAN, SONITALENTCORP [email protected] Reply to: [email protected] Job Title Lead Big Data Engineer Job Location -Remote -Atlanta, GA Duration - 12 Months+ Visa -USC,GC Only Mode Of Interview - Phone/Skype Note - Need LinkedIn and no Job Hoffers Job Description Skills Kafka, Python, Spark Streaming, Linux/Unix shell scripting Job Description Big Data Engineer Seeking an experienced hands-on enterprise Data Engineer lead. The successful candidate must have Big Data engineering experience and must demonstrate an affinity for working with others to create successful solutions. They must be a great communicator, both written and verbal, and have some experience working with business areas to translate their business data needs and data questions into project requirements. The candidate will participate in all phases of the Data Engineering life cycle and will work independently and collaboratively write project requirements, architect solutions, and perform data ingestion development and support duties. This Data Engineer will be responsible for creating new data flows into AWS at Norfolk Southern including coordinating with business and functional areas to establish and communicate our data fabric best practices, framework, and tools. Together we will establish a new data fabric for Norfolk Southern that will help create a common view of data and provide a centralized mechanism for its aggregation, cleansing, transformation, augmentation, validation, and syndication. Required: 6+ years of overall IT experience 3+ years of experience with high-velocity high-volume stream processing: Apache Kafka and Spark Streaming Experience with real-time data processing and streaming techniques using Spark structured streaming and Kafka Deep knowledge of troubleshooting and tuning Spark applications 3+ years of experience with data ingestion from Message Queues (Tibco, IBM, etc.) and different file formats across different platforms like JSON, XML, CSV 3+ years of experience with Big Data tools/technologies like Hadoop, Spark, Spark SQL, Kafka, Sqoop, Hive, S3, HDFS, or 3+ years of experience building, testing, and optimizing Big Data data ingestion pipelines, architectures, and data sets 2+ years of experience with Python (and/or Scala) and PySpark/Scala-Spark 3+ years of experience with Cloud platforms e.g. AWS, GCP, etc. 3+ years of experience with database solutions like Kudu/Impala, or Delta Lake or Snowflake or BigQuery 2+ years of experience with NoSQL databases, including HBASE and/or Cassandra Experience in successfully building and deploying a new data platform on Azure/ AWS Experience in Azure / AWS Serverless technologies, like, S3, Kinesis/MSK, lambda, and Glue Strong knowledge of Messaging Platforms like Kafka, Amazon MSK & TIBCO EMS or IBM MQ Series Experience with Databricks UI, Managing Databricks Notebooks, Delta Lake with Python, Delta Lake with Spark SQL, Delta Live Tables, Unity Catalog Knowledge of Unix/Linux platform and shell scripting is a must Strong analytical and problem-solving skills Preferred: Strong SQL skills with ability to write intermediate complexity queries Strong understanding of Relational & Dimensional modeling Experience with GIT code versioning software Experience with REST API and Web Services Good business analyst and requirements gathering/writing skills Keywords: user interface message queue sthree information technology green card Georgia |
[email protected] View all |
Thu Nov 16 21:47:00 UTC 2023 |