Big Data Developer
We believe productive and efficiency cannot be rule-bound, nor can it be achieved without setting targets. We fuel our people dreams of high achievement with hope since we believe people will flourish when they have hope.
@ATG we provide healthy, real-time and enjoyable work environment since we believe “People give their best when they enjoy their work”. We provide opportunities wherein people can grow and perform.
Big Data Developer
Location: San Jose, CA.
Duration: 12 Months
Experience: 7+ yrs
Skills Java, Python, R, HDFS, MapReduce, Yarn, Spark, Hbase, Hive, Pig, Flume, Sqoop
Roles and Responsibilities
Participate in technical planning & requirements gathering phases including design, coding, testing, troubleshooting, and documenting big data-oriented software applications. Responsible for the ingestion, maintenance, improvement, cleaning, and manipulation of data in the business’s operational and analytics databases, and troubleshoots any existent issues.
Implements, troubleshoots, and optimizes distributed solutions based on modern big data technologies like Hive, Hadoop, Spark, Elastic Search, Storm, Kafka, etc. in both an on premise and cloud deployment model to solve large scale processing problems.
Define and build large-scale near real-time streaming data processing pipelines that will enable faster, better, data-informed decision making within the business.
Work inside the team of industry experts on the cutting edge Big Data technologiesto develop solutions for deployment at massive scale.
Study data, identify patterns, make sense out of it and convert it to algorithms.
Designs and plans BI, and other Visualization Tools capturing and analyzing data from multiple sources to make data-driven decisions, as well as debugs, monitors, and troubleshoots solutions.
Keep up with industry trends and best practices, advising senior management on new and improved data engineering strategies that will drive departmental performance leading to improvement in overall improvement in data governance across the business, promoting informed decision-making, and ultimately improving overall business performance.
Skills and Experiences
Programming languages: Java, Python MUST. Good to have a knowledge of C
Experience with Hadoop development and QA
Experience with NFS (preferably ONTAP) and shared file systems development
Experience with Hortonworks technologies and Cloudera technologies
Knowledge of Hadoop ecosystem components like MapReduce, Spark, HBase and, Ranger and Sentry.
Security framework experience with Kerberos, LDAP