Java Data Engineer/ETL Developer with StreamSets
Java Data Engineer/ETL DeveloperVisa status: U.S. Citizens and those authorized to work in the U.S. are encouraged to apply. Tax Terms: W2, 1099 Corp-Corp or 3rd Parties: YesLocation: San Antonio, TXThe Streamsets & Java Data Engineer/ETL Developer is responsible for the maintenance, synchronization, cleaning, and migration of transactional data in a hybrid environment with both on-prem and highly modern cloud based microservices environment. The Streamsets Data Engineer works with the product teams in order to understand, analyze, document and efficiently implement to deliver streaming as well as batch-oriented data for synchronizing legacy and modern data stores ensuring data integrity.Responsibilities:Responsible for designing and coding data ingestion pipeline framework with Streamsets and Java for batch and streaming dataUnderstand data needs and be able to construct data pipelines for automating event driven bi-directional selective data replication, along with micro-batch and batch-based data pipelinesWork with the product teams in order to understand, analyze, document and efficiently implement to deliver streaming as well as batch-oriented data for synchronizing legacy and modern data stores ensuring data integrity.Provide support to the Application database design to aid in eliminating data duplication and enabling selective event based and schedule-based data transfer to endpoints within the cloud and legacy environment as required.Drive towards programmatic pipeline generation and orchestration to enhance repeatability and rapid deployment, using out of the box thinkingUtilize CI/CD tools while utilizing established design patterns and methods.Rapidly develop technical solutions working closely with the integrated product teams and developers with minimal direction from senior or lead resourcesThe candidate will be expected to support a variety of structured, semi-structured and unstructured data in streaming and batch frameworks.Develop of different pipelines in the Stream sets according the requirements of the business owner.Manage Critical Data Pipelines that power analytics for various business units.Experience:BS/MS degree in Computer Science, Engineering or a related subjectOverall 12+ year experience in framework and product development4+ Proven hands-on Software Development experience using Java and StreamsetsExperience in understanding the Kafka topic and development experience to read the data from Kafka topics and implement the complex logic on Stream sets.Hands on experience in designing and developing applications using Java EE platformsObject Oriented analysis and design using common design patterns.4+ years of excellent knowledge of Relational Databases, SQL and ORM technologies (JPA2, Hibernate)4+ years of Experience in the Spring FrameworkExperience in developing web applications using at least one popular web framework (JSF, Wicket, GWT, Spring MVC)The candidate must have a successful track record in ETL job design, development and automation activities with minimal supervision.Good knowledge on Java to write customize code to implement on Streamsets pipelines.Investigate and develop skills in new technologies