Data Architect
Job Summary
Seeking a Senior Specialist with 7 to 11 years of experience in Data Architecture to provide datadriven solutions
Job Description Looking for someone who is familiar with the below areas and Design develop and maintain robust data architecture frameworks to support data requirements and translate them into scalable data solutions
HDFS Architecture
NameNode High Availability HA
HDFS Federation
Rack Awareness
HDFS ReadWrite Flow
YARN Architecture
Capacity Scheduler
Fair Scheduler
Resource Management Queue Design
MapReduce Architecture
Shuffle and Sort
Apache Spark Architecture
Spark DAG Execution Plan
Spark Performance Tuning
Memory Management in Spark
Adaptive Query Execution AQE
Hive Architecture
Hive Metastore
Hive Optimization Techniques
Partitioning Bucketing
ORC Parquet Internals
Impala Architecture
Apache Kafka Architecture
Kafka Performance Tuning
Kafka Security
Streaming Architectures
CDC Change Data Capture
Apache Airflow Architecture
Workflow Orchestration
Data Lake Architecture
Lakehouse Architecture
Delta Lake Fundamentals
Data Modeling for Big Data
ETLELT Framework Design
Data Governance
Data Lineage
Metadata Management
Kerberos Authentication
Apache Ranger
Apache Knox
Encryption Security Best Practices
Cloudera CDP Architecture
AWS EMR
AWS Glue
S3 Architecture Optimization
Databricks Architecture
CICD for Data Pipelines
Monitoring Observability
Logging Frameworks
Cluster Sizing Capacity Planning
Disaster Recovery DR
Backup Recovery Strategies
High Availability Design
MultiTenancy Design
Data Quality Frameworks
Performance Troubleshooting
Data Skew Handling
Join Optimization Techniques
Small File Problem
Cost Optimization
RealTime Data Pipeline Design
Batch Processing Architecture
EventDriven Architecture
Lambda Architecture
EndtoEnd Solution Architecture
Migration from OnPrem Hadoop to Cloud
Scalability Reliability Design Patterns
System Design for Big Data Platforms
Architecture Tradeoffs and Decision Making
SQL
PySparkScala Coding#J-18808-Ljbffr