SysTechCorp Inc
Website:
systechus.com
Job details:
Company Description SysTechCorp Inc delivers leading-edge technology services that help clients achieve and exceed their strategic business objectives. The company’s practices in AI, machine learning, mobile, cloud, and ERP are supported by years of experience and deep domain expertise. SysTechCorp Inc serves a broad range of industry verticals, offering solutions tailored to complex and evolving business needs. Team members collaborate on innovative projects that leverage advanced technologies to create measurable impact for clients.
Key Responsibilities
● Execute EMR major version upgrade and HBase 1.4.x to 2.6.x migration across 18 clusters using blue/green replacement — sequence from smallest to largest, each cluster promoted only after the previous passes full validation
● Design and execute the snapshot, seed, and sync strategy for each cluster migration — initial snapshot, incremental sync, final cutover sync, and data integrity validation before traffic switchover
● Validate Apache Phoenix SQL compatibility on the upgraded HBase version — run the full Phoenix SQL regression suite; document any query incompatibilities; coordinate application-side fixes if required
● Validate Apache Zookeeper version compatibility — confirm the upgraded EMR bundled Zookeeper version is compatible with all HBase client applications
● Execute RegionServer rebalancing and compaction checks after each cluster upgrade — ensure even region distribution, confirm major compaction has completed, validate heap configuration
● Align Apache Spark runtime with the upgraded EMR version — confirm Spark 3.x compatibility for long-running jobs; validate job submission and output correctness
● Conduct data integrity validation after each cluster upgrade — row count checks on large tables, checksum comparison, and application integration test execution
● Execute and document rollback rehearsals before production upgrades — demonstrate rollback is achievable within 4 hours for smaller clusters; document the approach for large clusters
● Produce per-cluster upgrade evidence packages and coordinate production maintenance window timing
Must-Have Skills & Experience
● 7+ years big data or cloud data platform engineering; 3+ years hands-on Amazon EMR and Apache HBase operations and upgrades
● Amazon EMR — version lifecycle (EMR 5.x and 7.x), cluster provisioning and decommission, instance fleet configuration, cluster health monitoring
● Apache HBase 1.x to 2.x migration — RegionServer management, major and minor compaction, balancer configuration, snapshot and clone operations
● Apache Phoenix — SQL regression testing on HBase 2.x, secondary index compatibility, data type migration
● Apache Zookeeper — version compatibility, client configuration migration, ensemble management
● Blue/green cluster replacement strategy — parallel cluster provisioning, snapshot/seed/sync pattern, traffic cutover, data integrity validation
● Data integrity validation at scale — row count comparison, checksum approaches for large tables, integration test execution post-cutover
● Apache Spark 2.x/3.x on EMR — job compatibility across EMR versions, YARN configuration, long-running job management
● Rollback planning and rehearsal — RTO documentation, rollback execution, time-box adherence
Nice-to-Have Skills
● HBase repair tooling (HBCK2) — orphaned region repair and procedure recovery
● Kerberos-secured HBase cluster management — keytab management, HBase principal configuration
● AWS EMR on EKS or EMR Serverless — awareness for platform modernisation context
● Apache Flink or Apache Kafka on EMR
● CloudWatch EMR metrics — HDFS capacity, HBase read/write latency, JVM tuning
Tools & Platforms
Amazon EMR (5.x, 7.x), Apache HBase (1.x, 2.x), Apache Phoenix, Apache Zookeeper, Apache Spark (2.x/3.x), YARN, HDFS, AWS SSM, CloudFormation, CloudWatch, S3, Jenkins, Bitbucket, Confluence, AWS CLI, Python/Boto3
Click on Apply to know more.