Infosys
Website:
infosys.com
Company:
https://www.linkedin.com/company/infosys
Seniority: Not Applicable
Industries: IT Services and IT Consulting
Job details:
- Experience leading a Snowflake ® Trino or Spark ® Trino migration at scale.
- Trino cluster administration and deployment on Kubernetes.
- Workflow orchestration (Airflow, Dagster, or similar) and dbt.
- Experience building data-validation / reconciliation frameworks for migration parity.
- Knowledge of Trino internals or connector development (contributing custom connectors/UDFs).
- Streaming/CDC ingestion into Iceberg (Kafka, Flink, Debezium).
- Data governance, lineage, and cost-optimization tooling.
Hands-on production experience with Trino (or PrestoSQL/Presto).
- Strong, deep SQL expertise — complex analytical queries, window functions, CTEs, query optimization.
- Hands-on experience with Apache Iceberg (or comparable open table formats — Delta Lake, Hudi) including schema/partition evolution and table maintenance.
- Practical experience with Snowflake and/or Apache Spark — enough to read, understand, and migrate existing workloads.
- Understanding of distributed query execution: MPP architecture, join distribution, memory/spill behavior, partition pruning, and predicate pushdown.
- Experience with cloud object storage and columnar file formats (Parquet, ORC).
- Proficiency in at least one programming language (Python, Java, or Scala) for tooling, UDFs, and automation.
- Version control (Git) and CI/CD for data pipelines. Strong analytical and problem-solving mindset for debugging correctness and performance issues.
- Clear communication — able to document migration decisions and work with analytics, platform, and business teams.
- Ownership mentality with attention to data correctness and reliability.
Click on Apply to know more.