Website:
Job details:
Data Asset Developer / Hadoop-Spark
Location: Hyderabad
Experience: 5–8 Years
Open Positions: 5
Notice Period: 15 Days
Requirement: LTIMindtree
Role Overview
We are looking for an experienced Data Asset Developer with strong expertise in Apache Spark, Python, PySpark and Spark SQL to design and develop scalable data architectures and pipelines. The candidate will work closely with data scientists, BI/analytics teams and cybersecurity stakeholders to build high-quality, secure and efficient data assets.
Key Responsibilities
- Design and implement scalable data models and data architectures for cybersecurity data assets.
- Develop robust and scalable data pipelines using Databricks, PySpark/Spark and Spark SQL.
- Work with large and multiterabyte datasets and optimize pipelines for performance and cost.
- Design and maintain Star and Snowflake schema data models.
- Develop ETL processes and support data warehouse/data lakehouse architectures.
- Work with Databricks Notebooks and Unity Catalog.
- Ensure data quality, security, governance and compliance.
- Prepare technical documentation, metadata and data asset specifications.
- Collaborate with Data Scientists, BI/Analytics Engineers and business stakeholders.
- Troubleshoot data pipeline and performance issues and implement improvements.
Mandatory Skills
- Apache Spark
- PySpark
- Python for Data
- Spark SQL / SparkSQL
- Databricks
- Unity Catalog
- Data Modelling
- ETL / Data Pipelines
- Data Warehouse concepts
Good to Have
- Azure cloud data solutions
- Lakehouse architecture
- MySQL
- Machine Learning / Advanced Analytics
- Cybersecurity domain knowledge
- Databricks or cloud certifications
Education
B.Tech / MCA / M.Tech in Computer Science, Information Technology, Engineering or a related field preferred.
Interested candidates can share their updated CV along with current CTC, expected CTC, location and notice period.
Click on Apply to know more.