PERMEVO
Website:
permevo.com
Job details:
๐๐ผ๐ฏ ๐ง๐ถ๐๐น๐ฒ: Sr. Data Engineer
๐๐ผ๐ฏ ๐๐ผ๐ฐ๐ฎ๐๐ถ๐ผ๐ป: Chennai, Tamil Nadu, India
๐ช๐ผ๐ฟ๐ธ ๐ ๐ผ๐ฑ๐ฒ: Hybrid (WFO: Tuesday, Wednesday & Thursday)
๐๐ผ๐ฏ ๐ง๐๐ฝ๐ฒ: Full-Time
๐๐
๐ฝ๐ฒ๐ฟ๐ถ๐ฒ๐ป๐ฐ๐ฒ ๐ฅ๐ฒ๐พ๐๐ถ๐ฟ๐ฒ๐ฑ: 5-8 Years
๐ง๐ต๐ฒ ๐๐ต๐ฎ๐น๐น๐ฒ๐ป๐ด๐ฒ
We are seeking a highly skilled Data Engineer to design, build, and optimize scalable data platforms and pipelines that support analytics, reporting, and business intelligence initiatives.
In this role, you will work with diverse data sources, real-time streaming platforms, cloud-based data warehouses, and large-scale datasets to develop reliable and high-performance data solutions. You will collaborate closely with cross-functional teams to transform business requirements into robust data architectures that drive informed decision-making.
This opportunity is ideal for professionals passionate about modern data engineering, cloud technologies, distributed processing, and real-time data systems.
๐ฅ๐ผ๐น๐ฒ๐ & ๐ฅ๐ฒ๐๐ฝ๐ผ๐ป๐๐ถ๐ฏ๐ถ๐น๐ถ๐๐ถ๐ฒ๐
- Design, develop, and maintain scalable data pipelines for data ingestion, transformation, and integration across multiple data sources.
- Build and optimize ETL/ELT workflows to ensure efficient and reliable data movement.
- Work with relational and NoSQL databases to manage large-scale datasets and support business applications.
- Develop solutions leveraging cloud data warehouses such as Snowflake and Amazon Redshift.
- Process and analyze structured, semi-structured, and unstructured data stored in Amazon S3 using services such as AWS Athena and AWS Glue.
- Implement and manage real-time data streaming architectures using Apache Kafka.
- Utilize Apache Spark for distributed data processing and large-scale analytics when required.
- Monitor, troubleshoot, and optimize data pipeline performance to ensure scalability, reliability, and high availability.
- Ensure data quality, consistency, governance, and integrity across data ecosystems.
- Collaborate with business stakeholders, analysts, and engineering teams to translate requirements into technical solutions.
- Participate in architecture discussions and contribute to best practices in data engineering and platform design.
- Stay current with emerging data technologies and industry trends to continuously improve data infrastructure.
๐๐๐๐ฒ๐ป๐๐ถ๐ฎ๐น ๐ฆ๐ธ๐ถ๐น๐น๐ & ๐ฅ๐ฒ๐พ๐๐ถ๐ฟ๐ฒ๐บ๐ฒ๐ป๐๐
๐ ๐๐๐-๐๐ฎ๐๐ฒ ๐ค๐๐ฎ๐น๐ถ๐ณ๐ถ๐ฐ๐ฎ๐๐ถ๐ผ๐ป๐
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
- 5-8 years of hands-on experience in Data Engineering.
- Strong experience designing and managing data pipelines and data integration solutions.
- Expertise in relational databases, including:
o MySQL
o PostgreSQL
- Experience working with NoSQL databases such as:
o MongoDB
o Cassandra
- Hands-on experience with cloud data warehouses:
o Snowflake
o Amazon Redshift
- Experience processing data stored in Amazon S3 using:
o AWS Athena
o AWS Glue
- Strong experience implementing real-time data streaming solutions using Apache Kafka.
- Solid understanding of ETL/ELT frameworks, data transformation, and pipeline optimization.
- Proficiency in at least one programming language:
o Python
o Java
o Scala
- Strong analytical, troubleshooting, and problem-solving skills.
- Excellent communication and collaboration abilities.
๐ฃ๐ฟ๐ฒ๐ณ๐ฒ๐ฟ๐ฟ๐ฒ๐ฑ ๐ค๐๐ฎ๐น๐ถ๐ณ๐ถ๐ฐ๐ฎ๐๐ถ๐ผ๐ป๐
- Experience using Apache Spark for large-scale data processing and analytics.
- Exposure to cloud platforms such as:
o AWS
o Google Cloud Platform (GCP)
o Microsoft Azure
- Experience with big data technologies and distributed processing frameworks.
- Familiarity with data lake and cloud-native data architectures.
- Knowledge of performance tuning and optimization techniques for large datasets.
- Experience working in Agile development environments.
๐ง๐ฒ๐ฐ๐ต๐ป๐ถ๐ฐ๐ฎ๐น ๐ฆ๐ธ๐ถ๐น๐น๐
๐๐ฎ๐๐ฎ ๐๐ป๐ด๐ถ๐ป๐ฒ๐ฒ๐ฟ๐ถ๐ป๐ด
- ETL / ELT Development
- Data Integration
- Data Pipeline Design
- Data Warehousing
- Data Modeling
๐๐ฎ๐๐ฎ๐ฏ๐ฎ๐๐ฒ๐
- MySQL
- PostgreSQL
- MongoDB
- Cassandra
๐ฆ๐๐ฟ๐ฒ๐ฎ๐บ๐ถ๐ป๐ด & ๐ฅ๐ฒ๐ฎ๐น-๐ง๐ถ๐บ๐ฒ ๐ฃ๐ฟ๐ผ๐ฐ๐ฒ๐๐๐ถ๐ป๐ด
๐๐ถ๐ด ๐๐ฎ๐๐ฎ & ๐๐ป๐ฎ๐น๐๐๐ถ๐ฐ๐
- Apache Spark
- Distributed Data Processing
๐๐น๐ผ๐๐ฑ & ๐๐ฎ๐๐ฎ ๐ฃ๐น๐ฎ๐๐ณ๐ผ๐ฟ๐บ๐
- AWS
- Amazon S3
- AWS Athena
- AWS Glue
- Snowflake
- Amazon Redshift
๐ฃ๐ฟ๐ผ๐ด๐ฟ๐ฎ๐บ๐บ๐ถ๐ป๐ด ๐๐ฎ๐ป๐ด๐๐ฎ๐ด๐ฒ๐
Click on Apply to know more.