VySystems
Website:
vysystems.com
Company:
https://www.linkedin.com/company/vysystems
Seniority: Associate
Industries: Technology, Information and Media
Job details:
Position: Site Reliability Engineer (SRE) – Ab Initio
Experience: 5+ years
Location: Bangalore/ Hybrid
Employment Type: Full-time
Role Overview
We are looking for an experienced SRE with strong Ab Initio experience to support, monitor, maintain, and improve enterprise data-processing applications and production environments. The candidate will be responsible for ensuring high availability, reliability, and performance of Ab Initio applications and associated data pipelines.
Key Responsibilities
- Provide L2/L3 production support for Ab Initio applications and data pipelines.
- Monitor production jobs, workflows, schedules, and system health.
- Troubleshoot Ab Initio job failures, performance issues, data issues, and infrastructure-related incidents.
- Perform root-cause analysis (RCA) and implement permanent fixes.
- Develop and maintain monitoring, alerting, and operational dashboards.
- Automate repetitive operational tasks using Shell scripting/Python or similar tools.
- Manage incident, problem, and change processes in accordance with SRE/ITIL practices.
- Work with development, infrastructure, database, and application teams to resolve production issues.
- Analyze application and system logs to identify failures and performance bottlenecks.
- Support deployments, release activities, batch scheduling, and production validation.
- Participate in on-call/shift rotations and ensure timely incident resolution.
- Establish and improve SLIs, SLOs, SLAs, and operational metrics.
- Identify opportunities to improve system reliability, scalability, and operational efficiency.
- Maintain technical documentation, runbooks, troubleshooting guides, and knowledge articles.
Required Skills
- Strong hands-on experience with Ab Initio.
- Good understanding of Ab Initio GDE, Co>Operating System, plans, graphs, and components.
- Experience supporting ETL/data-processing batch jobs in production.
- Strong Unix/Linux skills and Shell scripting.
- Experience with SQL and relational databases such as Oracle, DB2, or PostgreSQL.
- Good understanding of production monitoring, incident management, and troubleshooting.
- Experience with scheduling tools such as Control-M, Autosys, or similar.
- Understanding of CI/CD and deployment processes.
- Experience with Git or other version-control systems.
- Strong analytical and problem-solving skills.
Click on Apply to know more.