SellersCommerce
Website:
sellerscommerce.com
Company:
https://www.linkedin.com/company/sellers-commerce
Industries: Software Development
Job details:
About SellersCommerce
SellersCommerce is a fast-growing eCommerce platform powering B2B, B2C, and B2E commerce for manufacturers, distributors, and brands. Our cloud-native platform runs the storefronts, catalogs, orders, and integrations that hundreds of merchants depend on every day — from small sellers to enterprise operations processing high daily order volumes. We're modernizing onto a microservices architecture on the cloud, and our customer base is expanding fast. Keeping that platform reliable, performant, and easy to onboard new customers onto is mission-critical — and that is exactly what this team owns.
The Role
We're looking for an Engineering Manager to lead our Service Delivery function — the team responsible for how our platform runs in production and how new customers go live on it. This is a hands-on operations and reliability leadership role, not a feature-team role.
You'll own the production environment, the on-call and incident-response function, the CI/CD and infrastructure-automation backbone, and the delivery workflows — onboarding, store/site builds, configuration, and integration cutovers — that turn a signed customer into a thriving one. You'll lead a team of DevOps/SRE, production-support, and onboarding engineers, and partner daily with our platform engineering, QA, Product, and Customer Success teams.
The platform is still scaling and maturing, so we want someone who can build the discipline — stand up the SLOs, the automation, the on-call, and the onboarding machine — not just maintain a finished system.
What You'll DoRun the platform
• Own production reliability — availability, performance, and customer-facing SLAs across a multi-tenant platform.
• Define and defend SLOs/SLIs — set error budgets, track them, and report platform health to leadership in language they can act on.
• Lead incident response — own the on-call rotation and escalation policy, run blameless post-incident reviews, and drive corrective actions to closure.
• Own disaster-recovery readiness — runbooks, failover drills, and backup verification against agreed RTO/RPO targets for the most critical services.
Deliver customers
• Own the delivery workflow — customer onboarding, store/site builds, configuration, and integration cutovers, made predictable, repeatable, and low-defect.
• Shorten time-to-live — standardize and automate onboarding so new customers go live faster and with fewer issues.
• Be the engineering escalation point for Customer Success and Support.
Automate and improve
• Own CI/CD and release engineering — pipelines, blue-green / canary deploys, and change management.
• Drive infrastructure-as-code — repeatable, reviewable environment provisioning.
• Mature observability — move the team from basic logging to full metrics, tracing, and actionable alerting.
• Partner on security and compliance — production controls behind frameworks such as PCI-DSS, SOC 2, GDPR, and ISO 27001.
Lead the team
• Manage, coach, and grow a team of 6–10 across DevOps/SRE, production support, and onboarding/delivery.
• Run the delivery cadence — planning, capacity management, and the balance between reactive support and proactive reliability work.
• Raise the bar — set hiring standards, run performance management, and build real career paths.
What You'll Bring
• 8+ years in software engineering or platform operations, including 3+ years managing engineering, DevOps, or SRE teams.
• Production SaaS at scale — demonstrated ownership of uptime, incident command, and SLA delivery for external customers.
• Cloud-native operations — hands-on experience running microservices on a major cloud with containers and Kubernetes in production.
• CI/CD and IaC — a track record automating delivery across multiple environments (dev → staging → pre-prod → prod).
• Incident leadership — strong on-call and incident-management experience, with post-incident reviews that actually change the roadmap.
• Compliance-aware — experience operating under frameworks such as PCI-DSS, SOC 2, GDPR, or ISO 27001.
• Clear communication — you can translate operational risk and reliability metrics for both executives and customer-facing teams.
• Builder's mindset — comfortable standing a function up, not only maintaining one.
• On-site — able to work from our Hyderabad office, 5 days a week.
Nice to Have
• Bachelor's/Master's in Computer Science or a related field.
• Background in e-commerce, retail tech, or multi-tenant B2B SaaS.
• Certifications such as Azure (AZ-104 / AZ-400) or AWS equivalents, CKA/CKAD, or ITIL.
• Experience with integration platforms (iPaaS) and high-volume data/catalog pipelines.
• A feel for cloud cost / FinOps — keeping spend honest as scale grows.
Technology You'll Work With
You won't write production features daily, but you'll lead the people who do — so you should be fluent enough to review an incident, challenge a design, and set operational standards with credibility.
Today — the platform you’ll run
The production system the team operates and modernizes is a mature, multi-tenant .NET Framework application powering B2B, B2C, Dropship & Payments storefronts:
Web & application: ASP.NET MVC 5.2.3 (Razor view engine), ASP.NET Web API 2, OWIN/Katana middleware; C# on .NET Framework 4.5.2 — 322 controllers and 2,800+ Razor views across the B2B, B2C, Dropship & Payments solutions.
Hosting & infra: IIS on Windows Server; multi-tenant host-header isolation via a custom IIS binding-automation service; Azure DevOps CI pipelines; Azure Key Vault & Azure Blob Storage.
Data & search: SQL Server with Entity Framework 6 (ORM); Elasticsearch (NEST) for catalog/search; Redis (StackExchange.Redis) for caching & sessions.
Background processing: Quartz.NET and Hangfire for scheduled jobs, order automation, exports & renewals.
Security & auth: ASP.NET Identity with OWIN OAuth; TLS; multi-tenant authentication; PCI-aware payments-service integration.
Front-end: jQuery, Bootstrap and Knockout.js — server-rendered Razor with client-side enhancement.
Where we’re heading — modernization
The target architecture the team is migrating toward, and the operational tooling you’ll stand up around it:
Cloud & infra: Microsoft Azure, Azure Kubernetes Service (AKS), Docker, Terraform, Azure DevOps, Helm, ArgoCD.
Reliability & observability: PagerDuty, OpenTelemetry, Prometheus, Grafana, Azure Monitor; SLO / error-budget practices; DR & backup.
Platform (to direct): .NET / C# microservices; REST, GraphQL & gRPC APIs; Kafka & service bus; SQL Server, MongoDB, Elasticsearch, Redis.
Security & compliance: WAF, zero-trust, TLS, encryption-at-rest; PCI-DSS, SOC 2, GDPR, ISO 27001.
What Success Looks Like (First 6–12 Months)
• A staffed, documented on-call rotation with a working incident-response and post-incident-review process, measured against agreed SLOs.
• Standardized, automated onboarding and site delivery — shorter cycle time, lower defect rate.
• Observability uplifted from basic logging to metrics and tracing on the most critical services.
• DR drills run and verified for critical services against RTO/RPO targets.
Why Join Us
• Real ownership — you'll define the reliability and delivery function, not inherit someone else's.
• High impact — your work is felt directly by every merchant on the platform.
• A modern stack and a platform in active modernization — plenty of greenfield to shape.
• A collaborative engineering culture that values blameless learning and continuous improvement.
Click on Apply to know more.