GOVERNANCE // 04 ENGINEERING SLA
OPERATIONAL SPECIFICATION · 2026 EDITION

RELIABILITY,
UPTIME AND COMMITMENTS.

We stand behind the stability and architectural endurance of the systems we engineer. This SLA details our incident classifications, response times, and engineering commitments.

SECTION 01

Scope & Applicability

This Service Level Agreement (“SLA”) governs software systems, custom web applications, APIs, and cloud infrastructure maintained by EthosCore under an active Engineering Retainer or Managed Infrastructure Agreement.

For project-based engagements during active delivery sprints prior to production deployment, milestone acceptance criteria defined in individual Statements of Work apply.

SECTION 02

Availability & Uptime Targets

EthosCore commits to delivering an architectural availability benchmark of 99.9% Uptime across production-grade microservices and database engines.

99.9%
Production Target

Maximum allowable monthly unscheduled downtime < 43 minutes.

24/7/365
Telemetry Probes

Continuous automated ping and synthetic endpoint monitoring.

≤ 60s
Alert Propagation

Automated on-call escalation via PagerDuty / Opsgenie.

SECTION 03

Incident Severity Matrix

Incidents are categorized by operational business impact to ensure immediate priority for mission-critical disruptions:

P1 · Critical / Blocker24/7 Urgent Response

Complete outage of the production application, core payment/POS billing failure, critical database corruption, or severe security breach preventing ongoing business transactions with no workaround.

P2 · Major / DegradedBusiness Hours + Extended

Major business feature impaired (e.g. inventory sync lag, reporting generation failure), but the primary customer transaction flow remains operational through an identifiable alternative workflow.

P3 · Moderate / Non-CriticalStandard Business Hours

Minor functionality defect, cosmetic UI rendering inconsistency, or administrative dashboard query timeout that does not prevent customer orders or data capture.

P4 · Minor / RequestNext Sprint Backlog

General technical questions, architectural advisory inquiries, scope refinement requests, or non-urgent configuration adjustments.

SECTION 04

Response & Remediation Windows

SeverityFirst Response SLAUpdate CadenceTarget Mitigation
P1 (Critical)< 60 Minutes (24/7)Every 60 MinutesContinuous until patched
P2 (Major)< 4 HoursEvery 4 Hours< 24 Hours
P3 (Moderate)< 1 Business DayDailyNext Sprint Release
P4 (Request)< 2 Business DaysAs requiredBacklog Prioritization
SECTION 05

Deployment & Rollback Protocol

Every production deployment executed by EthosCore engineers follows a zero-downtime, deterministic CI/CD release workflow:

  • Automated Preview Environments: All pull requests are verified on isolated staging sandbox instances prior to merging into production branches.
  • Instant Rollback Trigger: If automated health check endpoints or error-rate thresholds exceed 0.5% post-deployment, traffic is routed automatically to the previous immutable release image.
  • Database Migration Safety: All schema migrations follow expand-and-contract patterns to ensure backward compatibility during transitions.
SECTION 06

Scheduled Maintenance Windows

Routine infrastructure upgrades, major database version upgrades, and kernel security patches are conducted during pre-arranged maintenance windows:

• Advance Notice: Minimum 48 to 72 hours notice provided via email and dedicated support channels.
• Execution Window: Typically scheduled during lowest-traffic off-peak hours (e.g., 01:00 AM – 04:00 AM IST on weekends).
• Exclusion: Planned maintenance performed within notified windows is excluded from monthly uptime calculations.
SECTION 07

Root Cause Analysis (RCA)

Following any P1 incident, EthosCore provides a comprehensive, blameless engineering post-mortem document within 5 business days of incident closure.

1. Detailed timeline of failure detection, triage, and resolution.
2. Technical breakdown of underlying root trigger (architectural, code, or third-party dependency).
3. Immediate mitigations applied.
4. Actionable preventive engineering items added to the sprint backlog to prevent recurrence.
SECTION 08

Dedicated Escalation Channels

Clients on engineering retainers have access to private direct channels for escalation:

Primary Support Desk: support@ethoscore.com
General Inquiries: contact@ethoscore.com
Direct Retainer Hub: Dedicated Slack Connect / Discord technical workspace shared with EthosCore engineering leads.