Cloud Operations, SRE & Observability
Operational support for cloud-hosted and mission-critical environments, with emphasis on reliability, visibility, incident response, and clear operating procedures.
Support areas
- AWS cloud operations
- Monitoring, alerting, and dashboard support
- Incident and problem-management workflows
- Splunk search, analysis, and observability support
- Operational troubleshooting and escalation
- Runbooks, SOPs, and technical documentation
Typical roles
- Cloud Operations Engineer
- Site Reliability Engineer
- Observability / Splunk Engineer
- Production Support Engineer
- Systems Engineer
- Technical Operations Analyst