About Us:
As a Senior Operations Engineer at Kenility, you’ll join a tight-knit family of creative developers, engineers, and designers who strive to develop and deliver the highest-quality products to the market.
Technical Requirements:
- Bachelor’s degree in Computer Science, Software Engineering, or a related field.
- 5+ years of experience in infrastructure or IT operations, supporting complex hybrid environments.
- Strong hands-on expertise with Microsoft Azure gained through day-to-day operational ownership and support.
- Practical experience with Ansible, including developing, maintaining, and improving automation playbooks.
- Proficiency in scripting and operational automation using PowerShell, Python, or Bash.
- Hands-on experience with observability platforms such as Splunk, Datadog, or New Relic, including dashboard creation and alert configuration and optimization.
- Strong documentation practices, including maintaining runbooks, infrastructure diagrams, and change records.
- Experience supporting on-premises infrastructure involving networking, WiFi, servers, and endpoint devices is desirable.
- Familiarity with Cloudflare technologies, including WAF, DNS, and CDN services, is a plus.
- Experience working with ServiceNow and Azure DevOps is desirable.
- Previous exposure to PCI-compliant or other regulated environments is a plus.
- Minimum Upper Intermediate English (B2) or Proficient (C1).
Tasks and Responsibilities:
- Maintain the stability, availability, and overall health of hybrid infrastructure spanning on-premises systems and Microsoft Azure.
- Perform ongoing operational activities such as infrastructure monitoring, patch management, upgrades, and incident response.
- Take ownership of system uptime and day-to-day infrastructure reliability in an operations-focused environment.
- Identify repetitive operational activities and implement automation to improve efficiency and reduce manual intervention.
- Build and refine monitoring dashboards and alerts to provide actionable visibility into infrastructure health and performance.
- Respond effectively to outages, contribute to root cause analysis, and implement improvements to reduce the likelihood of recurring incidents.
- Create and continuously maintain operational documentation, including runbooks, diagrams, procedures, and change records.
- Provide clear and timely communication during incidents and other high-pressure operational situations.
- Collaborate within a financial-industry engagement alongside the partner team and report to the Infrastructure and Operations Manager.
Soft Skills:
- Responsibility
- Proactivity
- Flexibility
- Great communication skills