Systems Engineer
Remote or hybrid. If hybrid, position is based at our corporate office in Charlotte NC.
Position Summary
The Systems Engineer designs and implements infrastructure solutions to support application operations and site reliability. This role focuses on optimizing system performance, scalability, and reliability, ensuring seamless operation of application environments, and takes ownership of day-to-day operational health across supported systems, including participation in on-call support.
Roles and Responsibilities
- Design and deploy infrastructure solutions to support application scalability and reliability.
- Own day-to-day operational responsibilities for supported infrastructure, including monitoring health, responding to incidents, and ensuring systems remain stable and available.
- Participate in an on-call rotation to respond to and resolve infrastructure and application incidents outside of standard business hours.
- Implement and maintain site reliability engineering practices, such as monitoring, automation, and performance optimization.
- Proactively identify gaps in monitoring, coverage, capacity, or resilience, and address them before they result in incidents.
- Research emerging tools, technologies, and industry practices, and recommend and advocate for changes that improve system performance, reliability, or efficiency.
- Collaborate with Database Engineers and Release Engineers to ensure smooth integration of systems.
- Troubleshoot complex infrastructure and application environment issues.
- Update and regularly monitor Azure DevOps boards for projects and other non-operational work.
- Develop and maintain documentation for infrastructure and operational processes.
Required Qualifications
- 5+ years of experience in systems engineering, infrastructure engineering, or a related operational role.
- Experience with cloud infrastructure platforms (Azure preferred).
- Hands-on experience with monitoring/observability tools (e.g., Datadog).
- Experience with Azure DevOps or similar work-tracking/CI-CD platforms.
- Strong troubleshooting skills across networking, compute, storage, and application layers.
- Scripting/automation experience (PowerShell, Python, or similar).
- Demonstrated experience owning operational/production support responsibilities, not just project delivery.
- Willingness and ability to participate in a shared on-call rotation.
Preferred Qualifications
- Experience with in-memory data platforms or caching technologies.
- Experience with workload automation/job scheduling tools.
- Familiarity with site reliability engineering (SRE) principles and practices.
- Experience working in a regulated or enterprise environment (insurance, financial services, etc.).
Key Competencies
- Strong written and verbal communication, with the ability to work across technical and non-technical teams.
- Proactive and forward-looking mindset — able to spot risks, gaps, or emerging issues before they escalate into production problems.
- Curious and research-driven, staying current on new tools and approaches and comfortable making the case for change to peers and leadership.
- Ability to manage competing priorities across operational and project-based work.
- Detail-oriented documentation habits.
- Self-starter who takes initiative without needing constant direction and follows through on work independently.
- Comfortable working autonomously, exercising sound judgment on when to act versus when to escalate.
Pursuant to California regulation, the compensation range for this position is as stated and includes eligibility for performance-based bonuses.
California Pay Range: $110,000 USD - $150,000 USD