✨ AI Summary
As a Technology Support III team member, you will ensure the operational stability, availability, and performance of production application flows. This role involves providing end-to-end application or infrastructure service delivery, supporting day-to-day system maintenance, and participating in incident response activities. You will use enterprise-authorized AI capabilities for incident triage and trend analysis, execute incident, problem, and change management procedures, and contribute to post-incident reviews and root cause analysis to prevent recurrence. The role requires analyzing complex situations, anticipating issues, and supporting full stack technology systems. A minimum of 3 years of experience in IT services troubleshooting and maintenance is required, along with knowledge of large-scale technology environments, observability tools, and ITIL framework. Preferred qualifications include prior production support experience, strong critical thinking, problem-solving, public speaking, and facilitation skills.
🎁 Benefits & Perks
Comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching.
Requirements
Requires 3+ years of experience in troubleshooting, resolving, and maintaining IT services, with demonstrated experience using enterprise-authorized AI capabilities for production support. Must have knowledge of applications or infrastructure in large-scale environments (on-premises and public cloud), experience with observability and monitoring tools, and exposure to ITIL framework processes. Prior experience in a production or enterprise technology environment, strong critical thinking, problem-solving, public speaking, and facilitation skills are preferred.
Description
Propel operational success with your expertise in technology support and a commitment to continuous improvement.
As a Technology Support III team member in Employee Platforms Production Management, you will ensure the operational stability, availability, and performance of our production application flows. Encourage a culture of continuous improvement as you troubleshoot, maintain, identify, escalate, and resolve production service interruptions for all internally and externally developed systems, leading to a seamless user experience.
Job responsibilities
- Provides end-to-end application or infrastructure service delivery and supports day-to-day maintenance of the firm’s systems to ensure operational stability and availability
- Uses enterprise-authorized AI capabilities within the work environment to speed up incident triage and trend analysis from operational signals, validating outputs and handling operational data according to sensitivity and security requirements.
- Participate in incident response activities by gathering facts, assessing impact, documenting timelines, coordinating updates, and helping drive service restoration.
Execute assigned incident, problem, and change management procedures in alignment with operational, resiliency, audit, and control expectations.
Contribute to post-incident review, root cause analysis, action tracking, and recurrence prevention by documenting findings, validating evidence, and following through on assigned remediation items.
- Assist in the monitoring of production environments for anomalies, address issues utilizing standard observability tools, and identify issues for escalation and communication while providing solutions to business and technology stakeholders
- Analyze complex situations and trends to anticipate and solve issues while supporting incident, problem, and change management of full stack technology systems, applications, or infrastructure
- Applies reuse-first, AI-assisted practices within incident/problem/change routines to identify recurring interruption patterns and support validated remediation actions aligned to resiliency and security expectations.
Required qualifications, capabilities, and skills
- A. 3+ years of experience or equivalent expertise troubleshooting, resolving, and maintaining information technology services
Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support production support workflows with strong validation habits and awareness of data sensitivity. - Ability to review and validate AI-assisted incident recommendations before action, escalating when uncertain and following operational and security expectations.
- Demonstrated knowledge of applications or infrastructure in a large-scale technology environment both on premises and public cloud
- Experience in observability and monitoring tools and techniques
- Exposure to processes in scope of the Information Technology Infrastructure Library (ITIL) framework
Preferred qualifications, capabilities, and skills
Prior experience or expertise troubleshooting, resolving, and maintaining information technology services in a production or enterprise technology environment.
Strong critical thinking skills, including the ability to separate symptoms from root causes, ask effective questions, assess evidence, and make sound recommendations.
Demonstrated problem-solving ability with a practical, organized approach to diagnosing issues, documenting findings, and following through to resolution.
Comfort with public speaking, facilitation, and presenting technical or operational updates to groups in a clear, concise, and professional manner.
Curiosity, accountability, composure under pressure, attention to detail, and willingness to learn from incidents, peers, and operational feedback.
Working knowledge of incident, problem, and change management concepts, including escalation, impact assessment, post-incident review, and remediation tracking.