Escalations Engineer - Turkiye
About the role
About JumpCloud®
JumpCloud is Intelligent, Secure IT.
SmokeJumpers are JumpCloud's escalation engineers. We’re the team that gets called when a customer issue is too complex for Support and too urgent to wait in a product team's backlog. When an admin's users can't sign in, a device falls out of a group and loses its policies, or a sync to Google Workspace or Slack breaks, we find out why and get it fixed. We take issues from first report through root cause and resolution, working across logs, production data, and source code, and either fix them ourselves or hand product teams a clear, evidence-backed diagnosis. Additionally, we’re responsible for developing our internal applications and tooling that the entire organization relies on to understand customer profiles and troubleshoot their issues.
JumpCloud delivers cloud-based identity, device, and access management as a SaaS platform, so the problems we see span identity, authentication, security, endpoints, and cloud-scale architecture. Our stack includes Go, Kubernetes, Linux, Terraform, Python, PostgreSQL, MongoDB, and Kafka on AWS and GCP. If you love a hard puzzle, you'll fit right in, and we'll make sure you never run out of them.
What you’ll be doing:
- Most of your time will go to escalated customer issues and the privileged operational work that only SmokeJumpers are trusted to do.
- You'll drive triage, investigate across the whole JumpCloud platform, make careful fixes to production data, and act as the bridge between our global Support team and product engineering, communicating impact to customers and management along the way.
- When there's room, you'll also help build the internal tools that make Support, CSMs, and our own team faster, and you'll be empowered to fix and push code directly.
We’re looking for:
- 4+ years in escalation engineering, technical support engineering (L3), SRE, or software engineering on a SaaS platform
- A proven track record of root-causing complex issues in distributed systems, and the ability to explain clearly how you found the answer
- Strong SQL skills and experience working with complex schemas, as well as NoSQL databases such as MongoDB or Cassandra
- Experience programming in Java, Python, or Go, and comfort reading unfamiliar codebases to confirm how a system behaves
- Experience with logging and observability tools such as Datadog or ELK
- Working knowledge of APIs (REST, gRPC) and SDKs
- A solid understanding of identity and access concepts: SSO (SAML / OIDC), user provisioning (SCIM), MFA, and directory services
- Comfort with the Linux command line and basic container / Kubernetes operations (running pods, reading logs)
- Comfortable using AI-assisted tools (e.g. coding assistants, LLM-based agents) in day-to-day engineering work: troubleshooting, writing and reviewing code, querying data, and accelerating investigation of complex issues
- Strong written and verbal English communication skills, with the ability to coordinate incidents and communicate impact clearly to customers, management, and internal stakeholders across time zones
- Strong team player: this role is both Agile and interrupt-driven, and we're a team that's constantly working together to solve complex problems
Responsibilities:
-
Own escalated customer-found issues end to end: reproduce the problem, isolate the failing layer (endpoint agent, API, backend worker, or data), and drive it to resolution or a well-documented handoff to the owning product team
-
Investigate using Datadog logs and traces, production databases, endpoint agent logs, HAR files, and service source code, and build clear timelines of what happened and why
-
Perform safe, auditable production data operations in PostgreSQL, MongoDB, and MySQL for Support and Engineering teams: query, verify, apply changes in transactions, and clearly document results
-
Handle privileged operational requests, including administrator account recovery with identity verification, organization lifecycle changes, account configuration, and usage and billing reports
-
Serve as Incident Manager to coordinate the quick resolution of customer-impacting incidents, and contribute to post-incident reviews
-
Monitor alerts and on-call channels, and participate in an on-call rotation after your first 6 months of onboarding
-
Write clear, structured updates for Support, root-cause summaries suitable for customers, and precise requests for the data needed to move an investigation forward
-
Spot recurring issues and turn them into runbooks, knowledge base articles, monitors, or automation, and advocate with product teams for permanent fixes
-
Develop our internal SaaS application (JumpDesk), which gives CSE, CSM, TAM, Sales, and Finance teams critical customer data and safe self-service actions, along with other team tooling
-
Provide support for all Product and Engineering teams, and step into ongoing issues or discussions where your knowledge can speed up resolution
Preferred
-
Experience administering Mac, Linux, and Windows devices, including troubleshooting sign-in and authentication problems at the OS level
-
Experience with device management tools such as Apple MDM (ADE / Apple Business Manager / VPP), Microsoft Intune, or Microsoft Active Directory / Entra ID
-
Hands-on experience with identity providers and directory integrations such as Google Workspace, Microsoft 365 / Entra ID, Okta, or HRIS systems
-
Knowledge of RADIUS and LDAP
-
Prior escalation or support engineering experience at an identity, device management, or security SaaS company
-
Experience with AWS, message queues (Kafka, SQS), or Terraform
-
Front-end experience (React / TypeScript) for internal tool development