Base Career helps you apply smarter for this job.
Key skills for this role
Explore Your Next Role at Sapiens
Jobs
Senior
Bangalore, IN
Sapiens is on the lookout for a Lead Observability Engineer/ APM Engineer to become a key player in our Bangalore team, as we continue to scale observability across a large, multi-customer managed service environment. If you're an expert in Datadog APM who thrives inside traces, service maps, and flame graphs — and is equally comfortable in design forums shaping enterprise monitoring strategy — this role could be the perfect fit, offering high autonomy, deep technical ownership, and a seat at the table in defining how every application we support gets observed, understood, and kept healthy.
Sapiens International Corporation N.V. is a global leader in intelligent, SaaS-based software solutions. With Sapiens’ robust platform, customer-driven partnerships, and rich ecosystem, insurers are empowered to future-proof their organizations with operational excellence in a rapidly changing marketplace.
Our solutions help insurers harness the power of AI and advanced automation to support core solutions for property and casualty, workers’ compensation, and life insurance, including reinsurance, financial & compliance, data & analytics, digital, and decision management.
Sapiens boasts a longtime global presence, serving over 600 customers in more than 30 countries with our innovative offerings. Recognized by industry experts and selected for the Microsoft Top 100 Partner program, Sapiens is committed to partnering with our customers for their entire transformation journey and is continuously innovating to ensure their success.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
, USA
Title: Director of Software Development Location : Remote Position Summary The Director of Software Development is a key leadership role responsible for driving product vision, leading research and development initiative
, USA
Title: Account Executive, P&C Location: Remote Job Description The Account Executive is responsible for growing software and services revenue within an assigned portfolio of existing insurance clients. In this role, you
London, GBR
Sapiens is seeking a Regional Lead to direct customer service management and operational excellence for a regional enterprise SaaS organization. The role requires extensive customer success or service management leadersh
, USA
, USA
Bengaluru, IND
, USA
, USA
, USA
, USA
London, GBR
We are looking for an expert Datadog observability engineer whose primary depth is in Application Performance Monitoring. You will own the APM practice across a large, multi-customer managed service environment — defining the architecture, instrumenting applications, leading deep-dive performance investigations, and setting the standards that every application follows when it is onboarded.
This is a hands-on technical role. You will spend your time inside traces, service maps and flame graphs, and equally in design forums and steering committees where APM strategy, coverage and cost are decided. Alongside expert APM skills, we expect strong all-round Datadog observability capability spanning infrastructure, logs, database monitoring, network, Kubernetes, RUM, synthetics, ServiceNow integration and incident management.
Improve service reliability and customer experience across all supported applications.
Reduce mean time to detect and mean time to resolve through trace-led root cause analysis.
Raise alert quality by replacing symptom-based alerting with dependency-aware, service-level detection.
Establish enterprise APM standards that make onboarding repeatable, governed and cost-aware.
Design and implement Datadog APM across enterprise applications and microservices.
Lead APM onboarding for new and existing applications, end to end.
Define the APM architecture, instrumentation standards and deployment patterns.
Configure Datadog Agents and APM components across all application environments.
Implement and troubleshoot distributed tracing across multi-tier applications.
Trace end-to-end transactions across microservices and APIs, including asynchronous paths.
Pinpoint performance bottlenecks along the request path.
Analyse trace latency, errors, throughput and resource utilisation.
Use trace analytics to detect performance degradation and transaction bottlenecks early.
Hands-on experience instrumenting applications with Datadog APM across enterprise technologies, including:
Java
.NET and C#
Node.js
Python
Other enterprise application technologies and runtimes
Practical experience with common frameworks, application servers and microservices architectures is highly desirable, as is familiarity with automatic versus manual instrumentation, custom spans and library-level configuration.
Configure appropriate trace collection and sampling strategies per service and environment.
Configure error tracking and analytics.
Optimise APM telemetry volume and cost by tuning sampling and removing unnecessary trace collection, without losing diagnostic value.
Correlate application traces with database activity using Datadog Database Monitoring.
Identify slow SQL and database calls that contribute to application latency.
Analyse database query performance from the application perspective.
Troubleshoot connection pool exhaustion and database dependency issues.
Work closely with DBA teams across SQL Server, Oracle, PostgreSQL and other database platforms.
Build application dependency maps using Datadog APM and Universal Service Monitoring.
Identify upstream and downstream dependencies and critical transaction paths.
Establish and maintain service ownership through the Datadog Service Catalogue.
Support service-level observability and application health monitoring.
Implement and support APM for applications running on Azure, AWS and Kubernetes / AKS.
Correlate Kubernetes and container metrics with application traces.
Isolate whether a performance issue originates in the application, the container, the Kubernetes node, the database, the network or an external dependency.
Integrate Datadog APM events with ServiceNow ITOM and Event Management.
Define how APM events are converted into actionable events and incidents.
Ensure appropriate event aggregation, deduplication and suppression during maintenance windows.
Support CI and service mapping between Datadog and the ServiceNow CMDB.
Ensure APM-driven incidents are routed to the correct resolver groups.
Develop application-level dashboards covering availability, latency, error rate, throughput, dependency health and database performance, that answer: is the service healthy, what changed, what is the impact, where is the bottleneck, and what should the operator do next.
Define application health indicators and establish SLIs, SLOs and error budgets for critical services.
Establish enterprise APM standards and onboarding templates.
Define tagging, naming and ownership conventions across the estate.
Define monitoring requirements based on application architecture and business criticality.
Review APM coverage regularly and identify gaps.
Define and drive the APM maturity roadmap.
This role carries influence well beyond the tooling. We are looking for the following additional skills and attributes.
Problem solving — structured, evidence-led analysis under pressure; comfortable working from an ambiguous symptom to a proven root cause, and honest about what the data does and does not show.
Communication — explains complex performance behaviour clearly to engineers, application owners and executives; writes concise incident narratives and RCA documents that stand up to customer scrutiny.
Stakeholder management — builds credibility with application teams, DBAs, cloud engineers, service delivery managers and customers across distributed, multi-vendor teams; negotiates standards rather than imposing them.
Ownership and accountability — takes end-to-end responsibility for APM coverage and quality, and follows issues through to preventive action.
Continuous improvement — actively looks for opportunities to extend coverage, increase automation and cut manual effort, and measures the result.
Executive readiness — presents observability posture, risks, quick wins and roadmap in a form suitable for steering committee discussion.
Essential — Datadog certification, including at least one specialist certification relevant to APM.
Desirable — cloud certification such as Microsoft Azure Administrator / Solutions Architect or AWS Solutions Architect.
Desirable — Certified Kubernetes Administrator (CKA) or equivalent container platform certification.
Desirable — ITIL Foundation v4, and familiarity with SRE practices and frameworks.
Desirable — ServiceNow ITOM or Event Management accreditation.
Sapiens is an equal opportunity employer. We value diversity and strive to create an inclusive work environment that embraces individuals from diverse backgrounds.
Disclaimer: Sapiens India does not authorise any third parties to release employment offers or conduct recruitment drives via a third party. Hence, beware of inauthentic and fraudulent job offers or recruitment drives from any individuals or websites purporting to represent Sapiens. Further, Sapiens does not charge any fee or other emoluments for any reason (including without limitation, visa fees) or seek compensation from educational institutions to participate in recruitment events. Accordingly, please check the authenticity of any such offers before acting on them and were acted upon, you do so at your own risk. Sapiens shall neither be responsible for honouring or making good the promises made by fraudulent third parties, nor for any monetary or any other loss incurred by the aggrieved individual or educational institution.
If you come across any fraudulent activities in the name of Sapiens, please feel free report the incident at sapiens to sharedservices@sapiens.com
Sapiens International Corporation N.V. is a global leader of AI-centric, SaaS-based insurance software, delivering hyper-relevant experiences that are efficient, compliant, and innovative. With agile intelligence, Sapiens' solutions turn real-time data and human insight into precise action at every moment, across every risk.
The Sapiens platform includes agentic workflows accelerating every capability across policy, underwriting, claims, reinsurance, decisioning, and finance and compliance.
With more than 600 insurers in over 30 countries running on Sapiens, our deep industry expertise is the foundation of our long-term relationships, from initial implementation through to modernization and market transformation.
Sapiens is headquartered in London, serving customers in property and casualty, life, reinsurance, specialty, and workers’ compensation from offices across North America, Europe, the Middle East, and Asia Pacific.
We will keep you in the loop, as we focus on providing an inclusive screening and interview process. Each country has a local flavor, but here's what you can expect during our recruitment process:
Sapiens India does not authorize any third parties to release employment offers or conduct recruitment drives via a third party. Hence, beware of inauthentic and fraudulent job offers or recruitment drives from any individuals or websites purporting to represent Sapiens. Further, Sapiens does not charge any fee or other emoluments for any reason (including without limitation, visa fees) or seek compensation from educational institutions to participate in recruitment events. Accordingly, please check the authenticity of any such offers before acting on them and were acted upon, you do so at your own risk. Sapiens shall neither be responsible for honoring or making good the promises made by fraudulent third parties, nor for any monetary or any other loss incurred by the aggrieved individual or educational institution. In the event that you come across any fraudulent activities in the name of Sapiens, please feel free report the incident at sapiens to sharedservices@sapiens.com .
Privately held insurance software provider serving insurers with AI and SaaS systems for policy, underwriting, claims, billing.
Visit company websiteFull-time
Senior
Hybrid
Apply faster on company sites with our extension.