More from this employer
Bengaluru, IND
Workday Technical Resource Required Skills Workday Integrations – Design, build, and maintain integrations (EIB, Core Connectors, Studio) Workday Reporting – Create and modify custom reports, calculated fields, and repor
, IND
Required Skills Playwright Automation Strong hands-on experience with Playwright and TypeScript. Experience automating web applications across Chromium, Firefox, and WebKit browsers. Knowledge of Page Object Model (POM),
, IND
Role Overview We are seeking a skilled and motivated Data Engineer to design, develop, and maintain scalable data pipelines and data solutions. The ideal candidate will work closely with business stakeholders, data analy
Bengaluru, IND
Data Management & Governance Ensure data quality, integrity, and reliability through robust validation and monitoring frameworks. Implement best practices for data lineage, metadata management, and data cataloging using
Pune, IND
We are looking for experienced professionals with a strong background in AI and software development to join our dynamic team. As a GenAI Engineer, you will be responsible for designing, developing, and implementing AI-p
, IND
Key Responsibilities Design and develop enterprise RAG applications using LLMs, embeddings, vector databases, and hybrid search. Build end-to-end document ingestion and knowledge ingestion pipelines for structured and un
Hyderabad, IND
8+ years of hands-on experience in Ruby on Rails application development. Strong proficiency in Ruby programming language and Rails framework. Experience building scalable, secure, and high-performance web applications.
Pune, IND
J ob Description We are looking for a skilled Data Engineer with strong experience in data analysis, data modeling, and modern data platforms. The ideal candidate should have hands-on expertise in DBT and SQL, along with
Bengaluru, IND
, IND
, IND
Bengaluru, IND
Pune, IND
, IND
Hyderabad, IND
Pune, IND
Base Career helps you apply smarter for this job.
Site Reliability Engineer (SRE – SaaS Platform Operations)
Site Reliability Engineer (SRE – SaaS Platform Operations)
Experienced Azure-based SRE required to support a large-scale enterprise SaaS platform, with strong Microsoft SQL Server expertise, future PostgreSQL readiness, production operations, patching, automation, troubleshooting, and incident response capability.
Microsoft Azure, Microsoft SQL Server, PostgreSQL, Windows/Linux, Monitoring, Automation
Upcoming releases include database movement from MSSQL to PostgreSQL; PostgreSQL operational support is expected to become increasingly important.
4+ years relevant experience; strong MSSQL production support background; PostgreSQL exposure preferred
24x7 production support environment; US time zone support/night shifts as required
Client, delivery leadership, hiring panel, and internal stakeholders
We are seeking an experienced Site Reliability Engineer (SRE) to support a large-scale enterprise SaaS platform operating in cloud and high-availability environments.
This role is responsible for maintaining infrastructure reliability, availability, performance, and operational excellence across Microsoft Azure, Microsoft SQL Server, PostgreSQL readiness, Windows, Linux, monitoring, automation, and incident response functions.
This is not a first-line helpdesk role.
The role requires a hands-on engineer who can independently investigate complex technical issues, collaborate with engineering and platform teams, and provide clear technical communication to internal and client stakeholders.
Future Scope: In upcoming product releases, selected databases are expected to move from Microsoft SQL Server to PostgreSQL.
The role should therefore include PostgreSQL awareness and operational readiness in addition to current MSSQL responsibilities.
Manage and support Azure infrastructure, monitoring, storage, compute, and patching activities.
Stable, secure, and scalable platform operations.
Administer MSSQL workloads today and support PostgreSQL readiness for future releases, including tuning, maintenance, backups, and HA/DR.
Improved database performance, resiliency, and recoverability.
Plan and support SQL/database patching and operating system patching across Windows and Linux environments.
Improved security compliance and reduced operational risk.
Perform initial analysis, support P1/P2 triage, contribute to RCA, and improve runbooks.
Reduced MTTR and stronger production readiness.
Use PowerShell/Python, alert tuning, logging, dashboards, and Infrastructure-as-Code practices.
Better operational efficiency and proactive issue detection.
Reproduce issues, validate defects, and escalate with evidence to product engineering.
Faster defect resolution and improved customer experience.
Maintain highly available, reliable, and scalable cloud infrastructure in Microsoft Azure.
Monitor platform health, review technical logs, and proactively address performance and availability issues.
Improve infrastructure monitoring, alerting, and logging to support proactive reliability management.
Plan, coordinate, and support operating system patching and maintenance activities across Windows and Linux servers.
Ensure security updates, compliance patches, and platform upgrades are executed in line with change management processes while minimizing service disruption.
Drive operational excellence through standardization, automation, and continuous improvement.
Manage and support Microsoft SQL Server databases across production and non-production environments on Azure.
Support PostgreSQL operational readiness and future PostgreSQL database support as selected databases move from MSSQL to PostgreSQL in upcoming releases.
Troubleshoot and tune queries, SQL jobs, indexing, CPU, memory, I/O, and storage utilization.
Implement and maintain backup, restore, high availability, disaster recovery, and routine database maintenance processes.
Perform SQL Server and PostgreSQL database patching, upgrades, maintenance, and version lifecycle activities under enterprise change management processes.
Use T-SQL and scripts for investigations, reporting, and controlled production data fixes under change control.
Investigate complex infrastructure, application, database, Windows, Linux, and Azure-related issues.
Participate in incident response, perform initial technical analysis, and contribute to root cause analysis documentation.
Reproduce customer-reported issues in staging or lab environments where required.
Escalate verified product defects to engineering teams with clear technical evidence, logs, and impact analysis.
Create and maintain operational automation using PowerShell, Python, or equivalent scripting languages.
Support Infrastructure-as-Code and configuration management practices using tools such as Terraform and Puppet.
Collaborate on CI/CD and operational tooling improvements using platforms such as GitHub and Jenkins.
Enhance monitoring and alert management practices for SaaS product operations.
Collaborate closely with onsite DB SREs, application teams, platform teams, product support, and engineering.
Provide clear and timely communication on issue status, technical findings, risks, and next steps.
Join customer or onsite calls when required to explain technical findings and remediation actions.
Update runbooks, knowledge base articles, troubleshooting guides, and operational documentation.
Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Information Technology, or a related technical field.
4+ years of relevant experience in SRE, production operations, database administration, cloud infrastructure, or enterprise platform support.
2+ years of SaaS pipeline or SaaS product operations experience.
2+ years of hands-on experience managing Microsoft Azure cloud infrastructure and cloud monitoring tools.
4+ years in Microsoft SQL Server-centric roles such as DBA, Production Support Engineer, or Database SRE in 24x7 environments.
PostgreSQL exposure or willingness to support PostgreSQL environments as part of future platform modernization scope.
supporting SQL/database patching and operating system patching activities in controlled production environments.
Willingness and ability to work night shifts or US time zone aligned support coverage when required.
MSSQL 2016+, T-SQL, query tuning, SQL jobs, stored procedures, indexing, backup/restore, HA/DR
PostgreSQL administration, performance tuning, backup/recovery, replication, HA awareness, and migration/coexistence readiness
SQL/database patching, OS patching, version lifecycle management, maintenance planning, and controlled change execution
Microsoft Azure infrastructure, Azure VMs, Azure storage, Azure SQL or SQL on Azure VMs, Azure monitoring
Windows Server administration and working knowledge of Linux servers
Application performance monitoring, alert management, logging, incident response, RCA, runbooks
PowerShell, Python, scripting for investigations, reporting, and operational automation
Docker, Kubernetes, and microservices architecture familiarity
Experience setting up, configuring, and improving monitoring and tooling for SRE/DevOps operations of a SaaS product.
supporting PostgreSQL databases in enterprise production environments.
with database migration, modernization, or coexistence initiatives involving Microsoft SQL Server and PostgreSQL.
Knowledge of PostgreSQL performance tuning, backup/recovery, replication, and high availability architectures.
with PagerDuty or similar incident notification and escalation platforms.
with Elasticsearch is a strong advantage.
Ability to perform application debugging and collaborate effectively with Product Support, Platform, Database, and Engineering teams.
Ability to conduct research, evaluate technology options, make recommendations, and maintain a high level of expertise in systems software and operational tooling.
Strong understanding of compliance, policies, procedures, patching windows, and change control expectations in enterprise production environments.
Independent technical troubleshooting
Customer-focused communication
Production incident ownership
Analytical thinking and RCA mindset
Cross-functional collaboration
Operational discipline and documentation
Candidate screening should prioritize strong MSSQL production operations and Azure infrastructure experience, while also considering PostgreSQL exposure or readiness due to upcoming platform releases where selected databases are expected to move from MSSQL to PostgreSQL.
The role is best suited for an engineer who can combine database reliability, cloud operations, patching, automation, and incident response in a large-scale SaaS environment.
MSSQL DBA / production support background
Pure DevOps profile with limited database depth
Azure infrastructure operations experience
NOC/helpdesk-only background
PostgreSQL exposure or database modernization experience
Database developer with no production operations
SQL and OS patching experience in controlled environments
Cloud-only profile without MSSQL troubleshooting
Hands-on incident/RCA ownership
No exposure to controlled patching/change management
PowerShell/Python automation
Limited ownership in incident response
Comfortable with US time zone support
No willingness to support future PostgreSQL scope
This role provides the opportunity to contribute directly to reliability, performance, security, and operational maturity for a large-scale enterprise SaaS platform.
The selected candidate will support the current Microsoft SQL Server estate, help prepare for future PostgreSQL adoption, and work with modern cloud, database, automation, patching, and monitoring technologies while collaborating closely with experienced SRE, database, platform, support, and engineering teams.
Global technology firm providing digital transformation and infrastructure services.
Visit company websiteJobs and hiring trendsSkip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career