BETSOL logo

Senior SDET – Performance Engineering

BETSOL India

onsitefull-time
Posted Aug 12, 2026Apply by Feb 8, 2027

**Role & Seniority: ** Senior SDET (Senior Software Development Engineer in Test) specializing in Performance Engineering; 10+ years QA/automation with performance focus

**Stack/Tools: **

  • Performance/load testing: k6, JMeter, Gatling (plus custom framework/simulators)

  • Programming/scripting: Python, C#, Java (or similar)

  • Cloud/K8s: Azure, AKS, compute/storage/networking tuning

  • CI/CD: Jenkins, GitHub Actions, GitLab

  • Observability: Prometheus, Grafana, APM tools (Azure Monitor/Application Insights mentioned)

  • Data layer: Redis caching; DB tuning for MariaDB/MySQL/etc; messaging systems

  • Resilience: Chaos Monkey / Chaos engineering tools

  • Optional AI/AIOps tooling: Azure AI (GenAI/AIOps), Dynatrace, Datadog, New Relic, ELK/ML, etc.

  • Top 3 Responsibilities:

    • Design/implement end-to-end performance + load testing for microservices in Azure

    • Build automated performance test frameworks integrated into CI/CD and create simulators/mocks/traffic generators

    • Define/track performance KPIs and partner with engineering/SRE on bottleneck analysis, resilience/chaos testing, and RCA

  • Must-have Skills:

    • Strong performance testing tool experience (incl. custom framework development)

    • Deep knowledge of distributed systems + microservices

    • Hands-on Kubernetes (AKS preferred) and Azure ecosystem

    • SRE foundations (SLI/SLO, e

Full Description

Engineering the AI-powered enterprise. With AI and cloud-native solutions, BETSOL accelerates cloud transformation for enterprises across 17+ countries. BETSOL holds several engineering patents, and is recognized with industry awards. BETSOL maintains a net promoter score that is 2x the industry average.

BETSOL’s open source backup and recovery product line, Zmanda (Zmanda.com), delivers up to 50% savings in total cost of ownership (TCO) and delivers best-in-class performance.

BETSOL Global IT Services (BETSOL.com) builds and supports end-to-end enterprise solutions, reducing time-to-market for customers.

We take pride in being an employee-centric organization, offering comprehensive benefits and opportunities.

Learn more at betsol.com

Job Description

About the Role

We are looking for a Senior SDET specializing in Performance Engineering in a cloud-native Azure environment. This role focuses on driving scalability, reliability, and performance validation across distributed microservices systems. The candidate will design automated performance frameworks, build simulators and mocks, define KPIs, and partner with engineering, architecture, and SRE teams to ensure production-grade resilience.

Responsibilities

Design and implement end-to-end performance and load testing strategies for microservices-based systems Build custom simulators, traffic generators, and mocks for complex system dependencies Define, measure, and track performance KPIs (latency, throughput, error rate, saturation, scalability limits) Develop fully automated performance test frameworks integrated into CI/CD pipelines (Jenkins, GitHub Actions, GitLab) Execute load, stress, spike, endurance, and chaos testing Collaborate with architects, developers, product owners, and SRE teams to optimize system performance Analyze bottlenecks across application, database, and infrastructure layers Work closely with Azure services (AKS, compute, storage, networking) for performance tuning Implement observability using Prometheus, Grafana, and APM tools Optimize Redis caching, database queries (MariaDB, MySQL, etc), and messaging systems Support resilience engineering and chaos testing (Chaos Monkey or equivalent) Drive RCA for performance issues and production incidents Contribute to capacity planning and scalability strategy

Looking For

10+ years experience in QA, development and automation, with strong focus on performance engineering

Mandatory Skills

Technical Skills Strong experience in performance testing tools (K6, JMeter, Gatling, and creating custom frameworks) Proficiency in scripting (Python, C#, Java, or similar) Deep understanding of distributed systems and microservices architecture Hands-on experience with Kubernetes (AKS preferred) Strong knowledge of Azure cloud ecosystem Experience with CI/CD and DevOps practices Understanding of SRE principles (SLI/SLO, error budgets) Experience with observability and monitoring tools Strong database performance tuning expertise Soft Skills Excellent written and verbal English Calm, structured communication with both engineers and non-technical stakeholders. Strong problem-solving and systems thinking. Ownership mindset and clear accountability for outcomes.

Good To Have Skills

Experience in contact center / SaaS platforms Exposure to Kafka, RabbitMQ Knowledge of AIOps, AI-driven testing and anomaly detection Experience building custom performance tools or simulators

Qualifications

Bachelor’s/Master’s in Computer Science or related field

Additional Information

AI-Driven Performance Engineering (GenAI & AIOps)

Leverage Generative AI (GenAI) to auto-generate performance test scenarios, workloads, and synthetic datasets Implement AI-driven anomaly detection for identifying performance regressions and system bottlenecks Use machine learning models for predictive capacity planning and workload forecasting Integrate AIOps tools for intelligent alerting, noise reduction, and automated root cause analysis (RCA) Apply AI techniques for log analysis, pattern recognition, and failure prediction Build self-healing test systems with automated remediation triggers Enhance observability platforms (Prometheus, Grafana) with AI-based insights Utilize AI for dynamic test optimization based on real-time system behavior Collaborate with data science teams to implement advanced analytics for performance insights

AI & Observability Tooling (Real-World Examples)

Azure Monitor + Application Insights (with AI capabilities): Smart detection, failure anomaly detection, and auto-root cause insights

Azure OpenAI / GenAI integrations: Generate performance scenarios, synthetic workloads, and intelligent test data

Dynatrace (Davis AI): Automatic dependency mapping, causal AI for root cause analysis, and real-time anomaly detection

Datadog AI / Watchdog: Automated anomaly detection, performance regression identification, and alert correlation

New Relic AI: Predictive alerting and performance intelligence across distributed systems

Prometheus + Grafana (with ML plugins): Advanced metric analysis and anomaly detection extensions

Elastic Stack (ELK) with ML: Log anomaly detection, pattern recognition, and predictive insights

Chaos Engineering tools (Gremlin, Chaos Monkey): Integrated with observability platforms for resilience validation

k6 + AI-based extensions: Intelligent load modeling and performance insights

Custom AI/ML pipelines: Python-based models for predictive scaling, workload modeling, and anomaly detection

Working Hours

General shift with flexibility to support business and stakeholder requirements.

May require coordination with global stakeholders across different time zones as per business needs

Performance EngineeringK6JMeterGatlingPythonC#JavaKubernetesAzureCI/CDSREPrometheusGrafanaDatabase TuningMicroservicesChaos Engineering

Cookies & analytics consent

We serve candidates globally, so we only activate Google Tag Manager and other analytics after you opt in. This keeps us aligned with GDPR/UK DPA, ePrivacy, LGPD, and similar rules. Essential features still run without analytics cookies.

Read how we use data in our Privacy Policy and Terms of Service.