Software Development Engineer II - Production Reliability & Automation

Dentira
Dentira

Software Engineering

Bangalore Urban, Karnataka, India · Remote

Posted on Aug 6, 2026

About Dentira

Dentira is one of the fastest-growing and profitable healthcare startups, transforming clinic operations across multiple countries. Our flagship product revolutionized dental procurement with the world's first real-time price and product comparison engine for dental supplies across the US, Canada, Australia, and New Zealand.

Building on this success, our new Lab Platform streamlines case creation and management by connecting clinics, intraoral scanners, and labs. We are also leveraging advanced AI to solve complex challenges such as automatic invoice reconciliation and procedure cost prediction—solutions with broad applicability across healthcare.

As we expand into adjacent healthcare verticals, our goal is to achieve 10X growth over the next three years.

Role Overview

As an SDE II – Production Reliability & Automation, you'll own the reliability of Dentira's distributed backend platform that powers healthcare operations across multiple countries. This is not a technical support role—you'll investigate complex production incidents, debug distributed systems, automate recurring operational work, and build tooling that improves platform resilience.

You'll work with production systems spanning services, event brokers, databases, workflows, and cloud infrastructure while collaborating with engineering teams to eliminate recurring issues through automation rather than repeatedly fixing them.

Team: Backend / Distributed Systems
Location: Bengaluru, India
Shift: 5:00 PM – 2:00 AM IST (Monday – Friday)

Primary Responsibilities

  • Own end-to-end production incident investigation across distributed services, messaging systems, databases, workflows, and search platforms.
  • Identify root causes and implement permanent engineering solutions rather than temporary fixes.
  • Build automation tools including diagnostic utilities, reconciliation scripts, self-healing jobs, and AI-powered operational agents.
  • Develop internal tooling to improve incident response, troubleshooting, and production visibility.
  • Create and maintain technical documentation, flow diagrams, runbooks, and operational playbooks.
  • Collaborate with engineering teams during daily handovers to ensure smooth production operations.
  • Write clear incident timelines, technical analysis, and root cause documentation.
  • Continuously improve system reliability, observability, and operational excellence.

Required Skills & Qualifications

Experience

  • 4+ years of experience building and owning backend production systems.
  • Proven experience leading high-severity production incidents and driving root cause analysis.
  • Strong debugging skills with the ability to troubleshoot unfamiliar codebases under pressure.

Technical Skills

  • Strong programming experience in Node.js (TypeScript) or Go.
  • Solid Linux shell and command-line proficiency.
  • Hands-on experience with Kafka or RabbitMQ.
  • Strong understanding of SQL/NoSQL database performance and query optimization.
  • Experience with workflow orchestration platforms such as Temporal or similar technologies.
  • Experience with observability platforms like Prometheus, Grafana, Elasticsearch/OpenSearch, or equivalent.
  • Working knowledge of AWS and Kubernetes.
  • Strong understanding of distributed systems concepts including retries, idempotency, ordering guarantees, partial failures, and backpressure.

AI & Communication

  • AI-native engineering mindset with experience using coding agents, custom AI workflows, MCP servers, or developer automation tools.
  • Excellent written and verbal communication skills in English.
  • Strong ownership mindset with the ability to work independently during production support hours.

Good to Have

  • Experience with Terraform or other Infrastructure-as-Code tools.
  • Experience designing or establishing on-call and incident management practices.
  • Built internal engineering tools or AI agents adopted by other teams.
  • Experience in healthcare, fintech, or other highly regulated industries.
  • Passion for automation and continuous platform improvement.
  • Note: Interviews will be conducted offline (face-to-face) at our Bangalore office located in Garudacharpalya. Please apply only if you are available to attend the interview in person.