01Responsibilities
Design, develop and maintain backend infrastructure, workflows, and services with a focus on reliability, scalability, and operability (SRE principles) Develop solutions to support onboarding, partner integrations, managing, collecting, and analyzing data from large-scale deployments of home networks Embed within engineering teams to drive production readiness, service health, and continuous improvement of reliability metrics (SLIs/SLOs) Work closely with Cloud product owners to understand, analyze product requirements, provide feedback, and deliver a complete solution Technical leadership in software design to meet requirements of service stability, availability, scalability, and security Drive technical discussions across SDLC phases including requirements, design, peer reviews, and test strategy Own observability, monitoring, alerting, and incident response improvements; partner with TAC and operations teams for faster resolution Support test strategy and automation in both end-to-end solution and functional testing Customer-facing engineering role in debugging and resolving field issues Drive root cause analysis (RCA), post-incident reviews, and ensure systemic fixes and prevention mechanisms Qualifications: 10+ years of highly technical, hands-on development experience in any programming language Independent, self-driven, and able to work in a team environment Strong problem-solving skills with ability to abstract and communicate effectively Ability to drive technical discussions across cross-functional teams Proficient in design and implementation of microservices-based, API/endpoint architectures Strong background in event-based / pub-sub workflows & data ingestion solutions (Kafka or similar) Good understanding of cloud-based solutions (preferably AWS or GCP) Strong background in transactional databases and experience with NoSQL data stores Experience with monitoring/observability tools (Prometheus, Grafana, OpenTelemetry or similar) Experience with incident management, production support, and reliability engineering practices Understanding of scaling, resiliency patterns, and failure handling in distributed systems Experience with IoT/home gateway protocols (TR-069/TR-369 etc.) a plus Expert in Java; experience in Go/Python/NodeJS a plus Experience with streaming/data platforms (Kafka, Spark, Flink, etc.) Experience building scalable data solutions (Spanner, Elastic, etc.) Practical understanding of AWS/GCP cloud platform Education: BS degree in Computer Science, engineering, or equivalent experience Location: India (Flexible hybrid work model - work from Bangalore office for 20 days in a quarter) PLEASE NOTE: All emails from Calix will come from a '@calix.com' email address. Please verify and confirm any communication from Calix prior to disclosing any personal or financial information. If you receive a communication that you think may not be from Calix, please report it to us at talentandculture@calix.com.