Job Description: Job Title: Staff Engineer, AI/ML _RIO(OSS) Location: Bangalore (Hybrid) Why should you choose us? Rakuten Symphony is reimagining telecom, changing supply chain norms and disrupting outmoded thinking that threatens the industry’s pursuit of rapid innovation and growth. Based on proven modern infrastructure practices, its open interface platforms make it possible to launch and operate advanced mobile services in a fraction of the time and cost of conventional approaches, with no compromise to network quality or security.
Rakuten Symphony has operations in Japan, the United States, Singapore, India, South Korea, Europe, and the Middle East Africa region. For more information, visit: https://symphony.rakuten.com. Building on the technology Rakuten used to launch Japan’s newest mobile network, we are taking our mobile offering global. To support our ambitions to provide an innovative cloud-native telco platform for our customers, Rakuten Symphony is looking to recruit and develop top talent from around the globe.
We are looking for individuals to join our team across all functional areas of our business – from sales to engineering, support functions to product development. Let’s build the future of mobile telecommunications together! About Rakuten Group, Inc. (TSE: 4755) is a global leader in internet services that empower individuals, communities, businesses and society. Founded in Tokyo in 1997 as an online marketplace, Rakuten has expanded to offer services in e-commerce, fintech, digital content and communications to 2 billion members around the world.
The Rakuten Group has over 30,000 employees, and operations in 30 countries and regions. For more information visit https://global.rakuten.com/corp/ . About the RIO Team RIO (Rakuten Intelligent Operations) is Rakuten Symphony's AI-first operational intelligence platform and the OSS engineering team behind it. RIO replaces fragmented OSS tooling with a single, intelligent control layer for complex, multi-vendor telecom networks: unified real-time observability across RAN, Core, Transport and Cloud; AI-driven service assurance with anomaly detection, predictive analytics and proactive fault resolution; and closed-loop, intent-based automation.
What Do We Expect From You RIO's promise is AI-native network operations. As a Senior AI Engineer on the AI Model & Harness team, you will span two equally important halves: developing the ML models that detect anomalies, predict faults and localize root causes across telecom networks; and building the AI harness - the training, evaluation, inference and monitoring infrastructure - that takes those models from notebook to production with confidence.
You will work directly on top of RIO's ClickHouse and YugabyteDB (YBSQL) data platforms, in tight partnership with the Data Management Pipelines team. Key Responsibilities ML Modeling (approximately 50%) Anomaly detection: Develop and productionize multivariate time-series anomaly detection models over network KPIs and counters to surface degraded cells and network elements before outages.
Predictive fault & incident analysis: Build forecasting and classification models that predict faults and incidents ahead of failure, and correlate alarms across domains for root-cause analysis and noise reduction. Model portfolio: Own models across the operations lifecycle - KPI forecasting, event correlation, root cause localization, capacity prediction and optimization - from framing through production monitoring.
AI-defined KPIs: Move beyond vendor KPI formulas by learning leading/lagging indicators directly from raw counters, surfacing signals humans and rule-based systems miss. AI Harness & Inference Infrastructure (approximately 50%) Training & experimentation: Build reproducible training, experiment-tracking and hyperparameter-tuning workflows over large-scale telemetry datasets drawn from ClickHouse/YBSQL.
Evaluation frameworks: Design rigorous offline a