Description
The role is responsible for monitoring and analyzing platform performance, diagnosing and resolving production issues, conducting root cause analysis, fixing bugs, and implementing reliability and scalability improvements. It requires a bachelor’s degree in computer engineering or a related field, Python proficiency, experience with system monitoring, logging, observability tools, and enterprise-grade production systems, along with strong troubleshooting, ownership, and self-management skills.
