Robotic Systems Tester at Yandex (Yango) Tech Robotics (2024-06 – 2026-04)
- Tested AI-powered robotic systems in real-world user scenarios, ensuring stable, reliable, and predictable system operation.
- Identified and analyzed issues in both software and hardware components, validated fixes, and verified correct system behavior after changes were introduced.
- Used Python scripts, logs, and Grafana dashboards to monitor system behavior, diagnose failures, and check the status of key components.
- Supported machine learning-related processes, including data preparation, data quality checks, and validation of input data for further use in models.
- Set up Ubuntu 22.04-based workstations for robotic platforms, including Docker, NVIDIA drivers/toolkit, and Intel RealSense SDK.
- Deployed and tested software environments for robotic platforms, including Futuruka Server, video streaming, and remote robot control.
- Performed end-to-end system readiness checks, including cameras, configurations, containers, robot commands, and teleoperation workflows.
IT Monitoring Specialist at Paysend (2026-04 – Present)
- 24/7 monitoring of IT infrastructure and payment services using Zabbix, Grafana, OpenSearch, and DotCom.
- Onboarding infrastructure hosts into Zabbix monitoring, creating triggers, alerts, and monitoring rules.
- Analyzing alerts, logs, and system events in OpenSearch to identify incidents, errors, and abnormal service behavior.
- Using SQL queries for data analysis, service validation, and incident investigation.
- Working with GLPI, Jira, and Confluence within ITSM, incident, service request, and change management processes.
- Incident analysis, prioritization, and escalation to ensure system stability and service availability.
- Collaborating with DevOps, Network, DBA, and Application Support teams during critical incidents and service disruptions.
- Coordinating incident closure with subtask owners by confirming resolution status, investigation outcome, and documented root cause or closure rationale.
- Preparing incident reports and postmortems, contributing to incident response processes and service reliability improvements.