General
We are looking for Senior DevOps Engineers to operate and improve mission-critical Data Lake and Data Warehouse platforms for our banking client.
Responsibilities/Activities
-
Operate and continuously improve Data Lake and Data Warehouse platforms across production and non-production environments, with focus on stability, availability and performance
-
Investigate incidents, support major-incident recovery, perform root-cause analysis and implement permanent fixes for recurring issues
-
Manage Red Hat Enterprise Linux environments, including provisioning, patching, lifecycle tasks, capacity management, certificates, connectivity, firewall-related activities and vulnerability remediation
-
Perform operational Oracle administration on Exadata: monitor database health, investigate alerts and performance degradation, manage tablespaces, users and schemas, validate backup and recovery readiness, analyse SQL and run Data Pump schema refreshes
-
Work closely with senior DBAs on RMAN, AWR interpretation, complex performance tuning, deployments and other advanced database operations
-
Administer and support IBM DataStage, troubleshoot ETL jobs and platform components, support deployments and releases, and contribute to upgrades and maintenance activities
-
Build operational automation with Shell and, where useful, Python; reduce repetitive work and improve secure CI/CD pipelines with GitHub Actions and/or Azure DevOps
-
Improve monitoring, alerting and observability, and contribute to infrastructure upgrades plus GCP/MDPL modernization and cloud-readiness initiatives
-
Create and implement ServiceNow changes, support Incident, Major Incident, Problem and Change Management, maintain runbooks and support audit, risk, compliance and security activities
Requirements
Technical
-
At least 5 years of hands-on experience in DevOps, platform operations, infrastructure engineering or production support in mission-critical environments
-
Strong Red Hat Enterprise Linux administration skills and practical experience supporting production systems
-
Hands-on IBM DataStage administration and troubleshooting experience
-
Strong Shell scripting skills for operational support and automation
-
Experience administering Oracle Database environments on Exadata (Oracle 19c, 21c or 23c), including operational DBA activities and SQL investigations
-
Good understanding of database health monitoring, tablespace management, user/schema administration, backup and recovery readiness, and Data Pump schema refreshes
-
Experience designing, maintaining and troubleshooting CI/CD pipelines with GitHub Actions and/or Azure DevOps
-
Monitoring and observability experience (preferably with Grafana), plus strong incident troubleshooting and root-cause analysis skills
-
Experience with ServiceNow or a comparable IT service-management platform and working knowledge of Incident, Problem and Change Management
-
Strong understanding of reliability engineering, standardization and engineering best practices, with focus on availability, resilience, security and continuous improvement
Education
- Bachelor’s degree in Computer Science, Software Engineering or a related field, or equivalent practical experience.
Others
- Good oral and written communication skills (English and Romanian)
- Available for business-hours support, with occasional participation in major incidents and emergency overtime when needed
- Strong troubleshooting and ownership mindset, with a practical approach to operational excellence and disciplined execution
- Proactive in identifying risks and improvements, and effective under pressure during critical incidents and shifting priorities
Nice to have requirements
-
Experience operating large-scale Data Lake or Data Warehouse platforms
-
Google Cloud Platform (GCP), MDPL or broader cloud migration experience
-
Terraform and Infrastructure as Code experience
-
Python for operational automation
-
Hands-on vulnerability remediation experience
-
Exposure to AI-assisted engineering tools
-
Experience working in Agile/DevOps delivery models