DevOps Engineer (Middle+/Senior)
- Location
- Remote
- Employment
- Full-time
- Level
- Mid-level
- Category
- DevOps & SRE
- Posted
Description
Remote | Full-time | $3,000–4,000/month
We are looking for an experienced DevOps Engineer (Middle+/Senior) to take ownership of the infrastructure behind multiple production environments.
This is a hands-on position with responsibility for production stability, deployments, servers, backups, infrastructure security, troubleshooting, and incident response.
Our infrastructure currently includes approximately 20 separate production broker environments plus core/database infrastructure.
We are looking for someone who can first learn the existing infrastructure and deployment processes and then gradually take full ownership of them.
Environment
Our current environment includes technologies and infrastructure such as:
Linux / AlmaLinux
Nginx
Docker
GitHub
CI/CD
Cloudflare
Multiple production servers
Central/core database infrastructure
CRM applications
Web-based trading platforms
REST APIs
WebSockets / real-time services
Third-party data/API providers
The exact existing infrastructure will be reviewed during onboarding.
Responsibilities
You will be responsible for the reliability and operation of production infrastructure.
This includes:
Maintain approximately 20 production environments
Maintain core/database infrastructure
Provision new production servers from scratch
Configure environments for new brands/sub-brands
Deploy applications and infrastructure changes
Maintain and improve CI/CD processes
Manage application and infrastructure configuration
Configure and maintain Nginx/reverse proxy infrastructure
Work with Docker/containerized services
Manage GitHub-related deployment infrastructure, access, keys, and secrets
Maintain SSL certificates, domains, DNS, and Cloudflare-related infrastructure when required
Perform server and application migrations
Move environments between servers when required
Maintain database infrastructure
Manage database backups and restoration procedures
Implement and verify automated backups
Ensure that backups can actually be restored when needed
Maintain disaster-recovery procedures
Monitor server health, CPU, RAM, disk usage, services, and application availability
Maintain and improve existing monitoring and alerts
Troubleshoot production incidents
Investigate infrastructure, network, API, service, and real-time data issues
Investigate cases where external/internal APIs or data sources stop functioning correctly
Work with external hosting and infrastructure providers when required
Maintain infrastructure security
Manage SSH access, permissions, firewall rules, secrets, and production access
Perform safe production restarts and updates
Maintain rollback procedures
Minimize downtime during deployments and infrastructure changes
Document servers, services, environments, domains, dependencies, deployment processes, backup procedures, and recovery procedures
Work closely with Full-Stack Developers and QA
Production ownership
This position requires real ownership of production infrastructure.
We are not looking for someone who only executes predefined commands.
During the onboarding period, you will receive information about the existing environments and infrastructure.
As you become familiar with the system, we expect you to independently:
Diagnose infrastructure problems
Determine root causes
Decide on appropriate technical actions
Restore affected services
Coordinate with developers when the issue is application-related
Verify production health after changes
Document important incidents and infrastructure changes
Identify potential infrastructure risks before they become production incidents
Requirements
4+ years of commercial DevOps/System Engineering experience or equivalent strong hands-on experience
Strong Linux administration skills
Production experience with Linux-based environments
Strong Nginx knowledge
Docker experience
CI/CD experience
Git/GitHub experience
Experience managing multiple production environments
Experience provisioning production servers
Experience with database infrastructure, backups, and recovery
Understanding of networking, DNS, SSL/TLS, HTTP/HTTPS, reverse proxies, and firewalls
Experience troubleshooting production systems
Experience with monitoring and alerting
Understanding of infrastructure and application security
Experience managing SSH access, secrets, permissions, and deployment credentials
Ability to analyze logs and diagnose application/infrastructure failures
Experience working with third-party infrastructure/API providers
Ability to create clear infrastructure documentation
Ability to work independently after onboarding
English — Intermediate or higher
Experience with AlmaLinux/RHEL-compatible environments is a strong plus.
Experience with trading or FinTech infrastructure is not required.
Backup and recovery
We expect the DevOps Engineer to maintain a reliable backup and recovery process.
This means not only creating backups but also knowing:
What is backed up
How frequently backups are created
Where they are stored
How they can be restored
How a failed production environment can be rebuilt
How production can be recovered if a server becomes unavailable
Incident response
This is the only position in the team that includes responsibility for urgent production incidents outside standard working hours.
Production incidents outside working hours are not expected to be routine, but when a critical production service becomes unavailable, the DevOps Engineer must be reachable and able to respond.
We expect:
Reliable availability during normal working hours
Reasonable emergency availability outside working hours
Fast acknowledgement of critical production incidents
Ability to access and investigate production remotely when necessary
Clear communication during an incident
Notification to the team if you will be unavailable for a known period
The exact incident-response/on-call process will be agreed with the successful candidate.
First 1–3 months
We understand that the infrastructure contains multiple environments and cannot be fully learned in several days.
The initial period will focus on:
Understanding existing infrastructure
Mapping production environments
Understanding deployments
Understanding database and service dependencies
Reviewing access and security
Reviewing monitoring
Reviewing backup processes
Creating/updating infrastructure documentation
Identifying infrastructure risks and areas for improvement
After the onboarding period, we expect you to be able to manage standard infrastructure operations and troubleshoot most production issues independently.
Working conditions
Fully remote
Full-time
Monday–Friday
Standard working hours: 09:00–18:00 GMT+3 / 10:00–19:00 GMT+4
Emergency production availability outside working hours
Probation/onboarding period: 1–3 months
Compensation: $3,000–4,000/month, depending on experience and technical level
Higher compensation may be considered for an exceptionally strong candidate with significant production ownership experience
Long-term cooperation
English: Intermediate or higher
What is important to us
Reliability and communication are critical for this role.
We expect the DevOps Engineer to:
Be reachable during agreed working hours
Respond to critical infrastructure incidents
Notify the team about planned periods of unavailability
Take ownership of production problems
Investigate root causes rather than only restart services
Communicate clearly with developers, QA, and the Project Manager
Document important infrastructure knowledge
Proactively identify risks
Gradually reduce dependency on undocumented knowledge held by individual team members
Our goal is to build an infrastructure process where production knowledge, access, deployments, backups, and recovery procedures are documented and are not dependent on one person's memory.
Apply at the source
This role was published by Filiatix and listed via Djinni. Applications are handled there, not on this site.
Original posting: https://djinni.co/jobs/844317-devops-engineer-middle-senior/