We are seeking a highly experienced Linux engineer with over 10 years of hands-on expertise to take full ownership of our infrastructure. This role requires a deep understanding of Linux systems, Proxmox virtualization, and core network services, with a strong emphasis on automation, consistency, and reliability. You will be the primary expert responsible for managing and optimizing our environments, ensuring smooth operation and continuous improvement. This is a remote position ideal for a professional who thrives on ownership and proactive problem-solving.
Key Responsibilities
- Manage Proxmox environments comprehensively, including building and operating clusters and managing virtual machines, with all changes automated through Ansible.
- Administer and maintain Postfix mail servers to ensure reliable email delivery.
- Build and maintain WireGuard VPNs across multiple offices, servers, and environments to secure network communications.
- Operate and maintain Unbound DNS resolvers for efficient and secure DNS resolution.
- Design and maintain high-availability configurations using CARP to ensure service continuity.
- Use Ansible as the primary tool for provisioning, configuring, and maintaining servers, driving automation wherever possible.
- Monitor, tune, and troubleshoot Linux systems, converting recurring issues into automated solutions.
- Take full ownership of services from initial design through to production support, ensuring reliability and performance.
- Participate actively in production incident response, driving issues to timely resolution.
- Proactively identify and flag risks, misconfigurations, or technical debt before they escalate into incidents.
Required Qualifications
- Expert-level Linux administration skills with a proven track record in production environments.
- Extensive hands-on experience with Proxmox virtualization technology.
- Strong practical knowledge of Postfix, WireGuard, and Unbound.
- Experience designing and maintaining high-availability setups using CARP.
- Proficiency in Ansible for automation, with a mindset geared towards automating repetitive tasks.
- Excellent troubleshooting abilities and meticulous attention to detail.
Preferred Qualifications and Additional Skills
- Experience working with AWS cloud services, Terraform infrastructure as code, and GitLab CI for continuous integration and deployment.
- Familiarity with containerization technologies such as Docker.
- Understanding of Agile methodologies and collaborative development practices.
- Experience with observability and monitoring tools like Datadog, Grafana, and Prometheus to enhance incident detection and debugging.
Behavioral Attributes We Value
We highly value an ownership mindset, where you treat the systems you manage as your personal responsibility. Accountability is crucial—you should be comfortable owning your changes and leading fixes and retrospectives when issues arise. Proactive communication is essential; you will escalate problems early, keep stakeholders informed during outages, and ensure no issues remain unaddressed. Finally, a commitment to continuous improvement is important, using incidents and mistakes as opportunities to enhance monitoring, automation, and processes.
This role offers the opportunity to work remotely while playing a critical role in maintaining and evolving a robust Linux-based infrastructure. If you are passionate about automation, reliability, and taking full ownership of your work, this position is an excellent fit.