▹ Operate and maintain on-premises data center infrastructure, including physical and virtual servers, Linux systems, virtualization platforms, storage, and network infrastructure
▹ Administer Linux server environments, including configuration, patching, services, storage, networking, performance, and production troubleshooting
▹ Manage virtualized infrastructure and clusters, including provisioning, resource management, migrations, storage integration, and high-availability operations
▹ Support Kubernetes/RKE2 and containerized environments across cluster setup, workloads, networking, and platform operations
▹ Implement high-availability infrastructure, including load balancing, failover, and redundant service components
▹ Automate infrastructure configuration and repetitive tasks with Ansible and shell scripting to improve consistency and reduce manual work
▹ Deploy and operate internal web applications and platform services across development, testing, and production environments
▹ Configure and maintain Apache and Nginx web infrastructure, reverse proxies, certificates, and internal PKI-related services
▹ Maintain monitoring and alerting with Checkmk, investigating availability, performance, and capacity issues
▹ Manage backup and recovery procedures, including restore validation for operational readiness
▹ Troubleshoot issues across Linux, virtualization, networking, storage, application, and container layers using logs, monitoring data, and diagnostic tools
▹ Develop internal web applications and operational tooling that provide controlled interfaces to backend infrastructure and services
▹ Apply access controls, permissions, firewall policies, and operational safeguards to improve security and stability
▹ Work with external vendors and technical consultants on data center, Kubernetes, infrastructure, and platform-related projects
▹ Test infrastructure changes, upgrades, and deployments before production rollout, and support incident investigation and root-cause analysis
▹ Create and maintain technical documentation, operational procedures, runbooks, and infrastructure guides for daily operations and knowledge transfer