Description
Job Summary:
We are seeking a professional with a background in computer engineering to manage and optimize on-premises infrastructures, with emphasis on automation, observability, and technical support.
Key Highlights:
1. Proficiency in Kubernetes and on-premises Linux system administration
2. Experience with infrastructure automation and API Gateways
3. Focus on observability, monitoring, and continuous improvement
Description:
* Bachelor's degree in Computer Engineering, Information Systems, Networking, or related fields;
* Proficiency in Kubernetes, including cluster installation, configuration, maintenance, and scalability. (Certified Kubernetes Administrator \- CKA);
* Expertise in Linux system administration in on\-premises environments, including installation, configuration, and maintenance of physical and virtual servers. (Red Hat Certified Engineer RHCSE);
* Experience with infrastructure automation using tools such as Ansible, Terraform, Puppet, or similar;
* Proficiency in API Gateways (e.g., Kong, Apigee, etc.), with experience configuring and managing API traffic in on\-premises environments;
* Proficiency in observability (Zabbix, Dynatrace, Datadog, Prometheus, Grafana, ELK Stack);
* Experience with networking and communication protocols (TCP/IP, HTTP, DNS, SSL/TLS), and knowledge of API security and cryptography;
* Scripting skills in languages such as Python and Shell.
* Kubernetes implementation and maintenance: Configure, maintain, and scale Kubernetes clusters in on\-premises environments, ensuring high availability and performance;
* On\-premises infrastructure management and administration: Manage and optimize physical and virtualized server infrastructure, focusing on automation and environment reliability;
* Process automation and resource provisioning: Automate repetitive tasks related to provisioning, configuration, and monitoring of servers and applications using tools such as Ansible, Terraform, Puppet, etc;
* API management with API Gateway: Implement and manage API Gateway solutions to control traffic and optimize communication between microservices and systems.
* Observability and monitoring (Logs, Metrics, and Traces): Build and maintain monitoring and alerting systems to ensure real-time visibility into infrastructure and service status, using tools such as Dynatrace, Datadog, Prometheus, Grafana, ELK Stack, or similar;
* Technical support and problem resolution: Provide real-time infrastructure support and collaborate with development teams for incident diagnosis and resolution;
* Continuous infrastructure improvement: Propose and implement infrastructure improvements focused on automation, security, performance, and reduction of operational costs;
* Documentation: Create and maintain detailed technical documentation covering procedures, processes, and configurations.
2512040202181846409