Description
Job Summary:
We are seeking a professional to lead the operation, evolution, and reliability of the technology infrastructure, ensuring high availability, security, scalability, and observability.
Key Highlights:
1. Leadership of DevOps, SRE, or cloud infrastructure teams
2. Transforming businesses and careers with a focus on innovation and reliability
3. Evolution of container- and Kubernetes-based infrastructure architecture
**Open positions for those who want to **transform** businesses and their careers!**
---------------------------------------------------------------------
### **Desirable**
• Experience with IoT or telemetry systems
• Experience with real-time data platforms
• Experience with high-data-ingestion architectures
• Experience with production Kubernetes environments
• Knowledge of device communication protocols### **Job Requirements**
• Prior experience leading DevOps, SRE, or cloud infrastructure teams
• Experience with cloud-native and microservices environments
• Experience in critical production environments
Desired Knowledge
Infrastructure and Cloud
OCI, AWS, or Azure
Kubernetes / containers
cloud networking and architecture
DevOps
CI/CD
infrastructure automation
build and deployment pipelines
Observability
Grafana
Prometheus
ELK or equivalent tools
Databases
PostgreSQL or production relational databases
Messaging
MQTT or distributed messaging systems
Security
IAM
access control
cloud security best practices
Behavioral Profile
We seek a professional who has:
Strong technical leadership ability
Hands-on mindset
Reliability- and quality-oriented approach
Strong communication skills with engineering teams
Ability to structure processes and standardizations
Systemic view of software and infrastructure architecture### **Your Responsibilities**
• Lead the operation, evolution, and reliability of the technology infrastructure
• Evaluate the current platform architecture
• Structure operational and incident management processes
• Enhance monitoring and observability
• Organize infrastructure automation
• Increase production environment reliability
• Ensure that the platforms supporting the company’s products — including telemetry, IoT, data processing, and cloud-native applications — operate with high availability, security, scalability, and observability.
• Lead the DevOps/Infrastructure team, defining processes, architectures, and tools that support the full application lifecycle, from development through production operations.
• Develop and track department performance metrics
• Structure operational, incident, and continuous improvement processes
• Collaborate with development and product teams
• Manage public cloud environments (OCI, AWS, or Azure)
• Ensure infrastructure availability and scalability
• Lead the evolution of container- and Kubernetes-based infrastructure architecture
• Implement and enhance monitoring and observability systems
• Metrics, logs, tracing
• Ensure service levels and availability (SLA / SLO)
• Conduct incident analysis and post-mortems
• Automation and DevOps
• Evolve CI/CD pipelines
• Promote DevOps practices across development teams
• Automate infrastructure provisioning (Infrastructure as Code)
• Ensure infrastructure and application security practices
• Implement access control policies
• Work closely with information security teams
• Manage database infrastructure (e.g., PostgreSQL)
• Ensure backup, replication, availability, and performance
• Support high-ingestion data environments originating from telemetry and IoT### **It’s important to remember that ...**
At UDS, we hire competent people who want to **drive transformation using their knowledge.**
This is independent of your region, age, ethnicity or race, religion, gender identity, or sexual orientation.
Do your skills match the role? **That’s all that matters.**
Do your profile and values align with ours? **Join us in creating transformations.**