تپسل
تپسل

Senior Infrastructure Engineer

Tehran/Jordan
Full Time
Sunday to Thursday
-
Health insurance -Flexible working hours -Learning stipends -Game room -Coffee shop

این فرصت شغلی چقدر برای من مناسب است؟

51 - 200 employees
Marketing / Advertising
Iranian company dealing only with Iranian entities
1393
Privately held
توضیحات بیشتر

key Requirements

4 years experience in similar position
Linux - Intermediate
Docker - Basic
Kubernetes - Basic
Ceph - Basic
Prometheus - Basic
Gerafana - Basic
Ansible - Basic

Job Description

Description:
Pegah is the technology group driving a wide range of digital products and businesses, including Cafe Bazaar, Tapsell, Metrix, Bazaar Pay, Metis AI, Gapify, Athena AI, Bebin TV, Beeptunes, Footballi, and more, serving over 50 million active users.
The Technology Team builds and operates the shared infrastructure, platforms, services, and core capabilities that enable engineering teams across the group to operate reliably at massive scale every single day.
We are looking for a Senior Infrastructure Engineer to join our Technology team and help us build, operate, and improve the infrastructure foundations used across the group.

Position Summary:
As a Senior Infrastructure Engineer, you will work on the compute, networking, storage, and infrastructure services that power our production systems.
You will be responsible for operating and improving these systems and will work closely with Platform Engineers, Site Reliability Engineers, and engineering teams to build reliable, secure, observable, and automated infrastructure.
This role is a good fit for someone who enjoys solving complex systems problems, troubleshooting across multiple layers, automating repetitive work, and taking ownership of production infrastructure.

What You Will Do:
  • Design, operate, and improve shared compute, networking, and storage infrastructure.
  • Manage server provisioning, remote management, virtualization, and virtual machine lifecycle.
  • Operate and improve distributed storage systems across block, file, and object storage.
  • Design, configure, and troubleshoot network connectivity between infrastructure, platforms, services, and engineering environments.
  • Secure network access mechanisms, including VPN connectivity and infrastructure-level access controls.
  • Automate infrastructure provisioning, configuration, and operational processes using Infrastructure as Code and configuration management practices.
  • Troubleshoot complex production issues across Linux, networking, virtualization, storage, and distributed infrastructure layers.
  • Build and scale infrastructure with a focus on reliability, high availability, and recoverability.
  • Monitor infrastructure health and identify capacity, performance, availability, and operational risks.
  • Participate in incident response and scheduled on-call rotations for critical infrastructure.
  • Document infrastructure architectures, operational procedures, failure scenarios, and recovery processes.
  • Collaborate with engineers across platform, reliability, delivery, and security to continuously improve the overall technology environment.

What We Expect:
  • 4+ years of experience in Infrastructure Engineering, Systems Engineering, Network Engineering, or a similar production-focused role.
  • Strong problem-solving and systematic troubleshooting skills, especially in complex production environments.
  • Solid knowledge of Linux systems and underlying concepts.
  • Strong understanding of networking fundamentals, including TCP/IP, routing, DNS, VPNs, firewalls, and related concepts.
  • Hands-on experience operating virtualized infrastructure in production environments using technologies such as VMware vSphere/ESXi or OpenStack.
  • Hands-on experience with out-of-band server management technologies such as HPE iLO.
  • Experience with distributed storage technologies such as Ceph, including block, file, and S3-compatible object storage.
  • Experience with infrastructure automation, configuration management, and Infrastructure as Code tools such as Terraform and Ansible.
  • Hands-on experience with automated server provisioning and initialization using tools such as MAAS and cloud-init.
  • Hands-on experience with infrastructure security controls such as access management, secrets management, ACLs, and network segmentation.
  • Good understanding of high availability, fault tolerance, capacity planning, backup, recovery, and infrastructure resilience.
  • Proven experience owning and operating production infrastructure with a high degree of autonomy.
  • Strong automation mindset and ability to automate operational tasks.
  • Curiosity and willingness to learn technologies outside your primary area of expertise.

Nice to Have:
  • Experience working with Kubernetes and container orchestration environments.
  • Familiarity with monitoring and observability systems such as Prometheus, Grafana, VictoriaMetrics, or similar technologies.
  • Experience operating or supporting highly available databases, messaging systems, or distributed data platforms such as PostgreSQL, MySQL, or Kafka.
  • Practical experience using AI-assisted engineering tools such as Claude, Codex, Gemini, Cursor, etc. to accelerate troubleshooting, automation, documentation, and day-to-day technical workflows.

Benefits:
  • A dynamic working environment with an open, innovative, and result-oriented culture.
  • The opportunity to work on infrastructure serving multiple large-scale digital businesses.
  • Collaboration with experienced engineers across infrastructure, platform, reliability, data, AI, and product engineering teams.
  • Flexible working hours with a hybrid working model.
  • Structured on-call rotation program with compensation built into the overall package.
  • Supplementary health insurance.
  • Competitive compensation package.
  • Various on-site entertainment and workplace facilities.

Job Requirements

Gender
Men / Women
Military service
Military service must be done
Software
Linux| Intermediate Docker| Basic Kubernetes| Basic Ceph| Basic Prometheus| Basic Gerafana| Basic Ansible| Basic

ثبت مشکل و تخلف آگهی

ارسال رزومه برای تپسل