Site Reliability Engineer-1
Job type: Full Time · Department: Engineering · Work type: On-Site
Bengaluru, Karnataka, India
Digantara is a leading Space Surveillance and Intelligence company focused on ensuring orbital safety and sustainability. With expertise in space-based detection, tracking, identification, and monitoring, Digantara provides comprehensive domain awareness across all regimes, enabling end-users to gain actionable intelligence on a single platform. At the core of its infrastructure lies a sophisticated integration of hardware and software capabilities aligned with the key principles of situational awareness: perception(data collection), comprehension (data processing), and prediction(analytics). This holistic approach empowers Digantara to monitor all Resident Space Objects(RSOs) in orbit, fostering comprehensive domain awareness.
We are looking for an SRE-I with 1–2 years of experience to support infrastructure, automation, deployments, monitoring, and production reliability. The role involves working closely with engineering teams to automate processes, manage cloud infrastructure, and ensure reliable system operations.
Competitive incentives, galvanising workspace, blazing team, frequent outings—pretty much everything that you have heard about a startup + you get to work on space technology.
Hustle in a well-funded startup that allows you to take charge of your responsibilities and create your own moonshot
1–2 years of experience in SRE, DevOps, Infrastructure, or Cloud Engineering.
Strong hands-on experience with Docker and containerization.
Good knowledge of Linux, Python and Bash scripting.
Experience with Ansible and Terraform for automation and Infrastructure as Code.
Working knowledge of AWS.
Experience with Git and GitHub Actions for CI/CD.
Good working knowledge of Windows systems and administration for automation and tool integration.
Hands-on experience with Prometheus and Grafana for monitoring and observability.
Manage and troubleshoot Docker-based deployments.
Automate infrastructure and operational tasks using Ansible, Bash, and Terraform.
Manage and support AWS infrastructure.
Build and maintain CI/CD pipelines using GitHub Actions.
Implement monitoring, dashboards, metrics, and alerts using Prometheus and Grafana.
Troubleshoot production, infrastructure, and deployment issues.
Support automation across Linux and Windows environments.
Work with development teams to improve system reliability and deployment processes.
Visit client locations for deployments and integrations when required.
Bachelor's degree in Computer Science, IT, Engineering, or a related field.
1–2 years of relevant SRE/DevOps experience.
Strong knowledge of Docker, Ansible, Terraform, AWS, Git, CI/CD, Linux, and Windows.
Experience with Prometheus, Grafana, monitoring, and observability.
Good troubleshooting, analytical, and communication skills.
Python scripting experience.
Experience with Kubernetes or Nomad.
Knowledge of networking concepts and cloud-native technologies.
Autofill from resume
Save time by uploading your resume. (Only PDF or DOCX format supported)