Search by job, company or skills

DevOps Engineer L1

  • Posted a day ago
  • Over 50 applicants have applied

Job Description

About Rupeek:

Rupeek, established in 2015 and headquartered in Bangalore, stands as India's leading asset-backed digital lending fintech platform. Committed to making credit accessible to Indians in a fair and convenient manner, Rupeek pioneers innovative financial products focused on monetising India's $2 trillion gold market. Leveraging state-of-the-art technology and an automated asset-light supply chain, Rupeek is transforming the gold loan disbursal landscape across 40+ cities in India. With a customer base exceeding 5,00,000+, the company's strategic partnerships with top banks and financial institutions underscore its commitment to building gold-backed assets through low-risk, low-touch, and friction-free processes.

Rupeek's impressive journey is supported by key investors such as Sequoia Capital, Accel Partners, Bertelsmann, and GGV Capital. In Apr 2024, Rupeek turned profitable and raised $25 in equity capital from Manipal Group and Elevation Capital to fund its next phase of growth. Join us in redefining the future of finance through innovation, technology, and a commitment to financial inclusivity.

Job Title: DevOps Engineer L1

Education: B.Tech or Dual Degree (Computer Science/IT preferred)

Experience: 2-4 Years of relevant experience

Location: Bangalore

About the Role:

The DevOps Engineer – L1 is an entry - level, hands-on engineer responsible for independently owning

infrastructure components, maintaining CI/CD pipelines, managing observability systems, and supporting operational reliability across production environments.

This role is execution and ownership focused. The L1 is expected to operate with minimal guidance on day-to-day tasks, lead small-to-medium infrastructure improvements, and serve as the first point of escalation from junior engineers and interns.

The DevOps L1 grows into driving platform initiatives, leading incident response, and contributing to architecture decisions over time.

Key Responsibilities:

  • Own and operate EKS cluster health, node group scaling, and namespace-level resource management across Dev, Staging, and Prod.
  • Manage and troubleshoot AWS services including EC2, RDS, IAM, VPC, S3, ALB, and CloudWatch, cloudtrail.
  • Write, review, and apply Terraform changes for infrastructure provisioning and modification.
  • Maintain and extend Jenkins pipelines and shared libraries; enforce security and quality gates across CI/CD stages.
  • Manage image promotion across QA → Staging → Prod using Git and GitHub as the source of truth.
  • Build and maintain Grafana dashboards using Prometheus metrics, Loki logs, and Mimir for long-term storage.
  • Own L1/L2 on-call rotation via PagerDuty; lead incident triage, drive resolution, and write postmortems.
  • Tune alerting rules and recording rules to reduce noise and maintain reliable signal.
  • Investigate security events, triage CVEs, and drive remediation through the vulnerability management workflow.
  • Maintain runbooks and identify opportunities to automate repetitive operational tasks.
  • Mentor DevOps interns on tooling, debugging, and operational best practices.

Tools:

  • Version control: Git, GitHub
  • CI/CD: Jenkins
  • Cloud: AWS
  • Infrastructure as Code: Terraform
  • Container orchestration: Kubernetes (EKS)
  • Observability: Grafana, Prometheus, Loki, Mimir
  • Incident management: PagerDuty

Skill:

Must-Have Technical Skills

  • Hands-on experience with AWS core services (EC2, EKS, IAM, VPC, S3, ALB)
  • Kubernetes: deployments, services, namespaces, RBAC, HPA, and basic troubleshooting
  • Jenkins: building and maintaining pipelines and shared libraries
  • Terraform: writing, reviewing, and applying infrastructure changes
  • Git and GitHub: branching strategies, PR workflows, branch protection rules
  • Grafana and Prometheus: building dashboards and writing PromQL queries
  • Loki: log querying and aggregation (LogQL basics)
  • PagerDuty: on-call participation, escalation policies, and incident management
  • Linux administration and Bash scripting
  • Python scripting for automation (boto3, requests, API integrations)
  • Incident triage: log correlation, root-cause analysis, and postmortem writing
  • Networking fundamentals: DNS, TCP/IP, HTTP/S, load balancers, and firewalls

Bonus / Good-to-Have

  • Mimir: long-term metrics storage and query optimisation
  • Istio service mesh: AuthorizationPolicy, mTLS, Telemetry API
  • FinOps basics: cost anomaly detection, tagging, rightsizing

Persona:

  • Operationally Confident: Comfortable owning on-call shifts, diagnosing production issues independently, and making measured decisions under pressure.
  • Systems Thinker: Understands how components interact — can trace a failure from a PagerDuty alert through Grafana dashboards, Loki logs, and the application layer.
  • Automation-Minded: Defaults to scripting and tooling over manual repetition; contributes reusable solutions back to the team.
  • Security-Aware: Approaches infrastructure changes with a security mindset; aware of blast radius, least-privilege, and audit implications.
  • Process-Oriented: Understands the importance of documentation, change control, and structured workflows.
  • Collaborative: Comfortable working closely with backend, security, and product teams; communicates blockers early and clearly.
  • Eager to Grow: Actively upskills in areas of ownership — observability, cloud cost, or DevSecOps — without being prompted.

More Info

Job Type:
Industry:
Function:
Employment Type:

About Company

Job ID: 151999643

Beware of Scammers

We don’t charge money for job offers