Athena Executive Search & Consulting
See all jobs at Athena Executive Search & ConsultingDevOps Engineer
Posted 9 days ago
- Pay
- Not shared
- Location
- On-site · Gurugram
- Experience
- 3–5 yrs · Mid-level
- Type
- Full-time
Location: Gurugram, Haryana, India | Experience: Mid Level (3-5 years)
About the Company
Athena is a leading AI platform company helping businesses use intelligent technology to improve how they operate, serve customers, and make decisions. Built for teams that need dependable systems at scale, Athena combines advanced AI capabilities with practical product engineering to support measurable business outcomes. The company brings together engineering, product, and operations teams to build reliable platforms that can evolve with changing customer and market needs. This role will contribute to the infrastructure foundation behind that work, ensuring services remain secure, observable, resilient, and ready for growth. Based in Gurugram, the team values ownership, disciplined execution, and close collaboration across technical functions. The DevOps Engineer will join an environment where infrastructure decisions directly influence release velocity, platform reliability, developer productivity, and customer trust. Through automation and strong operational practices, Athena aims to help its teams ship confidently while maintaining the quality and availability expected from an AI-led technology business.
About the Role
As a DevOps Engineer, you will own the systems and practices that move software from development to dependable production operation. You will improve deployment automation, cloud infrastructure, monitoring, security, and incident readiness so engineering teams can release faster without compromising reliability. The role connects application development with platform operations, translating technical requirements into repeatable infrastructure and efficient workflows. Your work will reduce manual effort, shorten recovery times, improve system visibility, and create a stronger foundation for Athena’s AI platform. Success means developers can ship confidently, production services remain stable, and infrastructure scales with business demand.
Key Responsibilities
- Own CI/CD pipelines and deployment workflows through automation, reducing manual release effort and improving delivery speed, consistency, and rollback readiness across engineering teams.
- Manage cloud infrastructure and environments using infrastructure-as-code, improving scalability, repeatability, cost visibility, and operational control as platform usage grows.
- Build monitoring, logging, and alerting through appropriate observability tools, enabling faster detection, diagnosis, and resolution of service-impacting issues.
- Partner with developers to improve application operability, release design, configuration management, and production readiness before changes reach live environments.
- Strengthen platform security through access controls, secrets management, patching, vulnerability remediation, and practical compliance-oriented infrastructure standards.
- Participate in incident response and root-cause analysis, converting operational lessons into durable fixes that improve availability and reduce recurring failures.
- Maintain clear runbooks, architecture documentation, and operational procedures so teams can respond consistently and transfer knowledge effectively.
Essential Skills & Technologies
- Strong hands-on experience with Linux, Git, scripting, CI/CD tools, and cloud infrastructure, with the ability to automate repeatable engineering and operational workflows.
- Practical proficiency in infrastructure-as-code, containers, orchestration, monitoring, logging, networking, and secure configuration across production environments.
- Demonstrated ability to troubleshoot complex systems, collaborate with software engineers, and communicate operational risks, priorities, and trade-offs clearly.
Additional Plus
- Experience with AWS, Azure, or Google Cloud services and production workloads supporting scalable web applications or data-intensive platforms.
- Familiarity with Kubernetes, Terraform, Ansible, Docker, Prometheus, Grafana, ELK, or comparable modern platform engineering tools.
- Exposure to reliability engineering practices, SLOs, incident management, automated testing, and cloud cost optimization would strengthen your impact.
- Experience supporting AI, SaaS, or rapidly scaling technology products is an added advantage.
What You'll Bring
- You bring three to five years of hands-on experience building, automating, and operating reliable software infrastructure. You are comfortable working across development and operations, and you understand that good DevOps outcomes are measured through faster delivery, fewer failures, clearer visibility, and stronger recovery—not simply by the number of tools implemented.
- You can work confidently with Linux systems, version control, scripting, CI/CD, cloud services, containers, infrastructure-as-code, and observability platforms. You approach production issues methodically, communicate clearly during incidents, and follow problems through from immediate mitigation to durable resolution.
- You are also willing to challenge manual processes, simplify complex workflows, and introduce automation where it improves quality or developer productivity. A strong sense of ownership matters: you document what you build, make operational risks visible, and collaborate closely with engineers, product stakeholders, and other technical partners.
- You should be comfortable in a fast-moving AI platform environment where priorities evolve, decisions need practical judgment, and infrastructure must support both current reliability and future scale.
Why Join Us
- Build the infrastructure foundation for a leading AI platform, with direct influence on reliability, engineering velocity, scalability, and customer experience.
- Work closely with experienced product and engineering teams on meaningful platform challenges where automation and operational improvements create visible business impact.
- Grow your scope through ownership of cloud infrastructure, delivery systems, observability, security, and reliability practices in a scaling technology environment.
What We Offer
- A high-impact DevOps role in Gurugram with ownership across cloud infrastructure, automation, delivery, observability, and production reliability.
- The opportunity to work on an AI platform where infrastructure improvements directly strengthen product performance, engineering effectiveness, and customer trust.
- A collaborative, outcome-focused environment that values practical learning, cross-functional partnership, and disciplined technical execution.
Skills
- Linux
- CI/CD
- Cloud Infrastructure
- Scripting
- Git
- Infrastructure as Code
- Containerization
- Container Orchestration
- Monitoring
- Logging
- Networking
- Security Configuration
- Systematic Troubleshooting
- Incident Response