JOB DESCRIPTION
RoleSenior Cloud Infrastructure / DevOps Engineer (AWS)FunctionEngineering & Cloud Infrastructure
Reporting toCTOLocationRemote (India)
About XED
XED is a premier global executive education organization established in 2015, committed to empowering senior leaders across the Middle East, Far East, LATAM, and South Asia. We collaborate exclusively with Ivy League universities to design and deliver high-impact online and hybrid programs in strategy, leadership, entrepreneurship, innovation, digital, and finance. With a focus on immersive learning and technology, XED has empowered over 15,000 senior leaders from Fortune 500 companies and global brands.XED is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees, regardless of gender, background, age, or belief.Website: www.xedinstitute.org
Role Purpose
XED is launching Athena, our compute-intensive, AI-powered platform - built to handle thousands of concurrent users. This hands-on individual-contributor role owns the AWS infrastructure end to end — designing, building, and operating it — with a focus on rapid scalability, high availability, fault tolerance, security, and cost efficiency.Athena delivers its AI capabilities through managed third-party APIs (LLM and embeddings via hosted APIs; real-time voice and avatar via third-party providers) rather than self-hosted models. The core engineering challenge is therefore orchestrating expensive, rate-limited, latency-sensitive third-party services at high concurrency.The ideal candidate is a senior engineer (5+ years) who can both architect and operate production AWS infrastructure, available to join within two weeks, in a fully remote environment.
Detailed Responsibilities
  • Design, build, and operate end-to-end AWS infrastructure (compute, networking, storage, databases, security) for Athena and XED's wider product suite.
  • Architect for rapid scalability, high availability, and fault tolerance: multi-AZ deployments, auto-scaling, graceful degradation, and disaster recovery.
  • Orchestrate high-concurrency integration with rate-limited, real-time third-party APIs : queuing, backpressure, connection pooling, retries, circuit breakers, and fallbacks.
  • Build resilient real-time media streaming infrastructure (WebSocket / WebRTC, low-latency edges, media relays) to serve voice and avatar sessions at scale.
  • Implement and own infrastructure-as-code (Terraform, CloudFormation, or CDK) for reproducible, reviewable, auditable provisioning.
  • Extend and maintain CI/CD pipelines (Bitbucket Pipelines) enabling safe, frequent, low-risk deployments.
  • Establish end-to-end observability (metrics, logging, tracing, alerting); lead incident response and post-incident hardening.
  • Own security and compliance posture: IAM, network segmentation, secrets management, and encryption, supporting XED's GDPR data-processor obligations.
  • Drive cost efficiency across both AWS spend and third-party API usage: right-sizing, savings plans, and per-session / per-call cost instrumentation and controls.
  • Manage core AWS services (EC2, S3, RDS, Lambda, VPC, IAM, ECS/EKS, CloudFront, Route 53) alongside Athena's stack (Node.js/Next.js, MongoDB, Qdrant).
  • Conduct architecture reviews and security audits; track AWS and relevant tooling updates and recommend improvements to leadership.
Educational Qualification & Certifications
Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related field.AWS certifications are preferred, not mandatory. Demonstrated experience scaling and operating production systems outweighs certification. The most relevant credentials are the AWS Certified DevOps Engineer – Professional (DOP-C02) and AWS Certified Solutions Architect – Professional (SAP-C02); the Solutions Architect – Associate (SAA-C03) and Security – Specialty (SCS-C02) are also valued.
Experience
5+ years of hands-on experience both designing and operating AWS cloud infrastructure in production. The ideal candidate will have personally scaled a production system to thousands of concurrent users and operated latency-sensitive, API-dependent services in production. Strong infrastructure-as-code, containerisation, CI/CD, and observability practice is expected. Prior experience in a fully remote engineering team is preferred.Early joiners strongly preferred — candidates available to start within two weeks will be prioritised.
Professional Growth & Career Advancement
This role offers a compelling opportunity to architect and own the cloud backbone of a globally operating executive education platform at its defining launch-and-scale moment. High performers who demonstrate measurable impact on infrastructure reliability, cost efficiency, and engineering velocity may be considered for senior or lead roles within the Engineering function at XED.

Required Skills

Communication Skill Aws AI