Skip to content
← Back to job listings

Application Reliability Engineer

gravitonresearchcapital · Gurugram, Haryana, India

IT - Network / Systems / DB AdminSenior LevelExternal listingfull-time7 days ago

About The Role

  • Who we are
  • Graviton Research Capital is a privately funded quantitative trading firm striving for excellence
  • in financial markets research. We trade across a multitude of asset classes and trading venues
  • using a diverse range of concepts, from time series analysis and stochastic models to machine
  • learning and statistical inference. We analyse terabytes of data to identify pricing anomalies
  • and drive innovation in financial markets.

Key Responsibilities and Deliverables

  • The ideal candidate will possess a strong background in technical support, with a passion for
  • problem-solving and a commitment to excellence. As an Application Reliability Engineer, you

will be responsible for

  • Monitor production services and respond quickly to alerts, incidents, and outages to
  • ensure smooth operation and minimal downtime.
  • Monitor trading systems and infrastructure.,
  • Triage issues across trading support services, databases, and infra; escalate and
  • coordinate with the right owners, and drive root-cause analysis and ensure fixes are
  • implemented for long-term stability.
  • Serve as the first line of defense for trading operations.
  • Proactively identify, address recurring issues, and build automation to reduce manual
  • intervention.
  • Improve observability by enhancing monitoring, logging, and alerting systems.
  • Develop and maintain operational runbooks and SLO/SLA metrics.

Eligibility and Required Skills

  • Possess a degree in a highly analytical field, such as Engineering, or Computer Science
  • 2-5 years of experience in Python, Shell/Bash scripting.
  • Experience with Linux and shell/bash online tools.
  • Hands-on experience with databases (SQL, NoSQL)
  • Strong problem-solving and analytical skills
  • Excellent communication skills
  • Ability to remain calm and analytical under production pressure

Good to have

  • Familiarity with monitoring/alerting stacks (Prometheus, Grafana, ELK, etc.)
  • Familiarity with distributed messaging (Kafka) and caching systems (Redis)
  • Experience with CI/CD pipelines and deployment automation.
  • Prior experience in a support, SRE, or production engineering role.

This is an external listing. JobSpring does not represent or verify the employer. Report this listing