HighRadius brand banner

Site Reliability Engineer

HighRadiusHyderabad, IndiaPosted May 13
via Greenhouse

About Us 

HighRadius, a renowned provider of cloud-based Autonomous Software for the Office of the CFO, has transformed critical financial processes for over 800+ leading companies worldwide. Trusted by prestigious organizations like 3M, Unilever, Anheuser-Busch InBev, Sanofi, Kellogg Company, Danone, Hershey's, and many others, HighRadius optimizes order-to-cash, treasury, and record-to-report processes, earning us back-to-back recognition in Gartner's Magic Quadrant and a prestigious spot in Forbes Cloud 100 List for three consecutive years. 

With a remarkable valuation of $3.1B and an impressive annual recurring revenue exceeding $100M, we experience a robust year-over-year growth of 24%. With a global presence spanning 8+ locations and a recent addition in Poland, we're in the pre-IPO stage, poised for rapid growth. We invite passionate and diverse individuals to join us on this exciting path to becoming a publicly traded company and shape our promising future. 

Job Summary:
We are looking for a highly skilled and adaptable Site Reliability Engineer (7+ Years) to become a key member of our Cloud Engineering team. In this crucial role, you will be instrumental in designing and refining our cloud infrastructure with a strong focus on reliability, security, and scalability. As an SRE, you'll apply software engineering principles to solve operational challenges, ensuring the overall operational resilience and continuous stability of our systems. This position requires a blend of managing live production environments and contributing to engineering efforts such as automation and system improvements.

Key Responsibilities:
● Cloud Infrastructure Architecture and Management: Design, build, and maintain resilient cloud infrastructure solutions to support the development and deployment of scalable and reliable applications. This includes managing and optimizing cloud platforms for high availability, performance, and cost efficiency.
● Enhancing Service Reliability: Lead reliability best practices by establishing and managing monitoring and alerting systems to proactively detect and respond to anomalies and performance issues. Utilize SLI, SLO, and SLA concepts to measure and improve reliability. Identify and resolve potential bottlenecks and areas for enhancement.
● Driving Automation and Efficiency: Contribute to the automation, provisioning, and standardization of infrastructure resources and system configurations. Identify and implement automation for repetitive tasks to significantly reduce operational overhead. Develop Standard Operating Procedures (SOPs) and automate workflows using tools like Rundeck or Jenkins.
● Incident Response and Resolution: Participate in and help resolve major incidents, conduct thorough root cause analyses, and implement permanent solutions. Effectively manage incidents within the production environment using a systematic problem-solving approach.
● Collaboration and Innovation: Work closely with diverse stakeholders and cross-functional teams, including software engineers, to integrate cloud solutions, gather requirements, and execute Proof of Concepts (POCs). Foster strong collaboration and communication. Guide designs and processes with a focus on resilience and minimizing manual effort. Promote the adoption of common tooling and components, and implement software and tools to enhance resilience and automate operations. Be open to adopting new tools and approaches as needed.

Required Skills and Experience:

● Experience: We are looking for 7+ Years of industry experience .
● Cloud Platforms: Demonstrated expertise in at least one major cloud platform (AWS, Azure, or GCP). Extensive experience with containerization (Docker) and orchestration (Kubernetes) technologies.
● Automation & IaC: Proficiency in scripting languages (shell and Python). Experience with configuration management tools (Ansible or Puppet). Must have exposure to Infrastructure as Code (IaC) tools (Terraform or CloudFormation).
● Monitoring & Observability: Experience setting up and configuring monitoring tools (Prometheus, Grafana, or the ELK stack). Hands-on experience implementing OpenTelemetry for observability. Familiarity with monitoring and logging tools for cloud-based applications.
● Service Reliability Concepts: A strong understanding of SLI, SLO, SLA, and error budgeting.
● Infrastructure Management: Proven proficiency in on-premises hosting and virtualization platforms (VMware, Hyper-V, or KVM). Solid understanding of storage internals (NAS, SAN, EFS, NFS) and protocols (FTP, SFTP, SMTP, NTP, DNS, DHCP). Experience with networking and firewall technologies. Strong hands-on experience with Linux internals and operating systems (RHEL, CentOS, Rocky Linux). Experience with Windows operating systems supporting diverse environments.
● Soft Skills & Mindset: Excellent communication and interpersonal skills for effective teamwork. We value proactive individuals who are eager to learn and adapt in a dynamic environment. Must possess a pragmatic and adaptable mindset, with a willingness to step outside comfort zones and acquire new skills. Ability to consider the broader system impact of your work. Must be a change advocate for reliability initiatives.


Desired/Bonus Skills:
● Experience with DevOps toolchain elements like Git, Jenkins, Rundeck, ArgoCD, or Crossplane.
● Experience with database management, particularly MySQL and Hadoop.
● Knowledge of cloud cost management and optimization strategies.
● Exposure to Gen AI.
● Understanding of cloud security best practices, including data encryption, access controls, and
identity management.
● Experience implementing disaster recovery and business continuity plans.
● Familiarity with ITIL (Information Technology Infrastructure Library) processes

Similar roles

  • Nvidia logo

    Senior System Software Engineer, Modeling and Validation

    Nvidia·Shanghai, China

    NVIDIA is looking for a System Software Engineer to help build the modeling, validation, and bringup infrastructure that enables next-generation SoC and GPU platforms from early pre-silicon development through post-silicon debug. This role combines low-level platform software development with system-level validation and execution workflows across C-model, emulation, and silicon environments. We are looking for engineers with strong software fundamentals, hands-on…

    • Full-time
  • SAURABH SOFTSOURCING PRIVATE LIMITED logo

    Accountant / Assistant Accountant (CA)

    SAURABH SOFTSOURCING PRIVATE LIMITED·Ahmedabad, India

    The ideal candidate will be involved with preparing financial reports and statements, bank reconciliations, and conducting cyclical audits. Moreover, the candidate must have strong interpersonal skills and possess a strong business acumen. Responsibilities Create ad-hoc reports for various business needs Prepare tax documents Compile and analyze financial statements Manage budgeting and forecasting Qualifications Bachelor's degree in Accounting or related field…

    • Full-time
  • Grapes Worldwide logo

    Copywriter

    Grapes Worldwide·Delhi, India

    About us Grapes, India’s leading Integrated communications agency nurtures digital strategy and marketing approach across paid, earned, and owned platforms. With both brand and business impact in the forefront, Grapes offers full services in Digital and Communication Solutions – Strategy Wonks at head and Creative at heart, we are strong in setting KPIs, goals and executing innovative-creative campaigns. We combine…

    • Full-time
  • FoxTalent logo

    Software Test Engineer (m/w/d)

    FoxTalent·Nuremberg, Germany

    TASKS ▪ Du unterstützt das Team im Bereich Software-Testing und Software-Qualitätssicherung ▪ Du bist verantwortlich für die Sicherstellung unseres hohen Software-Qualitätsanspruchs hinsichtlich Funktionalität, UX, Sicherheit, Performance und Zuverlässigkeit ▪ Du führst manuelle und automatisierte Software-Tests (Unit, Integration, End-to-End, System, Regression) basierend auf Anforderungen sowie Risiken durch und übernimmst die eigenständige Planung des Testumfangs ▪ Du definierst und pflegst Testfälle, Testpläne…

    • On-site
    • Full-time
    • Blue Cardtied; settle 21–33mo
  • Worldquant logo

    C++ Software Engineer - Data Platform

    Worldquant·London, Canada

    <div class="content-intro"><p>WorldQuant develops and deploys systematic financial strategies across a broad range of asset classes and global markets. We seek to produce high-quality predictive signals (alphas) through our proprietary research platform to employ financial strategies focused on market inefficiencies. Our teams work collaboratively to drive the production of alphas and financial strategies – the foundation of a balanced, global investment…

    • On-site
    • Full-time
    • Express EntryPR day one, no employer
  • Worldquant logo

    Senior Python / C++ Software Engineer

    Worldquant·London, Canada

    <div class="content-intro"><p>WorldQuant develops and deploys systematic financial strategies across a broad range of asset classes and global markets. We seek to produce high-quality predictive signals (alphas) through our proprietary research platform to employ financial strategies focused on market inefficiencies. Our teams work collaboratively to drive the production of alphas and financial strategies – the foundation of a balanced, global investment…

    • On-site
    • Full-time
    • Express EntryPR day one, no employer
  • Auxilius.ai logo

    Full-Stack Software Engineer - AI Native Startup (m/f/d)

    Auxilius.ai·Munich, Germany

    You'll join an early-stage, AI-native startup with a product that has already proven market fit. We build cutting-edge AI solutions for Governance, Risk and Compliance (GRC) for enterprises around the world. Our customers are auditors, risk managers, and compliance teams, which means evaluation rigor, auditability, and EU AI Act readiness aren't afterthoughts for us. They're product requirements. TASKS As a…

    • On-site
    • Full-time
    • Blue Cardtied; settle 21–33mo
  • Akkodis logo

    Full Stack Developer

    Akkodis·Italy

    Akkodis è un'azienda leader che opera nel settore della consulenza e dei servizi tecnologici. È specializzata in soluzioni di ingegneria, IT e digitalizzazione, offrendo supporto a diverse industrie, tra cui automotive, aerospaziale, energia, life sciences e telecomunicazioni. Akkodis si concentra sull'innovazione e sulla trasformazione digitale, aiutando le aziende a ottimizzare i loro processi e a implementare nuove tecnologie. Con la…

    • Full-time