Site Reliability Engineer

apartmentThe Voleon Group placeNew York calendar_month 
Voleon is a technology company that applies state-of-the-art AI and machine learning techniques to real-world problems in finance. For nearly two decades, we have led our industry and worked at the frontier of applying AI/ML to investment management.

We have become a multibillion-dollar asset manager, and we have ambitious goals for the future.

As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor production-critical infrastructure and data pipelines. At Voleon, many SREs serve together on a Production Operations team tasked with improving shared production infrastructure.
Others are embedded with teams of software engineers to improve specific production systems owned by those teams. Voleon SREs work on important real-world problems and collaborate with passionate and talented colleagues in an empowering, results-driven environment.

This role is a way to make a real difference: your contributions will make our critical systems more reliable, lower operational risk, and increase the efficiency of our engineering effort.

Responsibilities
  • Improve fault-tolerance and maintainability of code in proprietary data pipelines and trading systems
  • Diagnose and fix bugs in code
  • Lead complex deployments
  • Automate manual workflows
  • Track and prioritize outstanding production-related issues
  • Share an on-call rotation responding to incidents to ensure the continuous operation of production-critical systems
Requirements
  • Experience with coding and debugging Python
  • Experience with Linux
  • Familiarity with Relational Databases & SQL
  • Sharp analytical and problem-solving skills and a persistent drive to make things work (better)
  • Strong growth mindset and a passion for learning
  • Strong technical communication skills
  • Attention to detail
  • 2 years of relevant industry experience
  • An undergraduate degree or comparable training in a quantitative field or equivalent, relevant industry experience
Preferred Qualifications
  • Familiarity with best practices concerning code maintainability, documentation, quality assurance, continuous integration and deployment
  • Experience supporting production systems
  • Experience with any of the following: gRPC microservices, Postgres, Pandas, Golang, R, Git, Jenkins, Bazel, Prometheus, Grafana, Airflow, Kubernetes

"Friends of Voleon" Candidate Referral Program

If you have a great candidate in mind for this role and would like to have the potential to earn $7,500 if your referred candidate is successfully hired and employed by The Voleon Group, please use this form to submit your referral. For more details regarding eligibility, terms and conditions please make sure to review the Voleon Referral Bonus Program.

Equal Opportunity Employer

The Voleon Group is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

local_fire_departmentUrgent

Senior Site Reliability Engineer II

apartmentReed Elsevier Technology ServicesplaceNew York
mitigation and Customer Data Management. You can learn more about LexisNexis Risk at the link below, ;br> About the Role: We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our...
business_centerHigh salary

Site Reliability Engineer/L3 Support

apartmentSS&C Technologies IncplaceNew York
the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology. Job Description Job Title: Site Reliability Engineer (SRE) / L3 Support Engineer REMOTE Getting to know us: As a leading financial services...
apartmentDTCCplaceJersey City, 2 mi from New York
the ITP and ECS business lines, ensuring the reliability, scalability, and performance of enterprise platforms. As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-critical systems. You will lead...