JJobsSonar

Site Reliability Engineer

The Voleon Group · United States

SRE / ReliabilityRemote

About this role

As a Site Reliability Engineer (SRE) at Voleon, you will work at the intersection of production operations and software development, improving, managing, and monitoring production-critical infrastructure and data pipelines. You will be responsible for enhancing fault-tolerance and maintainability of code in proprietary data pipelines and trading systems, diagnosing and fixing bugs, leading complex deployments, automating manual workflows, tracking and prioritizing production-related issues, and sharing an on-call rotation to ensure the continuous operation of critical systems. You will collaborate with passionate and talented colleagues in an empowering, results-driven environment, contributing to more reliable systems, lower operational risk, and increased engineering efficiency.

Skills & technologies

Must have

  • Python
  • Linux
  • Relational Databases
  • SQL
  • Git
  • Jenkins
  • Bazel
  • Prometheus
  • Grafana
  • Airflow
  • Kubernetes

Nice to have

  • gRPC microservices
  • Postgres
  • Pandas
  • Golang
  • R
  • Continuous Integration and Deployment
  • Code maintainability
  • Documentation
  • Quality Assurance

Read full description

About the job Voleon is a technology company that applies state-of-the-art AI and machine learning techniques to real-world problems in finance. For nearly two decades, we have led our industry and worked at the frontier of applying AI/ML to investment management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor production-critical infrastructure and data pipelines. At Voleon, many SREs serve together on a Production Operations team tasked with improving shared production infrastructure. Others are embedded with teams of software engineers to improve specific production systems owned by those teams. Voleon SREs work on important real-world problems and collaborate with passionate and talented colleagues in an empowering, results-driven environment. This role is a way to make a real difference: your contributions will make our critical systems more reliable, lower operational risk, and increase the efficiency of our engineering effort. Responsibilities Improve fault-tolerance and maintainability of code in proprietary data pipelines and trading systems Diagnose and fix bugs in code Lead complex deployments Automate manual workflows Track and prioritize outstanding production-related issues Share an on-call rotation responding to incidents to ensure the continuous operation of production-critical systems Requirements Experience with coding and debugging Python Experience with Linux Familiarity with Relational Databases & SQL Sharp analytical and problem-solving skills and a persistent drive to make things work (better) Strong growth mindset and a passion for learning Strong technical communication skills Attention to detail 2 years of relevant industry experience An undergraduate degree or comparable training in a quantitative field or equivalent, relevant industry experience Preferred Qualifications Familiarity with best practices concerning code maintainability, documentation, quality assurance, continuous integration and deployment Experience supporting production systems Experience with any of the following: gRPC microservices, Postgres, Pandas, Golang, R, Git, Jenkins, Bazel, Prometheus, Grafana, Airflow, Kubernetes “Friends of Voleon” Candidate Referral Program If you have a great candidate in mind for this role and would like to have the potential to earn $7,500 if your referred candidate is successfully hired and employed by The Voleon Group, please use this form to submit your referral. For more details regarding eligibility, terms and conditions please make sure to review the Voleon Referral Bonus Program. Equal Opportunity Employer The Voleon Group is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law. Compensation Range: $120K - $160K
$120,000–$160,000 / yearApply now

Similar SRE / Reliability jobs

All SRE / Reliability jobs

Site Reliability Engineer Sr

Dayforce · United States

SRE / ReliabilityRemoteEasy apply$80.5K/yr - $143.8K/yr2w ago