Retour à la recherche
S
SquarepointSource d’offres vérifiée

Software Developer - Data Reliability

Offre en anglais

The role involves maintaining and improving data infrastructure, automating operational workflows, and ensuring the stability of data pipelines. You will also lead incident resolution, perform root cause analysis, and manage system observability.

  • Sur place
  • Montréal, QC
  • Publié 28 juill. 2026
  • 1 poste

Résumé du poste

POSITION OVERVIEW As a Reliability Software Engineer on the Data team, you will be central to the infrastructure and operations that keep Squarepoint's data ecosystem running at scale. Our mandate is broad by design: we own the pipelines, platforms, and automation that our technology and investment teams depend on. The team operates at the intersection of engineering and operations. We build the observability tooling that surfaces problems before they become incidents, the job orchestration platform that runs production workloads at scale, the self-serve systems that let teams move faster without creating risk, and the production pipelines that onboard new data from various vendors. When incidents arise, we lead the response: root cause, remediation, and the follow-through that prevents recurrence. You will work closely with quant researchers, traders, and engineers across our organization, as well as the external vendors who supply our data, building strong working relationships while developing deep expertise in the systems that underpin their work. We value ownership, sharp analytical thinking, and engineers who take pride in leaving things better than they found them. If you thrive in an environment where your work has direct, visible impact on the reliability of critical systems and where no two days look the same, then this role is for you. Responsibilities * Observability: Keep Data Development infrastructure healthy, building the health monitoring platform with configurable monitors and custom checks that surfaces problems before they become incidents * System Automation: Build and improve automation that streamlines operational tasks and workflows, including the job orchestration platform that runs production workloads at scale * Self-Serve Request Automation: Build the Jira-driven platform that executes operational requests end to end, letting teams move faster without creating risk * Data Pipelines: Onboard new datasets from various vendors, and build the pipelines and frameworks that make that work scalable * Reliability Engineering: Own deployments, support users, plan for capacity and performance, and lead incident resolution end to end REQUIRED QUALIFICATIONS * Education: Bachelor's degree in Computer Science, Engineering, or a related subject * Experience: 4+ years of proven experience in Software Engineering, Software Reliability, or a similar role * Python: 3+ years of hands-on experience programming in Python, and familiarity with version control systems such as Git * Linux: Proficiency with the Linux command line * Databases: Experience with SQL and relational databases, primarily PostgreSQL * Data Transfer: Practical knowledge of common protocols and tools used to move data, such as SFTP, HTTP APIs, and cloud object storage * Communication: Excellent verbal and written skills for working effectively with a global team and external vendors * Mindset: Proactive, detail-oriented, and self-driven with a strong sense of ownership and accountability REQUIRED QUALIFICATIONS * Education: Bachelor's degree in Computer Science, Engineering, or a related subject * Experience: 4+ years of proven experience in Software Engineering, Software Reliability, or a similar role * Python: 3+ years of hands-on experience programming in Python, and familiarity with version control systems such as Git * Linux: Proficiency with the Linux command line * Databases: Experience with SQL and relational databases, primarily PostgreSQL * Data Transfer: Practical knowledge of common protocols and tools used to move data, such as SFTP, HTTP APIs, and cloud object storage * Communication: Excellent verbal and written skills for working effectively with a global team and external vendors * Mindset: Proactive, detail-oriented, and self-driven with a strong sense of ownership and accountability NICE TO HAVE * Experience with observability and monitoring tools such as Grafana, Kibana, or Prometheus * Experience developing automation tooling and implementing configuration management * Experience with cloud platforms such as Google Cloud or AWS * Experience operating job orchestration or workload scheduling systems at scale

Ce que vous ferez

The role involves maintaining and improving data infrastructure, automating operational workflows, and ensuring the stability of data pipelines. You will also lead incident resolution, perform root cause analysis, and manage system observability.

Exigences

Candidates must have a bachelor's degree in Computer Science or a related field and at least 4 years of experience in software engineering or reliability. Proficiency in Python, Linux, and relational databases like PostgreSQL is required.

Compétences indiquées

  • PostgreSQLSouhaitée
  • Amazon Web ServicesSouhaitée
  • LinuxSouhaitée
  • Google CloudSouhaitée
  • GitSouhaitée
  • PythonSouhaitée

Autres compétences pertinentes

Relevées dans la description du poste. Confirmez les exigences importantes ci-dessus.

  • Python
  • Linux
  • PostgreSQL
  • Git
  • Consul
  • Grafana
  • Kibana
  • Google Cloud
  • AWS
  • Automation
  • Observability
  • Incident Management
  • Data Pipelines
  • Infrastructure Management
  • Reliability Engineering

Domaines d’emploi

  • Software
  • Data & Analytics
  • Technology
  • Engineering

Renseignements supplémentaires

Formation minimale
Baccalauréat
Expérience minimale
2+ ans
Langue de l’offre
anglais
Heures de travail
40 heures par semaine