You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Early Window: Be the first to open itNo views yetCloses in
Company hidden

Senior Software Engineer

  • Hybrid
  • 6+ years

Salary

Not stated

Similar roles pay $170K - 235K a year · our estimate

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

At the company, you'll have the opportunity to create impact at scale — tackling meaningful challenges, growing your skills and seeing your contributions impact lives around the world.

Join the company and do work that matters – to you, to your community, and to the world. Progress starts with you.

As a Software Development Engineer on the Product Reliability Engineering (PRE) team, you won’t just watch those systems run- you’ll be one of the engineers building, automating, and evolving them.

PRE is not a traditional ops team. We are a software engineering organization that treats infrastructure as code, reliability as a product, and automation as a strategic advantage. You’ll write Python, build agentic AI tools, manage data platforms, and contribute to the distributed systems that process billions of real-time transactions. From day one, you are an engineer- and from day one, your work matters.

If you are endlessly curious about how large-scale systems stay resilient, obsess over elegant automation, and want to launch your career at the intersection of AI, infrastructure, and global financial technology — this role was built for you.

Build Automation That Scales

Design, develop, and deploy end-to-end automation for deployment pipelines, infrastructure provisioning, platform operations, and release orchestration across complex production environments.

Write clean, production-grade Python (and Go or Bash where it counts) to eliminate toil, reduce operational risk, and improve the reliability and scalability of critical engineering workflows.

Design and implement reusable frameworks for release scheduling, validation, rollback, reporting, and configuration management that support the software delivery lifecycle.

Drive automation initiatives that improve engineering efficiency, standardization, and operational excellence across teams.

Manage & Evolve Data Platforms

Design, build, operate, and continuously improve relational database platforms supporting critical payment systems and high-volume transaction processing.

Contribute to architecture decisions, platform enhancements, and engineering solutions that improve scalability, resiliency, and performance.

Lead database health and lifecycle operations including upgrades, patching, backup and recovery strategies, and platform modernization efforts.

Analyze and optimize database performance through index tuning, execution plan analysis, replication monitoring, and capacity management.

Develop automation for database operations, configuration management, and schema deployments using tools such as Ansible, Liquibase, and CI/CD pipelines.

Build proactive monitoring, observability, and reporting solutions that identify reliability risks before they impact production services.

Ship Agentic AI & ML-Powered Tools

Design and build GenAI-powered engineering solutions that automate deployment orchestration, operational workflows, release governance, and platform management.

Integrate LLM-driven capabilities into observability, incident response, troubleshooting, and developer productivity workflows to improve operational effectiveness.

Evaluate and implement emerging AI, automation, and machine learning technologies that improve reliability, efficiency, and engineering velocity.

Contribute to agentic automation strategies that help evolve PRE into an increasingly intelligent and autonomous engineering organization.

Own Observability & Platform Health

Design and build dashboards, alerts, telemetry pipelines, and health indicators using tools such as Prometheus, Grafana, Splunk, or ELK to provide visibility across globally distributed systems.

Analyze platform performance, reliability, utilization, and availability data to identify trends and implement long-term improvements.

Lead troubleshooting efforts across infrastructure, applications, databases, and platform services, performing root cause analysis and driving durable corrective actions.

Design and implement self-healing, automated remediation, and auto-scaling capabilities that improve system resilience and reduce operational overhead.

Engineer for Reliability & Security

Design and implement highly available, scalable infrastructure solutions that support business-critical payment systems operating at global scale.

Ensure platforms and services meet security, compliance, governance, and resiliency requirements across cloud-native and hybrid environments.

Drive vulnerability remediation, configuration hardening, patch management, and security automation efforts to improve platform security posture.

Partner with engineering teams to build reliability and security practices directly into the software development lifecycle.

Collaborate, Learn & Grow Fast

Partner with software engineers, product managers, platform teams, and global PRE peers to design, deliver, and operate reliable engineering solutions.

Participate in architecture reviews, design discussions, code reviews, and technical planning activities, contributing engineering expertise and best practices.

Create and maintain technical documentation, runbooks, operational procedures, and engineering standards that improve team effectiveness and knowledge sharing.

Participate in on-call rotations and incident response activities, driving operational improvements and helping teams learn from production events.

Take ownership of assigned initiatives from design through implementation, deployment, and operational support while continuously seeking opportunities to improve systems and processes.

&xa;the company requires at least 3 days in office, expectations of these days will be confirmed by your Hiring Manager.

Requirements

  • Basic Qualifications
  • Bachelor’s degree with 2+ years of relevant professional experience, OR an advanced degree with at least 2 years of relevant experience, OR 5+ years of relevant work experience.
  • Preferred Qualifications
  • Hands-on software engineering or automation experience using Python, Java, Go, JavaScript/TypeScript, or a comparable language.
  • Experience supporting Linux-based systems and troubleshooting distributed applications or infrastructure.
  • Hands-on experience with one or more logging or search platforms such as Splunk, ClickHouse, OpenSearch, or Elasticsearch.
  • Hands-on experience with metrics and visualization technologies such as Prometheus, Thanos, Grafana, or Bosun.
  • Experience developing backend services, APIs, command-line tools, integrations, or operational automation.
  • Experience with cloud platforms, preferably AWS or GCP, and with cloud-native architecture and services.
  • Experience with containers and orchestration technologies such as Docker, Kubernetes, or equivalent enterprise container platforms.
  • Experience with infrastructure as code, configuration management, CI/CD pipelines, Git, automated testing, and deployment tooling.
  • Understanding of telemetry pipelines, log collection and parsing, metrics collection, alerting, dashboards, data retention, access controls, and platform integrations.
  • Understanding of distributed systems, scalability, high availability, disaster recovery, performance tuning, capacity planning, and reliability engineering.
  • Experience in production incident troubleshooting, root cause analysis, problem management, and implementing preventive remediation.
  • Experience with vulnerability remediation, secure configuration, certificate management, patching, upgrades, and software lifecycle management.
  • Ability to translate user and platform requirements into maintainable engineering solutions and clear technical documentation.
  • Strong problem-solving, communication, and collaboration skills, with the ability to work effectively across globally distributed teams.
  • Information for US Applicants
  • The company has a comprehensive benefits package for which this position may be eligible that includes Medical, Dental, Vision, 401(k), FSA/HSA, Life Insurance, Paid Time Off, and Wellness Program.
  • Work Hours
  • Varies upon the needs of the department.
  • Travel Requirements
  • This position requires travel 5-10% of the time.
  • Mental/Physical Requirements
  • This position will be performed in an office setting. The position will require the incumbent to sit and stand at a desk, communicate in person and by telephone, frequently operate standard office equipment, such as telephones and computers.
  • The company is an EEO Employer

Where you’d work

Part of the week in the office

You can work from

  • United States

About the company

Company hidden

  • Industry: FinTech

Offices in Austin, United States, United States

Your chances

Still hiring, not crowded yet, and you'd be among the first.

  • 16 checks run
  • 6 good signs
  • 0 red flags

Still hiring?

11 checks

Actively hiring

In its favour3

  • Still on the company's own careers site, checked 2 h agoModerate evidence
  • Found in the last 48 hours, before the big job boardsModerate evidence
  • The company posted 54 roles in the last 2 weeksSlight evidence

How crowded?

5 checks

Low

In its favour3

  • In its Early Window: not on the big job boards yetStrong evidence
  • Senior level: far fewer people qualifySlight evidence
  • Asks for ClickHouse, which fewer than 1% of open roles doSlight evidence

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details12 facts · Role, Location, Compensation, Employment, Company
Tech stack
  • Python
  • Go
  • AWS
  • Kubernetes
  • Docker
  • Elasticsearch
  • Linux
  • CI/CD
  • Ansible
  • Prometheus
  • Grafana
  • Splunk
  • TypeScript
  • Java
  • JavaScript
  • ClickHouse
  • LLMs
Seniority
Senior
Type
Full-time
Industry
FinTech
Region
United States
Pay period
Annual
Show 6 more factsShow less

Role

Category
DevOps & Infrastructure
Seniority
Senior
Experience
2+ years
Tech stack
  • Python
  • Go
  • AWS
  • Kubernetes
  • Docker
  • Elasticsearch
  • Linux
  • CI/CD
  • Ansible
  • Prometheus
  • Grafana
  • Splunk
  • TypeScript
  • Java
  • JavaScript
  • ClickHouse
  • LLMs

Location

Work model
Hybrid
Region
United States
Offices
  • Austin, United States
  • United States
Remote from
  • United States

Compensation

Salary
Salary by agreement
Pay period
Annual

Employment

Type
Full-time

Company

Industry
FinTech

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Lead IT Engineer

    Salary by agreement

    • Remote
    • Lead & Manager
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Security Engineer, New Grad

    Salary by agreement

    • Office · Dublin
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Site Reliability Engineer

    Salary by agreement

    • Remote
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior Security Operations Analyst

    Salary by agreement

    • Hybrid · Wellington, Auckland
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason