DevOps / Site Reliability Engineer
Operations · Cincinnati, St. Louis, or Remote · Full Time
About PayRecs
PayRecs is the platform community and regional banks use to offer modern cross-border B2B payments to their business customers. International payments were one of the last corners of banking untouched by modern software, and the banks businesses depend on were being cut out of the fintech wave rather than powered by it. We fix both. Banks get a best-in-class payments product under their own brand and a new revenue line for their treasury team; businesses get international payments built on our three pillars: simplicity, transparency, and innovation. PayRecs is post-Series A and live with 20 bank partners across the US, moving money through direct integrations with payment providers including Visa and Corpay, and partnering with digital banking platforms like Q2. We are a small, senior team where every hire has visible impact.
The role
You will be our first dedicated reliability hire, reporting to the Director of Engineering. Our infrastructure practice is intentionally early for our stage: infrastructure as code and container orchestration run in our test environment today, observability runs on CloudWatch, and incident response and on-call are yours to design from a clean slate. You will mature the platform from that baseline, own the reliability practice end to end, and set the foundation a growing engineering team builds on. In a regulated, money-movement business, reliability is a first-class product concern, and this is a broad, hands-on ownership role with a path to grow into a reliability lead as we scale.
What you will own
- Production reliability: availability, performance, and the response when things break.
- Infrastructure as code across all environments: extend our CloudFormation footprint from the test environment to reproducible, auditable IaC everywhere, including production.
- Container orchestration: roll ECS out beyond the test environment to all environments on a sound, secure pattern.
- The deployment path: harden CI/CD in GitHub Actions for release safety so engineers ship without breaking production.
- Observability: evolve from CloudWatch toward a Grafana, Mimir, Loki, and Prometheus stack, owning the metrics, logging, tracing, and alerting that surface problems before customers feel them.
- Incident response and on-call, built from zero: tooling, rotation, runbooks, blameless postmortems, and the follow-through that stops repeat incidents.
- Change management and the reliability side of compliance: controls, audit trails, and deployment practices that satisfy SOC 2 and the standards of bank partners and auditors.
- Cloud cost and capacity as the platform and customer base grow.
- Partnership with the Director of Engineering and developers so reliability is built into how the team works, not bolted on afterward.
What this role is not
This role does not own feature development or product direction, and it does not set top-level architecture strategy; the CTO does, with your input on reliability and operability. You own the platform and the reliability practice, and you enable the product engineers rather than competing with them for feature work.
Your first six months
- 30 days: learn the system, the current deployment and incident reality, and the compliance obligations; stabilize the most acute reliability gaps.
- 60 days: stand up CI/CD safety, real observability and alerting, and an on-call and incident process, so reliability has a dedicated owner.
- 90 days: extend CloudFormation and ECS from the test environment toward production, and document change-management controls to audit standard.
- Six months: production under infrastructure as code and orchestration, a functioning on-call practice, and reliability and compliance controls that hold up under partner and auditor scrutiny.
What you bring
- 4 to 6 years in SRE, platform, DevOps, or infrastructure engineering, with real ownership of production systems.
- Strong AWS infrastructure-as-code experience (CloudFormation specifically, or Terraform with willingness to work in CloudFormation), including taking environments from manual or partial coverage to full IaC.
- Hands-on ECS (or comparable container orchestration) experience, including rolling it out to production.
- Deep observability and incident-response experience: you have carried a pager and improved the systems that page you, and you have stood up on-call practice rather than just inherited it.
- Comfort being the first and only reliability hire: able to build practice from a low baseline and make pragmatic, stage-appropriate calls.
- Experience in our stack: AWS, ECS, CloudFormation, GitHub Actions, and CloudWatch, with services in Node/TypeScript and Python and data in MySQL (RDS).
Nice to have
- Fintech, payments, or another regulated, high-stakes environment where correctness and uptime are non-negotiable.
- Direct experience with SOC 2, audit, and change-management controls.
- Security engineering exposure: identity and access, secrets management, network controls.
- Experience with a Grafana/Prometheus-class observability stack, since that is where we are taking observability.
Growth path
As engineering scales past a single team, this role can grow into a lead or head of platform and reliability scope, owning the function and eventually a team.
How we work
At PayRecs, we are a dynamic team that values:
- Empowerment: we bring our expertise to the table every day.
- Trust: we work with integrity and support one another.
- Inquisitiveness: we ask questions and love to learn.
- Transparency: we communicate openly, building trust.
- Challenger Mindset: we take risks thoughtfully, focusing on elevating our clients.
We prioritize a healthy work-life balance, offering professional growth opportunities, family-friendly policies, and plenty of time to recharge.
Compensation and benefits
Competitive base salary, on-call incentives, and equity.
- Open PTO: time to rest and recharge.
- Paid parental leave.
- Employee stock options: share in our growth.
- Health care: inclusive package with both HMO and PPO plans, plus dental and vision.
- Annual office stipend: the equipment and comfort you need.
- Flexible work environment: remote or in-office as suits you best.
PayRecs is an equal opportunity employer and does not discriminate in any employment opportunities or practices on the basis of race, color, creed, religion, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), gender identity or expression, marital or domestic partner status, age, national origin or ancestry, physical or mental disability, medical condition, genetic information, sexual orientation, political affiliation, military or veteran status, or any other characteristic protected by applicable federal, state, or local law. PayRecs also prohibits discrimination based on the perception that a person has any of these characteristics or is associated with a person who has or is perceived to have any of these characteristics.
PayRecs will provide the federal government with your Form I-9 information to confirm that you are authorized to work in the U.S.