Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Senior Site Reliability Engineer

tastytrade
Companytastytrade
CategoryEngineering
LocationChicago
RemoteOn-site (inferred)
EmploymentNot stated
LevelSenior
SalaryNot stated by the employer
Posted7 Aug 2026
Last verified8 Aug 2026
SourceEmployer ATS (greenhouse)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
Company Name: tastytrade Role: Senior Site Reliability Engineer Location: Chicago, IL (Hybrid, 3 days/week in office) Role Summary Come join tastytrade, part of IG Group, as we build the reliability practice behind the brokerage platform that active options, futures, and equities traders rely on every market day. As our first Senior Site Reliability Engineer, you'll define what reliability means at tastytrade, from customer-meaningful SLOs on order execution and market data delivery to the error-budget policy and observability standards that guide every engineering team that follows. You'll work embedded alongside our infrastructure and application engineering teams, contributing reliability patterns directly into our Ruby, Java, and Elixir services and helping them scale across our HashiCorp Nomad-based service fabric. This is a rare opportunity to shape a practice and a culture from day one, on a platform where every order, quote, and position has to be right, because real client capital is on the line. What You'll Do (Job Responsibilities) Define customer-meaningful SLOs and set error budgets with multi-window burn-rate alerting for critical brokerage flows, including order execution and market data delivery. Author tastytrade's first reliability standards, including SLO methodology, error-budget policy, an observability instrumentation guide, and a Production Readiness Review (PRR) checklist. Contribute reliability patterns, such as circuit breakers, retries with backoff, bulkheads, and load-shedding, directly into our Ruby, Java, and Elixir services. Extend our observability stack (Prometheus, Honeycomb, OpenTelemetry) and guide teams as they scale workloads across our HashiCorp Nomad service fabric. Design and run tabletop exercises and fault-injection testing to stress-test the platform against real-world failure and volatility scenarios. Mentor engineers across teams to build a culture of site reliability champions so the practice outlives any one person. Who You Are (Skills Needed) Production-quality coding experience in Ruby and/or Java, plus Python for automation. A track record embedding SRE practices within engineering teams, including SLOs, error budgets, and burn-rate alerting. Hands-on experience with OpenTelemetry, Prometheus, and Grafana, with the ability to instrument services directly. Strong Linux internals and networking fundamentals, including TCP/IP, UDP/multicast, packet capture, and flow analysis. On-call experience on production systems and comfort building a blameless post-incident review process. Experience influencing standards across teams you don't own; HashiCorp Nomad, Consul, or Vault experience is a strong plus. Company Perks + Benefits:      Performance Bonuses     Stock Purchase Options     Medical/Vision/Dental Benefits   401k Plan    20 Paid Vacation Days (plus an additional paid vacation day the month of your birthday!)     10 Paid Sick Days     Gym Membership Reimbursement     Commuter Benefits     Pet Insurance     Wellness & Mental Health Programs     Charitable Donation Matching     Two Paid Volunteer Days Off     Daily catered lunch when in the office     Full kitchen with snacks and beverages     In-building gym     Shuttle to/from Metra     Base Salary Range:  $180,000-$200,000 The actual salary offered will be based on the candidate's level of experience and qualifications.   Discretionary Performance Bonus: 15-20% of base salary based on individual and company performance.   About IGNA + tasty    IG North America is home to tastytrade, tasty live &  tastyfx—a family of brands built to democratize trading and empower individu