Oscilar

Sr./Staff - Infrastructure/Site Reliability Engineer (SRE)

Reposted Yesterday

Remote

Hiring Remotely in USA

Senior level

Remote

Hiring Remotely in USA

Senior level

Seeking a seasoned SRE to lead reliability for a cloud-native platform, overseeing infrastructure, CI/CD pipelines, observability, and mentoring engineers.

The summary above was generated by AI

Shape the future of trust in the age of AI
At Oscilar, we're building the most advanced AI Risk Decisioning™ Platform. Banks, fintechs, and digitally native organizations rely on us to manage their fraud, credit, and compliance risk with the power of AI. If you're passionate about solving complex problems and making the internet safer for everyone, this is your place.

Why join us:

Mission-driven teams: Work alongside industry veterans from Meta, Uber, Citi, and Confluent, all united by a shared goal to make the digital world safer.
Ownership and impact: We believe in extreme ownership. You'll be empowered to take responsibility, move fast, and make decisions that drive our mission forward.
Innovate at the cutting edge: Your work will shape how modern finance detects fraud and manages risk.

About the Role

Oscilar is growing fast, and so is the complexity of our systems. We’re looking for a experienced SRE to take ownership of reliability across our multi-region, cloud-native platform. You’ll have the mandate and autonomy to design, implement, and evolve systems that stay performant and resilient—through traffic spikes, dependency failures, and global deployments. You’ll be shaping how we scale, how we build observability, and how we run infrastructure that supports billions of events and large-scale data pipelines.

What You’ll Own

Architect and operate resilient cloud infrastructure (AWS, Pulumi, Kubernetes).
Lead initiatives to improve availability, latency, and performance at scale.
Design and evolve our CI/CD pipelines to optimize for speed, safety, and repeatability.
Define the metrics, alerts, and runbooks that form our observability backbone.
Run chaos experiments and failure simulations to harden the platform.
Mentor engineers and set best practices for SRE across the company.

What You Bring

Proven track record as a senior SRE or Infrastructure Engineer in high-scale environments.
Expert-level skills in AWS and Infrastructure as Code (Pulumi, Terraform).
Strong programming ability in Go or Python. We use Go.
Deep understanding of distributed systems (Kafka, ClickHouse) and microservices architecture.
Mastery of container orchestration (Kubernetes) and production debugging.
Strong sense of ownership, and the judgment to balance velocity with reliability.

Benefits

Compensation: Competitive salary and equity packages, including a 401k plan
Flexibility: Remote-first culture — work from anywhere
Health: 100% Employer covered comprehensive health, dental, and vision insurance with a top tier plan for you and your dependents (US)
Balance: Unlimited PTO policy
Technical: AI First company; both Co-Founders are engineers at heart; and over 50% of the company is Engineering and Product
Culture: Family-Friendly environment; Regular team events and offsites
Development: Unparalleled learning and professional development opportunities
Impact: Making the internet safer by protecting online transactions

Top Skills

AWS

Clickhouse

Java

Kafka

Kubernetes

Pulumi

Terraform

Similar Jobs

Newton.co

Site Reliability Engineer

3 Days Ago

In-Office or Remote

Mid level

Blockchain • Financial Services • Cryptocurrency • Web3

The Site Reliability Engineer will enhance operational reliability, managing incidents, improving system performance, and maintaining metrics to ensure scalability and resilience.

Top Skills: AWSJavaJavaScriptLinux ShellPython

Branch (branch.io)

Senior Site Reliability Engineer

5 Days Ago

In-Office or Remote

123K-160K Annually

Senior level

123K-160K Annually

Senior level

Mobile • Software • Analytics

The Senior Site Reliability Engineer will enhance system reliability, drive automation, and mentor teams in managing large-scale infrastructures and incident responses.

Top Skills: AerospikeAlertmanagerArgocdAWSBashCloudFormationFoundationdbGitopsGoGrafanaJavaKafkaKotlinKubernetesLinuxLokiPagerdutyPrometheusPythonSparkTerraform

Kraken Digital Asset Exchange

Site Reliability Engineer

6 Days Ago

Remote

Senior level

Blockchain • Financial Services • Cryptocurrency • Web3

As a Senior Site Reliability Engineer, you'll enhance the reliability and scalability of data infrastructure, manage cloud solutions, automate deployments, and ensure compliance and performance monitoring.

Top Skills: Apache AirflowSparkAWSBashDebeziumDockerKafkaKubernetesTerraform

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus