โ All Web3 jobs
Senior Site Reliability Engineer
January
New York CityHybrid$200k to $225k
About January At January, we're rebuilding one of the most broken parts of consumer finance: what happens after someone falls behind on debt. Our autonomous, compliant collections platform works directly for creditors, using data and AI to personalize outreach and offer consumers clearer, more compassionate ways to resolve what they owe. January has serviced $20B+ in debt and engaged 20M+ consumers. Backed by Series B funding, ~100+ people working out of Nolita, NYC. About the Role As a Senior SRE you will ensure the reliability, scalability, and performance of January's production and internal systems as we scale from thousands to millions of borrowers. You'll establish SRE practices from the ground up โ architecting resilient infrastructure, implementing proactive monitoring solutions, and building sustainable on-call processes. Key Responsibilities - Lead incident response and establish sustainable on-call practices, including comprehensive runbooks, blameless postmortems, and systematic improvements that reduce MTTR - Develop and maintain self-service observability solutions using modern monitoring tools - Create and maintain infrastructure as code (Terraform, CloudFormation) for consistent, scalable, secure cloud environments on AWS - Partner closely with feature teams to architect resilient infrastructure for critical components (databases, networking, async workflows, data pipelines) - Design and implement robust CI/CD pipelines with advanced deployment strategies (blue/green, canary) - Advocate for best practices early in feature design, ensuring we design with reliability in mind Requirements - Expertise leading incident response for high-availability production systems, thorough root cause analysis, and fostering blameless postmortem culture - Experience designing highly available deployment architectures across multiple targets (EC2, Fargate), with expertise in auto-scaling, health checks, and graceful degradation - Track record implementing effective monitoring and observability solutions (Datadog, Prometheus, ELK) - Strong knowledge of AWS cloud services and infrastructure-as-code using Terraform - Experience with CI/CD pipelines and automation - 5+ years SRE/DevOps OR 7+ years SWE with strong infrastructure focus
Apply on Deciml
Deciml matches Web3 professionals to roles with AI, free for candidates.