Cluster & Systems Capacity Engineer at Backblaze | Torre

Cluster & Systems Capacity Engineer

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Full-time

Legal agreement: Employment

Compensation
USD123k - 175k/year
location_on
Remote (for United States residents)
Match
skeleton-gauges
You have opted out of job matches in .
To undo this, go to the 'Skills and Interests' section of your preferences.
Review preferences
Shared by
Emma of Torre.ai
about 15 hours ago

Requirements and responsibilities


About BackblazeBackblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our partners, we’re helping customers break free from the restrictive, overpriced legacy solutions that hold them back, and blaze forward with the full power of the open cloud in their hands.Founded in 2007, we scaled the business with less than $3 million in outside funding until 2021, when we did a traditional IPO on the Nasdaq stock exchange. Today, Backblaze generates over $100m in revenue and is the leading specialized storage cloud - managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals.But while there is a lot to celebrate in our past, there is almost as much opportunity ahead of us. Backblaze is seeking a highly analytical, systems-oriented Cluster & Systems Capacity Engineer to drive the planning, forecasting, deployment, and optimization of hardware infrastructure across our global cloud storage platform. About the Role:This role ensures that Backblaze’s storage clusters, compute systems, and network infrastructure scale reliably, cost-efficiently, and ahead of demand. You will build and maintain predictive models, ensure consistent supply and demand alignment, and partner cross-functionally to inform strategic investment and deployment decisions. .This is a high-impact role within Cloud Operations, directly contributing to service availability, durability, performance, margin optimization, and long-term platform scalability.Key Responsibilities:Capacity Planning & ForecastingDevelop and maintain short, medium, and long-term capacity demand and hardware deployment forecasts across storage, compute, and network domains within the platformBuild predictive models that translate business demand signals into infrastructure requirements using historical utilization, growth trends, product sales plans, hardware lifecycle roadmaps, and other key business inputsPartner with Infrastructure, Production, and Network Engineering teams to align capacity plans with system design and scaling initiativesDevelop and automate forecasting pipelines, simulation calculators and tools, and capacity dashboards to improve data quality, reduce manual analysis, and provide stakeholders clear visibility into platform usage and cluster health metricsCluster Performance & Resource OptimizationMonitor and analyze cluster and system-level utilization and performance across CPU, memory, IOPS, and network resourcesAdjust deployment plans and recommended configurations in real-time to maintain adequate headroom and system stability in support of delivering a world-class customer experiencePartner with service and platform owners to develop headroom and live buffer policies, optimize hardware BoMs, leverage virtualized orchestration, and reduce product costCross-functional Organizational AlignmentWork in lockstep with Operations and Finance peers to align capacity plans and hardware requirements with capital budgets, cost targets, and financial outcomesSupport strategic optimization initiatives across infrastructure investments, engineering development, and operations processes, contributing to long-term infrastructure strategy and capital planningLead efforts to evaluate, procure, and provision requests for new or additional hardware, working with Systems and Network Engineering, SRE, NOC, and Data Center Operations teams to identify and deliver optimal solutionsMaintain alignment with Product and Sales to support customer onboarding, growth, and demand variabilityCommunicate complex capacity and infrastructure insights clearly to technical and non-technical stakeholdersRequired QualificationsBachelor’s degree in Computer Science, Engineering, Mathematics, Data Science, Information Systems, Statistics or a related, technical field (or equivalent experience).3-6+ years of experience in Site Reliability Engineering, Infrastructure Capacity Planning, Systems/Infrastructure Engineering, Production Engineering, Data Center Operations or similar Cloud Operations roleFamiliarity and experience working with Cloud Storage infrastructure, particularly highly-available, large-scale distributed systems supporting large amounts of data with high throughput and complex performance requirementsBackground in capacity modeling, performance analysis, scenario modeling, and/or infrastructure cost optimization, with an ability to quantify ideas within financial frameworks and forecasts.Proficiency in database and data analysis tools (preferably Snowflake, Metabase, Grafana, Python, SQL, Prometheus, Victoria Metrics, and Excel/Google Sheets)Demonstrated deep, creative, and logical thinking complimented by a strong data analysis skillsetExcellent communication and documentation skills, with the ability to share knowledge and explain concepts accurately and conciselyDesire to work on a highly-autonomous team that cares deeply about quality, cost, and the customer experienceBackblaze Perks: Healthcare for family, including dental and visionCompetitive compensation and 401K  RSU grants for full-time employees ESPP program  Flexible vacation policy Maternity & paternity leave MacBook Pro to use for work, plus a generous stipend to personalize your  workstation Childcare bonus (human children only) Fertility treatment and support Learning & development program Commuter benefits Culture that supports a healthy work-life balance To provide greater transparency to candidates, we share base pay ranges for all US-based job postings regardless of state. We set standard base pay ranges for all roles based on function, level, and country location, benchmarked against similar-stage growth companies. Final offer amounts are determined by multiple factors, including candidate location, skills, depth of work experience, and relevant licenses/credentials, and may vary from the amounts listed below.The expected salary range for this role is $123,000 - $175,000.