Infrastructure Engineer

Dexian

New York (NY)

On-site

USD 150,000 - 190,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Dexian, a global talent and technology solutions leader, seeks an experienced Infrastructure Engineer to drive capacity planning for a high-demand Azure environment. You will own the capacity management model, coordinate with SREs on IaC, and ensure SLAs, RTOs, and DR readiness while maintaining regulatory compliance across compute, storage, and network domains.

Dexian operates in 70+ locations with over 10,000 professionals and emphasizes Equal Opportunity employment across the United States.

Qualifications

  • Bachelor's degree in computer science or related field.
  • 10+ years in infrastructure capacity and performance engineering.
  • Experience in regulated environments and evidence collection.
  • Strong data analysis and telemetry forecasting abilities.
  • Proficient in change and configuration management.
  • Experience coordinating diverse engineering teams.
  • Experience with Azure services and RBAC.
  • Excellent stakeholder communication.

Responsibilities

  • Own the end-to-end Capacity Management model for Azure services.
  • Ensure capacity and buffers to meet SLAs, RTOs, RPOs, and regulatory needs.
  • Collaborate with SRE to implement capacity via IaC and autoscaling.
  • Contribute to documentation and evidence like security plans and audits.
  • Develop service-level capacity models and buffer standards.
  • Design and tune autoscaling policies with quotas and throttling.
  • Perform baseline and trend analyses and translate into actions.
  • Forecast future demand and translate into capacity plans.
  • Participate in change review processes documenting impacts.
  • Manage cryptographic configurations under change control.
  • Oversee external services and ensure standards.
  • Enforce region-restriction policies for processing and DR.
  • Balance performance, resilience, and cost through tuning.
  • Prepare dashboards and reports for audits and decisions.

Skills

Infrastructure capacity
Azure services
Data analysis
Change management
Executive communication
Cross-team coordination

Education

Bachelor's degree in Computer Science
Advanced degree is a plus

Tools

SQL
Networking
Monitoring tools

Job description

Key Responsibilities

We are seeking an experienced Infrastructure Engineer with expertise in cloud capacity management to join our team. This role focuses on leading capacity planning and optimization for a high-demand Azure public cloud environment. The ideal candidate will ensure the availability of resilient, scalable capacity across compute, storage, network, and platform services. They will collaborate with site reliability teams to meet performance and reliability targets, while maintaining compliance with rigorous program controls. This position offers an opportunity to shape critical cloud infrastructure capabilities in a regulated environment, emphasizing evidence-based practices and continuous monitoring.

  • Own and manage the end-to-end Capacity Management operating model for Azure services involved in high-criticality projects, including planning, modeling, forecasting, monitoring, tuning, and governance.
  • Ensure sufficient capacity and engineered buffers to meet service-level agreements (SLAs), recovery time objectives (RTOs), recovery point objectives (RPOs), and contractual or regulatory requirements, with particular attention to region-specific restrictions and ongoing monitoring.
  • Partner with site reliability engineers to implement capacity practices via infrastructure as code (IaC), gated change controls, performance baselines, autoscaling strategies, and resilience patterns.
  • Contribute to documentation and compliance evidence such as system security plans, control narratives, corrective action plans, and continuous monitoring artifacts.
  • Develop and maintain service-level capacity models across various Azure components, establishing buffer standards based on service criticality and validating against demand patterns and failover scenarios.
  • Design and tune autoscaling policies, setting guardrails on quotas and throttling to ensure performance stability.
  • Conduct baseline and trend analyses of utilization, throughput, and performance metrics, translating insights into tuning actions, reservations, savings plans, and architectural improvements.
  • Forecast future demand based on product roadmaps and business growth, translating forecasts into capacity plans and procurement strategies.
  • Participate in change review processes, ensuring capacity and security impacts are properly assessed and documented.
  • Manage cryptographic mechanisms and cryptography-related configurations under change control, maintaining versioned inventories and validation compliance.
  • Oversee external services supporting capacity, confirming they meet required standards and conduct ongoing oversight.
  • Enforce region-restriction policies for processing, storage, backups, and disaster recovery specific to high-impact systems.
  • Balance performance, resilience, and cost-efficiency through resource rightsizing, tiering, and scheduled scaling, ensuring proactive capacity adjustments.
  • Perform criticality analysis to prioritize capacity needs, aligning backup, monitoring, and security policies accordingly.
  • Validate disaster recovery (DR) capacity and ensure buffers are maintained for failover scenarios without impacting steady-state operations.
  • Define, measure, and report key capacity KPIs, including utilization, saturation, headroom, runway duration, scaling effectiveness, quota use, DR readiness, and cost-performance metrics.
  • Prepare dashboards and reports to monitor program compliance, support audits, and inform strategic decision-making.
Required Qualifications
  • Bachelor's degree in computer science or a related technical field; advanced degrees are a plus.
  • 10+ years of experience in infrastructure capacity and performance engineering across compute, storage, network, and platform services.
  • Demonstrated experience working within regulated environments and familiarity with high-assurance concepts and evidence collection standards.
  • Strong data analysis skills, with the ability to interpret telemetry and forecast data into actionable insights.
  • Proven process discipline with change and configuration management, and the ability to manage dependencies across IT operations and finance teams.
  • Experience coordinating diverse engineering teams and aligning delivery across multiple platforms and tools.
  • Proficiency with Azure services and concepts such as identity management, SQL, storage, networking, security policies, and role-based access control.
  • Excellent communication skills with the ability to manage stakeholder relationships and present technical information to executive audiences.
  • Capacity to translate complex technical requirements into clear plans, milestones, and measurable outcomes.
Preferred Qualifications
  • Experience developing capacity models for multi-region architectures with strict geographic restrictions.
  • Proven success in disaster recovery planning and execution, validated failover capacity, and documented evidence.
  • Experience managing corrective action and continuous monitoring submissions within regulated environments.
  • Strong collaboration with site reliability teams on service level objectives, error budgets, and reliability engineering patterns.
Why This Opportunity May Be IfAppeling

This position provides the chance to lead critical cloud infrastructure initiatives within a high-regulation environment, influencing how capacity and resilience are managed at an enterprise scale. You will work with cutting-edge tools and methodologies to ensure service availability and compliance, contributing directly to the success of complex, high-stakes projects. The role supports professional growth through involvement in strategic planning, cross-functional collaboration, and continuous improvement in cloud operations.

Dexian stands at the forefront of Talent + Technology solutions with a presence spanning more than 70 locations worldwide and a team exceeding 10,000 professionals. As one of the largest technology and professional staffing companies and one of the largest minority-owned staffing companies in the United States, Dexian combines over 30 years of industry expertise with cutting-edge technologies to deliver comprehensive global services and support.
Dexian connects the right talent and the right technology with the right organizations to deliver trajectory-changing results that help everyone achieve their ambitions and goals.
To learn more, please visit https://dexian.com/.
Dexian is an Equal Opportunity Employer that recruits and hires qualified candidates without regard to race, religion, sex, sexual orientation, gender identity, age, national origin, ancestry, citizenship, disability, or veteran status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Windows SME/ Lead Potential
Windows SME/ Lead Potential

Dexian • New York (NY)

On-site
USD 140,000 - 190,000
Digital Services Technical Manager
Digital Services Technical Manager

Dexian • Washington

On-site
USD 140,000 - 210,000
Sr. Program Manager, Cloud Migration
Sr. Program Manager, Cloud Migration

Dexian • Tallahassee (FL)

On-site
USD 90,000 - 150,000
Cloud Capacity Manager
Cloud Capacity Manager

IntePros • New York (NY)

On-site
USD 150,000 - 190,000
Sr. Linux Infrastructure Engineer
Sr. Linux Infrastructure Engineer

Dexian • Spring (TX)

Hybrid
USD 120,000 - 180,000
Full Stack Engineer
Full Stack Engineer

Dexian • Town of Texas (WI)

On-site
USD 120,000 - 150,000
Snowflake Platform Support
Snowflake Platform Support

Dexian • Coppell (TX)

On-site
USD 90,000 - 130,000
DevOps Engineer
DevOps Engineer

Dexian • Bedford (NH)

On-site
USD 140,000 - 190,000
Cloud Infrastructure Manager
Cloud Infrastructure Manager

IDEX Corporation • United States

On-site
USD 127,000 - 191,000
Health benefits
401(k) with company match
PTO
Senior Azure/ Microsoft 365 Infrastructure Manager
Senior Azure/ Microsoft 365 Infrastructure Manager

NIPRO Corporation - Global • Miami (FL)

On-site
USD 140,000 - 190,000