Employment Type: Long-term W2 contract with renewals typically occurring every 3–6 months. Contract extensions and conversion to a full-time position are possible but are not guaranteed.
Pay Range: $70–$75/hour, depending on experience
Start Date: ASAP
Location: Remote eligible during the contract term. Preference will be given to candidates who are open to working onsite 1–2 days per week in one of the following locations:
- St. Louis, Missouri
- Charlotte, North Carolina
- Stamford, Connecticut
If the position converts to full-time employment, the expected schedule would be four days onsite and one day remote.
Position Overview
The Infrastructure Intelligence and Analytics team builds and operates the data platform and AI-agent infrastructure that supports proactive network monitoring and autonomous investigation across large-scale network operations.
The Platform Engineer IV will design, build, automate, and maintain the AWS infrastructure supporting the organization’s data lake, data pipelines, AI-agent runtime environments, CI/CD processes, graph database systems, and production applications.
This position combines cloud infrastructure engineering, DevOps, automation, application development, and production support. The selected candidate will help ensure environments are stable, scalable, secure, and capable of supporting data science and agentic AI workloads at enterprise scale.
Responsibilities may span several focus areas depending on team priorities and the selected candidate’s strengths. Candidates are not expected to have experience with every technology listed below.
Key Responsibilities
AWS Infrastructure and Data Platforms
- Design, build, and manage AWS infrastructure supporting an enterprise data lake, including Amazon S3, AWS Glue, Athena, and EMR.
- Manage secure cross-account connectivity and data movement between internal teams, upstream data providers, and downstream consumers.
- Configure and support VPC networking, security groups, IAM roles and policies, PrivateLink, VPC peering, and transit gateway connectivity.
- Build and maintain infrastructure supporting AI-agent runtime environments, including compute resources for agents deployed through enterprise AI platforms.
- Design event-driven infrastructure using triggers, queues, messaging platforms, and publish/subscribe frameworks.
- Support the deployment and operation of AWS Neptune for network-topology and digital-twin use cases.
- Implement and maintain infrastructure as code using Terraform or CloudFormation.
- Manage secrets, access-key rotations, service accounts, and security controls using tools such as AWS Secrets Manager and Delinea.
- Support secure cloud and on-premises integrations, including Splunk and related platforms.
CI/CD and Deployment Automation
- Build and maintain CI/CD pipelines using GitLab CI/CD, Docker, Artifactory, and related deployment technologies.
- Manage container builds, image versioning, deployment promotion, and rollback processes.
- Automate the movement of applications and AI agents from development through production.
- Support cross-account deployments, AI-gateway integrations, and connectivity with internal platform teams.
- Improve deployment reliability, repeatability, auditability, and release speed.
Application Development and Internal Tooling
- Develop and maintain Python-based scripts, command-line tools, automation utilities, small services, and internal applications.
- Build integration code connecting internal systems with upstream data providers, downstream consumers, and external platforms.
- Support and enhance existing ETL pipelines written in Scala and Apache Spark.
- Assist with troubleshooting data pipelines and onboarding new data sources.
- Write maintainable, reusable, and well-documented code using established software-engineering practices.
- Develop unit and integration tests and incorporate automated testing into CI/CD pipelines.
Production Operations
- Maintain production stability through monitoring, alerting, troubleshooting, and incident response.
- Support service-level expectations for data-pipeline availability and AI-agent uptime.
- Implement health checks and monitoring for application errors, latency, resource utilization, and system performance.
- Coordinate with data, cloud, networking, Splunk, and platform teams to resolve access, firewall, connectivity, and integration issues.
- Support the data-engineering team with storage provisioning, access controls, pipeline compute resources, and data-source onboarding.
- Assist with workflow orchestration and scheduler development using tools such as Airflow and event-driven triggers.
- Identify opportunities to improve system reliability, scalability, security, and operational efficiency.
- Perform other duties as required.
Required Qualifications
- Five or more years of platform engineering, cloud infrastructure, DevOps, systems engineering, or related experience.
- Strong hands-on experience with AWS and several of the following services: EC2, S3, IAM, VPC, Glue, Athena, EMR, Secrets Manager, and CloudWatch.
- Experience implementing infrastructure as code using Terraform or CloudFormation.
- Experience supporting cloud networking, including VPCs, DNS, CIDR ranges, NAT, endpoints, firewalls, and security groups.
- Understanding of IAM policy design, least-privilege access, and service-account management.
- Experience with Docker, containerized environments, and container lifecycle management.
- Experience building or maintaining CI/CD pipelines.
- Proficiency in Python or another general-purpose programming language used for automation, tooling, or application development.
- Experience with Linux-based operating systems and shell scripting.
- Experience with monitoring, alerting, troubleshooting, and production support.
- Understanding of event-driven systems or messaging technologies such as Kafka, SNS, SQS, or EventBridge.
- Experience using Git-based development workflows, code reviews, dependency management, and modular software design.
- Experience integrating systems through REST APIs, AWS SDKs, or tools such as boto3.
- Strong communication and collaboration skills, including the ability to explain infrastructure decisions to technical and nontechnical stakeholders.
- Ability to work across teams and coordinate with cloud, networking, security, data, and application owners.
Preferred Qualifications
Experience in any of the following areas is helpful but not required:
- AWS Neptune, Neo4j, or other graph databases
- Apache Kafka or comparable streaming and messaging platforms
- Apache Spark, PySpark, or Scala
- Spark Streaming or Structured Streaming
- Airflow or another workflow-orchestration platform
- GitLab CI/CD and Artifactory
- Cross-account AWS architecture, PrivateLink, VPC peering, or transit gateways
- Prometheus, Grafana, CloudWatch, or similar observability tools
- FastAPI, Flask, or other production API frameworks
- AI/ML infrastructure, model serving, compute management, or artifact-management tools such as MLflow
- LangGraph, LangSmith, or other agentic AI platforms
- Splunk integrations, Edge Processor, or MCP connectivity
- Telecommunications, network operations, or another large-scale infrastructure environment
- AWS certifications such as Solutions Architect or DevOps Engineer
Education
Bachelor’s degree in Computer Science, Information Technology, Systems Engineering, Engineering, or a related discipline is preferred. Equivalent professional experience will also be considered.
Typical experience guidelines include:
- Bachelor’s degree with five or more years of relevant experience
- Master’s degree with three or more years of relevant experience
Equal Employment Opportunity Statement
Brooksource is an equal opportunity employer that does not discriminate on the basis of actual or perceived race, color, creed, religion, national origin, ancestry, citizenship status, age, sex or gender, including pregnancy, childbirth, lactation, and related medical conditions; gender identity or gender expression; sexual orientation; marital status; military service and veteran status; physical or mental disability; protected medical condition as defined by applicable state or local law; genetic information; or any other characteristic protected by applicable federal, state, or local laws and ordinances.
Pay Disclaimer
The pay range for this job level is a general guideline only and is not a guarantee of compensation or salary. Additional factors considered when extending an offer include, but are not limited to, the responsibilities of the position, education, experience, knowledge, skills, abilities, internal equity, alignment with market data, applicable bargaining agreements, and other applicable laws.
Benefits and Perks
Brooksource offers competitive medical, dental, and vision coverage, as well as Health Savings Account, Dependent Care FSA, and supplemental benefit options designed to fit employees’ individual needs.
Employees may also be eligible for:
- A 401(k) plan with a company match and full vesting after eligibility requirements are met
- Paid time off and sick time
- Paid company holidays, subject to hourly requirements and completion of the 13-week new-hire period
- An Employee Assistance Program offering resources such as virtual counseling, financial services, legal services, and life coaching