Principal AI Infrastructure Engineer (Part-time -> Full time)

Active Parks

Boston (MA)

Hybrid

USD 344,400 - 688,800

Part time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

An innovative AI company is seeking a Principal-level AI Infrastructure Engineer to enhance their architecture and systems. The role involves reviewing current systems, designing scalable architectures, and supporting implementations. Applicants should have significant experience with LLM-agnostic systems and scalability under high loads. This position starts as a fractional advisory role with an opportunity to transition into a full-time position. Located in the Boston area for occasional meetings, this role offers a chance to significantly influence the company's technology direction.

Qualifications

  • Experience in building and scaling LLM-agnostic systems.
  • Strong background in API-heavy systems during production loads.
  • Expertise in token-efficient architecture design.

Responsibilities

  • Review and pressure-test current architecture.
  • Design next-generation scalable architecture.
  • Support implementation and guide technical decisions.

Skills

Built and scaled LLM-agnostic systems
Scaled AI or API-heavy systems
Experience operating at billion-token-per-day scale
Deep expertise in rate limits and batching
Designed token-efficient architectures
Worked with closed and open-model providers

Job description

Engagement Type: Fractional (Advisory) to start, opportunity to move into full time role following initial engagement.

Compensation: $250–$500 per hour (depending on experience)

Time Commitment: 5–10 hours per week to start

Initial Duration: 4–12 weeks

Location: United States only. Boston area preferred (within a few hours’ drive) for occasional in-person meetings with founder.

About Us

We are a bootstrapped AI company building a high-throughput research and intelligence engine for the public-sector market.

We ingest and analyze public records at scale to identify government agencies entering active buying cycles for our clients’ solutions.

  • $100K revenue in first 6 months
  • 80%+ retention rate
  • Annual agreements with brand-name companies
  • Projecting $500K revenue this year
  • On track for profitability
  • Lean, senior team

Our initial internal AI research platform was built by one senior engineer and has successfully supported early customer growth.

Now, as customer volume increases and use cases diversify, we are adding senior talent and need to architect the next generation of our internal and external tools to support 100x current capacity.

The Role

We are seeking a Principal-level AI Infrastructure Engineer to:

  • Review and pressure-test our current architecture
  • Design a next-generation, LLM-agnostic system capable of 100x scale
  • Help guide and support implementation of that architecture

This is an engineering and systems role — not a management position.

There is a clear opportunity to evolve into a full-time lead engineer role after the initial engagement for the right person.

Engagement Phases
Phase 1 – Architecture & Codebase Review
  • Review current system architecture and codebase
  • Evaluate LLM usage patterns and token efficiency
  • Assess API orchestration, rate limiting, batching, queuing, and retry logic
  • Identify bottlenecks, fragility points, and scaling risks
  • Deliver a structured architectural assessment
Phase 2 – Next-Generation Architecture Design (100x Scale)
  • Design a scalable, LLM-agnostic AI architecture
  • Plan for 100x current throughput
  • Architect for:
    • Token and inference cost control
    • Provider abstraction (closed + open models)
    • Resilience and fallback routing
    • Distributed job orchestration
    • High-concurrency environments
    • Advise on local vs hosted inference strategy
    • Evaluate GPU cost, latency, and inference tradeoffs
Phase 3 – Implementation Support
  • Guide implementation of the new architecture
  • Review critical technical decisions during buildPressure-test scaling assumptions
  • Help prevent structural technical debt
Required Experience
  • Built and scaled LLM-agnostic systems
  • Scaled AI or API-heavy systems under real production load
  • Experience operating at billion-token-per-day scale (or comparable throughput environments)Deep expertise in rate limits, retries, batching, queuing, and distributed failure modes
  • Designed token-efficient architectures
  • Worked with both closed-model providers and open-source models
  • Deployed models locally or within controlled infrastructure
  • Evaluated GPU cost, latency, and inference tradeoffs
Preferred Background
  • Former CTO, Principal Engineer, or Staff Engineer
  • Experience at a VC-backed startup with a successful outcome or a major public technology company
  • History of scaling AI-native or API-intensive systems
  • Comfortable collaborating closely with a technically involved founder and senior engineer
  • Systems-oriented, pragmatic, and product-aware
Why This Is Interesting
  • Strong early product-market fit
  • Real production workload and scaling pressure
  • High ownership and architectural influence
  • Lean team with meaningful upside
  • Clear path to deeper involvement for the right person
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal AI Engineer
Principal AI Engineer

People In AI • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Principal Staff Software Engineer
Principal Staff Software Engineer

Harnham • Oakland (CA)

Hybrid
USD 400,000
Strong compensation
Meaningful equity
Opportunity to build from scratch
Principal Forward Deployed Engineer
Principal Forward Deployed Engineer

Elios AI • Boston (MA)

Hybrid
USD 250,000 - 350,000
Competitive compensation
Autonomy and impact
Hybrid work model
Principal Software Engineer (Applied AI)
Principal Software Engineer (Applied AI)

Standard Template Labs • New York (NY)

On-site
USD 130,000 - 160,000
Competitive salary
Equity
Collaborative work environment
+1
Staff Engineer
Staff Engineer

Recruiting From Scratch • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive equity
Ownership in infra
Direct collaboration with founders
+1
AI Engineer
AI Engineer

Latitude • Miami (FL)

Hybrid
USD 150,000 - 230,000
Founding AI Engineer: Mission-Critical LLM Systems (Hybrid)
Founding AI Engineer: Mission-Critical LLM Systems (Hybrid)

Open Talent • San Francisco (CA)

On-site
Senior Software Engineer, Infrastructure
Senior Software Engineer, Infrastructure

AI Talent Now • San Francisco (CA)

On-site
USD 250,000 - 300,000
4% gross profit share
Bonus potential up to $600k
Principal AI Engineer
Principal AI Engineer

IMR Soft LLC • New York (NY)

On-site
USD 180,000 - 260,000
Senior/Staff AI Engineer
Senior/Staff AI Engineer

AI Talent Now • San Mateo (CA)

Hybrid
USD 120,000 - 160,000