Josh Pros is hiring a Senior Fullstack AI Engineer for a long-term, W-2 only contract based in Richardson, TX with hybrid work (2 days onsite per week). In this role, you will own the end-to-end design, architecture, and delivery of enterprise AI platforms and applications, including LLM and agentic capabilities.
Working across the full lifecycle from concept through production, you will shape technical direction for model strategy, scalable AI system design, and measurable evaluation practices. The position emphasizes building reusable AI services and establishing engineering standards for reliability, safety, and cost-aware performance.
What you’ll do
- Own the end-to-end design of AI platforms and applications
- Define model strategy using OpenAI, Google, and an approach spanning small vs. large models
- Design scalable AI systems using RAG, multi-agent frameworks, and workflow orchestration
- Apply knowledge graphs, hybrid search, and semantic caching to improve retrieval and performance
- Evaluate and implement emerging AI technologies, including LLMs, multimodal models, and real-time systems
- Build reusable AI services, SDKs, and frameworks
- Set best practices for prompt engineering and versioning, model evaluation, and observability
- Implement and oversee logging, monitoring, and incident response
- Develop modular, scalable backend systems for agentic workflows
- Lead and mentor AI/ML engineers and data scientists
- Drive design reviews, architecture decisions, and coding standards
- Partner with product and business leaders to identify high-impact AI use cases
- Own delivery of complex AI initiatives from concept through production
- Define evaluation frameworks for quality, safety, and business impact
- Ensure responsible AI practices, including bias, privacy, and compliance
- Optimize performance and cost through model routing, caching, and infrastructure design
- Establish SLAs/SLOs and ensure system scalability and reliability
Required experience
- 8+ years in Software Engineering, ML Engineering, or Data Science
- 3+ years working with Applied AI / LLM systems
- Proven experience in a technical leadership role
- Strong production experience with Python and/or TypeScript
- Cloud-native architecture experience with AWS, Azure, or GCP
- Experience with distributed systems, Docker, Kubernetes, and CI/CD
- Hands-on LLM application development, including prompting, tool use, and multi-agent systems
- Experience with RAG architectures, vector databases, and retrieval optimization
- Knowledge of AI evaluation, monitoring, and observability
- Proven ability to architect and scale AI systems in production
- Proven ability to lead cross-functional initiatives and mentor teams
- Ability to communicate complex AI concepts to technical and business stakeholders
Technologies you may work with
- OpenAI, Google, LLMs, multimodal models, real-time systems
- RAG (retrieval-augmented generation), multi-agent frameworks, workflow orchestration
- Knowledge graphs, hybrid search, semantic caching, vector databases
- Python, TypeScript
- AWS, Azure, GCP, Docker, Kubernetes, CI/CD
- LangSmith, OpenAI Evals
- Cursor, Codex, Windsurf
Location and contract details
- Location: Richardson, TX
- Work mode: hybrid, 2 days onsite per week
- Duration: long-term contract
- Employment: W-2 only (no C2C or 1099)
- Rate: depends on experience
Nice to have
- Experience with coding assistants such as Cursor, Codex, or Windsurf
- Multimodal or real-time AI applications (voice, streaming, UI agents)
- Exposure to AI evaluation frameworks such as LangSmith or OpenAI Evals
- Background in highly regulated or privacy-focused environments
- Experience shaping AI strategy or building AI capabilities at scale