Test Engineer (AI)

Sixsentix AG

Zürich

Vor Ort

CHF 105.000 - 130.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Flexible working hours
Attractive fringe benefits
Opportunities for professional development
Team events

Zusammenfassung

A leading quality assurance firm based in Zurich is looking for an AI System Validator. In this role, you will design testing frameworks for AI applications, ensure compliance of AI outputs, and manage continuous monitoring solutions. The ideal candidate is passionate about AI safety, possesses a solid software testing background, and has excellent communication skills. Join a diverse and international team committed to driving innovation and quality in digital transformations.

Qualifikationen

  • Familiarity with RAG concepts and multi-agent systems required.
  • Strong background in GenAI challenges including hallucinations and output variability.
  • Solid experience in enterprise software testing and API frameworks expected.

Aufgaben

  • Design and implement testing frameworks for AI applications.
  • Evaluate AI outputs for factual correctness and compliance.
  • Set up continuous monitoring for production AI systems.

Kenntnisse

Familiarity with RAG concepts
Strong understanding of GenAI challenges
Solid software testing background
Proficiency in API testing frameworks
Data handling expertise
Outstanding communication skills

Tools

Postman
REST Assured
pytest
Git
MLflow
Weights & Biases

Jobbeschreibung

At Sixsentix, we don’t just test software – we accelerate innovation. As a leading provider of quality assurance & software testing services, we empower global enterprises to deliver high-quality solutions continuously faster and smarter. Our secret? A fearless blend of cutting-edge tools, proven methodologies and an approach that thrives on bold ideas.
Across all industries, we are the driving force behind some of the most successful digital transformations and platform innovation in Europe and beyond.

  • Impact from Day One: Your work will directly shape the digitalization of leading companies
  • Expert Collaboration: Work alongside some of the brightest minds in quality assurance and testing
  • Global Reach, Local Support: Be part of a diverse, international team with consultants in Switzerland, Serbia, Austria, Germany, Poland and the UAE
  • Continuous Learning: We champion growth—through training, mentorship and a culture that celebrates curiosity
  • Inclusive Culture: We value diversity and believe that different perspectives make us stronger
Who we’re looking for

You are a driven, team-oriented professional who thrives in dynamic environments. You enjoy solving complex problems, collaborating across borders and delivering exceptional value to clients as consultant. At Sixsentix, you will find a place where your ideas matter and your growth is a priority.

What we offer
  • An environment where you can leverage your previous experience to its fullest and apply new way of working by performing our services using AI and innovative approaches (Agent Business Model)
  • Significant professional development in the rapidly evolving AI governance space
  • Chance to work with cutting-edge AI technologies
  • Exposure to emerging regulatory frameworks and industry standards
  • Opportunities to influence AI strategy at the enterprise level
  • Engaging projects across various sectors
  • A community of skilled professionals in Quality Assurance and IT
  • Flexible working hours with the option to work from home
  • Team events designed to help you connect and share knowledge with your colleagues
  • Attractive fringe benefits, such as extra holidays, sports support, tenure awards and more
Your responsibilities
AI system validation & testing
  • Design and implement testing frameworks for LLM-based applications, multi-agent systems and RAG pipelines
  • Evaluate AI outputs for conciseness, completeness, semantic accuracy, factual correctness, contextual relevance, tone consistency, bias, safety and compliance
  • Implement governance workflows, risk scoring methods and compliance reporting as part of the overall AI Governance and Risk Management framework at clients
  • Set up continuous monitoring for production AI systems, with automated alerts for anomalies
  • Validate LLM APIs across multiple providers using tools like Postman, REST Assured and pytest
  • Integrate AI testing suites into CI/CD pipelines for regression testing, benchmarking and deployment gates
  • Set up AI platforms and components in client environments with close collaboration to their IT Security, I&AM and infrastructure teams
  • Manage version control for prompts and test cases using Git and leverage MLOps tools such as MLflow or Weights & Biases
  • Manage diverse data formats (JSON, CSV, Parquet, JSONL) and build evaluation datasets from real-world scenarios
  • Design automated metrics and human-in-the-loop workflows, including inter-annotator agreement processes
  • Understand and explain to stakeholders concepts around knowledge bases, vectorization and ingestion
  • Create and maintain domain-specific evaluation benchmarks and ground truth datasets
Further responsibilities
  • Build a high-performing team and serve as the subject matter expert towards the client
  • Cultivate and strengthen long-term client relationships, focusing on retention and business growth
  • Contribute with your knowledge and experience within the Sixsentix test consulting community
Your profile
  • Familiarity with RAG concepts, vector databases, embedding strategies and multi-agent systems
  • Strong understanding of GenAI challenges such as hallucinations, prompt sensitivity, consistency issues and output variability
  • Solid software testing background in enterprise environments, with the ability to communicate technical concepts to varied stakeholders
  • Proficiency in API testing frameworks, CI/CD integration and Git-based collaboration
  • Data handling expertise and knowledge of evaluation methodologies for AI performance and risk assessment
  • Self-starter mindset, comfortable working in ambiguous environments to define new testing methodologies and quality standards
  • Passionate about ensuring AI systems are safe, reliable and beneficial for end users
  • Demonstrate outstanding communication and presentation skills beyond pure technical expertise
Not the right fit? Explore our other open roles and find one that matches you.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Test Manager
Test Manager

Sixsentix AG • Zürich

Hybrid
CHF 90.000 - 120.000
Flexible working hours
Team events
Extra holidays
+2
Applied AI Engineer
Applied AI Engineer

Cyber Resilience Shield • Zürich

Hybrid
CHF 110.000 - 140.000
CHF 4'000 yearly for work-related equipment
Team events including snowboarding and go-karting
Flexible work environment
Senior AI Engineer
Senior AI Engineer

RepRisk • Zürich

Hybrid
CHF 90.000 - 130.000
Flexible working hours
Paid training and volunteering days
Health & fitness subsidy
+1
AI Engineer
AI Engineer

LHH • Lausanne

Vor Ort
CHF 110.000 - 150.000
Senior AI Consultant
Senior AI Consultant

ProActys GmbH • Schweiz

Hybrid
CHF 100.000 - 130.000
Flexible work model
Continuous learning opportunities
Flat hierarchy with big autonomy
+2
Senior Consultant Generative AI & Agentic AI 60-100%
Senior Consultant Generative AI & Agentic AI 60-100%

Eraneos • Zürich

Vor Ort
CHF 120.000 - 160.000
Customer Success Manager
Customer Success Manager

Enterprise Bot • Zürich

Vor Ort
Vertraulich
Fast-growing AI company
Global team
High ownership
+1
AI Software Engineer
AI Software Engineer

Avaloq1 • Zürich

Hybrid
CHF 140.000 - 200.000
Hybrid work model
QA Engineer II
QA Engineer II

Swiss Re - Schweizerische Rückversicherungs-Gesellschaft • Zürich

Hybrid
MXN 778.000 - 1.125.000
Inclusive work environment
Career development opportunities
AI Trainer (AI Tools)
AI Trainer (AI Tools)

Dreamleap • Schweiz

Hybrid
CHF 80.000 - 100.000
Flexible working model
High trust environment
Real use cases and problems