We're hiring interns for quality assurance and operations across verticals for the summer internship programme. This is a 12-week, full-immersion internship. You'll both contribute to and oversee the work produced by our business evaluators across multiple European languages and make sure it meets the quality bar that AI lab clients expect. Around 20% of your time goes toward structured AI training: how large language models work, how RLHF fits into the development pipeline, and what quality standards look like for business and cross-lingual AI evaluation data. No prior AI knowledge is required. You'll learn everything on the job. The other 80% is hands-on QA, operations, and product testing, the other 20 percent is the curriculum we created.
KEY RESPONSIBILITIES 1. Quality-checking business evaluation output. Our business evaluators produce RLHF preference rankings on AI-generated business strategy content, cross-lingual business evaluations, and general business fact-checking. Each evaluation includes a rubric score and a written rationale. Your job is to audit that output. You'll check whether evaluators are making real business judgments (not just grammar checks), whether their assessments of strategy memos and competitive analyses reflect actual professional standards, and whether cross-lingual evaluations are catching cultural and contextual issues, not just translation accuracy. 2. Calibration across the business team. Business evaluation has more subjective territory than finance or law. What counts as "sound strategy advice" depends on context. You'll run inter-annotator agreement checks, identify where evaluators are diverging, build consensus on quality standards, and calibrate new evaluators during onboarding. The cross-lingual work adds another layer: you'll make sure evaluators are judging native-language outputs to native-speaker standards. 3. Testing Sovrano AI's products and tools. We're building internal tooling and AI-powered products (interview preparation agents, evaluation platforms, workflow automation). You'll test them as a real user would, finding bugs, edge cases, and UX issues, and documenting everything clearly for the product team. 4. Process improvement. You'll bring your experience and perspective to how we run our business evaluation pipeline: onboarding, quality feedback, client deliverables, and cross-language coordination. 5. Contributing to industry benchmarks. The evaluation work you do here feeds directly into the benchmarks top-tier AI labs use to assess and publicly report the factual capabilities of their models. When OpenAI, Google DeepMind, or Anthropic measure how well their models reason about business, datasets built by evaluators like you are part of what they test against. IDEAL QUALIFICATIONS Currently enrolled in an MBA/Master's/Bachelor program (or recently completed) Fluent in English plus at least one other European language Comfortable working independently in a fully remote setup Reliable internet connection and a quiet workspace NICE TO HAVE Prior experience in consulting, operations, project management, business development, or strategy Experience working across cultures or in international teams Comfort with structured data, spreadsheets, and basic analytics