We're hiring marketing master's students to evaluate AI-generated marketing and communications content across European languages. This is a 12-week, full-immersion internship. The work is technical, hands-on, and paid. Around 20% of your time goes toward structured AI training: how large language models work, how RLHF fits into the development pipeline, what good evaluation looks like and why it matters. No prior AI knowledge is required. You'll learn everything on the job. The other 80% is live evaluation work on production AI models. This is the core of the internship. Your marketing knowledge and your languages are the reason you're here.
KEY RESPONSIBILITIES 1. Evaluating AI-generated marketing copy across languages. You'll compare AI-generated marketing content (email campaigns, ad copy, social media posts, product descriptions) and judge which version is better. Not just for grammatical accuracy, but for persuasiveness, tone, cultural appropriateness, and whether it would actually perform in-market. An AI-generated ad that's technically correct in Spanish but reads like it was translated from English is a failure. You'll be the one catching that. This is called RLHF (reinforcement learning from human feedback), and it's how AI labs train their models to produce content that actually works across markets. 2. Fact-checking AI in your native language. AI models hallucinate. They make up statistics, cite campaigns that never ran, and confidently state things that are wrong. You'll catch them. If you speak Italian and English, you might review AI outputs in both languages, flagging factual errors, culturally tone-deaf claims, and phrasing that no native speaker in a professional setting would use. Most AI evaluation today only covers English. European languages are massively underserved, and that gap is where your profile fits. 3. Red-teaming AI for brand safety. AI content tools are being adopted across every marketing department in Europe. But these models sometimes produce outputs that are biased, offensive, or off-brand in ways that are hard to predict. You'll test them. That means deliberately trying to get the model to produce problematic content, documenting exactly how it fails, and scoring the severity. Brands need this testing done by people who understand communications, not just engineers. 4. Evaluating AI-generated market research. AI labs building tools for consulting and market intelligence need evaluators who understand methodology. You might review an AI-generated market analysis brief and score it on data accuracy, whether the conclusions follow from the evidence, whether it reflects real European market dynamics, and whether a professional would consider it actionable. Your academic training in marketing research methods is directly applicable here. 5. Contributing to industry benchmarks. The evaluation work you do here feeds directly into the benchmarks top-tier AI labs use to assess and publicly report the factual capabilities of their models. When OpenAI, Google DeepMind, or Anthropic measure how well their models reason about marketing, datasets built by evaluators like you are part of what they test against. IDEAL QUALIFICATIONS Currently enrolled in a Master's in Marketing, Digital Marketing, Communications, or a closely related program Fluent in English plus at least one other European language (German, Spanish, French, Italian, and Polish are in highest demand) Comfortable working independently in a fully remote setup Reliable internet connection and a quiet workspace NICE TO HAVE Prior experience in digital marketing, brand management, content strategy, advertising, or market research Experience running campaigns across multiple European markets Experience working across cultures or in international teams CONTRACT & PAYMENT TERMS
Depending on experience (paid)