Evaluate and annotate AI-generated responses while providing structured feedback. Collaborate with teams and clients to maintain high-quality standards in AI performance.
Responsibilities
Review AI-generated responses for clarity, accuracy, correctness, relevance, and overall quality.
Annotate and label language content according to detailed project guidelines.
Generate or evaluate prompts based on assignment requirements.
Identify linguistic errors, dataset concerns, and structural inconsistencies.
Provide clear, structured feedback to support improvements in AI performance.
Communicate complex language requirements clearly to both linguistic and non-linguistic stakeholders.
Collaborate with QA Leads and clients to apply feedback and maintain quality standards.
Participate in required client-facing meetings with your camera on.
Submit daily work reports and consistently meet productivity and quality expectations.
Requirements
Professional fluency in **Italian** with strong written and verbal communication skills.
A bachelor’s degree in Linguistics, Languages, Computer Science, or a related field, or equivalent professional experience.
**A degree or certification in Italian, where required. **
Strong analytical skills and exceptional attention to detail.
The ability to identify patterns, errors, inconsistencies, and subtle language issues.
Experience in annotation, evaluation, translation, localization, linguistics, research, education, or quality assurance is helpful but not required.
An interest in Artificial Intelligence, Machine Learning, language technology, or data annotation.
The ability to learn new tools, follow detailed guidelines, and adapt to changing project requirements.
Strong organization and task-management skills.
Professional communication skills and confidence participating in client-facing meetings.
The ability to work independently and collaborate effectively with a global team.
Video Specialist training SpaceXAI’s Grok AI to interpret and generate video content. Producing annotated, curated video data using editing, motion graphics, and VFX expertise.
Part - time AI workflow consultant evaluating ChatGPT and Claude across realistic business processes. Scoring outputs, documenting model behavior, and providing actionable feedback for 24 - MAG’s AI projects.
AI Evaluation Specialist assessing AI - generated outputs, reasoning, and tool use for 24 - MAG LLC’s remote consulting platform. Providing rubric - based quality evaluations and actionable written feedback.
Freelance AI Visual Artist experimenting with generative image, video, music, and text tools. Creating innovative visual concepts for Paintgun’s performance advertising campaigns.
AI Success Architect helping Meltwater customers transform media, social, and AI signals into actionable workflows. Designing LLM agents, MCP integrations, and repeatable customer deliverables.
AI agent engineer building and managing autonomous agents for Sticker Mule’s software, manufacturing, and AI commerce platform. Connecting agents to tools, measuring results, and advising teams.
Senior AI Analyst prototyping and deploying cloud - based AI solutions for CBC/Radio - Canada media systems. Integrating LLMs, intelligent agents and AI capabilities across digital platforms.
Audit Innovation Principal building AI solutions for Caseware's audit and accounting software platform. Translating live practitioner workflows into reliable tools with engineering and applied science teams.
Strategy and AI operations leader partnering with Parallelz executives on planning, analysis, business operations, and fundraising. Building AI - native systems as Parallelz scales native mobile apps into web experiences.
Global AI and data risk audit leader overseeing RBC's AI governance, model risk, and data controls. Providing regulatory assurance across the bank's global operations.