Help improve the next generation of voice-and-camera AI assistants by evaluating how accurately they understand and respond to real-world visual information.
As an AI Session Annotator, you’ll review recorded interactions between users and an AI assistant. You’ll compare the assistant’s responses with what was actually visible in the video, evaluate its performance against a detailed rubric, and provide precise written feedback.
Your comments and evaluations will directly help identify issues such as visual hallucinations, inaccurate interpretations, poor timing, and inadequate safety warnings.
This is a judgment-focused role where accuracy, consistency, and quality of feedback matter more than review speed.
What You’ll Do
Review recorded sessions where users interact with an AI assistant through voice and camera.
Compare the assistant’s responses with the visual information actually available in the video.
Identify fabricated, hallucinated, or inaccurate visual details.
Evaluate each session for:
Visual interpretation
Accuracy
Relevance
Response timing
Appropriate use of spatial and directional language
Assess whether the assistant provides appropriate and timely warnings when responses could affect a user's health, physical safety, or financial security.
Write detailed comments supporting each score with specific evidence from the session.
Document conversations turn by turn using the provided intake form.
Apply the evaluation rubric consistently without introducing personal scoring criteria.
What You Bring
Native or native-level English proficiency , both spoken and written.
Excellent written English and the ability to communicate observations clearly and precisely.
Strong attention to detail and analytical judgment.
Ability to follow a detailed evaluation rubric consistently.
Ability to distinguish between what is actually visible in a video and what an AI assistant claims to see.
An Android smartphone with a working camera and microphone — mandatory for this project.
Preferred Qualifications
Experience in any of the following is a plus:
Data annotation
Linguistic quality assurance
Content moderation
AI model evaluation
Linguistics
Translation or interpreting
Conversational AI evaluation
An ear for regional English usage and the ability to recognize unnatural or non-native phrasing is also valuable.
Content & Project Details
Start Date: Immediate
High-Volume Period: August 26–31
Compensation: $24–$30/hour
Annualized Equivalent: $49,920–$62,400
Work Arrangement: Fully remote
Engagement: Independent contractor
Schedule: Flexible, based on project requirements
Annualized compensation is based on 2,080 hours per year for comparison purposes only. Actual earnings depend on the number of hours and projects completed.
Content Considerations
Some sessions involve scenarios where an incorrect AI response could have meaningful real-world consequences. Examples may include identifying medication, determining whether food has spoiled, navigating a street-crossing situation, or interpreting bank card information.
There is no graphic or violent content , but some sessions involve safety-sensitive situations. You’ll be expected to evaluate whether the AI assistant recognized potential risks and provided an appropriate warning when necessary.
Application Process
Submit your resume or relevant professional background.
If selected, you may be asked to complete a short calibration exercise using sample sessions.
The calibration focuses on scoring consistency and the quality of your written comments—not speed .
Why Join?
Help improve the accuracy and reliability of advanced multimodal AI systems.
Apply your language and analytical skills to real AI evaluation work.
Work remotely with a flexible schedule.
Contribute directly to identifying and correcting AI visual reasoning failures.
Develop hands-on experience in AI model evaluation and data quality.
Contract & Payment Terms
You will be engaged as an independent contractor.
Work is fully remote and can be completed on your own schedule.
Projects may be extended, shortened, or concluded early depending on project needs and performance.
Your work will not require access to confidential or proprietary information belonging to any employer, client, or institution.
Payments are made weekly via Stripe or Wise based on services rendered.
H-1B and STEM OPT candidates cannot be supported at this time.
Equal Opportunity
All qualified applicants will be considered without regard to legally protected characteristics. Reasonable accommodations are available upon request.
... workflows, and model optimisation. - Practical knowledge of Explainable AI (XAI) and model interpretability techniques. - Strong understanding of Responsible AI , AI governance, and ethical model-development practices. - Experience implementing AI guardrails, safety controls, content filtering, and controlled-generation ...
... Computer Science, Engineering, or a related field. A Ph.D. in AI or Machine Learning is a plus.- Proven experience as an AI Engineer, Machine Learning Engineer, or AI Architect, with a track record of successful AI solution design and implementation.- Streamlit - Python- Generative AI- HuggingFace, OpenAi and any Custom Model ...
... deliver end-to-end technology solutions. - Conduct code reviews, write unit and integration tests, and ensure strong standards for code quality, documentation, maintainability and automated testing. - Support regulatory and control-focused initiatives by delivering solutions aligned to enterprise data management objectives ...
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. AWS AI Services Minimum 3 Year(s) ...
... Develop and maintain automation scripts using Python and other scripting tools to automate repetitive tasks and reporting workflows. Evaluate, test, and deploy new AI tools and automation solutions tailored for internal use, ensuring they align with organizational goals. Assist with system monitoring, reporting, analytics, and ...
... Azure Generative AI (GenAI) We are seeking a highly skilled and innovative Data Scientist with hands-on experience in Azure AI services , particularly Azure OpenAI and Generative AI technologies . The ideal candidate will have a strong foundation in machine learning, natural language processing (NLP), and cloud-based AI solutions. ...
... and collaborative development workflows. AI & Agentic Frameworks: LangGraph, CrewAI, Microsoft Semantic Kernls, AutoGen, OpenAI Agents Framework, Anthropic APIs, or Azure OpenAI Services. Cloud & Tooling: Azure (Functions, AI Services, AKS), Docker, CI/CD pipelines, GitHub Copilot or other AI-assisted development tools.
... rollback mechanisms using LangChain/LangGraph. Build and maintain the tool and data integration layer connecting agents to external data sources, APIs, and domain-specific infrastructure systems. Collaborate closely with agent, data, and domain teams to ensure platform abstractions support evolving research and product ...
... (MCP).- Experience with cloud-based AI platforms.- CI/CD and DevOps https://jobeax.com/link/wG4xWp4WUrWMLMb3 Candidate Profile :- Candidates with experience as GenAI Engineer, AI Engineer, LLM Engineer, Agentic AI Developer, Applied AI Engineer, AI Application Engineer, or Generative AI Engineer are preferred.- Important: Candidates ...
... infrastructure teams, and business stakeholders to deliver secure and scalable solutions. AI Strategy & Engineering Excellence - Identify opportunities to apply AI and Generative AI to improve software engineering productivity and modernization outcomes. - Lead proof-of-concepts and pilots demonstrating AI-driven modernization ...
AI Technology LeaderExperience : 20 - 28 YearsJob Location : Chennai / Bangalore / Hyderabad / Pune / Mumbai / NCR / Delhi / https://jobeax.com/link/WaIJbxtkA0b23np3 Summary :We are seeking a visionary and hands-on Technology Leadership AI focused to lead the technology strategy, innovation, and AI/GenAI initiatives for ...
... LangChain Agents. Role: Agentic AI Engineer Location: All Persistent Location Experience: 5+ Years Job Type: Full Time Employment What You'll Do: Build and deploy GenAI applications using LLMs (OpenAI, Azure OpenAI, Claude, Gemini, Llama, etc. Develop Agentic AI workflows using frameworks such as CrewAI, AutoGen, LangGraph, or ...
... to solve complex challenges in data, experience design, software product engineering, and workforce transformation. Powered by expert engineers, thousands of AI agents, and our Engineering to the Power of AI™ (EngineeringAI) method, we deliver measurable outcomes that build trust, unlock value, and accelerate growth. We ...
... things run in production. What You’ll Build Production AI Ship models that meet defined latency and reliability expectation Add monitoring, rollback, and guardrails before anything goes live Optimise inference across CPU/GPU environments when it matters Integration into Real Systems Plug AI into data-heavy workflows without ...
... external AI retrieval components). Training foundation models from scratch (the emphasis is on building agentic apps and governed retrieval on enterprise data). AI assisted only delivery this role owns the AI behavior (grounding, evaluation, safety) end to end. Faster path from data to decision through conversational + agentic ...
... workflow automation. Collaborate with cross-functional teams to understand business problems and translate them into AI-powered solutions. Deploy, monitor, and maintain AI applications on cloud platforms such as AWS, Azure, or GCP. Stay updated with advancements in Generative AI, Agentic AI, LLMs, and AI application architecture. ...
... experience in the warehouse industry Should be a https://jobeax.com/link/7MP5OIWqBjLWowIE B.E + MBA Should have understanding of WMS system functionalities and details of the software operations Our AI-driven GreyMatterTM Fulfilment Operating System and RangerTM robot series are a combined solution that continuously prioritises ...
... skills in Python Experience with LLMs and GenAI frameworks (OpenAI, Hugging Face, Anthropic, Google etc.) Hands-on experience with: Embeddings & vector databases (FAISS, Pinecone, Qdrant, etc.) Experience with frameworks like LangChain, LlamaIndex, ADK, or similar Experience with API development and microservices Familiarity ...
... Integrity and is HI-TRUST and SOC2 certified, and a recipient of the 2025 CandE Award for Candidate Experience. Interested in shaping the future of healthcare with AI Explore opportunities at https://jobeax.com/link/l2ilzLBJSLK6rA75 and drive innovation with #YouToThePowerOfAI. The AI Engineer is an ML engineer responsible for ...