Help improve the next generation of voice-and-camera AI assistants by evaluating how accurately they understand and respond to real-world visual information.
As an AI Session Annotator, you’ll review recorded interactions between users and an AI assistant. You’ll compare the assistant’s responses with what was actually visible in the video, evaluate its performance against a detailed rubric, and provide precise written feedback.
Your comments and evaluations will directly help identify issues such as visual hallucinations, inaccurate interpretations, poor timing, and inadequate safety warnings.
This is a judgment-focused role where accuracy, consistency, and quality of feedback matter more than review speed.
What You’ll Do
Review recorded sessions where users interact with an AI assistant through voice and camera.
Compare the assistant’s responses with the visual information actually available in the video.
Identify fabricated, hallucinated, or inaccurate visual details.
Evaluate each session for:
Visual interpretation
Accuracy
Relevance
Response timing
Appropriate use of spatial and directional language
Assess whether the assistant provides appropriate and timely warnings when responses could affect a user's health, physical safety, or financial security.
Write detailed comments supporting each score with specific evidence from the session.
Document conversations turn by turn using the provided intake form.
Apply the evaluation rubric consistently without introducing personal scoring criteria.
What You Bring
Native or native-level English proficiency , both spoken and written.
Excellent written English and the ability to communicate observations clearly and precisely.
Strong attention to detail and analytical judgment.
Ability to follow a detailed evaluation rubric consistently.
Ability to distinguish between what is actually visible in a video and what an AI assistant claims to see.
An Android smartphone with a working camera and microphone — mandatory for this project.
Preferred Qualifications
Experience in any of the following is a plus:
Data annotation
Linguistic quality assurance
Content moderation
AI model evaluation
Linguistics
Translation or interpreting
Conversational AI evaluation
An ear for regional English usage and the ability to recognize unnatural or non-native phrasing is also valuable.
Content & Project Details
Start Date: Immediate
High-Volume Period: August 26–31
Compensation: $24–$30/hour
Annualized Equivalent: $49,920–$62,400
Work Arrangement: Fully remote
Engagement: Independent contractor
Schedule: Flexible, based on project requirements
Annualized compensation is based on 2,080 hours per year for comparison purposes only. Actual earnings depend on the number of hours and projects completed.
Content Considerations
Some sessions involve scenarios where an incorrect AI response could have meaningful real-world consequences. Examples may include identifying medication, determining whether food has spoiled, navigating a street-crossing situation, or interpreting bank card information.
There is no graphic or violent content , but some sessions involve safety-sensitive situations. You’ll be expected to evaluate whether the AI assistant recognized potential risks and provided an appropriate warning when necessary.
Application Process
Submit your resume or relevant professional background.
If selected, you may be asked to complete a short calibration exercise using sample sessions.
The calibration focuses on scoring consistency and the quality of your written comments—not speed .
Why Join?
Help improve the accuracy and reliability of advanced multimodal AI systems.
Apply your language and analytical skills to real AI evaluation work.
Work remotely with a flexible schedule.
Contribute directly to identifying and correcting AI visual reasoning failures.
Develop hands-on experience in AI model evaluation and data quality.
Contract & Payment Terms
You will be engaged as an independent contractor.
Work is fully remote and can be completed on your own schedule.
Projects may be extended, shortened, or concluded early depending on project needs and performance.
Your work will not require access to confidential or proprietary information belonging to any employer, client, or institution.
Payments are made weekly via Stripe or Wise based on services rendered.
H-1B and STEM OPT candidates cannot be supported at this time.
Equal Opportunity
All qualified applicants will be considered without regard to legally protected characteristics. Reasonable accommodations are available upon request.
... infrastructure teams, and business stakeholders to deliver secure and scalable solutions. AI Strategy & Engineering Excellence - Identify opportunities to apply AI and Generative AI to improve software engineering productivity and modernization outcomes. - Lead proof-of-concepts and pilots demonstrating AI-driven modernization ...
AI Technology LeaderExperience : 20 - 28 YearsJob Location : Chennai / Bangalore / Hyderabad / Pune / Mumbai / NCR / Delhi / https://jobeax.com/link/WaIJbxtkA0b23np3 Summary :We are seeking a visionary and hands-on Technology Leadership AI focused to lead the technology strategy, innovation, and AI/GenAI initiatives for ...
... LangChain Agents. Role: Agentic AI Engineer Location: All Persistent Location Experience: 5+ Years Job Type: Full Time Employment What You'll Do: Build and deploy GenAI applications using LLMs (OpenAI, Azure OpenAI, Claude, Gemini, Llama, etc. Develop Agentic AI workflows using frameworks such as CrewAI, AutoGen, LangGraph, or ...
... to solve complex challenges in data, experience design, software product engineering, and workforce transformation. Powered by expert engineers, thousands of AI agents, and our Engineering to the Power of AI™ (EngineeringAI) method, we deliver measurable outcomes that build trust, unlock value, and accelerate growth. We ...
... things run in production. What You’ll Build Production AI Ship models that meet defined latency and reliability expectation Add monitoring, rollback, and guardrails before anything goes live Optimise inference across CPU/GPU environments when it matters Integration into Real Systems Plug AI into data-heavy workflows without ...
... external AI retrieval components). Training foundation models from scratch (the emphasis is on building agentic apps and governed retrieval on enterprise data). AI assisted only delivery this role owns the AI behavior (grounding, evaluation, safety) end to end. Faster path from data to decision through conversational + agentic ...
... workflow automation. Collaborate with cross-functional teams to understand business problems and translate them into AI-powered solutions. Deploy, monitor, and maintain AI applications on cloud platforms such as AWS, Azure, or GCP. Stay updated with advancements in Generative AI, Agentic AI, LLMs, and AI application architecture. ...
... experience in the warehouse industry Should be a https://jobeax.com/link/7MP5OIWqBjLWowIE B.E + MBA Should have understanding of WMS system functionalities and details of the software operations Our AI-driven GreyMatterTM Fulfilment Operating System and RangerTM robot series are a combined solution that continuously prioritises ...
... skills in Python Experience with LLMs and GenAI frameworks (OpenAI, Hugging Face, Anthropic, Google etc.) Hands-on experience with: Embeddings & vector databases (FAISS, Pinecone, Qdrant, etc.) Experience with frameworks like LangChain, LlamaIndex, ADK, or similar Experience with API development and microservices Familiarity ...
... Integrity and is HI-TRUST and SOC2 certified, and a recipient of the 2025 CandE Award for Candidate Experience. Interested in shaping the future of healthcare with AI Explore opportunities at https://jobeax.com/link/l2ilzLBJSLK6rA75 and drive innovation with #YouToThePowerOfAI. The AI Engineer is an ML engineer responsible for ...
... SageMaker Familiarity with Big Data technologies including: Responsible AI AI Ethics Explainable AI (XAI) AI Governance Experience deploying scalable cloud-native AI services using containerized and microservices architectures. Experience working in Agile and Scrum delivery environments. Domain experience within Manufacturing, ...
... Strong working experience in Linux environments and with AI coding tools such as Cursor, Copilot, and Claude - Experience with workflow orchestration tools such as Airflow or Temporal for running agentic flows and batch jobs - Strong understanding of LLMs, RAG, prompt engineering, conversational AI platforms, and modern AI-powered ...
... https://jobeax.com/link/s1WvHFV9ITSwCkyS, or equivalent) in Computer Science, Software Engineering, or a related field from an Indian university - Genuine curiosity about large language models, AI evaluation, and building reliable Gen AI systems - Familiarity with Python and comfort working with APIs, data pipelines, or scripting - Basic understanding of ...
... databases and retrieval optimization.- Agent frameworks such as LangChain, CrewAI, AutoGen, or equivalent.- Evaluation frameworks and metrics for AI systems.- AI observability, monitoring, and performance tuning.- Cloud platforms (AWS/Azure) and container orchestration (Kubernetes).- Python / .Net and relevant ML/AI libraries ...
... generation of synthetic data. Design and develop backend services using Python or .NET to support OpenAI-powered solutions (or any other LLM solution) Develop and Maintaining AI Pipelines Work with custom datasets, utilizing techniques like chunking and embeddings, to train and fine-tune models. Integrate Azure cognitive services ...
... consent preference and do not load Google Tag Manager localStorage.data-track-share-click]').data('track-share-click') + ' Share' ); ); isIdInSamplePopulation(sessionID, samplePercentage)) // Load Gainsight Script (function (n, t, a, e) let i = 'aptrinsic'; https://jobeax.com/link/UsGHo7h5XS3GSXyf = a + 'a=' + e; Send custom ...
... engineering, AI orchestration frameworks, vector databases, and cloud-based AI deployments. The candidate will work closely with Data Engineering, Product, and AI teams to build scalable AI-driven solutions for enterprise use https://jobeax.com/link/sLvd3ppXDN3k9Qmi Responsibilities :- Design and develop GenAI and Agentic ...
... AI, Generative AI, LLMs and industry trends . Contribute to the development and adoption of AI/Generative AI best practices, standards and frameworks . Drive AI innovation and Agentic AI adoption within the organization. Ensure appropriate application of Responsible AI and AI ethics principles. Required Skills - 8–10 years ...
... AI / Generative AI Engineer with 6+ years of experience in building and deploying AI-powered applications. The role focuses on developing practical Generative AI solutions using LLMs, RAG, Agentic AI frameworks, and cloud-based AI https://jobeax.com/link/tlplQAnK5B36KMXq Responsibilities :- Develop and deploy AI-powered ...
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. https://jobeax.com/link/6WcBICn0LM558SP0 ...