jobs in SolveIT Consultant Sdn Bhd

Work from Home Freelance AI Evaluation Assistant Jobs, Salary up to SGD 12 in SolveIT Consultant Singapore - Maukerja

Freelance AI Evaluation Assistant jobs

Freelance AI Evaluation Assistant

SGD10 - SGD12 Per Hour

Singapore, Singapore

Fresh Graduates
Be an early applicant!
Posted 8 hours ago • Closing 13 Aug 2027
Be an early applicant!
Share
Save

Working Location

  • St Andrew's Rd Singapore Singapore Singapore 17

Job Description

Requirements

Job role: Agentic Evaluation Specialist (Mandarin)

Engagement Type: Freelancing

Location: Singapore

Language: Mandarin / Malay/Tamil/Hindi/Bengali (native proficiency)

Priority languages – Mandarin and Malay

Hourly Commitment: 4-5 hours daily (flexible)

Hourly Compensation: $ 12

Project Duration: 1 month (extendable)

Educational qualification: Master’s degree / Bachelor's degree with relevant experience can apply

Roles & Responsibilities:

• Review agentic traces from AgentX, including tool calls, tool results, and interactions with its data environment, to understand the full sequence of steps the agent took, not just its final answer.

• Read two candidate responses for a given task and select the one that better completes the task, follows instructions, and behaves appropriately along the way.

• Verify each response against the trace: check that numbers, facts, and conclusions match what the tools actually returned, and catch hallucinations, misreadings, or skipped steps.

• Write short, specific justifications (typically 2–3 sentences) explaining why one response is better.

• Apply consistent judgment criteria (accuracy, helpfulness, safety, and adherence to task instructions) across many tasks.

• Evaluate language quality and cultural appropriateness for your Singapore language market.

• Flag unclear, incomplete, or ambiguous cases according to guidelines.

Requirements :

• Native proficiency in any of the above listed language with strong English comprehension for task instructions.

• Familiarity with agentic AI workflows: comfortable with tool calling, multi-step task execution, and how an AI agent interacts with external systems and data.

• Ability to read structured technical output such as JSON, tables, basic SQL queries, and API responses, enough to tell whether a response matches the data.

• Strong attention to detail and the ability to apply evaluation criteria consistently across many tasks.

• Clear written English for rating justifications.

• Based in Singapore, with cultural and market familiarity relevant to the language.

Preferred skills:

• Prior experience in AI data labelling, content rating, RLHF, or model evaluation, especially side-by-side (pairwise) comparison tasks.

• Background in linguistics, localisation, QA, data analysis, or related fields.

• Experience working with AI assistants or agent-based tools in a professional or personal capacity.

Selection process :

Shortlisted applicants will complete a short screening assessment made up of sample evaluation items. Each item includes a user request, an agentic trace, and two responses. You'll choose the better response and write a brief justificati

Responsibilities

Job role: Agentic Evaluation Specialist (Mandarin)

Engagement Type: Freelancing

Location: Singapore

Language: Mandarin / Malay/Tamil/Hindi/Bengali (native proficiency)

Priority languages – Mandarin and Malay

Hourly Commitment: 4-5 hours daily (flexible)

Hourly Compensation: $ 12

Project Duration: 1 month(extendable)

Educational qualification: Master’s degree / Bachelor's degree with relevant experience can apply

Roles & Responsibilities:

• Review agentic traces from AgentX, including tool calls, tool results, and interactions with its data environment, to understand the full sequence of steps the agent took, not just its final answer.

• Read two candidate responses for a given task and select the one that better completes the task, follows instructions, and behaves appropriately along the way.

• Verify each response against the trace: check that numbers, facts, and conclusions match what the tools actually returned, and catch hallucinations, misreadings, or skipped steps.

• Write short, specific justifications (typically 2–3 sentences) explaining why one response is better.

• Apply consistent judgment criteria (accuracy, helpfulness, safety, and adherence to task instructions) across many tasks.

• Evaluate language quality and cultural appropriateness for your Singapore language market.

• Flag unclear, incomplete, or ambiguous cases according to guidelines.

Requirements :

• Native proficiency in any of the above listed language with strong English comprehension for task instructions.

• Familiarity with agentic AI workflows: comfortable with tool calling, multi-step task execution, and how an AI agent interacts with external systems and data.

• Ability to read structured technical output such as JSON, tables, basic SQL queries, and API responses, enough to tell whether a response matches the data.

• Strong attention to detail and the ability to apply evaluation criteria consistently across many tasks.

• Clear written English for rating justifications.

• Based in Singapore, with cultural and market familiarity relevant to the language.

Preferred skills:

• Prior experience in AI data labelling, content rating, RLHF, or model evaluation, especially side-by-side (pairwise) comparison tasks.

• Background in linguistics, localisation, QA, data analysis, or related fields.

• Experience working with AI assistants or agent-based tools in a professional or personal capacity.

Selection process :

Shortlisted applicants will complete a short screening assessment made up of sample evaluation items. Each item includes a user request, an agentic trace, and two responses. You'll choose the better response and write a brief justification.

Benefits

  • WORK FROM HOME
  • REMOTE
  • FREELANCE

Important Information

Never provide your bank or credit card details when applying for jobs. Do not transfer any money or complete unrelated online surveys. If you see something suspicious, Report this Job ad.

Learn More