jobs in SolveIT Consultant Sdn Bhd

居家办公 Freelance Agentic Evaluation Specialist 工作, 薪水 up to SGD 12, SolveIT Consultant Singapore 公司招聘中 - Ricebowl

Freelance Agentic Evaluation Specialist jobs

Freelance Agentic Evaluation Specialist

SGD10 - SGD12 每小时

Singapore, Singapore

Fresh Graduates
最后机会申请此工作。
Posted 2 hours ago • Closing 13 Aug 2027
最后机会申请此工作。
分享
保存

工作地点

  • St Andrew's Rd Singapore Singapore Singapore 17

职位描述

任职资格

Job role: Agentic Evaluation Specialist (Mandarin)

Engagement Type: Freelancing

Location: Singapore

Language: Mandarin / Malay/Tamil/Hindi/Bengali (native proficiency)

Priority languages – Mandarin and Malay

Hourly Commitment: 4-5 hours daily (flexible)

Hourly Compensation: $ 12

Project Duration: 1 month (extendable)

Educational qualification: Master’s degree / Bachelor's degree with relevant experience can apply

Roles & Responsibilities:

• Review agentic traces from AgentX, including tool calls, tool results, and interactions with its data environment, to understand the full sequence of steps the agent took, not just its final answer.

• Read two candidate responses for a given task and select the one that better completes the task, follows instructions, and behaves appropriately along the way.

• Verify each response against the trace: check that numbers, facts, and conclusions match what the tools actually returned, and catch hallucinations, misreadings, or skipped steps.

• Write short, specific justifications (typically 2–3 sentences) explaining why one response is better.

• Apply consistent judgment criteria (accuracy, helpfulness, safety, and adherence to task instructions) across many tasks.

• Evaluate language quality and cultural appropriateness for your Singapore language market.

• Flag unclear, incomplete, or ambiguous cases according to guidelines.

Requirements :

• Native proficiency in any of the above listed language with strong English comprehension for task instructions.

• Familiarity with agentic AI workflows: comfortable with tool calling, multi-step task execution, and how an AI agent interacts with external systems and data.

• Ability to read structured technical output such as JSON, tables, basic SQL queries, and API responses, enough to tell whether a response matches the data.

• Strong attention to detail and the ability to apply evaluation criteria consistently across many tasks.

• Clear written English for rating justifications.

• Based in Singapore, with cultural and market familiarity relevant to the language.

Preferred skills:

• Prior experience in AI data labelling, content rating, RLHF, or model evaluation, especially side-by-side (pairwise) comparison tasks.

• Background in linguistics, localisation, QA, data analysis, or related fields.

• Experience working with AI assistants or agent-based tools in a professional or personal capacity.

Selection process :

Shortlisted applicants will complete a short screening assessment made up of sample evaluation items. Each item includes a user request, an agentic trace, and two responses. You'll choose the better response and write a brief justificati

岗位职责

Job role: Agentic Evaluation Specialist (Mandarin)

Engagement Type: Freelancing

Location: Singapore

Language: Mandarin / Malay/Tamil/Hindi/Bengali (native proficiency)

Priority languages – Mandarin and Malay

Hourly Commitment: 4-5 hours daily (flexible)

Hourly Compensation: $ 12

Project Duration: 1 month(extendable)

Educational qualification: Master’s degree / Bachelor's degree with relevant experience can apply

Roles & Responsibilities:

• Review agentic traces from AgentX, including tool calls, tool results, and interactions with its data environment, to understand the full sequence of steps the agent took, not just its final answer.

• Read two candidate responses for a given task and select the one that better completes the task, follows instructions, and behaves appropriately along the way.

• Verify each response against the trace: check that numbers, facts, and conclusions match what the tools actually returned, and catch hallucinations, misreadings, or skipped steps.

• Write short, specific justifications (typically 2–3 sentences) explaining why one response is better.

• Apply consistent judgment criteria (accuracy, helpfulness, safety, and adherence to task instructions) across many tasks.

• Evaluate language quality and cultural appropriateness for your Singapore language market.

• Flag unclear, incomplete, or ambiguous cases according to guidelines.

Requirements :

• Native proficiency in any of the above listed language with strong English comprehension for task instructions.

• Familiarity with agentic AI workflows: comfortable with tool calling, multi-step task execution, and how an AI agent interacts with external systems and data.

• Ability to read structured technical output such as JSON, tables, basic SQL queries, and API responses, enough to tell whether a response matches the data.

• Strong attention to detail and the ability to apply evaluation criteria consistently across many tasks.

• Clear written English for rating justifications.

• Based in Singapore, with cultural and market familiarity relevant to the language.

Preferred skills:

• Prior experience in AI data labelling, content rating, RLHF, or model evaluation, especially side-by-side (pairwise) comparison tasks.

• Background in linguistics, localisation, QA, data analysis, or related fields.

• Experience working with AI assistants or agent-based tools in a professional or personal capacity.

Selection process :

Shortlisted applicants will complete a short screening assessment made up of sample evaluation items. Each item includes a user request, an agentic trace, and two responses. You'll choose the better response and write a brief justification.

好处

  • WORK FROM HOME
  • REMOTE
  • FREELANCE

重要安全守则

申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。

了解更多