Skip to main content
ResumeKart
← Back to Jobs

ML Researcher — Agentic Reinforcement Learning

KAISHI PARTNERS PTE. LTD.•Singapore
Full-time3-7
S$10k - S$20k
per year
👁️ 0 views•📝 0 applications•Posted 9/20/2026•Expires 10/20/2026
Tailor Resume for This JobCheck ATS Score

Get alerts for roles like this

More ML Researcher — Agentic Reinforcement Learning roles in Singapore — straight to your inbox. No account needed.

Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.

Job Description

About the company Our client is an early-stage, venture-backed AI startup building a personal shopping agent that understands users’ preferences, helps them discover relevant products, and supports them in taking action.

The team brings together AI agents, personalisation, and commerce to build a service that becomes more useful through ongoing interactions and real customer outcomes. They are hiring in Singapore, with an opportunity for early team members to shape the research agenda and learning systems behind the product.

The opportunity You will investigate how an AI agent can make better decisions for a user over time. In commerce, an immediate action is an incomplete measure of success: a purchase may later be returned, preferences can change, and the most useful recommendation may be to buy nothing.

Working closely with engineers, you will help define learning objectives, build rigorous evaluations, and develop methods for improving agent behaviour as usable data becomes available. What you’ll do Investigate reward design and credit assignment for delayed outcomes such as satisfaction, returns, and repeat use.

Develop approaches to user modelling, memory, and adaptation as preferences and needs change. Explore policies for deciding when an agent should ask, recommend, act, or wait. Build reproducible experiments, meaningful baselines, and ablation studies.

Evaluate policy improvements critically, including uncertainty, misleading proxies, and unintended behaviours. Collaborate with engineers on the data and instrumentation needed for research.

Translate promising findings into behaviours that can be tested in the product, introducing more sophisticated learning methods when the evidence supports them. What you’ll bring Deep knowledge of reinforcement learning and hands-on research or implementation experience.

Strong Python and machine learning engineering skills. Experience designing experiments and evaluating results critically. The ability to translate an

Required Skills

Reinforcement LearningPythonMachine Learning EngineeringExperiment DesignCritical Evaluation of ResultsAI AgentsPersonalizationCommerce SystemsUser ModelingMemory MechanismsPolicy Decision AlgorithmsReproducible ExperimentsBaseline DevelopmentAblation StudiesUncertainty QuantificationData InstrumentationProduct Testing

Partner picks for ML Researcher — Agentic Reinforcement Learning in Singapore

Matched to the skills this page calls for and the candidate's location.

Partner
  • Partner course provider

    Structured programmes for software engineers, data science and DevOps.

    covers python
  • edXVerified partner
    Partner course provider

    Courses and programmes from universities and institutions worldwide.

    covers python
  • UdemyVerified partner
    Partner course provider

    A marketplace of instructor-created courses across technology, business and creative skills.

    covers python

Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns

The best-paying roles in your field. Every week. Free.

Join 10,000+ professionals getting job alerts and salary insights in their inbox

We respect your privacy. Unsubscribe anytime with one click.