76 Research Jobs - October 2026 - Urgent Hiring

Showing 76 jobs results for "research"
Never miss any updates for Research jobs

Singapore

  • Design of Reward Models: In the RL process, designing an effective reward model is crucial. It must accurately reflect the effectiveness of the reasoning process and guide the model to iteratively improve its reasoning ability. This involves not only setting appropriate evaluation criteria across different tasks, but also ensuring the reward model to adapt dynamically during training to match the evolving model performance.
  • Stability of the Training Process: In the absence of high-quality SFT data, ensuring stable training in RL becomes a major challenge. RL often involves extensive exploration and trial-and-error, which may lead to unstable training or even performance degradation. Developing robust training strategies is essential to ensure the reliability and effectiveness of the training process for models.
  • Expanding from Mathematics and Code Tasks to Natural Language Tasks: Current RL reasoning methods are primarily applied to mathematics and code tasks, where CoT data is more abundant. However, natural language tasks are more open and complex. Expanding from successful RL strategies to natural language processing tasks requires in-depth research and innovation in both data design and RL methodology to enable cross-task general reasoning capabilities. ...
Posted
a month ago

Singapore

  • Design of Reward Models: In the RL process, designing an effective reward model is crucial. It must accurately reflect the effectiveness of the reasoning process and guide the model to iteratively improve its reasoning ability. This involves not only setting appropriate evaluation criteria across different tasks, but also ensuring the reward model to adapt dynamically during training to match the evolving model performance.
  • Stability of the Training Process: In the absence of high-quality SFT data, ensuring stable training in RL becomes a major challenge. RL often involves extensive exploration and trial-and-error, which may lead to unstable training or even performance degradation. Developing robust training strategies is essential to ensure the reliability and effectiveness of the training process for models.
  • Expanding from Mathematics and Code Tasks to Natural Language Tasks: Current RL reasoning methods are primarily applied to mathematics and code tasks, where CoT data is more abundant. However, natural language tasks are more open and complex. Expanding from successful RL strategies to natural language processing tasks requires in-depth research and innovation in both data design and RL methodology to enable cross-task general reasoning capabilities. ...
Posted
8 days ago

Singapore

  • Co-create with the team: Creativity is at the core of TikTok. Whether it's building the product or shaping the team, we strive to spark imagination and deliver impact — for ourselves, the platform, the communities we serve, and society as a whole.
  • Grow through challenges: At TikTok, you’ll tackle highly challenging projects that drive industry breakthroughs and global influences. With hundreds of millions of users, there’s always an opportunity to introduce new technologies and ideas that shape better user experiences. Every challenge is a chance to learn, innovate, and grow.
  • Work style and culture: We value practical problem-solving and the pursuit of excellence in everything, encouraging everyone to work with the mindset of “Always Day 1.” Our company culture is diverse and inclusive, where everyone collaborates as equal and operates in an agile and flexible environment that empowers creativity. ...
Posted
a month ago

Singapore

  • Co-create with the team: Creativity is at the core of TikTok. Whether it's building the product or shaping the team, we strive to spark imagination and deliver impact — for ourselves, the platform, the communities we serve, and society as a whole.
  • Grow through challenges: At TikTok, you’ll tackle highly challenging projects that drive industry breakthroughs and global influences. With hundreds of millions of users, there’s always an opportunity to introduce new technologies and ideas that shape better user experiences. Every challenge is a chance to learn, innovate, and grow.
  • Work style and culture: We value practical problem-solving and the pursuit of excellence in everything, encouraging everyone to work with the mindset of “Always Day 1.” Our company culture is diverse and inclusive, where everyone collaborates as equal and operates in an agile and flexible environment that empowers creativity. ...
Posted
9 days ago

Singapore

  • Research on foundation models and frontier generative technologies: Explore the capability boundaries of large language models (LLMs) and multimodal large models (MLLMs), especially the evolution of foundation capabilities for agentic scenarios. Conduct in-depth research and propose innovative post-training paradigms, tackling challenges in complex reasoning such as CoT reasoning and long-horizon planning, tool use, reinforcement learning alignment such as innovative applications of PPO, GRPO, and other new RLHF algorithms for agent behavior alignment, reward model design, and AI safety. Build model capabilities for autonomous decision-making and reflection from the underlying architecture. Explore and break through the underlying architectures of frontier visual generation algorithms such as diffusion and flow matching, building a strong perception and environment simulation foundation for multimodal agents.
  • Frontier breakthroughs in large model agents: Lead underlying algorithmic innovation for Agentic AI. Focus on academic challenges such as agentic reinforcement learning, complex logical reasoning and planning, long-term memory mechanisms, and knowledge injection. Explore the theoretical foundations of emergent multi-agent collaboration and optimal interaction strategies in highly complex, open-ended environments.
  • Forward-looking technical evaluation and theoretical construction: Define and build the evaluation system for next-generation AIGC and agent systems. Identify fundamental problems with core academic value in real-world, large-scale commercial scenarios such as advertising and e-commerce content ecosystems, and design new algorithmic frameworks to address long-tail generalization, interpretability, and multimodal alignment challenges. ...
Posted
a month ago

Singapore

  • Research on foundation models and frontier generative technologies: Explore the capability boundaries of large language models (LLMs) and multimodal large models (MLLMs), especially the evolution of foundation capabilities for agentic scenarios. Conduct in-depth research and propose innovative post-training paradigms, tackling challenges in complex reasoning such as CoT reasoning and long-horizon planning, tool use, reinforcement learning alignment such as innovative applications of PPO, GRPO, and other new RLHF algorithms for agent behavior alignment, reward model design, and AI safety. Build model capabilities for autonomous decision-making and reflection from the underlying architecture. Explore and break through the underlying architectures of frontier visual generation algorithms such as diffusion and flow matching, building a strong perception and environment simulation foundation for multimodal agents.
  • Frontier breakthroughs in large model agents: Lead underlying algorithmic innovation for Agentic AI. Focus on academic challenges such as agentic reinforcement learning, complex logical reasoning and planning, long-term memory mechanisms, and knowledge injection. Explore the theoretical foundations of emergent multi-agent collaboration and optimal interaction strategies in highly complex, open-ended environments.
  • Forward-looking technical evaluation and theoretical construction: Define and build the evaluation system for next-generation AIGC and agent systems. Identify fundamental problems with core academic value in real-world, large-scale commercial scenarios such as advertising and e-commerce content ecosystems, and design new algorithmic frameworks to address long-tail generalization, interpretability, and multimodal alignment challenges. ...
Posted
8 days ago

Singapore

  • Conduct cutting-edge research in machine learning, computer vision or natural language processing.
  • Support and incubate products with machine learning, especially state-of-the-art AIGC techniques such as GPT, Stable Diffusion, etc.
  • Final year Ph.D or recent Ph.D graduates in Computer Science, engineering or quantitative field ...
Posted
a month ago

Singapore

  • Conduct cutting-edge research in machine learning, computer vision or natural language processing.
  • Support and incubate products with machine learning, especially state-of-the-art AIGC techniques such as GPT, Stable Diffusion, etc.
  • Final year Ph.D or recent Ph.D graduates in Computer Science, engineering or quantitative field ...
Posted
9 days ago

Singapore

  • Assist in completing various user research and analysis tasks, evaluate research requests, scope research projects, manage research suppliers and execute projects, and write professional research reports;
  • Quickly collect and synthesize information, have strong user empathy and insight, and can flexibly scan market information through various means, quickly providing valuable information and suggestions for the business;
  • Have good communication and project management skills, able to think independently to solve problems, and willing to actively participate in the full life-cycle of research projects. ...
Posted
a month ago

Singapore

  • Assist in completing various user research and analysis tasks, evaluate research requests, scope research projects, manage research suppliers and execute projects, and write professional research reports;
  • Quickly collect and synthesize information, have strong user empathy and insight, and can flexibly scan market information through various means, quickly providing valuable information and suggestions for the business;
  • Have good communication and project management skills, able to think independently to solve problems, and willing to actively participate in the full life-cycle of research projects. ...
Posted
8 days ago

Singapore

  • Recommendation Large Models: Address challenges such as gradient convergence and representation drift in ultra-long behavioral sequences, enabling the system to achieve true ""logical reasoning"" capabilities.
  • Unified Multimodal Semantic Space: Explore alignment across video, image-text content, and user intent, constructing a fully multimodal semantic space that goes beyond text.
  • Agentic Rec:Develop recommendation agents with capabilities such as self-reflection, tool invocation, and long-horizon planning, driving a transformation of recommendation and content distribution experiences. ...
Posted
a month ago

Singapore

  • Drive core technology development for large language model code direction, continuously optimizing code comprehension, reasoning, and generation capabilities.
  • Focus on improving code comprehension, reasoning, and generation capabilities in real-world production codebases, improving TikTok service code performance and privacy compliance.
  • Explore Code Agent capabilities suitable for actual business production environments, and improve TikTok R&D efficiency. ...
Posted
a month ago

Singapore

  • Recommendation Large Models: Address challenges such as gradient convergence and representation drift in ultra-long behavioral sequences, enabling the system to achieve true ""logical reasoning"" capabilities.
  • Unified Multimodal Semantic Space: Explore alignment across video, image-text content, and user intent, constructing a fully multimodal semantic space that goes beyond text.
  • Agentic Rec:Develop recommendation agents with capabilities such as self-reflection, tool invocation, and long-horizon planning, driving a transformation of recommendation and content distribution experiences. ...
Posted
9 days ago

Singapore

  • Drive core technology development for large language model code direction, continuously optimizing code comprehension, reasoning, and generation capabilities.
  • Focus on improving code comprehension, reasoning, and generation capabilities in real-world production codebases, improving TikTok service code performance and privacy compliance.
  • Explore Code Agent capabilities suitable for actual business production environments, and improve TikTok R&D efficiency. ...
Posted
9 days ago

Singapore

  • a PhD or Master’s degree in Computer Science, Machine Learning, NLP/CV, Systems, or related fields (or equivalent research experience).
  • Strong programming and debugging skills in Python and/or C/C++; solid data structures & algorithms foundation.
  • Hands-on experience in at least one: ...
Posted
8 days ago

Singapore

  • Native training and inference architecture redesign for LLMs
  • Etreme performance optimization and AI infrastructure innovation
  • End-to-End generative paradigm innovation ...
Posted
a month ago

Singapore

  • Native training and inference architecture redesign for LLMs
  • Etreme performance optimization and AI infrastructure innovation
  • End-to-End generative paradigm innovation ...
Posted
9 days ago

Singapore

  • a PhD or Master’s degree in Computer Science, Machine Learning, NLP/CV, Systems, or related fields (or equivalent research experience).
  • Strong programming and debugging skills in Python and/or C/C++; solid data structures & algorithms foundation.
  • Hands-on experience in at least one: ...
Posted
a month ago

Singapore

  • Insufficient understanding of underground industry variants, AIGC, and other adversarial content by general large models
  • Challenges in long-context comprehension, information extraction, and instruction adherence.
  • Integrating fragmented risk control knowledge into agent-usable skills ...
Posted
a month ago

Singapore

  • Insufficient understanding of underground industry variants, AIGC, and other adversarial content by general large models
  • Challenges in long-context comprehension, information extraction, and instruction adherence.
  • Integrating fragmented risk control knowledge into agent-usable skills ...
Posted
9 days ago

Singapore

  • Native training and inference architecture redesign for LLMs
  • Etreme performance optimization and AI infrastructure innovation
  • End-to-End generative paradigm innovation ...
Posted
a month ago

Singapore

  • Native training and inference architecture redesign for LLMs
  • Etreme performance optimization and AI infrastructure innovation
  • End-to-End generative paradigm innovation ...
Posted
8 days ago

Singapore

  • Education & Foundation: Ph.D. in Computer Science, AI, Mathematics, or a related field, with a strong foundation in data structures, algorithms, and mathematical modeling.
  • AI/ML Expertise: Solid understanding and research experience in Deep Learning, NLP, CV, Reinforcement Learning, Generative Models, or Multimodal Learning.
  • Coding & Engineering: Proficient in major programming languages and machine learning frameworks (e.g., PyTorch, TensorFlow), combined with excellent problem-solving, self-learning, and teamwork skills. ...
Posted
a month ago

Singapore

  • Education & Foundation: Ph.D. in Computer Science, AI, Mathematics, or a related field, with a strong foundation in data structures, algorithms, and mathematical modeling.
  • AI/ML Expertise: Solid understanding and research experience in Deep Learning, NLP, CV, Reinforcement Learning, Generative Models, or Multimodal Learning.
  • Coding & Engineering: Proficient in major programming languages and machine learning frameworks (e.g., PyTorch, TensorFlow), combined with excellent problem-solving, self-learning, and teamwork skills. ...
Posted
9 days ago

Singapore

  • Insufficient understanding of underground industry variants, AIGC, and other adversarial content by general large models
  • Challenges in long-context comprehension, information extraction, and instruction adherence.
  • Integrating fragmented risk control knowledge into agent-usable skills ...
Posted
a month ago

Singapore

  • Insufficient understanding of underground industry variants, AIGC, and other adversarial content by general large models
  • Challenges in long-context comprehension, information extraction, and instruction adherence.
  • Integrating fragmented risk control knowledge into agent-usable skills ...
Posted
8 days ago

Singapore

  • Currently pursuing PhD in Computer Science, AI, Mathematics, or a related technical discipline, with a strong foundation in data structures, algorithms, and mathematical modeling.
  • AI/ML Expertise: Solid understanding and research experience in Deep Learning, NLP, CV, Reinforcement Learning, Generative Models, or Multimodal Learning.
  • Priority will be given to candidates with publications in international AI/CS conferences or journals (e.g., NeurIPS, ICML, ICLR, CVPR, ACL, KDD, SIGIR, WWW) or top rankings in recognized algorithmic competitions. ...
Posted
a month ago

Singapore

  • Currently pursuing PhD in Computer Science, AI, Mathematics, or a related technical discipline, with a strong foundation in data structures, algorithms, and mathematical modeling.
  • AI/ML Expertise: Solid understanding and research experience in Deep Learning, NLP, CV, Reinforcement Learning, Generative Models, or Multimodal Learning.
  • Priority will be given to candidates with publications in international AI/CS conferences or journals (e.g., NeurIPS, ICML, ICLR, CVPR, ACL, KDD, SIGIR, WWW) or top rankings in recognized algorithmic competitions. ...
Posted
8 days ago

Singapore

  • Explore the integration of foundation models with ranking algorithms to improve personalized ranking accuracy and user experience.
  • Explore end-to-end generative search models based on multimodal pre-training.
  • Explore LLM-based agent technology to improve user satisfaction under complex ambiguous queries and multi-turn search scenarios. ...
Posted
a month ago