jobs in ByteDance

ByteDance

50-200 Employees · IT / Software

47 Followers

412 Job Vacancies

Be the first to hear about new jobs at ByteDance!

ByteDance Careers and Job Vacancies

Active Jobs Expired Jobs

Singapore

  • Bridge the gap between cutting-edge Generative AI model research and real-world industrial applications by solving the “last mile” of AI deployment, ensuring models are performant, grounded, and effectively integrated into complex agent solutions that deliver direct business impact.
  • Design and prototype AI-native agent solutions based on client requirements, including agent workflows, system integrations, and LoRA fine-tuned models tailored to specific industry use cases, such as automated credit risk summaries for fintech or style-specific image and video generation.
  • Deliver production-ready implementations by writing high-quality code to integrate AI model APIs into agent frameworks such as ADK and LangChain, together with supporting components such as RAG, vector databases, memory systems, cache management, and skills. ...
Posted
20 days ago

Singapore

  • Responsible for the iteration of the underlying architecture of the large model inference engine and end-to-end GPU performance optimization, through means such as operator fusion and compilation optimization, deeply optimizing GPU memory access, computing pipeline, and Stream asynchronous scheduling, eliminating inference computing bottlenecks, improving single-card inference throughput, and reducing inference latency.
  • Adapt to all series of GPU/NPU hardware architectures, refine the universality of the inference engine and hardware adaptability, and build a high-performance, low-loss underlying base for large model inference.
  • Lead the design, development, and optimization of distributed parallel solutions for large model inference scenarios, with a focus on implementing multi-dimensional parallel strategies such as tensor parallelism (TP), pipeline parallelism (PP), sequence parallelism, and MoE expert parallelism, to address core issues such as multi-card splitting and deployment of ultra-large models, high cross-card communication overhead, load imbalance, and low parallel efficiency. ...
Posted
20 days ago

Singapore

  • Responsible for the computational performance optimization of ByteDance's recommendation mid-platform models, conduct in-depth tuning for inference/training bottlenecks in business scenarios, and improve computing utilization.
  • Lead the design and development of high-performance kernel libraries, covering general-purpose and business-customized kernels, including general computation and communication parallelism, to ensure the ultimate performance of kernels.
  • Deeply cultivate model compilation optimization technology, and based on directions such as graph optimization, kernel fusion, computation scheduling, and code generation, build and improve the mid-platform model compilation system. ...
Posted
20 days ago

Singapore

  • Responsible for building and evolving Lagrange - the central console platform that supportsAl-related engineering workflows across the organization
  • Analyse user needs and develop software solutions, applying principles and techniques of computer science, engineering, and mathematical analysis.
  • Work on problems with real technical depth and long-term impact: platform architecture, engineering productivity, system scalability, operational intelligence, and the practical integration of Al into engineering and operations workflows. ...
Posted
20 days ago

Singapore

  • Assisting in the design and development of software and tools to enhance system automation, monitoring, and operational efficiency.
  • Participating in troubleshooting and resolving system issues, analyzing root causes, and implementing preventative measures.
  • Contributing to the enhancement of existing software by updating capabilities and supporting testing and validation procedures. ...
Posted
20 days ago

Singapore

  • Responsible for the overall architecture design and implementation of model inference services, building a high-performance, highly available, and scalable enterprise-level inference system for large-parameter, high-complexity AI models, overcoming various architectural challenges in the implementation of complex model inference, and supporting the efficient launch of models across all business scenarios.
  • Responsible for the R&D and optimization of the core modules of the inference framework, covering core capabilities such as inference engine scheduling, monitoring and alerting, canary release, etc., continuously iterating on the framework performance, and resolving performance bottlenecks, resource bottlenecks, and stability issues in high-concurrency and large-model inference scenarios.
  • Keep track of the latest inference technologies in the industry, conduct technology selection and innovation in combination with business scenarios, accumulate distributed high-concurrency service architecture solutions, and promote the upgrade and standardization of the team's technical system. ...
Posted
20 days ago

Singapore

  • Architect and implement solutions that enable internal and external customers to leverage ByteDance’s globally scaled content delivery network.
  • Build metrics, tools, automation, visualizations, and monitoring systems to support the operation and optimization of edge services.
  • Develop procedures and workflows that improve efficiency, build trust, and ensure compliance across operational processes. ...
Posted
20 days ago

Singapore

  • Define and execute go-to-market and sales strategies for GenAI and cloud solutions.
  • Drive new business acquisition while expanding strategic relationships within existing enterprise accounts.
  • Build and maintain executive-level relationships with C-suite stakeholders and key decision-makers across target industries. ...
Posted
20 days ago

Singapore

  • For training track, develop the Volcano Ark training platform, enabling both internal and external users to perform serverless post-training (e.g., Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL)) on the Ark platform.
  • Design elastic training solutions for complex multi-tenant workloads, supporting mixed-tenant training across multiple data centers and heterogeneous hardware while optimizing training throughput, resource utilization, and system stability.
  • Build next-generation reinforcement learning infrastructure to improve training efficiency, while designing intuitive and developer-friendly APIs for RL training workflows. ...
Posted
20 days ago

Singapore

  • Supporting the end-to-end recruitment lifecycle, from sourcing to screening and engaging top-tier technical talent.
  • Assisting with talent mapping and market research to provide actionable insights that support business objectives and growth.
  • Optimizing recruitment processes and supporting new initiatives aimed at enhancing efficiency and effectiveness. ...
Posted
20 days ago