About this role
About the Team The Applied Machine Learning (AML) - Ark team provides machine learning platform products on VolcanoEngine with cloud native resource scheduling system which intelligently orchestrates different tasks and jobs with minimised costs of every experiment and maximised resource utilisation, rich modelling tools including customised machine learning tasks and web IDE, and multi-framework high performance model inference services. In 2021, through VolcanoEngine, we released this machine learning infrastructure to the public, to provide more enterprises with reduced costs of computation power, lower barriers to machine learning engineering and deeper developments in AI capabilities. Responsibilities -Design and develop core pipeline components for the Ark MaaS platform (Model as a Service), supporting API capabilities such as text conversations, multimodal understanding, and multimodal generation. -Participate deeply in the evolution of cloud-native architectures, including Service Mesh, load balancing (LB), and intelligent routing. Design and implement highly available solutions such as full-link canary releases, traffic degradation, circuit breaking, and rate limiting. -Optimize system performance for large-model multimodal inference scenarios involving long connections, high throughput, streaming output, and low-latency requirements. -Ensure system stability under large-scale model invocation scenarios and resolve architectural bottlenecks caused by sudden traffic surges. Minimum Qualifications - B. Sc or higher degree in Computer Science or related fields from accredited and reputable institutions with at least 3 years of relevant experience. - Familiar with developments and operations of distributed systems under Linux platform. - Proficient with at least 2 or more programming languages such as Golang / Python / C / C++ / Java / Scala / Javascript. ACM ICPC / Codeforces winners are preferred - Excellent in technical design and coding skills. Able to balance technical perspectives with product sense, hardware performance & stability and team cooperation. Preferred Qualifications -Understanding of Agent-related concepts and technologies such as Function Calling and MCP; hands-on experience building Agents is a plus. -Familiarity with cloud-native and Service Mesh technologies such as Kubernetes, Docker, Istio, and Envoy. -Strong understanding of multimodal data processing and storage, including image and video/audio data. -Interest in or practical experience with LLM inference pipelines and large-model engineering systems.
Frequently asked questions
What does a Backend Engineer, ARK Large Model Platform (Singapore) at ByteDance do?
About the Team The Applied Machine Learning (AML) - Ark team provides machine learning platform products on VolcanoEngine with cloud native resource scheduling system which intelligently orchestrates different tasks and jobs with minimised costs of every experiment and maximised resource utilisation…
How much does a Backend Engineer, ARK Large Model Platform (Singapore) at ByteDance pay?
The employer did not list a salary for this role. Most similar Singapore roles publish their band on the job page.
Is this Backend Engineer, ARK Large Model Platform (Singapore) role remote, hybrid, or on-site?
The listing is based in Singapore. Check the posting for remote or hybrid options.
How do I apply for this Backend Engineer, ARK Large Model Platform (Singapore) role?
You can apply directly on ByteDance's careers page. ApplyLah can tailor your résumé and cover letter to this exact role in seconds first.