Backend Engineer Lead - ARK Large Model Platform (Singapore)

Singapore

ByteDance

ByteDance is a technology company operating a range of content platforms that inform, educate, entertain and inspire people across languages, cultures and geographies.

View all jobs at ByteDance

Apply now Apply later

Responsibilities

ByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.

Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.

Why Join Us
Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible.
Together, we inspire creativity and enrich life - a mission we aim towards achieving every day.
To us, every challenge, no matter how ambiguous, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always.
At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve.
Join us.

About the Team
The Applied Machine Learning (AML) - Enterprise team provides machine learning platform products on VolcanoEngine with cloud native resource scheduling system which intelligently orchestrates different tasks and jobs with minimised costs of every experiment and maximised resource utilisation, rich modelling tools including customised machine learning tasks and web IDE, and multi-framework high performance model inference services.

In 2021, through VolcanoEngine, we released this machine learning infrastructure to the public, to provide more enterprises with reduced costs of computation power, lower barriers to machine learning engineering and deeper developments in AI capabilities.

Responsibilities
Responsible for Ark Large Model Platform development on Volcano Engine, researching systematic solutions on large model solution implementations and applications in various industries, striving to reduce the IT cost of large model applications, meeting the users' ever-growing demand for intelligent interaction and improving the lifestyle and communications of users in the future world.

- Research, design, and develop computer and network software or specialised utility programs.
- Analyse user needs and develop software solutions, applying principles and techniques of computer science, engineering, and mathematical analysis.
- Update software, enhances existing software capabilities, and develops and direct software testing and validation procedures.
- Work with computer hardware engineers to integrate hardware and software systems and develop specifications and performance requirements.

Qualifications

Minimum Qualifications
- B. Sc or higher degree in Computer Science or related fields from accredited and reputable institutions.
- Familiar with developments and operations of distributed systems under Linux platform.
- Proficient with at least 2 or more programming languages such as Golang / Python / C / C++ / Java / Scala / Javascript. ACM ICPC / Codeforces winners are preferred
- Excellent in technical design and coding skills. Able to balance technical perspectives with product sense, hardware performance & stability and team cooperation.
- Experienced or interested in at least one of the following topics:
1. Machine Learning Application: had experience in developments or implementations in data, model training, inference, application in various machine learning domains such as LLM / CV / NLP / Speech / Recommendation / Risk Control, etc.
2. LLM Application: dataset construction (conversations, RLHF, etc.), high performance finetuning (LoRA / P-Tuning / RLHG), model inference and deployment, applications (prompt engineering, retrieval augmentation, LangChain), new model exploration (LLama / Falcon / miniGPT4)
3. Cloud Computing: kubernetes application development (such as Operator), micro-service & service mesh, flow control, cloud storage, exploration of technology commercialization, Terraform and etc.
- At least 3 years of relevant experience.

ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

Apply now Apply later

* Salary range is an estimate based on our AI, ML, Data Science Salary Index 💰

Job stats:  0  0  0

Tags: Computer Science Distributed Systems Engineering Golang Java JavaScript Kubernetes LangChain Linux LLaMA LLMs LoRA Machine Learning ML infrastructure Model inference Model training NLP Prompt engineering Python Research RLHF Scala Terraform Testing

Perks/benefits: Career development

Region: Asia/Pacific
Country: Singapore

More jobs like this