San Jose, California, United StatesFull TimeEntry-level254k–480k USDPosted Today
About the Team
Established in 2023, the ByteDance Seed team is dedicated to pioneering new paths toward artificial general intelligence. We aspire to advance the frontier of intelligence to drive progress for both technology and society.
With a long-term vision for the AI sector, the Seed team's research spans MLLM, GenMedia, AI for Science, and Robotics. We maintain a global presence with laboratories and career opportunities across China, Singapore, and the United States. To date, we have launched industry-leading general foundation models and cutting-edge multimodal capabilities. Our technology powers over 50 application scenarios — including Doubao, Jimeng, TRAE, Dola and Dreamnia — and serves enterprise customers through Volcano Engine and BytePlus. Third-party data shows that the Doubao App ranks first in user volume in the Chinese market, while Doubao foundation models lead the industry in average daily token consumption.
The Seed Multimodal Interaction and World Model team is dedicated to developing models that boast human-level multimodal understanding and interaction capabilities. The team is working also aspires to advance the exploration and development of multimodal assistant products
Responsibilities
- Drive research and engineering to advance models that enhance understanding of multimodal data and enlarge reasoning capabilities.
- Explore research ideas that optimize both the model's performance and efficiency.
- Establish scaling laws, design and conduct systematic ablations that result in transferrable conclusions.
The base salary range for this position in the selected city is $254400 - $480000 annually.
Established in 2023, the ByteDance Seed team is dedicated to pioneering new paths toward artificial general intelligence. We aspire to advance the frontier of intelligence to drive progress for both technology and society.
With a long-term vision for the AI sector, the Seed team's research spans MLLM, GenMedia, AI for Science, and Robotics. We maintain a global presence with laboratories and career opportunities across China, Singapore, and the United States. To date, we have launched industry-leading general foundation models and cutting-edge multimodal capabilities. Our technology powers over 50 application scenarios — including Doubao, Jimeng, TRAE, Dola and Dreamnia — and serves enterprise customers through Volcano Engine and BytePlus. Third-party data shows that the Doubao App ranks first in user volume in the Chinese market, while Doubao foundation models lead the industry in average daily token consumption.
The Seed Multimodal Interaction and World Model team is dedicated to developing models that boast human-level multimodal understanding and interaction capabilities. The team is working also aspires to advance the exploration and development of multimodal assistant products
Responsibilities
- Drive research and engineering to advance models that enhance understanding of multimodal data and enlarge reasoning capabilities.
- Explore research ideas that optimize both the model's performance and efficiency.
- Establish scaling laws, design and conduct systematic ablations that result in transferrable conclusions.
The base salary range for this position in the selected city is $254400 - $480000 annually.
