Jobs

Courses

Resources

Companies

Placements

Community

Home >

Jobs >

Research Scientist Graduate (Foundation Model, Vision and Language) - 2025 Start (PhD)

California, United States (On-site)

Research Scientist Graduate (Foundation Model, Vision and Language) - 2025 Start (PhD)

3 Months ago • All levels • Research & Development • $250,000 PA - $430,000 PA

Job Summary

Job Description

ByteDance's Doubao (Seed) Team is seeking a Research Scientist Graduate (Foundation Model, Vision and Language) to join their team in 2025. This role will focus on cutting-edge research and development in computer vision and natural language processing, particularly in the areas of multi-modality, vision and language, and more. The candidate will enhance multimodal understanding and reasoning through various stages of the development process, including data acquisition, model evaluation, pre-training, SFT, reward modeling, and reinforcement learning. The candidate will also synthesize large-scale, high-quality multi-modal data and investigate robust evaluation methodologies to assess model performance.

Must have:

Research and engineering experience in computer vision and natural language processing.
Experience in multi-modal understanding, vision and language, such as multimodal pre-training, visual instruction tuning, alignment learning.
Work with very large-scale datasets and build very large-scale datasets to scale up foundation models.
Experience with language models and apply them in various downstream tasks.
Highly competent in algorithms and programming; Strong coding skills in Python and popular deep learning frameworks.
Ability to work independently; Strong communication skills.

Good to have:

Publications in top-tier venues such as CVPR, ECCV, ICCV, NeurIPS, ICLR, ICML, EMNLP, ACL, NAACL.
Impactful open-source projects on GitHub and a demonstrated engineering ability to quickly solve new challenges.

Perks:

100% premium coverage for employee medical insurance
75% premium coverage for dependents
Health Savings Account(HSA) with a company match
Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life and AD&D insurance plans
Flexible Spending Account(FSA) Options like Health Care, Limited Purpose and Dependent Care
10 paid holidays per year
17 days of Paid Personal Time Off (PPTO)
10 paid sick days per year
12 weeks of paid Parental leave
8 weeks of paid Supplemental Disability
Mental and emotional health benefits through our EAP and Lyra
401K company match
Gym and cellphone service reimbursements

8 skills required

8 skills required for this role

Add these skills to join the top 1% applicants for this job

algorithms

github

reinforcement-learning

computer-vision

deep-learning

python

foundation

talent-acquisition

Job Details

Responsibilities

Established in 2023, the ByteDance Doubao (Seed) Team is dedicated to building industry-leading AI foundation models. We aim to do world-leading research and foster both technological and social progress. With a long-term vision and a strong commitment to the AI field, the Team conducts research in a range of areas including natural language processing (NLP), computer vision (CV), and speech recognition and generation. It has labs and researcher roles in China, Singapore, and the US. Leveraging substantial data and computing resources and through continued investment in these domains, our team has built a proprietary general-purpose model with multimodal capabilities. In the Chinese market, Doubao models power over 50 ByteDance apps and business lines, including Doubao, Coze, and Dreamina, and was launched to external enterprise clients through Volcano Engine. The Doubao app is the most used AIGC app in China. Why Join Us Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible. Together, we inspire creativity and enrich life - a mission we aim towards achieving every day. To us, every challenge, no matter how ambiguous, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always. At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve. Join us. About the Team Welcome to the Doubao-Vision team, where we spearhead multi-modality foundation models on visual understanding and visual generation. Our mission is to solve the visual intelligence problem for AI. We conduct cutting-edge research on areas like vision and language, large vision models, and generative foundation models. The team is a mix of experienced research scientists and engineers, aiming to advance the research boundaries in foundation models and apply our technologies to our rich application scenarios, whereas a feedback loop is created to help further improve our foundation technologies. Join us in shaping the future of AI technologies and revolutionizing our product experience for global users. We are looking for talented individuals to join our team in 2025. As a graduate, you will get unparalleled opportunities for you to kickstart your career, pursue bold ideas and explore limitless growth opportunities. Co-create a future driven by your inspiration with Bytedance. Successful candidates must be able to commit to an onboarding date by end of year 2025. We will prioritize candidates who are able to commit to these start dates. Please state your availability and graduation date clearly in your resume. Applications will be reviewed on a rolling basis. We encourage you to apply early. Candidates can apply for a maximum of TWO positions and will be considered for jobs in the order you applied for. The application limit is applicable to Bytedance and its affiliates' jobs globally. Responsibilities - Conduct cutting-edge research and development in computer vision and natural language processing, especially in the areas of multi-modality, vision and language, etc. - Enhance multimodal understanding and reasoning (images and videos etc), throughout the entire development process, encompassing data acquisition, model evaluation, pre-training, SFT, reward modeling, and reinforcement learning, to bolster overall performance. - Synthesize large-scale, high-quality multi-modal data through methods such as rewriting, augmentation, and generation to improve the abilities of foundation models in various stages (pretraining, SFT, RLHF). - Investigate and implement robust evaluation methodologies to assess model performance at various stages (ranging from covering diverse multimodal skills to improving user preference alignment), unravel the underlying mechanisms and sources of their abilities, and utilize this understanding to drive model improvements.

Qualifications

Minimum Qualifications: - Research and engineering experience in one or more areas of computer vision and natural language processing, including but not limited to: Experience in multi-modal understanding, vision and language, such as multimodal pre-training, visual instruction tuning, alignment learning, and other related topics. - Work with very large-scale datasets, and build very large-scale datasets to scale up foundation models. - Experience with language models and apply them in various downstream tasks. - Highly competent in algorithms and programming; Strong coding skills in Python and popular deep learning frameworks. - Work and collaborate well with team members. - Ability to work independently; Strong communication skills. Preferred Qualifications: - Candidates with publications in top-tier venues such as CVPR, ECCV, ICCV, NeurIPS, ICLR, ICML, EMNLP, ACL, NAACL, etc - Candidates with impactful open-source projects on GitHub and a demonstrated engineering ability to quickly solve new challenges. ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too. ByteDance Inc. is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://shorturl.at/cdpT2

Job Information

【For Pay Transparency】Compensation Description (annually)

The base salary range for this position in the selected city is $250000 - $430000 annually.​
Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units.​
Our company benefits are designed to convey company culture and values, to create an efficient and inspiring work environment, and to support our employees to give their best in both work and life. We offer the following benefits to eligible employees: ​
We cover 100% premium coverage for employee medical insurance, approximately 75% premium coverage for dependents and offer a Health Savings Account(HSA) with a company match. As well as Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life and AD&D insurance plans. In addition to Flexible Spending Account(FSA) Options like Health Care, Limited Purpose and Dependent Care. ​
Our time off and leave plans are: 10 paid holidays per year plus 17 days of Paid Personal Time Off (PPTO) (prorated upon hire and increased by tenure) and 10 paid sick days per year as well as 12 weeks of paid Parental leave and 8 weeks of paid Supplemental Disability. ​
We also provide generous benefits like mental and emotional health benefits through our EAP and Lyra. A 401K company match, gym and cellphone service reimbursements. The Company reserves the right to modify or change these benefits programs at any time, with or without notice.​
For Los Angeles County (unincorporated) Candidates:​
Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment:​
1. Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues;​
2. Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; and​
3. Exercising sound judgment.​

Similar Jobs

Customer Onboarding Manager - Indonesia

USE Insider

Jakarta, Jakarta, Indonesia (On-Site)

• 3 Months ago

Senior Software Engineer, Site Reliability Engineering, Google Cloud

Google

Warsaw, Masovian Voivodeship, Poland (On-Site)

• 3 Months ago

Software Engineer, Systems ML - SW/HW Co-design

Lead Data Scientist

Ascent Health

Karnataka, India (On-Site)

• 3 Months ago

C++ TEAM LEAD (MARKETS EXPANSION)

Equivalent Jobs

(Remote)

• 2 Months ago

Physical Design Engineer

UST

Karnataka, India (On-Site)

• 4 Months ago

CPLD/ FPGA Design Engineer

Intel Corporation

Guadalajara, Jalisco, Mexico (Hybrid)

• 1 Month ago

Design Verification Leader (MIPI / USB3 / Ethernet)

SiliconAuto India

Bengaluru, Karnataka, India (On-Site)

• 4 Months ago

Software Engineer III, Machine Learning, Pixel Camera

Google

New Taipei, New Taipei City, Taiwan (On-Site)

• 1 Month ago

Senior Formal Verification Engineer, Google Cloud

Google

Tel Aviv-Yafo, Tel Aviv District, Israel (On-Site)

• 1 Month ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Software Engineer, Computer Vision (Technical Leadership)

Software Engineer III, Infrastructure, Google Cloud

Google

Bengaluru, Karnataka, India (On-Site)

• 3 Months ago

Staff Engineer, CPU Microarchitecture

Samsung Semiconductor

San Jose, California, United States (Hybrid)

• 2 Days ago

Principle Software Developer - Montreal

Snowed In Studios

Quebec, Canada (Remote)

• 3 Months ago

Principal Engineer

Druva

Hyderabad, Telangana, India (On-Site)

• 4 Months ago

Java Backend Developer (Remote OK)

Red Point Labs

Argentina (Remote)

• 8 Months ago

Senior Data Scientist

Lifelancer

Bengaluru, Karnataka, India (On-Site)

• 3 Months ago

Senior Unity Developer

Easy Brain

Limassol, Limassol, Cyprus (Hybrid)

• 4 Months ago

Manager, Software Engineering(Scala)

The Walt Disney Company

San Francisco, California, United States (On-Site)

• 2 Months ago

Graphics Programmer

Larian Studios

Guildford, England, United Kingdom (On-Site)

• 3 Months ago

Get notifed when new similar jobs are uploaded

Jobs in San Jose, California, United States

CONTRACT - Events Marketing Specialist (LatAm)

Nintendo

Redmond, Washington, United States (Hybrid)

• 2 Months ago

Senior Software Engineer

Azra Games

Austin, Texas, United States (Hybrid)

• 2 Months ago

Experienced Software Engineer - Traffic Platform

ByteDance

San Jose, California, United States (On-Site)

• 3 Months ago

STEP Intern

Patel greene

Sarasota, Florida, United States (On-Site)

• 3 Months ago

Lead Engine Systems Engineer

Tencent

Irvine, California, United States (On-Site)

• 4 Months ago

Operations Associate

DraftKings

Pueblo, Colorado, United States (On-Site)

• 1 Week ago

Staff Engineer, BI Reporting

Nagarro

California, United States (On-Site)

• 3 Months ago

Sr. Product Manager

My Fitness Pal

United States (Remote)

• 2 Months ago

Principal Engineer - Project Technical Lead

Hypixel Studios

Seattle, Washington, United States (Remote)

• 3 Months ago

Helpdesk Support Technician

Intrepid Studios, Inc

San Diego, California, United States (On-Site)

• 5 Months ago

Get notifed when new similar jobs are uploaded

Research & Development Jobs

Mechanical Engineering Intern

Regent Craft

North Kingstown, Rhode Island, United States (On-Site)

• 4 Months ago

Software Engineer (Leadership) - Machine Learning

Data Parallel Accelerator Performance Intern

Rivos

Hsinchu, Hsinchu City, Taiwan (Hybrid)

• 3 Months ago

Senior Software Engineer, TPU, Google Cloud Platform

Google

Taipei City, Taiwan (On-Site)

• 1 Month ago

Publishing Tech PM

Krafton

Seoul, South Korea (On-Site)

• 4 Weeks ago

Machine Learning Engineer - Machine Learning Infrastructure

ByteDance

Seattle, Washington, United States (On-Site)

• 3 Months ago

Antenna Design Engineer- Pico- San Jose

ByteDance

San Jose, California, United States (On-Site)

• 4 Weeks ago

Design Verification Engineer

Intel Corporation

Penang, Malaysia (Hybrid)

• 1 Month ago

Embedded C IRC238457

GlobalLogic

Chennai, Tamil Nadu, India (Hybrid)

• 4 Months ago

Executive Recruiting Partner

Riot Games

Los Angeles, California, United States (On-Site)

• 1 Month ago

Get notifed when new similar jobs are uploaded

About The Company

ByteDance

1167 Active Jobs

Where imagination meets innovation, delivering limitless gaming experiences.

Get notified when new jobs are added by ByteDance

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

A global community of game builders. Helping people upskill and land jobs in the best gaming studios.

Company

Key Links

hello@outscal.com

Made in INDIA 💛💙

Research Scientist Graduate (Foundation Model, Vision and Language) - 2025 Start (PhD)

Job Summary

Job Description

8 skills required

8 skills required for this role

Job Details

Similar Jobs

Customer Onboarding Manager - Indonesia

Senior Software Engineer, Site Reliability Engineering, Google Cloud

Software Engineer, Systems ML - SW/HW Co-design

Lead Data Scientist

C++ TEAM LEAD (MARKETS EXPANSION)

Physical Design Engineer

CPLD/ FPGA Design Engineer

Design Verification Leader (MIPI / USB3 / Ethernet)

Software Engineer III, Machine Learning, Pixel Camera

Senior Formal Verification Engineer, Google Cloud

Similar Skill Jobs

Software Engineer, Computer Vision (Technical Leadership)

Software Engineer III, Infrastructure, Google Cloud

Staff Engineer, CPU Microarchitecture

Principle Software Developer - Montreal

Principal Engineer

Java Backend Developer (Remote OK)

Senior Data Scientist

Senior Unity Developer

Manager, Software Engineering(Scala)

Graphics Programmer

Jobs in San Jose, California, United States

CONTRACT - Events Marketing Specialist (LatAm)

Senior Software Engineer

Experienced Software Engineer - Traffic Platform

STEP Intern

Lead Engine Systems Engineer

Operations Associate

Staff Engineer, BI Reporting

Sr. Product Manager

Principal Engineer - Project Technical Lead

Helpdesk Support Technician

Research & Development Jobs

Mechanical Engineering Intern

Software Engineer (Leadership) - Machine Learning

Data Parallel Accelerator Performance Intern

Senior Software Engineer, TPU, Google Cloud Platform

Publishing Tech PM

Machine Learning Engineer - Machine Learning Infrastructure

Antenna Design Engineer- Pico- San Jose

Design Verification Engineer

Embedded C IRC238457

Executive Recruiting Partner

About The Company

Accounts Payable Analyst

Senior XR Strategy Expert

Hardware Engineering Lab Manager - Pico

Network Engineer, Optical Long-Haul and Submarine

Tech Lead Manager - Global E-Commerce Logistics

Tech Lead, Camera Algorithms Engineer

Tech Lead Manager - Global E-Commerce Logistics

Tech Lead - Global E-Commerce Logistics

Backend Software Engineer - Global E-Commerce Logistics

Backend Software Engineer - Global E-Commerce Logistics

Level Up Your Career in Game Development!