ByteDance logo
ByteDance
Posted 162 days agoVerified live 11h ago

Research Scientist in Vision Foundation Model - Seed - Graduates - 2027 Start (BS/MS)

Brief overview

San Jose, CAIn-person
UndergradOr in progress
$213k–$388k/yrStated range
2,631 H-1B approvalsDept. of Labor
463 green cardsCertified filings
Coding abilityData structuresFundamental algorithm skillsProficiency in C/C++ or PythonComputer visionMultimodal learningMachine learningProblem-solvingCommunication skills

About the company

ByteDance logo
ByteDancebytedance.com

ByteDance is a technology company that develops content creation platforms and services.

Visa sponsorship history

4 years sponsoring, last filed FY2026H-1B dependent

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
2,631H-1B approved
99%approval rate
1,036new H-1B hires
463PERM certified
$204,340median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
2023605
2024997
2025933
202696
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
2023234
2024196
2025193
202680
Green Card (PERM) FilingsCertified green card filings: a long-term commitment to international hires.
2023113
2024140
2025157
202653
Top sponsored roles
Software EngineerProduct ManagerData ScientistBackend Software EngineerResearch Scientist
Sponsored employees from
ChinaIndiaCanadaTaiwanHong Kong

Job description

Summary

ByteDance is dedicated to pioneering new paths toward artificial general intelligence, and they are seeking a Research Scientist in Vision Foundation Model. The role involves developing and scaling vision foundation models, optimizing model architectures, and exploring real-world applications of vision models in multimodal systems.

Responsibilities

  • Develop and scale vision foundation models across image and video modalities
  • Design data pipelines, pre-training strategies, and post-training methods for vision tasks
  • Improve core capabilities such as perception, reasoning, and multimodal understanding
  • Optimize model architectures, training efficiency, and evaluation frameworks
  • Explore real-world applications of vision models in multimodal systems

Skills

  • Currently pursuing a Bachelor's or Master's degree in computer science, mathematics, engineering, or a related field, with an expected graduation date in 2027 and the ability to commit to an onboarding date by the end of 2027
  • Excellent coding ability, data structures, and fundamental algorithm skills, proficient in C/C++ or Python, etc
  • Demonstrated interest or project experience in relevant areas
  • Experience with computer vision, multimodal learning, or machine learning through internships is preferred
  • Strong problem-solving and communication skills

Qualifications

Must Haves

  • Currently pursuing a Bachelor's or Master's degree in computer science, mathematics, engineering, or a related field, with an expected graduation date in 2027 and the ability to commit to an onboarding date by the end of 2027
  • Excellent coding ability, data structures, and fundamental algorithm skills, proficient in C/C++ or Python, etc
  • Demonstrated interest or project experience in relevant areas

Nice to Haves

  • Experience with computer vision, multimodal learning, or machine learning through internships is preferred
  • Strong problem-solving and communication skills

Benefits

  • Medical, dental, and vision insurance
  • A 401(k) savings plan with company match
  • Paid parental leave
  • Short-term and long-term disability coverage
  • Life insurance
  • Wellbeing benefits
  • 10 paid holidays per year
  • 10 paid sick days per year
  • 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure)

More jobs like this