Skip to content
← Back to job listings

Principal Interactive Vision Model Researcher

Tencent America LLC · US-California-Los Angeles, United States

LeadQuick applyfull-timeabout 10 hours ago

About The Role

LIGHTSPEED STUDIOS is made up of passionate players who advance the art & science of game development through great stories, great gameplay, and advanced technology. We are focused on bringing next generation experiences to gamers who want to enjoy them anywhere, anytime, across multiple genres and devices.

About the Hiring Team

Lightspeed Tech Center is a R&D department under Lightspeed Studios which develop PUBG Mobile and other high-quality games. Our Tech Center leads the research, exploration, and discovery of innovative technologies and provides technical services for all games during all phases of life cycle, including engine, audio, QA, AI, next generation game, technical cooperation, etc.

What the Role Entails

  • 1.Design and iterate on the next-generation video generation foundation architecture.
  • 2.Explore core technologies for long video generation, including long-context attention, KV Cache compression, and memory mechanisms.
  • 3.Research high-compression-ratio video tokenizers and advance unified modeling capabilities for multi-resolution and multi-frame-rate videos.
  • 4.Lead video pre-training at the large token scale, defining data mixture and curriculum learning strategies.
  • 5.Explore Scaling Laws, define scientific scale-up paths, and continuously improve key model capabilities.
  • 6.Track industry state-of-the-art , lead comparative experiments, drive technical roadmap decisions, and publish academic papers.

Who We Look For

  • 1.Ph.D. in AI-related fields with first-author papers at top-tier conferences.
  • 2.Proficient in the principles and engineering implementation of diffusion models and autoregressive generation.
  • 3.Experience training video/image generation models from scratch; highly proficient in PyTorch and large-scale distributed training.
  • 4.Deep understanding of the design trade-offs in Video VAE/Tokenizers, with 3+ years of relevant research experience.
  • 5.Publications related to video generation or diffusion models.
  • 6.Experience in core industry product R&D or hands-on experience with the latest technologies is preferred.
  • 7.Experience leading end-to-end video foundation model projects is preferred.

Location State(s)

US-California-Los Angeles

The expected base pay range for this position in the location(s) listed above is $158,300.00 to $326,000.00 per year. Actual pay may vary depending on job-related knowledge, skills, and experience.
Employees hired for this position may be eligible for a sign on payment, relocation package, and restricted stock units, which will be evaluated on a case-by-case basis.
Subject to the terms and conditions of the plans in effect, hired applicants are also eligible for medical, dental, vision, life and disability benefits, and participation in the Company’s 401(k) plan. The Employee is also eligible for up to 15 to 25 days of vacation per year (depending on the employee’s tenure), up to 13 days of holidays throughout the calendar year, and up to 10 days of paid sick leave per year.
Your benefits may be adjusted to reflect your location, employment status, duration of employment with the company, and position level. Benefits may also be pro-rated for those who start working during the calendar year.

Equal Employment Opportunity at Tencent

As an equal opportunity employer, we firmly believe that diverse voices fuel our innovation and allow us to better serve our users and the community. We foster an environment where every employee of Tencent feels supported and inspired to achieve individual and common goals.

This listing was posted by a verified recruiter at Tencent America LLC. Report this listing