Senior Machine Learning - AI & GPU Performance

🇳🇱 Netherlands - Remote
📊 Data🟣 Senior

Job description

Who are we?

From your everyday PowerPoint presentations to Hollywood movies, AI will transform the way we create and consume content.

Today, people want to watch and listen, not read — both at home and at work. If you’re reading this and nodding, check out our brand video.

Despite the clear preference for video, communication and knowledge sharing in the business environment are still dominated by text, largely because high-quality video production remains complex and challenging to scale—until now….

Meet Synthesia

We’re on a mission to make video easy for everyone. Born in an AI lab, our AI video communications platform simplifies the entire video production process, making it easy for everyone, regardless of skill level, to create, collaborate, and share high-quality videos. Whether it’s for delivering essential training to employees and customers or marketing products and services, Synthesia enables large organizations to communicate and share knowledge through video quickly and efficiently. We’re trusted by leading brands such as Heineken, Zoom, Xerox, McDonald’s and more. Read stories from happy customers and what 1,200+ people say on G2.

In February 2024, G2 named us as the fastest growing company in the world. Today, we’re at a $2.1bn valuation and we recently raised our Series D. This brings our total funding to over $330M from top-tier investors, including Accel, Nvidia, Kleiner Perkins, Google and top founders and operators including Stripe, Datadog, Miro, Webflow, and Facebook.

About the role

As a ML Performance Engineer you will join a team of 40+ Researchers and Engineers within the R&D Department working on cutting edge challenges in the Generative AI space, with a focus on creating highly realistic, emotional and life-like Synthetic humans through text-to-video. Within the team you’ll have the opportunity to work on the applied side of our research efforts and directly impact our solutions that are used worldwide by over 60,000 businesses.

This is an opportunity to work for a company that is impacting businesses at a rapid pace across the globe.

What will you be doing?

As a ML Performance Engineer in the AI & GPU Performance team you will contribute to the design and development of high performance solutions.You will own one or more projects for computationally optimizing large-scale model training and inference pipelines.By partnering with researchers and research teams you’ll identify high-impact initiatives and push the boundaries of model performance. You will work on re-implementing models in an efficient manner by using PyTorch and underlying technologies like CUDA/Triton, Torch compilation, etc.

This would include:

  • Evaluating, profiling and optimising compute resource usage (e.g., Hopper & Blackwell GPUs) for cost and time efficiency at training and inference times.
  • Developing customized efficient solutions for inference pipelines (CUDA/Triton kernels) as well as Introducing or enhancing tooling for achieving optimal computational performance (e.g. DL compilers, ONNX, TensorRT)
  • Driving the adoption of best practices for large-model training, including checkpointing, gradient accumulation, and memory optimisation among others.
  • Introducing or enhancing tooling for distributed training, performance monitoring, and logging (e.g., DeepSpeed, PyTorch Distributed).
  • Designing and implement techniques for model parallelism, data parallelism, and mixed-precision training
  • Keeping updated on the latest research in model compression (e.g., quantization, pruning) and advanced optimisation methods.

Who are you?

  • You are a ML engineer passionate about high performance computing.
  • You have a background in Computer Science / Engineering  and 3+ years of industry experience. (PhD preferred)
  • You have worked on optimising large models for over 2 years
  • You have experience developing CUDA/Triton kernels and optimizing models with DL compilers (torch.compile)
  • You have great coding skills in Python and C++ and you care about writing clean, and efficient code.
  • You have experience with optimising distributed systems and distributed tools like DDP, Deepspeed, Accelerate or similar
  • You have some experience in the video space (Diffusion models / GAN’s).
  • You are interested in doing research, trying new things and pushing the boundaries, going beyond what’s already known.

The good stuff…

  • Attractive compensation (salary + stock options + bonus)
  • Private Health Insurance in London
  • Hybrid work setting with an office in London
  • 25 days of annual leave + public holidays
  • Work in a great company culture with the option to join regular planning and socials at our hubs.
  • A generous referral scheme when you know people that are amazing for us
  • Strong opportunities for your career growth

You can see more about Who we are and How we work here: https://www.synthesia.io/careers

#LI-MD1

Share this job:
Please let Synthesia know you found this job on Remote First Jobs 🙏

Similar Remote Jobs

Benefits of using Remote First Jobs

Discover Hidden Jobs

Unique jobs you won't find on other job boards.

Advanced Filters

Filter by category, benefits, seniority, and more.

Priority Job Alerts

Get timely alerts for new job openings every day.

Manage Your Job Hunt

Save jobs you like and keep a simple list of your applications.

Search remote, work from home, 100% online jobs

We help you connect with top remote-first companies.

Search jobs

Hiring remote talent? Post a job

Frequently Asked Questions

What makes Remote First Jobs different from other job boards?

Unlike other job boards that only show jobs from companies that pay to post, we actively scan over 20,000 companies to find remote positions. This means you get access to thousands more jobs, including ones from companies that don't typically post on traditional job boards. Our platform is dedicated to fully remote positions, focusing on companies that have adopted remote work as their standard practice.

How often are new jobs added?

New jobs are constantly being added as our system checks company websites every day. We process thousands of jobs daily to ensure you have access to the most up-to-date remote job listings. Our algorithms scan over 20,000 different sources daily, adding jobs to the board the moment they appear.

Can I trust the job listings on Remote First Jobs?

Yes! We verify all job listings and companies to ensure they're legitimate. Our system automatically filters out spam, junk, and fake jobs to ensure you only see real remote opportunities.

Can I suggest companies to be added to your search?

Yes! We're always looking to expand our listings and appreciate suggestions from our community. If you know of companies offering remote positions that should be included in our search, please let us know. We actively work to increase our coverage of remote job opportunities.

How do I apply for jobs?

When you find a job you're interested in, simply click the 'Apply Now' button on the job listing. This will take you directly to the company's application page. We kindly ask you to mention that you found the position through Remote First Jobs when applying, as it helps us grow and improve our service 🙏

Apply