Skyfall AISkyfall AI

Member of Technical Staff

Toronto, Ontario, CanadaOn-site or hybridFULL_TIME7w ago

About Skyfall We're building something the enterprise world doesn't have yet: an Enterprise World Model. While the rest of the field races to scale LLMs as the primary path to intelligence, we see a more principled opportunity in a hybrid architecture by combining the semantic richness of LLMs with the structural reaso

PythonPytorch

About Skyfall

We're building something the enterprise world doesn't have yet: an Enterprise World Model.

While the rest of the field races to scale LLMs as the primary path to intelligence, we see a more principled opportunity in a hybrid architecture by combining the semantic richness of LLMs with the structural reasoning capabilities of world models. As such, we are developing Enterprise World Models (EWMs) that can understand, simulate, and act across complex enterprise systems. This isn't our first time building something that matters. Skyfall was founded by the original Maluuba team, pioneers of the deep learning revolution, who worked closely with leaders such as Yoshua Bengio and Richard Sutton, before Maluuba’s $160M acquisition by Microsoft, where it became Microsoft’s AI research center in Canada. We've built companies before. We've made research breakthroughs before. Now we're doing it again, this time with a clear target: the next $5B enterprise AI company.We're VC-backed, moving fast, and assembling a world-class team to match our ambitions.

We're not building another AI wrapper instead we are building the next foundation layer for enterprise AI.

Job Overview:

Skyfall AI is seeking a Research Engineer (Member of Technical Staff) to join our team to work on our journey towards an autonomous enterprise. We are looking for strong engineers who have hands-on experience in the following areas: GPU and inference-time acceleration and optimization, LLM post-training, parallelization, prompt tuning, and bandit-based tree search.

Requirements

Job Responsibilities:

  • Develop tools and infrastructure for running models in interactive environments
  • Optimize GPU/inference performance and parallelization
  • Optimize code refinement & tree search performance and parallelization
  • Develop methods for code-based world modeling
  • Contribute to publications and open source projects

Minimum Qualifications:

  • Master's degree in Computer Science or other relevant technical field
  • Proven programming experience in Python
  • Experience with cloud GPU environments such as AWS, Lambda, …
  • Hands-on experience hosting open source LLMs
  • Hands-on experience with Pytorch

Preferred Qualifications

  • Published papers at a top conference (i.e NeurIPS, ACL, ICML, ICLR etc)

**Location -**Toronto (hybrid)

Benefits

✨ The perks:

  • Competitive salary with 80% health and dental coverage
  • 15 vacation days and 4 paid sick days annually
  • Regular socials and annual team retreats
  • Hybrid setup with a team of sharp, curious people who genuinely love what they build

Frequently asked

Is this Member of Technical Staff role remote?

This role is based in Toronto, Ontario, Canada.

Related