About the Role
The ML Models team at Tenstorrent works at the intersection of AI research and high-performance hardware, bringing state-of-the-art machine learning models to life on custom AI accelerators. This role involves shaping the development and deployment of future AI models, from training large language models to optimizing inference performance at scale.
Responsibilities
- Lead research and development efforts focused on LLM training and inference optimization.
- Train, evaluate, and optimize state-of-the-art AI models on Tenstorrent hardware.
- Improve performance through techniques such as speculative decoding, quantization, kernel fusion, flash attention, and distributed training.
- Investigate system bottlenecks and collaborate cross-functionally to drive performance improvements.
- Translate cutting-edge ML research into scalable, production-ready solutions.
Requirements
- 4+ years of industry and/or academic experience in ML research and LLM development.
- PhD, published research, or experience with speculative decoding is highly valued.
Skills
- Strong Python and PyTorch experience developing and training deep learning models.
- Deep understanding of ML architectures, LLM training, and inference optimization.
- Hands-on experience training large-scale machine learning models.
Location
- Toronto, ON
Work Type
- Hybrid
Experience Level
- Various experience levels
Education Level
- PhD
Salary/Compensations
- $100k - $500k
Benefits
- Highly competitive compensation package and benefits
About the Company
- Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency.
- Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible.
- We value collaboration, curiosity, and a commitment to solving hard problems.
- Tenstorrent is building next-generation AI systems that push the boundaries of model training, inference, and large-scale distributed compute.
Equal Opportunity
- We are an equal opportunity employer.
- This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2). These requirements apply to persons located in the U.S. and all countries outside the U.S. As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency. If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.
