Чтобы адаптировать резюме под вакансию или составить сопроводительное письмо, загрузите резюме
описание
This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology. Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2).
Tenstorrent develops high-performance AI platforms that combine software models, compilers, platforms, networking, and semiconductors. Its AI Models team brings advanced LLMs and vision models to custom AI hardware and accelerators.
задачи
Bring up, run, and debug modern ML models such as transformers using PyTorch or TensorFlow;
Analyze model behavior and performance and identify bottlenecks across the stack;
Improve the efficiency, correctness, and scalability of model execution in real systems;
Work closely with compiler, kernel, and hardware teams to drive performance and system-level improvements;
Help translate state-of-the-art model architectures into production-grade, high-performance deployments.
требования
Strong experience building and working with ML models in PyTorch or TensorFlow;
Strong understanding of modern ML model architectures, such as transformers;
Solid software engineering fundamentals with strong debugging and problem-solving skills;
Comfort working in a fast-moving, research-meets-engineering environment;
Nice to have: experience with profiling or performance tuning, familiarity with quantization, FlashAttention, kernel fusion, memory hierarchies, C++, CUDA, or systems programming.
условия
Highly competitive compensation package and benefits;
Employment is contingent upon eligibility to access U.S. export-controlled technology; the offer may depend on citizenship, permanent residency status, or obtaining prior license approval from the U.S. Commerce Department or applicable federal agency.