Engineer joining Cerebras to rapidly deploy AI models on proprietary CSX systems. Focused on debugging and enhancing model performance in a dynamic, innovative environment.
Responsibilities
Contribute to the end-to-end bring up of ML models on Cerebras CSX systems.
Work across the stack: model architecture translation, graph lowering, compiler optimizations, runtime integration, and performance tuning.
Debug performance and correctness issues spanning model code, compiler IRs, runtime behavior, and hardware utilization.
Propose and prototype improvements across tools, APIs, or automation flows to accelerate future bring ups.
Requirements
Bachelor’s, Master’s, or PhD in Computer Science, Engineering, or a related field
Comfort navigating the full AI toolchain: Python modeling code, compiler IRs, performance profiling, etc.
Strong debugging skills across performance, numerical accuracy, and runtime integration.
Experience with deep learning frameworks (e.g., PyTorch, TensorFlow) and familiarity with model internals (e.g., attention, MoE, diffusion)
Proficiency in C/C++ programming and experience with low-level optimization
Proven experience in compiler development, particularly with LLVM and/or MLIR
Strong background in optimization techniques, particularly those involving NP-hard problems.
Benefits
Competitive salary and benefits package
Opportunities for professional growth and career advancement
A dynamic and innovative work environment
The chance to work on cutting-edge technologies and make a significant impact on the future of AI.
Staff Generative AI Engineer developing production - grade AI applications for business value across the organization. Collaborating with engineers, scientists, and product teams to deliver scalable solutions.
Senior NLP/LLM Engineer exploring and analyzing LLM capabilities while collaborating with cross - functional teams at Social Discovery Group. Enhancing AI models' effectiveness and optimizing their performance.
Design and implement AI - powered applications for wealth management using GenAI and LLMs, collaborating with business teams to enhance client engagement and operational efficiency.
Senior Gen AI Developer (LLM) role in Toronto (Hybrid) for 12 months. Requires 6 - 8 yrs experience with LLMs, GenAI, Java, Spring, AI Agents, MLOps, CI/CD.
Lead AI solution design & development, integrate LLMs into Java/Spring apps for NLP & predictive analytics. Collaborate cross - functionally on AI strategy & mentor junior developers.
AI Developer for full lifecycle LLM solutions. Build production - ready AI systems, develop data pipelines, deploy models in cloud environments, and collaborate with stakeholders.
Senior Generative AI Engineer role requiring strong GenAI, Python & LLM experience. Onsite position in Mississauga, Ontario with contract/full - time options.
Machine Learning Engineer specializing in large language models for John Snow Labs. Working on AI model training and optimization for healthcare applications in a fully remote setting.
VP Analyst providing strategic guidance and insights on AI infrastructure for Gartner clients. Engaging senior IT leaders and facilitating enterprise AI adoption through innovative content.