jobs
Full Stack LLM Engineer
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv
- Company
- Cerebras
- Location
- Toronto, CAN
- Status
- Open
- Posted
- 2025-07-18T20:38:09.402+00:00
Cerebras is hiring a full stack LLM engineer for its Inference Core Model Bringup team in Toronto, working hybrid. The role involves bringing open-source and customer models such as LLaMA and Qwen onto Cerebras CSX systems, covering model architecture translation, graph lowering, compiler optimization, runtime integration and performance tuning. Applicants should hold a degree in computer science or a related field and have C/C++ proficiency, deep learning framework experience, familiarity with model internals and proven LLVM or MLIR compiler work. The posting states no salary range.
Original job posting