Skip to content
AI.info

jobs

Full Stack LLM Engineer

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference serv

Company
Cerebras
Location
Toronto, CAN
Status
Open
Posted
2025-07-18T20:38:09.402+00:00

Cerebras is hiring a full stack LLM engineer for its Inference Core Model Bringup team in Toronto, working hybrid. The role involves bringing open-source and customer models such as LLaMA and Qwen onto Cerebras CSX systems, covering model architecture translation, graph lowering, compiler optimization, runtime integration and performance tuning. Applicants should hold a degree in computer science or a related field and have C/C++ proficiency, deep learning framework experience, familiarity with model internals and proven LLVM or MLIR compiler work. The posting states no salary range.

Original job posting