Inference Compiler Engineer Job at Cerebras Systems (Dubai)

Inference Compiler Engineer

Cerebras Systems

Location:
United Arab Emirates , Dubai

Category:
IT - Software Development

Contract Type:
Not provided

Salary:

Not provided

Save Job

Apply Position

Job Description:

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. Would you like to participate in creating the fastest Generative Models inference in the world? Join the Cerebras Inference Team to participate in development of unique Software and Hardware combination that sports best inference characteristics in the market while running largest models available. Cerebras wafer scale inference platform allows running Generative models with unprecedented speed thanks to unique hardware architecture that provides fastest access to local memory, ultra-fast interconnect and huge amount of available compute. You will be part of the team that works with latest open and closed generative AI models to optimize for the Cerebras inference platform. Your responsibilities will include working on model representation, optimization and compilation stack to produce the best results on Cerebras current and future platforms.

Job Responsibility:

Analysis of new models from generative AI field and understanding of impacts on compilation stack
Implementation of compiler and frontend features to support new models, improve inference characteristics and Cerebras user experience
Collaboration with other teams throughout feature implementation
Research on new methods for model optimization to improve Cerebras inference

Requirements:

Degree in Engineering, Computer Science, or equivalent in experience and evidence of exceptional ability
Strong experience working with Python and C++ languages
Experience working with PyTorch and HuggingFace Transformers library
Knowledge and experience working with Large Language Models (understanding Transformer architecture variations, generation cycle, etc.)

Nice to have:

Knowledge of MLIR based compilation stack is a plus

What we offer:

Build a breakthrough AI platform beyond the constraints of the GPU
Publish and open source their cutting-edge AI research
Work on one of the fastest AI supercomputers in the world
Enjoy job stability with startup vitality
Our simple, non-corporate work culture that respects individual beliefs

Additional Information:

Job Posted:
February 17, 2026

Cerebras Systems - All Job Offers

Job Link Share:

Inference Compiler Engineer

Cerebras Systems

Location:
United Arab Emirates , Dubai

Category:
IT - Software Development

Contract Type:
Not provided

Salary:

Job Description:

Job Responsibility:

Requirements:

Nice to have:

Additional Information:

Job Posted:
February 17, 2026

Looking for more opportunities? Search for other job offers that match your skills and interests.

Similar Jobs for Inference Compiler Engineer

Python / PyTorch Developer — Frontend Inference Compiler

LLM Inference Performance & Evals Engineer

Performance Engineer - Inference

Software Engineer, Systems ML - Compilers / Backend

Research Engineer, Scaling

AI Research Engineer, Scaling

Senior Research Engineer

Engineering Manager, GPU Kernel

Inference Compiler Engineer

Cerebras Systems

Location:United Arab Emirates , Dubai

Category:IT - Software Development

Contract Type:Not provided

Salary:

Job Description:

Job Responsibility:

Requirements:

Nice to have:

Additional Information:

Job Posted:February 17, 2026

Looking for more opportunities? Search for other job offers that match your skills and interests.

Similar Jobs for Inference Compiler Engineer

Python / PyTorch Developer — Frontend Inference Compiler

LLM Inference Performance & Evals Engineer

Performance Engineer - Inference

Software Engineer, Systems ML - Compilers / Backend

Research Engineer, Scaling

AI Research Engineer, Scaling

Senior Research Engineer

Engineering Manager, GPU Kernel

Location:
United Arab Emirates , Dubai

Category:
IT - Software Development

Contract Type:
Not provided

Job Posted:
February 17, 2026