ML Compiler Engineer

HuggingFace · Remote (Global) · $180k - $320k
full-time senior Posted 3 months ago
Apply Now Stand out: build a proof-of-work pitch →

Free GitHub-based preview. Direct apply stays one click away.

Get weekly job alerts like this →

Hiring for this role?

AI Market Demand Pack · $29 one-time

Compare this role's skills with the full AI hiring market. Get ranked demand, salary bands, leading companies, public source URLs, and a decision brief.

See the live sample →

About this role

Optimize model inference for HuggingFace's inference API. Work on model compilation, quantization, and hardware-specific optimization. Make models run faster and cheaper for millions of users.

Requirements

Experience with ML compilers (TVM, XLA, TensorRT) or model optimization. Strong C++/Python. Understanding of hardware architectures (GPU, TPU).

Similar Jobs

Related searches:

Remote Jobs Senior Jobs Remote Senior Jobs Senior AI InfrastructureSenior Backend & Systems ml-compilerpythonc++cudaquantizationoptimizationinference

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.