aijobs.net

ML Systems Engineer — Inference Acceleration

Paris Offices

EUR 80K-150K (estimate) Mid-level Full Time

Apply Save
Found 1d ago
Tasks
Perks/Benefits
Skills/Tech-stack

Benchmarking | C++ | CUDA | Communication and Computation Overlap | Compiler optimization | Computer Architecture | Continuous batching | Distributed Training | Expert parallelism | GPU Programming | Graph Execution | HIP | High Performance | High-Performance Computing | KV cache | MLIR | Machine Learning | Memory Management | Operator fusion | Paged Attention | Parallelism | Performance Computing | Pipeline parallelism | Prefill and Decode | Prefill and Decode Scheduling | Profiling | Python | ROCm | Scheduling | Sequence parallelism | Tensor Parallelism | TensorRT-LLM | Tiling | Transformer Inference | Triton

Education

N/A

Roles

Engineer | ML Systems Engineer | Systems Engineer

Regions

Europe

Countries

France

States

Île-de-France, FR

Cities

Paris, Île-de-France, FR

Apply Save
Language: en Views: 0 Clicks: 1 Saves: 0

Related jobs