aijobs.net

Machine Learning Engineer — Inference Optimization

Australia

A AUD 140K-180K (estimate) Senior-level Full Time

Apply Save
Found 18h ago
Tasks
Perks/Benefits
Skills/Tech-stack

Attention Mechanisms | Batching | CUDA | CUDA kernel | CUDA kernel tuning | Compute Graphs | Distributed Systems | GPU Performance | GPU Performance Optimization | Inference benchmarking | KV cache | Kernel tuning | Language Models | Large Language Models | Low Latency | Low-Latency Systems | Machine Learning | Memory Optimization | Model Serving | Neural Networks | ONNX Runtime | Performance optimization | PyTorch | Quantization | ROCm | Speculative decoding | Streaming | TensorRT | Triton | VLLM

Education

N/A

Roles

Engineer | Learning Engineer | Machine Learning Engineer

Regions

Asia/Pacific

Countries

Australia

Apply Save
Language: en Views: 0 Clicks: 0 Saves: 0

Related jobs