Find jobs in AI/ML, Data Science and Big Data
12 results
for GPU Kernel
(Skill/Tech stack)
-
MLOps Engineer (JAX, PyTorch, Pallas/Triton) USD 140K-220KDistributed Training | Evaluation Methodologies | GPU Kernel | GPU kernel programming | Infrastructure OptimizationFlexible project duration based on performance | Remote work | Weekly payments via Stripe or WiseMid-level Full TimeUnited States - Remote R12d ago
-
AI infrastructure | ATS | Ashby | C++ | CUDARemote workSenior-level Full TimeRemote R13d ago
-
Staff Engineer, Inference Optimizations USD 191K-239KAttention Mechanisms | BF16 | CUDA | CUDA kernels | Distributed SystemsConference reimbursement | Education reimbursement | Employee assistance program | Flexible time off | LinkedIn LearningSenior-level Full TimeBoston R17d ago
-
Apache Kafka | Apache Spark | AutoGPT | BitsAndBytes | C#Senior-level Full Time2586 Fort Meade MD, United States25d ago
-
Senior-level Full TimeCupertino27d ago
-
Research MLE (Training Optimization)大模型训练优化工程师 CNY 144K-240KC++ | CUDA | DeepSpeed | Distributed Training | FSDPSenior-level Full TimeBeijing, Beijing, China28d ago
-
Machine Learning Engineer, Connectomics USD 82K-130KAffinity prediction | BigDataViewer | C++ | CAVE | Cloud infrastructureMid-level Full TimeSan Francisco1mo ago
-
Containerization | Debugging | DeepSpeed | Distributed Systems | GPU KernelMid-level Full TimeAbu Dhabi1mo ago
-
Inference Optimization Intern – Performance Modeling USD 40K-142KC++ | CUDA | GPU Architecture | GPU Kernel | GPU Kernel DevelopmentEntry-level InternshipSunnyvale, CA1mo ago
-
Application Software Engineer, Inference USD 135K-185KAgent Orchestration | Agent SDK | Auto Scaling | Batch scheduling | C++401k plan | Employee stock purchase plan | Long-term incentives | Medical, dental & vision coverage | Onsite Palo AltoEntry-level Full TimePalo Alto, CA1mo ago
-
Inference Optimization Manager USD 229K-286KCloud infrastructure | Distributed Systems | GPU Kernel | GPU kernel programming | Inference engine401k matching | Flexible paid time off | Health insurance | Remote work options | Team onsite eventsMid-level Full TimeUnited States / Canada1mo ago
-
LLM Inference Frameworks and Optimization Engineer USD 160K-230KC++ | CUDA | CUDA graph | Cluster scheduling | CompilerEquity | Health insuranceMid-level Full TimeSan Francisco, Singapore, Amsterdam1mo ago