Find jobs in AI/ML, Data Science and Big Data
15 results
for GPU Kernel
(Skill/Tech stack)
-
Apache Kafka | Apache Spark | AutoGPT | BitsAndBytes | C#Senior-level Full Time2586 Fort Meade MD, United States5d ago
-
C# | C++ | CPU architecture | CUDA | GPU ArchitectureComprehensive benefits package | Equity | Mentoring opportunitiesSenior-level Full TimeUS, CA, Santa Clara R7d ago
-
Senior-level Full TimeCupertino7d ago
-
Research MLE (Training Optimization)大模型训练优化工程师 CNY 144K-240KC++ | CUDA | DeepSpeed | Distributed Training | FSDPSenior-level Full TimeBeijing, Beijing, China8d ago
-
Solutions Architect, Agentic Optimization USD 152K-241KAttention kernels | C++ | CUDA | Containers | Data PipelinesEquity | Remote work | Travel for customer visits and conferencesSenior-level Full TimeUS, CA, Santa Clara R8d ago
-
Machine Learning Engineer, Connectomics USD 82K-130KAffinity prediction | BigDataViewer | C++ | CAVE | Cloud infrastructureMid-level Full TimeSan Francisco11d ago
-
Containerization | Debugging | DeepSpeed | Distributed Systems | GPU KernelMid-level Full TimeAbu Dhabi23d ago
-
Inference Optimization Intern – Performance Modeling USD 40K-142KC++ | CUDA | GPU Architecture | GPU Kernel | GPU Kernel DevelopmentEntry-level InternshipSunnyvale, CA26d ago
-
AllGather | AllReduce | Artificial Intelligence | Asynchronous pipelines | BenchmarkingSenior-level Full TimeSeattle, United States R27d ago
-
Application Software Engineer, Inference USD 135K-185KAgent Orchestration | Agent SDK | Auto Scaling | Batch scheduling | C++401k plan | Employee stock purchase plan | Long-term incentives | Medical, dental & vision coverage | Onsite Palo AltoEntry-level Full TimePalo Alto, CA1mo ago
-
Inference Optimization Manager USD 229K-286KCloud infrastructure | Distributed Systems | GPU Kernel | GPU kernel programming | Inference engine401k matching | Flexible paid time off | Health insurance | Remote work options | Team onsite eventsMid-level Full TimeUnited States / Canada1mo ago
-
CUDA | CUDNN | Cutlass | Deep learning | GPU ArchitectureMid-level Full TimeUS-WA-Bellevue1mo ago
-
LLM Inference Frameworks and Optimization Engineer USD 160K-230KC++ | CUDA | CUDA graph | Cluster scheduling | CompilerEquity | Health insuranceMid-level Full TimeSan Francisco, Singapore, Amsterdam1mo ago
-
Staff Compiler Engineer - PyTorch + Kernel DSLPLATE USD 163K-253KAutotuning | Collective Primitives | Cost Based Compilation | Custom ISA | Cutlass401k | Adoption support stipend | Charitable giving match | Fertility care stipend | Flexible work environmentSenior-level Full TimeSan Jose, California, United States1mo ago
-
Entry-level Full TimeNew York, NY, United States1mo ago