Training / AI Infrastructure at Genesis AI
- Company: Genesis AI
- Location: Remote / Paris
- Employment type: Full-time
- Posted: 2026-04-30
- Technologies: PyTorch, CUDA, Python
About the role
WHAT YOU’LL DO - Drive down wall-clock time to convergence by profiling and eliminating bottlenecks across the foundation model training stack stack, from data pipelines to GPU kernels - Design, build, and optimize distributed training systems (PyTorch) for multi-node GPU clusters, ensuring scalability, robustness, and high utilization - Implement efficient low-level code (CUDA, cuDNN, Triton, custom kernels) and integrate it seamlessly into high-level training frameworks - Optimize workloads for hardware efficiency: CPU/GPU compute balance, memory management, data throughput, and networking - Develop monitoring and debugging tools for large-scale runs, enabling rapid diagnosis of performance regressions and failures WHAT YOU’LL BRING - Deep experience in distributed systems, ML infrastructure, or high-performance computing (8+ years) - Production-grade expertise in Python - Low-level pe…
Apply on Genesis AI's official careers page: https://jobs.ashbyhq.com/genesis/b9ce5163-c4e1-4395-814b-358e3aea04f2