Vai al contenuto

Machine Learning Engineer | Python | PyTorch | Distributed Training | GPU | Hybrid

Questa offerta fa per te?

Crea il tuo CV e scopri la tua percentuale di corrispondenza con questa posizione — e con tutte le altre.

Crea il mio CV

La posizione

Productize and optimize models from research into reliable, performant, and cost-efficient services with clear SLOs.
Scale training across nodes/GPUs (DDP/FSDP/ZeRO, pipeline/tensor parallelism) and own throughput/time-to-train via profiling and optimization.
Implement model-efficiency techniques (quantization, distillation, pruning, KV-cache, Flash Attention) for training and inference without materially degrading quality.
Build and maintain model-serving systems (vLLM/Triton/ONNX/TensorRT/AITemplate) with batching, streaming, caching, and memory management.
Integrate with vector/feature stores and data pipelines (FAISS/Milvus/Pinecone/pgvector; Parquet/Delta) as needed for production.
Define and track performance and cost KPIs; run continuous improvement loops and capacity planning; partner with ML Ops and scientists on handoffs and evaluations.

Vedi l'annuncio completo

Mansioni, profilo, competenze e vantaggi — crea il tuo account gratuito.

Almeno 6 caratteri. Più è lunga, più è sicura.
o

Hai già un account?

Offerte simili

Altre posizioni che potrebbero interessarti.

Vedi tutto →

La tua località

Offerte e aziende saranno filtrate su questo paese.

Suggeriti

Tutti i paesi 68