Zum Inhalt springen

Member of Technical Staff (Research Engineer - LLM Systems & Performance)

Ist diese Stelle etwas für Sie?

Erstellen Sie Ihren Lebenslauf und entdecken Sie Ihre Übereinstimmung mit dieser Stelle — und mit allen anderen.

Lebenslauf erstellen

Die Stelle

Lead and optimize end-to-end LLM systems, from supervised fine-tuning and reinforcement learning pipelines to high-throughput production inference. Develop and enhance SFT/RL components (e.g., Verl, SkyRL), covering data loading, training loops, logging, and evaluation. Contribute to LLM inference infrastructure (e.g., vLLM, SGLang) with batching, KV-cache management, scheduling, and serving optimizations. Profile and optimize end-to-end performance (throughput, latency, memory bandwidth) with tools like Nsight and profilers to identify bottlenecks. Work across multi-GPU clusters using NCCL, NVLink and various parallelism strategies (data/tensor/pipeline/expert/context) and explore quantization (INT8/FP8/FP4, mixed precision). Collaborate with researchers to move ideas from paper to prototype to scaled experiments and production, delivering clean, well-tested, well-documented code for multiple teams.

Die vollständige Anzeige sehen

Aufgaben, Anforderungen, Kompetenzen und Vorteile — mit Ihrem kostenlosen Konto.

Mindestens 6 Zeichen. Je länger, desto sicherer.
oder

Bereits ein Konto?

Ihr Ort

Stellen und Unternehmen werden nach diesem Land gefiltert.

Vorschläge

Alle Länder 66