Senior AI/ML Engineer - LLMs, RAG, Real-Time Aerospace (NYC)
応募状況をスマホで確認 Whileresume の無料アプリ(iPhone・Android)。
仕事内容
Own and scale a production-grade LLM pipeline with end-to-end ownership from retrieval to evaluation to fine-tuning.
Tackle retrieval quality, latency bottlenecks, and eval-driven fine-tuning for real-time aerospace operations using hybrid search (BM25 + vector) and reranking.
Implement and operate RAG architectures with LangChain / LlamaIndex, LoRA/QLoRA, and production tooling (PyTorch, FastAPI, Redis, Postgres).
Build and monitor metrics (recall@k, precision@k, groundedness, hallucination) and automate evaluation and continuous improvement loops.
Collaborate directly with founders and domain experts to translate mission requirements into reliable, scalable GenAI systems.
Requirements include shipping production GenAI systems in fast-moving environments; strong CS or Applied ML background; experience with LLMs (GPT-4, Claude, Llama) and related stacks.
Tackle retrieval quality, latency bottlenecks, and eval-driven fine-tuning for real-time aerospace operations using hybrid search (BM25 + vector) and reranking.
Implement and operate RAG architectures with LangChain / LlamaIndex, LoRA/QLoRA, and production tooling (PyTorch, FastAPI, Redis, Postgres).
Build and monitor metrics (recall@k, precision@k, groundedness, hallucination) and automate evaluation and continuous improvement loops.
Collaborate directly with founders and domain experts to translate mission requirements into reliable, scalable GenAI systems.
Requirements include shipping production GenAI systems in fast-moving environments; strong CS or Applied ML background; experience with LLMs (GPT-4, Claude, Llama) and related stacks.
求人の全文を見る
業務内容、求める人物像、スキル、待遇 — 無料アカウントの作成で閲覧できます。
または
すでにアカウントをお持ちですか?
ログインこちらの求人もおすすめです
近い求人はまだありません — 最新の求人をご紹介します。