Remote Inference Engine Engineer - LLMs & Diffusion
Inferact
Inferact is seeking an inference runtime engineer to push the boundaries of LLM and diffusion model serving. You will optimize how models execute across diverse hardware and architectures, contributing to the core of vLLM and related projects. This fully remote role...
This job was verified from Jooble US. Applications are completed on the original source.
Apply on the original listing ↗
Something wrong with this job?