vllm-project/tpu-inference

[Feature]: Add embedding model functionality to tpu-inference

Open

#899 opened on Oct 20, 2025

View on GitHub
 (6 comments) (0 reactions) (0 assignees)Python (205 forks)github user discovery
enhancementgood first issuehelp wanted

Repository metrics

Stars
 (348 stars)
PR merge metrics
 (PR metrics pending)

Description

🚀 The feature, motivation and pitch

copied from feature request in vllm upstream

Great first issue!

Should enable for both torchax and jax (if needed).

Alternatives

No response

Additional context

No response

Before submitting a new issue...

  • Make sure you already searched for relevant issues, and asked the chatbot living at the bottom right corner of the documentation page, which can answer lots of frequently asked questions.

Contributor guide