Perplexity: Embed V1 4B

1

Create an API key from your OpenRouter dashboard and set it as an environment variable:

2

Use perplexity/pplx-embed-v1-4b with the OpenRouter API:

OpenRouter provides an OpenAI-compatible embeddings API that you can call directly, or using the OpenAI SDK.

In the examples below, the OpenRouter-specific headers are optional. Setting them allows your app to appear on the OpenRouter leaderboards.

Using third-party SDKs

For information about using third-party SDKs and frameworks with OpenRouter, please see our frameworks documentation.

Name	Type	Default	Description
`max_tokens`	integer	—	This sets the upper limit for the number of tokens the model can generate in response.
`temperature`	float	`1`	This setting influences the variety in the model's responses.
`top_p`	float	`1`	This setting limits the model's choices to a percentage of likely tokens: only the top tokens whose probabilities add up to P.
`top_k`	integer	`0`	This limits the model's choice of tokens at each step, making it choose from a smaller set.
`frequency_penalty`	float	`0`	This setting aims to control the repetition of tokens based on how often they appear in the input.
`presence_penalty`	float	`0`	Adjusts how often the model repeats specific tokens already used in the input.
`web_search_options`	map	—	Configures native web search options for models and providers that support web-connected answers.