Working with Prompts | Confident AI Docs

Overview

You can pull a prompt version from Confident AI like how you would pull a dataset. It works by:

Providing Confident AI with the alias and optionally version of the prompt you wish to retrieve
Confident AI will provide the non-interpolated version of the prompt
You will then interpolate the variables in code

You should pull prompts once and save it in memory instead of pulling it everytime you need to use it.

You can pull and manage your prompts in any project by configuring a CONFIDENT_API_KEY.

For default usage, set CONFIDENT_API_KEY as an environment variable.
To target a specific project, pass a confident_api_key directly when creating the Prompt object.

1 from deepeval.prompt import Prompt, PromptMessage
2 
3 prompt = Prompt(
4   alias="YOUR-PROMPT-ALIAS",
5   confident_api_key="confident_us...",
6 )

When both are provided, the confident_api_key passed to Prompt always takes precedence over the environment variable.

Using Prompt Versions

Pull prompt with alias

Pull your prompt version by providing the alias you’ve defined:

Python

TypeScript

curL

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull()

By default, Confident AI will return the latest version of your prompt. However, you can also specify the version to override this behavior.

1 prompt.pull(version="00.00.01")

Interpolate variables

Now that you have your prompt template, interpolate any dynamic variables you may have defined in your prompt version.

Python

TypeScript

curL

1 interpolated_prompt = prompt.interpolate(name="Joe")

For example, if this is your prompt version:

Messages

Text

1 {
2   "role": "system",
3   "content": "You are a helpful assistant called {{ name }}. Speak normally like a human."
4 }

And your interpolation type is {{ variable }}, interpolating the name (e.g. “Joe”) would give you this prompt that is ready for use:

1 {
2   "role": "system",
3   "content": "You are a helpful assistant called Joe. Speak normally like a human."
4 }

And if you don’t have any variables, you must still use the interpolate() method to create a copy of your prompt template to be used in your LLM application.

Use interpolated prompt

By now you should have an interpolated prompt version, for example:

Messages

Text

1 {
2   "role": "system",
3   "content": "You are a helpful assistant called Joe. Speak normally like a human."
4 }

Which you can use to generate text from your LLM provider of choice. Here are some examples with OpenAI:

Python

TypeScript

curL

main.py

1 from deepeval.prompt import Prompt
2 from openai import OpenAI
3 
4 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
5 prompt.pull()
6 interpolated_prompt = prompt.interpolate() # interpolate prompt
7 
8 response = OpenAI().chat.completions.create(
9     model="gpt-4o-mini",
10     messages=interpolated_prompt
11 )
12 
13 print(response.choices[0].message.content)

Accessing Tools from Prompts

After pulling a prompt, you can access any tools that were defined in the prompt version via the tools property. Each tool contains:

name: The name of the tool
description: A description of what the tool does
input schema: The JSON schema defining the tool’s input parameters

Python

TypeScript

main.py

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull()
5 
6 # Access tools from the prompt
7 for tool in prompt.tools:
8     print(f"Tool Name: {tool.name}")
9     print(f"Description: {tool.description}")
10     print(f"Input Schema: {tool.input_schema}")

Pull Prompts By Label

You can pull a prompt from Confident AI using its alias, you can also pull specific versions of prompts using verion and label.

Version

Label

main.py

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull(version="00.00.01")

How Are Prompts Pulled?

Confident AI automatically caches prompts on the client side to minimize API call latency and ensure prompt availability, which is especially useful in production environments.

Cache

No Cache

Customize refresh rate

By default, the cache is refetched every 60 seconds, where DeepEval will automatically update the cached prompt with the up-to-date version from Confident AI. This can be overridden by setting the refresh parameter to a different value. Fetching is done asynchronously, so it will not block your application.

Python

TypeScript

main.py

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull(refresh=60)
5 interpolated_prompt = prompt.interpolate(name="Joe")

Disable Caching

To disable caching, you can set refresh=0. This will force an API call every time you pull the prompt, which is particularly useful for development and testing.

Python

TypeScript

main.py

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull(refresh=0)
5 interpolated_prompt = prompt.interpolate(name="Joe")

Log Prompts During Evals

You can associate a prompt with your evals to get detailed insights on how each prompt and their versions are performing. It works by:

Pulling or creating prompts in deepeval using the Prompt object
Logging prompts as hyperparameters in your evaluations

Pull prompts

You should first pull a prompt from Confident AI using alias, version or label.

Python

TypeScript

curL

1 from deepeval.prompt import Prompt
2 
3 prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
4 prompt.pull(version="00.00.01")

You can now interpolate and use this prompt in your LLM app.

Logging prompt in `evaluate`

Now, simply add this prompt as a free-form key-value pair to the hyperparameters argument in the evaluate() function

Python

Typescript

curL

1 evaluate(
2     ...
3     hyperparameters={
4         "Model": "YOUR-MODEL",
5         "Prompt": prompt,
6     },
7 )

This will automatically attribute the prompt used during this test run, which will allow you get detailed insights in the Confident AI platform.

Log Prompts During Tracing

Associating prompts with LLM traces and spans is a great way to determine which prompts performs best in production.

Setup tracing

Attach the @observe decorator to functions/methods that make up your agent, and specify type llm for your LLM-calling functions.

main.py

1 from deepeval.tracing import observe
2 
3 @observe(type="llm", model="gpt-4.1")
4 def your_llm_component():
5     ...

Specifying the type is necessary because logging prompts is only available for LLM spans.

Pull and interpolate prompt

Pull and interpolate the prompt version to use it for LLM generation.

main.py

1 from deepeval.tracing import observe
2 from deepeval.prompt import Prompt
3 from openai import OpenAI
4 
5 @observe(type="llm", model="gpt-4.1")
6 def your_llm_component():
7     prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
8     prompt.pull()
9     interpolated_prompt = prompt.interpolate(name="Joe")
10     response = OpenAI().chat.completions.create(model="gpt-4o-mini", messages=interpolated_prompt)
11     return response.choices[0].message.content

Execute your function

Then simply provide the prompt to the update_llm_span function.

main.py

1 from deepeval.tracing import observe, update_llm_span
2 from deepeval.prompt import Prompt
3 from openai import OpenAI
4 
5 @observe(type="llm", model="gpt-4.1")
6 def your_llm_component():
7     prompt = Prompt(alias="YOUR-PROMPT-ALIAS")
8     prompt.pull()
9     interpolated_prompt = prompt.interpolate(name="Joe")
10     response = OpenAI().chat.completions.create(model="gpt-4o-mini", messages=interpolated_prompt)
11     update_llm_span(prompt=prompt)
12     return response.choices[0].message.content

Remember to pull the prompt before updating the span, otherwise the prompt will not be logged.

This will automatically attribute the prompt used to the LLM span.

Overview

Using Prompt Versions

Pull prompt with alias

Python

TypeScript

curL

Interpolate variables

Python

TypeScript

curL

Messages

Text

Use interpolated prompt

Messages

Text

Python

TypeScript

curL

Accessing Tools from Prompts

Python

TypeScript

Pull Prompts By Label

Version

Label

How Are Prompts Pulled?

Cache

No Cache

Customize refresh rate

Python

TypeScript

Disable Caching

Python

TypeScript

Log Prompts During Evals

Pull prompts

Python

TypeScript

curL

Logging prompt in evaluate

Python

Typescript

curL

Log Prompts During Tracing

Setup tracing

Pull and interpolate prompt

Execute your function

Logging prompt in `evaluate`