AI

Pro

From setting up your default chat engine and deciding what kind of permissions you give it when searching or working with your Mac, the AI view is where you set these options. For creating AI-based images, see the Image Generation view. And if you need to detect or convert speech to text in images and media files, see the Transcription view.

Chat

Choose your AI model and settings specific to it, as needed. Also set from where the model can get information, if it can effect changes to your database, and what kind of summaries you'd like it to return.

Chat Setup: Specify what large language model (LLM) you want to use and set up any required parameters for it. Note several of the controls here are dynamic and the options will change depending on what LLM you've chosen.

Assistant: Certain AI models have access to "tooling" and may be able to accept DEVONthink-related commands. You need to decide what behaviors you will allow it to use on your Mac and with your databases.

Max. Recent Chats: Set the number of remembered chats you can revisit in the Recent Chats popup in the Chat assistant.

Library

This pane is where you can define your own custom repository of roles, prompts, and skills to use with external AI. From creating "personalities" like subject matter experts to having a ready-made set of prompts you often use to going beyond simple inquiries with skills built from more complex instruction sets, there are things to explore and opportunities to push the limits with AI and DEVONthink. The interface is divided into a "bookshelf" of several built-in options and – where you'll add your own – and the controls to create and modify them.

Function List: On the left is the list of all your AI functions. Select one to see its controls and parameters on the right. Add or remove functions with the plus and minus buttons at the bottom of the pane. The button has commands for creating new functions of each type: Prompt, Role, or Skill. There are also these commands: Duplicate, Rename, Import, Export, and Delete. If you delete some of the built-in functions, you can restore them via Add Default Items.

Controls: When editing an AI function, you will use the controls on the righthand side. These let you define what you want to happen, when you want it to happen, and what level of autonomy you give AI in accomplishing the tasks. Here are the other available controls:

If you're interested in expanding your AI integration, you can read more about using this pane in the Automation > AI Prompts, Roles, and Skills section.

MCP

The MCP settings provide controls for setting up DEVONthink's MCP server, allowing connecting to and working with your databases while outside our application.

You can read more about using the MCP server in the AI in Practice > Assistants section.

Summarization

Model: Choose a specific AI provider and model for summarization or Default to use the Chat model.

Language: Choose a specific language to use for the summary or leave it as Automatic to use the system's primary language.

Style: Determine what summary format you'd like in response to asking chat to summarize a document. The choices are:

Custom Prompt: Create your own prompt defining what kind of response you'd like, including how you'd like the summary to be structured. Use the special placeholder %@ to refer to the information being summarized.

Image Generation

Choose and set up a text-to-image AI model. These controls are dynamic and their options change depending on the model you choose.

Image Generator Setup:

Transcription

AI speech-to-text processes incoming media files and processes them per these settings. For example, an .mp3 file could be transcribed into a separate annotation file for future use.

For each type of text recognition, images or audio/video, choose where to store the recognized text via the Destination dropdown. The available options are:

Images: Decide what live OCR engine you want to process images added to your database:

Audio & Video: Choose the transcription engine you want to process media files added to your database:

Add timestamps to transcription: Examines the speech and inserts timestamps at certain points.

Transcription Language: Choose the language of the media file to be transcribed. Only used with OpenAI's Whisper.

API Key: Enter the API key you received from your AI transcription provider, e.g., OpenAI.