Skip to main content

Overview

Knowledge Base stores the content AI Puffer can search before it generates an answer or a piece of content. Use it for support answers, product details, documentation, policies, posts, pages, WooCommerce products, uploaded documents, and other source text you want AI Puffer to use as context.

Providers

Use Local, OpenAI or Google managed stores, Pinecone, Qdrant, or Chroma.

Manage Vector Stores

Create, select, or delete vector targets.

Add Data

Add data and manage source records.

Settings

Configure chunking, batches, content rules, semantic search, and other settings.

Semantic Search

Publish and configure a frontend vector search form.

Troubleshooting

Fix missing targets, dimension errors, and empty results.

Providers

For Local, Pinecone, Qdrant, and Chroma, the store dimension must match the embedding model’s output dimension. Use the same model and dimension for both indexing and querying. For example, if your Pinecone index, Qdrant collection, or Chroma collection is 3072 dimensions, use a 3072-dimension embedding model when adding data and when searching that data later.
If the dimension does not match, the vector provider can reject the data or return unusable search results.

Local

Local stores source chunks, metadata, and embedding vectors in your WordPress database. No Pinecone, Qdrant, or Chroma account is required. Your WordPress hosting stores the data; AI Puffer Cloud does not host it. Local still needs an embedding model for indexing and search. Use a Cloud embedding model or a compatible model from your own provider. The text model generating answers can be a different provider. To create a store:
  1. Open AI Puffer > Knowledge Base > Stores.
  2. Click Create store.
  3. Select Local.
  4. Enter the Name and Dimension.
  5. Click Create store.
  6. Use Add source in Data to choose this store and the embedding model before adding content.
Local store creation with name and dimension fields
Creating the store sets its dimension, not its embedding model. When adding or searching data, use the same model and dimension. Models with equal dimensions are not necessarily compatible. If you switch embedding models, re-index the content into a store configured for the new model. Local indexing and semantic search through a hosted embedding model use that provider’s billing or Cloud credits. Local storage itself does not consume Cloud storage credits, but uses your WordPress database and hosting resources.

OpenAI

OpenAI Vector Stores use your OpenAI account directly.
  1. Go to AI Puffer > Settings > AI.
  2. Select OpenAI as the AI provider.
  3. Enter your OpenAI API key.
  4. Sync models if needed.
  5. Go to AI Puffer > Knowledge Base > Stores to create or refresh OpenAI vector stores.
OpenAI Vector Stores do not require a separate embedding model selection in Knowledge Base. OpenAI handles file storage, chunking, embedding, and vector search on its side.
OpenAI API key settings

Pinecone

Pinecone is configured from the Integrations settings.
  1. Go to AI Puffer > Settings > Integrations.
  2. Select Pinecone.
  3. Enter your Pinecone API Key.
  4. Click Sync Indexes to load indexes from Pinecone.
  5. Go to AI Puffer > Knowledge Base > Stores to create, refresh, or delete indexes.
When you create a Pinecone index in AI Puffer, enter the dimension that matches the embedding model you plan to use.
Pinecone API key

Qdrant

Qdrant requires both an endpoint URL and an API key.
  1. Go to AI Puffer > Settings > Integrations.
  2. Select Qdrant.
  3. Enter your Qdrant URL.
  4. Enter your Qdrant API Key.
  5. Click Sync Collections to load collections from Qdrant.
  6. Go to AI Puffer > Knowledge Base > Stores to create, refresh, or delete collections.
When you create a Qdrant collection in AI Puffer, enter the dimension that matches the embedding model you plan to use.
Qdrant API key

Chroma

Chroma uses endpoint, tenant, and database settings.
  1. Go to AI Puffer > Settings > Integrations.
  2. Select Chroma.
  3. Enter your Chroma URL. For Chroma Cloud, you can use https://api.trychroma.com
  4. Enter your Chroma API Key if you use Chroma Cloud or an authenticated server.
  5. Enter the Tenant.
  6. Enter the Database.
  7. Click Sync Collections to load collections from Chroma.
  8. Go to AI Puffer > Knowledge Base > Stores to create, refresh, or delete collections.
For local Chroma, the default tenant is default_tenant and the default database is default_database.
Chroma API key

Embedding Providers

Local, Pinecone, Qdrant, and Chroma store vectors that AI Puffer creates with a selected embedding model. Before adding data to these providers, configure the embedding provider you want to use in AI Puffer > Settings > AI. Supported embedding providers include AI Puffer Cloud, OpenAI, Google, Azure, OpenRouter, and Ollama where available. For Cloud, connect the account and refresh its models; you do not enter a provider API key. The selector shows actual model names. Supported output dimensions come from the model’s capabilities; a store is not limited to 512 dimensions. The selected embedding model must match the dimension of the Pinecone index, Qdrant collection, or Chroma collection. xAI is not an embedding provider or vector store provider in the current integration. xAI chatbots, forms, and text workflows can still use retrieved Knowledge Base context from Local, OpenAI, Google, Pinecone, Qdrant, or Chroma because AI Puffer sends that context as text.

Manage Vector Stores

Use AI Puffer > Knowledge Base > Stores to create, refresh, inspect, or delete vector targets. The Stores tab manages Local stores, OpenAI and Google managed stores, Pinecone indexes, Qdrant collections, and Chroma collections. The Data tab uses these targets when you add content.
Knowledge Base provider selector

OpenAI Vector Stores

  1. Add your OpenAI API key in AI Puffer > Settings > AI.
  2. Go to AI Puffer > Knowledge Base > Stores.
  3. Select OpenAI as the provider.
  4. Click Create Store.
  5. Enter a store name.
  6. Click Create.
OpenAI handles the vector store search on its side. AI Puffer stores a local source record so you can see what was added.
OpenAI Create Vector

Pinecone Indexes

  1. Add your Pinecone API key in AI Puffer > Settings > Integrations.
  2. Go to AI Puffer > Knowledge Base > Stores.
  3. Select Pinecone as the provider.
  4. Select the embedding model you plan to use.
  5. Click Create Store.
  6. Enter an index name.
  7. Enter the dimension for the selected embedding model.
  8. Click Create.
Use the same embedding model when you add data to the index and when a module searches that index.
Pinecone Create Index

Qdrant Collections

  1. Add your Qdrant URL and API key in AI Puffer > Settings > Integrations.
  2. Go to AI Puffer > Knowledge Base > Stores.
  3. Select Qdrant as the provider.
  4. Select the embedding model you plan to use.
  5. Click Create Store.
  6. Enter a collection name.
  7. Enter the dimension for the selected embedding model.
  8. Click Create.
Use the same embedding model when you add data to the collection and when a module searches that collection.
Qdrant Create collection

Chroma Collections

  1. Add your Chroma endpoint, tenant, database, and API key in AI Puffer > Settings > Integrations.
  2. Go to AI Puffer > Knowledge Base > Stores.
  3. Select Chroma as the provider.
  4. Click Create Store.
  5. Enter a collection name.
  6. Click Create.
Chroma collections do not require a dimension when they are created in AI Puffer. Use the same embedding model when you add data to the collection and when a module searches that collection.
Chroma Create collection
Use Refresh when you need AI Puffer to fetch the latest stores, indexes, or collections from the selected provider. To delete a target, use the available action in the Stores table.

Add Data

Use AI Puffer > Knowledge Base > Data to add new source data and manage existing source records. Before adding data:
  1. Go to AI Puffer > Knowledge Base > Data.
  2. Select a provider.
  3. Select the target vector store, index, or collection.
  4. Click Add source.
  5. Choose the target store and, for Local, Pinecone, Qdrant, or Chroma, its embedding model.
  6. Choose Website, Q&A, Text, or Files and add the content.
Knowledge Base Add data panel

Q&A

Use Q&A for short answers that should be easy to retrieve later.
  1. Select Q&A.
  2. Enter the question.
  3. Enter the answer.
  4. Click Add Q&A.
AI Puffer stores the pair as text:
Knowledge Base Q&A tab

Text

Use Text for policies, instructions, product notes, support snippets, or any source text that does not already exist as WordPress content.
  1. Select Text.
  2. Paste the source text.
  3. Click Add Text.
Knowledge Base Text tab

Files

Use Files when the source is already in a document.
  1. Select Files.
  2. Click Choose files.
  3. Select one or more files.
Files start uploading and training after selection. Supported file extensions:
For Local, Pinecone, Qdrant, and Chroma, AI Puffer extracts text, splits large files into chunks, creates embeddings, and stores each chunk in the selected index or collection. File chunks can be embedded in batches to reduce the number of embedding API requests. File size is limited by your WordPress/PHP upload settings. OpenAI Vector Store uploads also use OpenAI’s file limits.
Knowledge Base Files tab

Website

Use Website when the source is WordPress content.
  1. Select Website.
  2. Choose All or Choose items.
  3. Select the content types.
  4. If using Choose items, select the individual published items.
  5. Click Add Items.
Website training uses published content. Posts and pages are selected by default. WooCommerce products appear when WooCommerce is active. Public custom post types can also appear. When WordPress content is indexed, AI Puffer builds the source text from the URL, title, excerpt, content, public custom fields, public taxonomies, and available WooCommerce product data.
Knowledge Base Website all mode

Manage Data

The source table in the Data tab shows the local records created while adding data.
Knowledge Base source table
Available actions:
Knowledge Base source table
Knowledge Base source preview

Settings

Open AI Puffer > Knowledge Base > Settings to configure Knowledge Base behavior. Knowledge Base has Data, Stores, Search, and Settings tabs. The sections below describe the settings controls; some screenshots show their earlier tab layout.

Chunking

Document chunking controls how AI Puffer splits large uploaded files and WordPress Website content before embedding them for Local, Pinecone, Qdrant, or Chroma. Use smaller chunks when an embedding provider rejects long input. Keep some overlap for long documents where meaning continues across sections. Some embedding models have lower hard limits than the visible maximum, so AI Puffer may apply a safer model-specific cap during indexing. OpenAI Vector Store file uploads use OpenAI File Search chunking instead of the Pinecone, Qdrant, and Chroma chunking settings above.
Knowledge Base Batches settings

Batches

The Batches tab contains Embedding Batches, which controls how many file chunks AI Puffer sends to the embedding provider in one request.
  1. Go to AI Puffer > Knowledge Base > Settings.
  2. Open Batches.
  3. Adjust the batch size for the embedding provider you use.
  4. Wait for the settings autosave to finish.
If the Batches tab shows Upgrade, activate Pro before saving batch settings.
Knowledge Base Batches settings
For example, a batch size of 50 means AI Puffer sends up to 50 prepared file chunks to the embedding API at once. Larger batches can make file upload training much faster because they reduce repeated API calls.
Embedding batch settings apply to the supported providers’ chunked file uploads for Local, Pinecone, Qdrant, and Chroma. Cloud applies its own batch limits. Website training uses document chunking, but these batch-size settings do not change Q&A, Text, Website training, semantic search queries, or OpenAI Vector Store file uploads.
If a provider returns rate limit errors such as HTTP 429, lower that provider’s batch size and try again. AI Puffer can pause and retry file upload processing when the provider sends a retry delay, but lowering the batch size is usually better for accounts with stricter quotas.

Content Rules

Content Rules define which WordPress fields are included when Website training or list-screen indexing sends WordPress content to a vector target.
  1. Go to AI Puffer > Knowledge Base > Settings.
  2. Open Content Rules.
  3. Click Configure.
  4. Select a post type.
  5. Adjust Basic Labels if you want different labels for source URL, title, excerpt, or content.
  6. Enable or disable custom fields.
  7. Enable or disable taxonomies.
  8. For WooCommerce products, enable or disable product data such as SKU, price, stock, dimensions, and attributes.
  9. Save.
If the Save button shows Upgrade, activate Pro before saving indexing rules.
Knowledge Base indexing controls

Others

Open AI Puffer > Knowledge Base > Search for Semantic Search. Other Knowledge Base controls appear under Settings. The Others tab contains Semantic Search and admin visibility settings. Semantic Search publishes a search form that queries a Local store, Pinecone index, Qdrant collection, or Chroma collection from the frontend. Open the Search tab and configure the search target.
  1. Select the storage provider: Local, Pinecone, Qdrant, or Chroma.
  2. Select the index or collection.
  3. Select the embedding model.
  4. Set Number of Results.
  5. Set No Results Text.
  6. Test a query in Try semantic search.
  7. Copy the shortcode.
Semantic Search uses the global settings from this panel. It does not use OpenAI Vector Stores in the current UI. Use the same embedding model and dimension used when the Local, Pinecone, Qdrant, or Chroma data was added.
Knowledge Base Semantic Search settings

Admin Visibility

The post-list Index button is managed from AI Puffer > Settings > Utilities. See WordPress Utilities.

Troubleshooting

Configure the provider credentials, then sync or create the vector target again.
Confirm the embedding model dimension matches the index or collection dimension.
Check Knowledge Base > Settings > Content Rules for that post type.
Enable AI Puffer > Settings > Utilities > Index button and confirm the user role can access the vector content indexer module.
Confirm the selected target contains trained data and the same embedding model is selected.