> AI agents: this is one page from PostHog's docs. Full index of Markdown docs for LLMs: https://posthog.com/llms.txt

# Linking OpenAI as a source - Docs

Copy page

# Linking OpenAI as a source - Docs

![](https://res.cloudinary.com/dmukukwp6/image/upload/texture_tan_9608fcca70)

![](https://res.cloudinary.com/dmukukwp6/image/upload/texture_tan_dark_a92b0e022d)

Let AI connect your sources for you

Skip the manual setup — run this in your project and the wizard auto-detects your databases and APIs and connects them to PostHog.

`npx @posthog/wizard warehouse`

[Learn more](/wizard.md)

![PostHog Wizard hedgehog](https://res.cloudinary.com/dmukukwp6/image/upload/wizard_3f8bb7a240.png)

![](https://res.cloudinary.com/dmukukwp6/image/upload/wizard_3f8bb7a240.png)Let AI connect your sources for you

**Alpha release**

This source is currently in **alpha**. The interface and available tables may change.

The OpenAI connector syncs your organization's API usage, spend, and admin data into the PostHog Data Warehouse, so you can analyze your AI usage and costs per project, model, user, and API key alongside your product data.

## Prerequisites

You need an OpenAI organization and an Admin API key. Only organization owners can create Admin API keys, and a regular project API key can't read organization usage or costs.

## Adding a data source

1.  In PostHog, go to the [Sources tab](https://app.posthog.com/data-management/sources) of the data pipeline section.
2.  Click **\+ New source** and click **Link** next to this source.
3.  Enter your credentials (see [Configuration](#configuration) below) and click **Next**.
4.  Select the tables you want to sync, choose a sync method and frequency, then click **Import**.

Once the syncs are complete, you can start querying this data in PostHog.

When linking OpenAI, you'll need:

-   **Admin API key** – create one in your [OpenAI organization settings](https://platform.openai.com/settings/organization/admin-keys) by clicking **Create admin key**. The key starts with `sk-admin`.

## Sync modes

Each table can be synced in one of several modes, depending on what the source supports:

-   **Webhook** (when available) – the source pushes changes to PostHog in real time. Fastest freshness, lowest ongoing cost, and the only mode that reliably captures updates and deletes.
-   **Incremental** – only new or updated rows are synced on each run, using a cursor field (such as an `updated_at` timestamp). Cheaper than a full refresh, but deletes aren't captured.
-   **Append only** – new rows are appended using a cursor field; existing rows are never updated. Ideal for immutable, append-only tables like event logs.
-   **Full refresh** – the whole table is reloaded on every sync. Use it when a table has no reliable cursor or when you need deletions reflected.

See [sync methods](/docs/cdp/sources.md#sync-methods) for a full explanation of how each mode works and how to choose between them.

The usage and costs tables sync incrementally on `start_time`, and audit logs on `effective_at`. Each incremental sync re-reads the trailing day of usage and cost buckets, since OpenAI keeps restating a day's bucket as late usage lands; the overlap is deduplicated automatically.

Usage and cost rows are daily buckets broken down by project, API key, model, and the other dimensions OpenAI supports for each endpoint. The `id` column is a synthesized key derived from the bucket start and those dimensions.

## Configuration

| Option | Type | Required |
| --- | --- | --- |
| Admin API key | password | Yes |

## Supported tables

| Table | Description | Sync method | Incremental field | Primary key |
| --- | --- | --- | --- | --- |
| usage_completions | Chat and text completions token usage aggregated into daily buckets, broken down by project, user, API key, model, batch, and service tier. | Incremental, Full refresh | start_time | — |
| usage_embeddings | Embeddings token usage aggregated into daily buckets, broken down by project, user, API key, and model. | Incremental, Full refresh | start_time | — |
| usage_moderations | Moderations token usage aggregated into daily buckets, broken down by project, user, API key, and model. | Incremental, Full refresh | start_time | — |
| usage_images | Image generation usage aggregated into daily buckets, broken down by project, user, API key, model, image size, and source. | Incremental, Full refresh | start_time | — |
| usage_audio_speeches | Text-to-speech usage aggregated into daily buckets, broken down by project, user, API key, and model. | Incremental, Full refresh | start_time | — |
| usage_audio_transcriptions | Speech-to-text usage aggregated into daily buckets, broken down by project, user, API key, and model. | Incremental, Full refresh | start_time | — |
| usage_vector_stores | Vector store storage usage aggregated into daily buckets, broken down by project. | Incremental, Full refresh | start_time | — |
| usage_code_interpreter_sessions | Code interpreter session usage aggregated into daily buckets, broken down by project. | Incremental, Full refresh | start_time | — |
| usage_web_search_calls | Web search tool-call usage aggregated into daily buckets, broken down by project, user, API key, model, and context level. | Incremental, Full refresh | start_time | — |
| usage_file_search_calls | File search tool-call usage aggregated into daily buckets, broken down by project, user, API key, and vector store. | Incremental, Full refresh | start_time | — |
| costs | Daily spend for your OpenAI organization, broken down by project, line item, and API key. | Incremental, Full refresh | start_time | — |
| projects | Projects in your OpenAI organization (includes archived projects). | Full refresh | — | — |
| users | Members of your OpenAI organization. | Full refresh | — | — |
| invites | Pending and historical invitations to join your OpenAI organization. | Full refresh | — | — |
| admin_api_keys | Organization-level Admin API keys. | Full refresh | — | — |
| project_users | Membership rows mapping users to the projects they belong to, one row per (project, user). | Full refresh | — | — |
| project_service_accounts | Service accounts (bot users not tied to a person) in each project. | Full refresh | — | — |
| project_api_keys | API keys provisioned in each project. | Full refresh | — | — |
| project_rate_limits | Per-model rate limit configuration for each project. | Full refresh | — | — |
| audit_logs | User actions and configuration changes within your OpenAI organization. | Incremental, Full refresh | effective_at | — |

## Troubleshooting

-   **401 Unauthorized** – your Admin API key is invalid or has been revoked. Create a new one and reconnect.
-   **403 Forbidden** – your key isn't an Admin API key. Ask an organization owner to create one in the [organization settings](https://platform.openai.com/settings/organization/admin-keys).

If your sync is failing or data looks wrong, see the [Data warehouse troubleshooting guide](/docs/data-warehouse/troubleshooting.md). If that doesn't help, [contact support](https://us.posthog.com/#panel=support%3Asupport%3Adata_warehouse%3A%3Atrue) – we're happy to help.

### Was this page useful?

HelpfulCould be better