All docs
/
Gen AI
/
AI Tokens

AI Tokens

Understand AI Tokens

Overview

Understanding AI Tokens

In the world of Large Language Models (LLMs), tokens are the fundamental building blocks for processing and understanding text.

Think of tokens as small chunks, each typically representing 3-4 characters. A 100-word passage generally breaks down into approximately 125–135 tokens.

When working with AI models, you'll encounter two types:

  • Input tokens: The prompts you send to the AI.
  • Output tokens: The AI’s generated response.
Note: In Clay, you can control usage by setting a maximum output length in your model configuration.

What Is TPM?

TPM (Tokens Per Minute) refers to how many tokens a model can process within one minute—across both input and output. Different features and providers have different TPM requirements, often tied to your API tier.

API tier requirements

ChatGPT Generate Text

  • Requirement: 30,000 TPM
  • Access: Works with any paid tier as long as the API key has access to GPT-4 or GPT-4 Turbo
  • (Note: Free-tier or unpaid OpenAI keys may not have access)

ClayGent Web Research

  • OpenAI: Requires Tier 2 or higher (≥450,000 TPM)
  • Anthropic: Requires Tier 4 or higher (≥400,000 TPM)
  • Gemini:
    • Requires Tier 2, but only works with the Gemini 2.0 Flash model
    • Gemini 1.5 models may require additional access through Vertex AI or custom tiers

Upgrade API tier

If you need higher token limits or access to advanced features, you'll need to upgrade your API tier with your chosen provider.

Here's how to find more information on upgrading, with each major provider.

Monitoring API Usage

Each platform provides tools to track your usage:

AI pricing in Clay

When using AI features in Clay, you'll consume both Actions and Data Credits:

  • Actions: Each AI enrichment consumes 1 Action (platform orchestration work)
  • Data Credits: Cost varies by model—Clay offers both fixed and variable AI pricing

Clay uses two pricing structures for AI models:

  • Fixed pricing: A flat number of data credits per task (applies to most models, including Clay's own models like Neon, Helium, and Argon)
  • Variable pricing: Data credits based on actual token usage plus a 20% premium (applies to advanced reasoning models used for sophisticated web research)

You can control AI spending by:

  • Setting maximum output length in your model configuration
  • Setting custom budgets for each run
  • Choosing more cost-effective models for simpler tasks
  • Selecting fixed-price models when cost predictability is important

To learn more about how AI is priced in Clay, see our guide on how AI is priced.

Clay credits vs. personal API keys

When you're using Clay's AI tools, you have two options for managing your API access: using Clay credits or connecting your own personal API keys.

Each option has different benefits and considerations in terms of cost, convenience, and management requirements. Let's explore these options:

Clay credits

  • Clay manages rate limits, tier access, and scaling for you
  • No need to worry about upgrading API tiers manually
  • Pricing varies by model (fixed or variable depending on which model you select)
  • Cost visibility in Clay UI shows data credit consumption

Personal API Keys

  • Can reduce data credit costs (you only pay Actions, not Data Credits)
  • Requires meeting provider-specific tier requirements
  • You manage your own API billing directly with the provider

Note: When using a personal API key, price breakdowns won't appear in the Clay UI. You'll need to monitor your usage and upgrade tiers manually.

Explore other docs

Gen AI

Guide: Using Clay plugin

Go-to-market plays you can run from a coding agent with the Clay plugin — source your market, score it for fit, and prioritize who gets worked first.

View article
Gen AI

Audiences for agents and the CLI

Read, segment, and manage your workspace's people, companies, and deals from the Clay CLI — and from the Clay agents that run it for you.

View article
Getting started

Clay MCP for Reps vs. Agent Plugin

Choose between MCP for Reps and the Clay Agent Plugin based on who will be using it and how much your ops team needs to approve in advance.

View article
Enrich

Enrichment in Workflows

Enrich an Audiences segment through a guided workflow that opens with its trigger and write-back step already configured.

View article
Export

Getting started with Sequencer

Sequencer runs cold email campaigns at scale from your Audiences data, handling sending infrastructure, copy, and deliverability in one place.

View article
Export

Connect your own email accounts

Connect email accounts you already own — Google Workspace, Microsoft 365, or any mailbox that supports SMTP — and use them as sending accounts for your campaigns.

View article
Export

Email warmup

What email warmup does, how it differs on purchased and self-connected accounts, and why you'll see unfamiliar emails while it runs.

View article

Other popular resources

Experts

Find a Clay Expert

Explore our network of Clay experts and agencies.

View experts
Community

Join our slack community

Find help in our slack community, and support channels.

Go to slack
Cohorts

Join a cohort, learn Clay fast!

The faster way to master Clay. Sign in if you're enrolled in a cohort (current or past) or apply!

Learn more about cohorts
Talents

Hire GTME Talent

Find and connect with GTM talent who've demonstrated expertise in building advanced workflows

Explore GTME talents