Pay for what you use.
Credits included on every plan.

Every plan comes with a monthly balance of credits. Calls draw from it at the same rates whichever plan you are on, and only unique content is ever billed.

Tried almost everything — structured memory files, qmd etc. The only thing that works reliably is supermemory.

Not facing any memory issues after setting it up.

Harshil MathurFounder, Razorpay

100B+tokens a month187msmedian recall

Plans

  1. Pro

    $19per month

    Get Pro

    $20included per month

    For developers building with AI memory.

    Everything in Free, plus

    • Pay as you go (auto-top ups with spend limits)
    • 3 team seats, unlimited end users
    • Priority support
  2. Max

    Popular

    $100per month

    Get Max

    $130included per month

    For developers who need more headroom than Pro.

    Everything in Pro, plus

    • 6× the credits of Pro
    • Gmail connector
  3. Scale

    $399per month

    Get Scale

    $600included per month

    For teams and production workloads.

    Everything in Max, plus

    • S3 and web crawler connectors
    • Unlimited team seats, per-tag access
    • Dedicated support
    • SOC 2 · HIPAA BAA · self-hosted option

Free

$0 per month · $5 of credits included, renewed every month

Start free

Enterprise

Custom deployments with dedicated engineering.

For organizations with committed spend, custom deployments and their own security requirements. Unlimited usage, metered your way.

Talk to sales

Security and compliance

  • Air-gapped self-hosting or a dedicated managed instance
  • SOC 2 · HIPAA · GDPR
  • Custom contracts and DPA
  • SSO and custom integrations
  • Enterprise MCP

Scale and performance

  • Unlimited usage on committed-spend pricing
  • Custom metering and billing
  • Custom rate limits and throughput
  • Dedicated infrastructure

Support

  • Dedicated account manager
  • Forward-deployed engineer
  • 1:1 onboarding and integration
  • Uptime SLA
  • Priority Slack channel

Rate card

Billed in SM tokens, the unique tokens we actually ingest. Repeats and unchanged content are free, and the rates are the same on every plan.

MemoryA memory graph per user: profiles and fact hierarchies that agents learn from in real time. Powered by our own model.

Plain text$5/ 1M SM tokens

Rich content$10/ 1M SM tokens

SuperRAGMultimodal extraction, contextual chunking and retrieval. No embeddings or vectors to manage. Also available as a filesystem. smfs.ai

Text$1/ 1M SM tokens

Rich: images, PDFs, audio, video$2/ 1M SM tokens

Search and traversalSemantic search and graph traversal across stored content, in one call. 187ms median server time, built for agent loops.

Per query$5/ 1M queries

OperationsRe-ranking, aggregation, query rewriting and other building blocks for richer queries.

Per operation$100/ 1M operations

Startups and research

The Scale plan free for three months for qualifying early-stage startups and academic research teams, with its included usage and dedicated support.

Apply

Questions

How does the credit balance work?
Every plan comes with a monthly dollar balance. Each API call, whether storing memories, searching or indexing, draws from that balance at the rates above. Top-ups add to it whenever you like.
What is an SM token?
SM tokens are the unique tokens supermemory actually ingests and embeds. Repeats and unchanged content are not billed again, which for most workloads means a large discount on the raw token count.
What happens if I store the same content twice?
Nothing is billed. supermemory deduplicates at the token level, so re-uploading a document, re-syncing a connector or pushing the same conversation history costs nothing the second time.
Do unused credits roll over?
Subscription credits reset monthly. Credits you buy as a top-up never expire; they stay in your balance until you use them.
What happens when I run out?
On Free, usage pauses until the next month's credits arrive; there is no pay as you go. On Pro, Max and Scale, usage beyond the included credits is pay as you go at the same rates: auto top-up refills the balance before it runs dry, and spend limits keep a runaway agent from surprising you.
Can I self-host?
Self-hosted deployments are available on Scale and Enterprise. Enterprise also supports fully air-gapped deployments, where LLM inference may be the only outbound call, or none at all.
Do you offer startup or research credits?
Yes. Qualifying early-stage startups and academic research teams get the Scale plan free for three months, including its included usage, plus dedicated support.