Loading startup

DealFlow OS
Market indexPublic companiesMorning briefingTerminal
Request accessSign in

DealFlow

Terminal access

IndexTop moversSignals

Public markets

Public companiesMorning briefing
Browse index
T

The Token Company

Pre-SeedFunding signal
thetokencompany.com/B2BCA, USA
Market index

Market data is refreshed once per day from public sources. Information may be incomplete or outdated — verify independently before making decisions. This is not investment advice.

Investor read

Evidence-bound summary — expand sections for movement, risks, and signals.

Memo snapshot · May 20, 2026, 6:12 PM

Worth a meetingEvidence-bound analyst verdict
  • •Funding: Raised $500K across 1 funding round. Latest: $500K Pre-seed (Jan 2026). Investors: Y Combinator.
  • •Named backers in indexed sources: Y Combinator.
  • •Hiring: 1 hiring‑related row(s); role‑spam risk if mostly generic boards
  • •Product/news: 8 product/news‑styled row(s); headline risk without filings

DealFlow OS uses public web data and automated enrichment. Research may be incomplete, outdated, or incorrect. Verify important information before making investment or outreach decisions.

TL;DR

Seed (YC)

The Token Company

Funding

Raised $500K across 1 funding round. Latest: $500K Pre-seed (Jan 2026). Investors: Y Combinator. (High).

Quick read

  • •The Token Company
  • •Reported angle: Why Your RAG App's Token Bill Is So High

Key signals

Funding

Raised $500K across 1 funding round. Latest: $500K Pre-seed (Jan 2026). Investors: Y Combinator. (High).

Hiring

1 hiring‑related row(s); role‑spam risk if mostly generic boards (Low).

Product / news

8 product/news‑styled row(s); headline risk without filings (High).

Evidence summary

Verified facts

  • •The Token Company
  • •Reported angle: Why Your RAG App's Token Bill Is So High

Recent movers

  • •May 11, 2026 · Blog / news — Why Your RAG App's Token Bill Is So High
  • •May 11, 2026 · Blog / news — The Token Company - LLM Input Compression with bear-1 & bear-1.1
  • •May 11, 2026 · Blog / news — How input compression enabled world class performance for long running agents

+5 more in Recent movement below

  • •May 11, 2026 · Blog / news — Why Your RAG App's Token Bill Is So High
  • •May 11, 2026 · Blog / news — The Token Company - LLM Input Compression with bear-1 & bear-1.1
  • •May 11, 2026 · Blog / news — How input compression enabled world class performance for long running agents
  • •May 11, 2026 · Blog / news — Cut Your LLM API Costs by 70% Without Losing Quality
  • •May 11, 2026 · Blog / news — bear-1.1: Improved LLM Compression Model - The Token Company
  • •May 11, 2026 · Blog / news — bear-1: LLM Input Compression Model - The Token Company
  • •May 11, 2026 · Blog / news — Blog - The Token Company
  • •May 20, 2026 · ycombinator.com — The Token Company | Y Combinator
  • •Y Combinator

Suggested next steps

  • ▸Open a founder conversation this month; validate traction claims against the indexed evidence.
  • ▸Add to the active watchlist so new signals trigger alerts.

Funding & hiring signals

funding_articleConfidence: low

The Token Company | Y Combinator

Winter 2026 (W26) batch. LLM input compression middleware (bear-1 / bear-1.1) cutting tokens ~66% in <100ms.

Affected score: noObserved: May 20, 2026

Weak funding evidence: mention does not include a parsed raise amount/round and may refer to non-funding context.

Open roles (indexed)

No open roles indexed yet.

Public index
59Activity 98/100Strong public activity signal
7D+17 (+40.5%)
30D+16 (+37.2%)
High Confidence

The index price and activity score are algorithmic estimates based on observed public company-level signals. They may be incomplete, stale, or inaccurate and are not investment, legal, tax, or business advice.

Source health

  • githubok
    Last checked Mon, Jun 29, 10:26 AM
  • public_page:homeok
    Last checked Mon, Jun 29, 10:26 AM
  • public_page:careersok
    Last checked Mon, Jun 29, 10:26 AM
  • public_page:pricingok
    Last checked Mon, Jun 29, 10:26 AM
  • public_page:blogok
    Last checked Mon, Jun 29, 10:26 AM
  • public_market_enrichmentok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_careersok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blog_pax-historiaok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blog_helonicok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blog_coqaok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blog_bear-2-safetyok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blogok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_ok
    Last checked Mon, Jun 29, 09:16 AM
  • public_page:_blog_why-rag-token-costs-are-highok
    Last checked Mon, May 11, 04:39 AM
  • public_page:_blog_cut-llm-api-costsok
    Last checked Mon, May 11, 04:39 AM
  • public_page:_blog_bear-1-1ok
    Last checked Mon, May 11, 04:39 AM
  • public_page:_blog_bear-1ok
    Last checked Mon, May 11, 04:39 AM

Signal timeline

21 dated public signals · newest on the right

Funding
Hiring
Product
Code
Press
Company
Other
Dec 25, 2025Apr 2, 2026Jul 9, 2026
Funding (1)Hiring (2)Product (1)Code (1)Press (10)Company (4)Other (2)

Signal breakdown

Latest momentum signal per category. Expand a card to inspect raw payloads.

Public source summary

Total evidence rows
21
Latest evidence
Thu, Jul 2, 04:02 PM

Source types found

blogcareers_pagefunding_articlegithubofficial_siteotherpressproduct

Strongest / recent news-style rows

  • The Token Company | Y Combinator

    Wed, May 20, 06:12 PM · confidence 88%high quality

  • The Token Company - LLM Input Compression with bear-1 & bear-1.1

    Pricing · Sun, May 10, 10:48 PM · confidence 90%high quality

  • The Stoker Company

    wikipedia · Mon, Jun 29, 10:26 AM · confidence 50%medium quality

Public signal timeline

Newest first · 21 event(s)

1
Thu, Jul 2, 04:02 PM · blog · 90% · publichigh qualityPrompt Compression API — Cut LLM Token Costs

Source: Blog

Cut OpenAI, Anthropic, and Gemini API costs with accuracy held flat — or take a smaller cut and lift accuracy by several points instead. The bear-2 prompt compression API strips low-signal tokens from your inputs before they hit the LLM. Works with GPT, Claud…

Source ↗
2
Thu, Jul 2, 04:02 PM · official_site · 90% · publichigh qualityPrompt Compression API — Cut LLM Token Costs

Source: Homepage

Cut OpenAI, Anthropic, and Gemini API costs with accuracy held flat — or take a smaller cut and lift accuracy by several points instead. The bear-2 prompt compression API strips low-signal tokens from your inputs before they hit the LLM. Works with GPT, Claud…

Source ↗
3
Mon, Jun 29, 10:26 AM · official_site · 75% · publichigh qualityThe Token Company Jobs

Source: official_site

The Token Company Jobs You need to enable JavaScript to run this app.

Source ↗
4
Mon, Jun 29, 10:26 AM · official_site · 75% · publichigh qualityPrompt Compression API — Cut LLM Token Costs

Source: official_site

Prompt Compression API — Cut LLM Token Costs | The Token Company Dark Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent) 512x512 PNG (transparent) 1024x1024 PNG (transparent) Light Mode 1280x128 SVG 1280x128 PNG (transparent) 2…

Source ↗
5
Mon, Jun 29, 10:26 AM · official_site · 75% · publichigh qualityPrompt Compression API — Cut LLM Token Costs

Source: official_site

Pricing | The Token Company Dark Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent) 512x512 PNG (transparent) 1024x1024 PNG (transparent) Light Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent)…

Source ↗
6
Mon, Jun 29, 09:16 AM · blog · 90% · publichigh qualityOne of the biggest token consumers globally improved quality by removing context bloat

Source: Blog / news

Pax Historia, processing 193B tokens/month on OpenRouter, ran a 268K-vote model arena with bear-1.1 compression. Compressed models scored higher and A/B tests showed +5% purchase amount lift.

Source ↗
7
Mon, Jun 29, 09:16 AM · blog · 90% · publichigh qualityHow input compression enabled world class performance for long running agents

Source: Blog / news

Helonic runs AI agents on construction drawings at near million-token prompts. bear-1.2 compression trims tokens while preserving every critical detail.

Source ↗
8
Mon, Jun 29, 09:16 AM · blog · 90% · publichigh qualityCompressing Conversational Context Without Losing the Thread

Source: Blog / news

Bear-2 compression improved CoQA accuracy from 93.3% to 95.3% while cutting tokens by 8.2%. Removing filler helps the model focus.

Source ↗
9
Mon, Jun 29, 09:16 AM · blog · 90% · publichigh qualityIntroducing Bear-2-Safety

Source: Blog / news

Bear-2-Safety compresses input to safety classifiers by up to 30% while preserving or improving F1 across Llama-Guard, ShieldGemma, and Gemini. Unsafe content is preserved at 95-100% retention.

Source ↗
10
Wed, May 20, 06:12 PM · funding_article · 88% · verified_publichigh qualityThe Token Company | Y Combinator

Winter 2026 (W26) batch. LLM input compression middleware (bear-1 / bear-1.1) cutting tokens ~66% in <100ms.

Source ↗
11
Mon, May 11, 04:39 AM · careers_page · 90% · publichigh qualityThe Token Company - LLM Input Compression with bear-1 & bear-1.1

Source: Careers

Join The Token Company. We're hiring ML engineers, software engineers, and more to build the compression layer for LLMs.

Source ↗
12
Mon, May 11, 04:39 AM · blog · 90% · publichigh qualityWhy Your RAG App's Token Bill Is So High

Source: Blog / news

Most RAG applications over-fetch context by 3-5x. Break down where your retrieval tokens go and how compression, reranking, and chunk hygiene cut costs.

Source ↗
13
Mon, May 11, 04:39 AM · blog · 90% · publichigh qualityCut Your LLM API Costs by 70% Without Losing Quality

Source: Blog / news

A practical framework for reducing LLM API costs across system prompts, conversation history, RAG context, tool schemas, and output. Real pricing math included.

Source ↗
14
Mon, May 11, 04:39 AM · blog · 90% · publichigh qualitybear-1.1: Improved LLM Compression Model - The Token Company

Source: Blog / news

bear-1.1 is the latest LLM input compression model from The Token Company. An improved version of bear-1 with better accuracy preservation and faster compression speeds. Reduce AI costs by 3x.

Source ↗
15
Mon, May 11, 04:39 AM · blog · 90% · publichigh qualitybear-1: LLM Input Compression Model - The Token Company

Source: Blog / news

bear-1 is an LLM input compression model by The Token Company that reduces tokens by 66% while maintaining or improving accuracy. Reduce AI costs by 3x with a single API call.

Source ↗
16
Sun, May 10, 10:48 PM · product · 90% · publichigh qualityThe Token Company - LLM Input Compression with bear-1 & bear-1.1

Source: Pricing

Simple pricing for LLM input compression. Free tier to get started, Pro for production, Enterprise for scale.

Source ↗
17
Thu, Jul 2, 04:02 PM · careers_page · 70% · publicmedium qualityThe Token Company Jobs

Source: Careers page

The Token Company Jobs

Source ↗
18
Mon, Jun 29, 10:26 AM · github · 55% · publicmedium qualitythe-token-company

Source: npm_registry

Node.js SDK for The Token Company — compress LLM prompts to reduce costs and latency

Source ↗
19
Mon, Jun 29, 10:26 AM · other · 50% · publicmedium qualityThe Stoker Company

Source: wikipedia

Source ↗
20
Mon, Jun 29, 10:26 AM · other · 50% · publicmedium qualityThe Timken Company

Source: wikipedia

Source ↗
21
Wed, May 20, 06:12 PM · press · 85% · verified_publicmedium qualityLaunch YC: The Token Company

Compresses 100k tokens in under 100ms.

Source ↗

Official / company site

4 row(s)

The company's own site — the authoritative description of what they sell and to whom. Marketing-controlled, so treat claims as positioning rather than verified traction.

official_site·Thu, Jul 2, 04:02 PM·Confidence 90%high qualitypublic
Prompt Compression API — Cut LLM Token Costs

Cut OpenAI, Anthropic, and Gemini API costs with accuracy held flat — or take a smaller cut and lift accuracy by several points instead. The bear-2 prompt compression API strips low-signal tokens from your inputs before they hit the LLM. Works with GPT, Claud…

Why it matters: Primary source — the company's own positioning; best read for what they sell and to whom, not for traction claims.

Open source ↗
official_site·Mon, Jun 29, 10:26 AM·Confidence 75%high qualitypublic
The Token Company Jobs

The Token Company Jobs You need to enable JavaScript to run this app.

Why it matters: Primary source — the company's own positioning; best read for what they sell and to whom, not for traction claims.

Open source ↗
official_site·Mon, Jun 29, 10:26 AM·Confidence 75%high qualitypublic
Prompt Compression API — Cut LLM Token Costs

Prompt Compression API — Cut LLM Token Costs | The Token Company Dark Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent) 512x512 PNG (transparent) 1024x1024 PNG (transparent) Light Mode 1280x128 SVG 1280x128 PNG (transparent) 2…

Why it matters: Primary source — the company's own positioning; best read for what they sell and to whom, not for traction claims.

Open source ↗
official_site·Mon, Jun 29, 10:26 AM·Confidence 75%high qualitypublic
Prompt Compression API — Cut LLM Token Costs

Pricing | The Token Company Dark Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent) 512x512 PNG (transparent) 1024x1024 PNG (transparent) Light Mode 1280x128 SVG 1280x128 PNG (transparent) 256x256 SVG 256x256 PNG (transparent)…

Why it matters: Primary source — the company's own positioning; best read for what they sell and to whom, not for traction claims.

Open source ↗

News

4 row(s)

Third-party press coverage. Independent reporting corroborates company claims; repeated coverage across outlets is a momentum signal.

product·Sun, May 10, 10:48 PM·Confidence 90%high qualitypublic
The Token Company - LLM Input Compression with bear-1 & bear-1.1

Simple pricing for LLM input compression. Free tier to get started, Pro for production, Enterprise for scale.

Why it matters: Independent coverage — third-party corroboration of company claims; recurring coverage indicates rising visibility.

Open source ↗
other·Mon, Jun 29, 10:26 AM·Confidence 50%medium qualitypublic
The Stoker Company

Why it matters: Independent coverage — third-party corroboration of company claims; recurring coverage indicates rising visibility.

Open source ↗
other·Mon, Jun 29, 10:26 AM·Confidence 50%medium qualitypublic
The Timken Company

Why it matters: Independent coverage — third-party corroboration of company claims; recurring coverage indicates rising visibility.

Open source ↗
press·Wed, May 20, 06:12 PM·Confidence 85%medium qualityverified_public
Launch YC: The Token Company

Compresses 100k tokens in under 100ms.

Why it matters: Independent coverage — third-party corroboration of company claims; recurring coverage indicates rising visibility.

Open source ↗

Funding / news

1 row(s)

Funding announcements and investor-database records. The strongest public signal of capitalization: round, amount, and syndicate quality when disclosed.

funding_article·Wed, May 20, 06:12 PM·Confidence 88%high qualityverified_public
The Token Company | Y Combinator

Winter 2026 (W26) batch. LLM input compression middleware (bear-1 / bear-1.1) cutting tokens ~66% in <100ms.

Why it matters: Funding signal — Pre-Seed per this source; verify against the linked original before relying on it.

Open source ↗

Hiring

2 row(s)

Open roles and careers pages. Active hiring implies runway to spend and shows where the company is investing (engineering vs GTM vs ops).

careers_page·Mon, May 11, 04:39 AM·Confidence 90%high qualitypublic
The Token Company - LLM Input Compression with bear-1 & bear-1.1

Join The Token Company. We're hiring ML engineers, software engineers, and more to build the compression layer for LLMs.

Why it matters: Hiring signal — open roles imply runway to spend and show where the company is investing.

Open source ↗
careers_page·Thu, Jul 2, 04:02 PM·Confidence 70%medium qualitypublic
The Token Company Jobs

The Token Company Jobs

Why it matters: Hiring signal — open roles imply runway to spend and show where the company is investing.

Open source ↗

GitHub

1 row(s)

Public engineering activity. Sustained commits, releases, and stars indicate real product development and, for dev tools, developer adoption.

github·Mon, Jun 29, 10:26 AM·Confidence 55%medium qualitypublic
the-token-company

Node.js SDK for The Token Company — compress LLM prompts to reduce costs and latency

Why it matters: Engineering signal — public repo activity evidences active development and possible developer adoption.

Open source ↗

Blog

9 row(s)

Company blog and newsletters. Shipping cadence and technical depth of posts hint at product velocity and team quality.

blog·Thu, Jul 2, 04:02 PM·Confidence 90%high qualitypublic
Prompt Compression API — Cut LLM Token Costs

Cut OpenAI, Anthropic, and Gemini API costs with accuracy held flat — or take a smaller cut and lift accuracy by several points instead. The bear-2 prompt compression API strips low-signal tokens from your inputs before they hit the LLM. Works with GPT, Claud…

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, Jun 29, 09:16 AM·Confidence 90%high qualitypublic
One of the biggest token consumers globally improved quality by removing context bloat

Pax Historia, processing 193B tokens/month on OpenRouter, ran a 268K-vote model arena with bear-1.1 compression. Compressed models scored higher and A/B tests showed +5% purchase amount lift.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, Jun 29, 09:16 AM·Confidence 90%high qualitypublic
How input compression enabled world class performance for long running agents

Helonic runs AI agents on construction drawings at near million-token prompts. bear-1.2 compression trims tokens while preserving every critical detail.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, Jun 29, 09:16 AM·Confidence 90%high qualitypublic
Compressing Conversational Context Without Losing the Thread

Bear-2 compression improved CoQA accuracy from 93.3% to 95.3% while cutting tokens by 8.2%. Removing filler helps the model focus.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, Jun 29, 09:16 AM·Confidence 90%high qualitypublic
Introducing Bear-2-Safety

Bear-2-Safety compresses input to safety classifiers by up to 30% while preserving or improving F1 across Llama-Guard, ShieldGemma, and Gemini. Unsafe content is preserved at 95-100% retention.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, May 11, 04:39 AM·Confidence 90%high qualitypublic
Why Your RAG App's Token Bill Is So High

Most RAG applications over-fetch context by 3-5x. Break down where your retrieval tokens go and how compression, reranking, and chunk hygiene cut costs.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, May 11, 04:39 AM·Confidence 90%high qualitypublic
Cut Your LLM API Costs by 70% Without Losing Quality

A practical framework for reducing LLM API costs across system prompts, conversation history, RAG context, tool schemas, and output. Real pricing math included.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, May 11, 04:39 AM·Confidence 90%high qualitypublic
bear-1.1: Improved LLM Compression Model - The Token Company

bear-1.1 is the latest LLM input compression model from The Token Company. An improved version of bear-1 with better accuracy preservation and faster compression speeds. Reduce AI costs by 3x.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗
blog·Mon, May 11, 04:39 AM·Confidence 90%high qualitypublic
bear-1: LLM Input Compression Model - The Token Company

bear-1 is an LLM input compression model by The Token Company that reduces tokens by 66% while maintaining or improving accuracy. Reduce AI costs by 3x with a single API call.

Why it matters: Company publishing — post cadence and depth hint at product velocity.

Open source ↗

Private workspace

Sign in as an active team member to view private notes, watchlist controls, transcript evidence, and interaction history.

Sign in
DealFlow OS · Public market terminal
Privacy PolicyTerms & Conditions