TWITTER_POST

Otso Veistera is promoting thetokenco, a YC W26 company offering an API for…

Brief

Otso Veistera is promoting thetokenco, a YC W26 company offering an API for compressing LLM inputs before inference. He claims the approach cuts token consumption and latency while improving output quality, and cites customer case studies showing a 5% increase in purchases driven by stronger user preference for responses produced from compressed prompts.

Source evidence

title: @OtsoVeistera: You're wasting half your context window. We’re launching [@thetokenco](https://t...
author: OtsoVeistera
contenttype: twitterpost
published: 2026-03-03T19:00:16+00:00
source_url: https://x.com/OtsoVeistera/status/2028908334657749332

word_count: 60

Tweet by @OtsoVeistera

You're wasting half your context window. We’re launching @thetokenco (YC W26) today. We compress LLM inputs before they reach the model. Fewer tokens, lower cost, faster inference. Models also perform better. In customer case studies we’ve seen a +5% lift in user purchases due to higher preference for outputs from compressed prompts. The API is live. Link in the comments


Posted: 2026-03-03T19:00:16.000Z
Engagement: 509 likes, 80 retweets, 76 replies