By scaling compute on user context, we reduce token spend. But it's about more than lowering cost. To develop expertise is to reduce the energy it takes to solve a problem, freeing capacity to solve harder problems yet
Was great chatting about this with the amazing @LM_Braswell!
Kleiner Perkins (@kleinerperkins)
The models we use every day are brilliant strangers. They forget your organization the moment a chat ends, then relearn it on the next query.
@EngramLab fixes that. It learns your world once and reuses that memory, matching frontier systems on 1-10% of the tokens.
@Microsoft, @NotionHQ, and @Harvey are already testing it within their organizations.
Congratulations to the team, and hear directly from @danbiderman (CEO and co-founder) and Sabri Eyuboglu (CTO and co-founder) with @LMBraswell ⬇️
Video
— https://nitter.net/kleinerperkins/status/2069788333937688818#m