9 October 2026

ttok 1.0 released, fixing a default tokenizer problem

First reported

Simon Willison ran this on .

  • A new version, ttok 1.0, has been released as a tool for counting text in units the AI model reads.
  • The update was prompted by a bug where the tool defaulted to a tokenizer (the piece that splits text into chunks for the model) from GPT-4, an OpenAI model.
  • The author noticed the problem after upgrading to the earlier ttok 0.4 version.

How it was covered

Simon WillisonDaily notes and links

The ttok 1.0 release and a default-tokenizer problem found in 0.4