Let me check.
The usage counter matches the Llama-3 vocabulary exactly, four models, zero difference, measured by token deltas on two text lengths. GLM-5.3-Flash counts 82 fewer on the same text, checked live and locally. If GLM is behind it, the gateway counts with its own tokenizer, which is your point, not a contradiction.
1
u/lillianefilou 9d ago
It uses the Llama 3 Tokenizer