Hackernews
new
show
ask
jobs
GLM-5.3-FlashX: Delivering inference speeds of 200 tokens/s
3 points
posted 6 hours ago
by theanonymousone
(docs.z.ai)
No comments yet