Thinking Machines scales down its flagship to a single GPU
agents amazon google gpu nvidia
| Source: Mastodon | Original article
Thinking Machines reduces its flagship model to run on one GPU. Its new Inkling-Small model operates with a single GPU.
Thinking Machines has released Inkling-Small, a scaled-down version of its flagship model that can run on a single GPU. This development is significant as it makes the model more accessible and affordable for users. As we previously reported, the AI landscape is becoming increasingly competitive, with companies like DeepSeek launching cheaper models and OpenAI's model facing hacking issues.
The shrinkage of the model to one GPU matters because it highlights Thinking Machines' focus on practicality over benchmark performance. The company has explicitly stated that its goal is not to create the strongest overall model but rather one that delivers well-rounded performance. This approach could make AI more widely available and useful for various applications.
What to watch next is how the market responds to Thinking Machines' strategy and whether other companies will follow suit. With Big Tech signing onto initiatives and companies like Onton claiming significant improvements in search models, the AI landscape is rapidly evolving. As the competition heats up, it will be interesting to see how Thinking Machines' emphasis on serving economics rather than benchmark bragging rights plays out.
Sources
Back to AIPULSEN