Thinking Machines Lab released the open-source large model Inkling on July 15, featuring a 975 billion parameter MoE architecture that activates approximately 41 billion parameters during inference. It supports text, image, and audio inputs, as well as a context window of 1 million tokens. In the MCP Atlas agent tool evaluation, Inkling achieved a task completion rate of 74.1%, and a completion rate of 77.6% in SWE-Bench Verified. Inkling is open-sourced under the Apache 2.0 license on Hugging Face and is now available on OpenRouter, priced at $1 per million input tokens and $4.05 per million output tokens. It is suitable for agent workflows based on OpenRouter, such as Hermes and OpenClaw, primarily targeting institutional users who cannot use Chinese models and prioritize compliance and customizability.
This content is provided for general informational purposes only and doesn't constitute financial, investment, legal, or tax advice. Any events, rewards, online promotions, or related information mentioned herein should not be considered a recommendation, solicitation, or invitation to purchase, sell, trade, or otherwise deal in any crypto assets. Crypto assets are highly volatile and may result in loss. The availability of WEEX services, products, and related events may vary by region. You are responsible for ensuring that your participation is in accordance with applicable local laws and regulations.





























