How An Inkling From Thinking Machines Could Change AI’s Course
AIThis post was created with the assistance of artificial intelligence (AI).

TL;DR

Thinking Machines has released Inkling, a large open-weight multimodal AI model, openly available on Hugging Face under Apache 2.0. This move challenges norms around model openness and ownership, with implications for AI development and regulation.

Thinking Machines has released its first foundation model, Inkling, openly available on Hugging Face under Apache 2.0 license. This marks a departure from typical industry practices, emphasizing open access over proprietary control, and could influence future AI development and ownership models.

The Inkling model is a Mixture-of-Experts transformer with 975 billion parameters, supporting a 1-million-token context window. It was trained on 45 trillion tokens of multimodal data, including text, images, audio, and video, with a native multimodal input design. The model’s weights are publicly available on Hugging Face under Apache 2.0, allowing users to download, modify, and deploy independently.

Unlike typical model releases, Thinking Machines explicitly stated that Inkling is not the strongest model available but prioritized transparency and open access. The company also indicated that a separate Model Acceptable Use Policy restricts certain applications, such as surveillance or deceptive practices, which complicates the notion of “open source.” The model was trained with a hybrid optimizer and involved over 30 million reinforcement learning rollouts, with some training data generated by open-weight models like Kimi K2.5.

At a glance
reportWhen: announced April 2024
The developmentThinking Machines publicly released its first foundation model, Inkling, under an open license, marking a significant shift in AI model distribution and ownership practices.

Implications of Open-Weight Model Release for AI Ownership

The public release of Inkling under open licensing challenges the industry norm of proprietary models and raises questions about ownership, control, and regulation. By making the weights freely available, Thinking Machines enables organizations to fine-tune, inspect, and deploy the model independently, potentially accelerating innovation but also complicating oversight and safety measures.

This move could influence industry standards around transparency and open access, especially as concerns grow over AI safety, misuse, and accountability. However, the presence of a separate use policy layered on top of the open license introduces ambiguity about the scope of permissible uses, which could impact adoption and trust.

AI Systems Performance Engineering: Optimizing Model Training and Inference Workloads with GPUs, CUDA, and PyTorch

AI Systems Performance Engineering: Optimizing Model Training and Inference Workloads with GPUs, CUDA, and PyTorch

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Industry Norms and the Shift Toward Open Models

Traditionally, large foundation models are either proprietary or released with limited access, often through APIs. Recent efforts, like Meta’s Llama 2, have moved toward open weights, but with restrictions. Inkling’s release under Apache 2.0, combined with candid details about training and performance, marks a notable shift toward transparency and open ownership.

Thinking Machines, founded by former OpenAI CTO and staffed with team members involved in ChatGPT’s development, aims to challenge existing paradigms by prioritizing open access and honesty about model capabilities. The company’s decision to publish full weights immediately, rather than a closed API, underscores a different approach to AI deployment and ownership.

“Our goal is to empower developers and organizations with full access, while maintaining responsible use through our policies.”

— Thinking Machines spokesperson

Amazon

multimodal AI development kits

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Unresolved Questions About Inkling’s Use and Control

It remains unclear how the separate Model Acceptable Use Policy will be enforced and whether it will effectively limit misuse. The distinction between open weights and restrictions layered on top raises questions about legal enforceability and practical control.

Additionally, the full scope of the training data and the potential for fine-tuning or repurposing by third parties are still to be examined. The impact of this release on industry standards and regulatory responses is also uncertain.

Amazon

large language model fine-tuning tools

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Next Steps for Adoption and Industry Response

Expect independent researchers and organizations to test and benchmark Inkling’s capabilities further, especially in safety and bias. The company may also publish more details on the training data and use policy.

Regulators and industry groups could respond by reconsidering licensing norms and developing standards for open-weight models, especially regarding safety and misuse prevention. The model’s adoption and real-world impact will become clearer as it is integrated into applications.

Domain-Specific Small Language Models: Efficient AI for local deployment

Domain-Specific Small Language Models: Efficient AI for local deployment

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Key Questions

What makes Inkling different from other large language models?

Inkling is a 975-billion-parameter multimodal model released openly under Apache 2.0, allowing independent use and modification, unlike most proprietary models.

Does open weights mean the model is entirely open source?

No. While the weights are openly available, the training data and full training pipeline are not published, and a separate use policy may impose restrictions.

What are the potential risks of releasing such a large open model?

Risks include misuse for harmful applications, bias amplification, and challenges in enforcing use policies. The layered restrictions add complexity to managing these risks.

How might this release influence the AI industry?

It could set a precedent for more open models and shift norms toward transparency, but also prompts debate over ownership, safety, and regulation.

Source: ThorstenMeyerAI.com

You May Also Like

Apple Is Reaching for Chinese Memory. Europe Doesn’t Even Have That Option.

Apple is lobbying US authorities to buy memory chips from China’s CXMT, exposing Europe’s lack of options in the global chip supply chain and its vulnerabilities.

Cybersecurity operations signal monitor: A backdoor in a LinkedIn job offer

Cybersecurity operations signal monitor identifies a backdoor in a LinkedIn job posting, raising concerns about targeted cyber threats and organizational security.

The Safety Card, Played From Every Side: David Sacks, Anthropic, and the Fable Standoff

White House adviser David Sacks claims Anthropic refused to fix a cybersecurity flaw, leading to model bans. Anthropic disputes this, highlighting ongoing safety debates.

Galaxy Announces The Appointment Of Steven Bandrowczak To Board Of Directors

Galaxy announces the appointment of Steven Bandrowczak to its Board of Directors, strengthening its leadership team. The move is effective immediately.