Discussion about this post

User's avatar
Maxorgel's avatar

Very important and timely discussion and I completely agree that Europe needs to be more independent and building its own frontier model is not a realistic option.

However, I think it's worth differentiating using strong harnesses set-ups from using open weight models and being specific about the risk of frontier labs stealing companies' business models.

First of all any serious company should already be operating with a setup that strongly relies on a harness and I believe most do. That's independent of what model is in use.

Open weight models might often be the right solution, but we should also talk about the risks. For many industries and tasks even being 3 month behind can be critical. That's especially true for adversarial areas like cybersecurity, trading, … Also we can't know for sure the distance between the frontier and the best open weight models will remain. Also there are safety consideration around open weights that are at least worth evaluating.

Additionally, enterprise users already have contractual agreement with frontier labs that forbid them using their data for training. While the question of some indirect data leakage is probably complex and labs might use demand data I think it's unlikely they would outright steal knowledge of how the firm runs. I don't think Anthropic required insight knowhow of Figma to build Claude Design or of pharma companies to start drug research. Whether one wants to pay a potential competitor is another question, but we've had similar situations with big tech for decases.

All of these questions have to be looked at case by case and they can be with a good harness. There might often be situations were companies chose to use the frontier model for a time if price / value is right and they don't have a deep dependency and high switching cost.

From everything I've read it looks like the US military also often chose to at least partially go with the frontier models. The fight with Anthropic makes no sense if they were happy to use an open weight model.

Scenarica's avatar

The PSD2 analogy is the strongest section and it deserves stress-testing on the point where it breaks. In banking, the data being ported is inert. A transaction history doesn't change when you move it from one bank to another. The file is the file.

In AI, the "data" being ported includes fine-tuning datasets, prompt architectures, evaluation frameworks, agent configurations, and performance baselines that are inherently coupled to the model they were built for. A prompt that works on one model produces different output on another. An evaluation benchmark calibrated to one model's behaviour isn't portable without recalibration. Porting a bank account is copying a file. Porting an AI harness is transplanting a living system that behaves differently on the new substrate.

The portability mandate is right. The implementation is an order of magnitude harder than PSD2 because the thing being ported is model-dependent in ways a bank balance was never bank-dependent. The standard that solves this isn't a data-export format. It's an abstraction layer that separates the firm's institutional logic from the model's specific behaviour. That's what the harness is supposed to do. Whether any current harness actually achieves it in practice is the testable claim underneath the whole strategy.

3 more comments...

No posts

Ready for more?