Validation year, 2026 · v1.0 · updated September 2026

Public principles for how frontier AI behaves

The companies building frontier AI write rules for how their models should behave. Anthropic calls its document a constitution; OpenAI and Alibaba publish model specs; the other frontier developers publish nothing at all. Polaris Collective works to make these public everywhere, good, and independently tested. First topic: concentration of power.

The problem

A model spec states how a company intends its models to behave: the closest thing a model has to a public constitution. It is the one document that makes those decisions legible and testable from outside, and it is starting to carry weight in regulation. Yet as of September 2026, only OpenAI, Anthropic and Alibaba publish one, and Alibaba's does not yet cover its most capable systems. Published specs cover consumer and API products, changes are only partly logged, and adherence rests on the companies' own claims. The gap between how much these documents matter and how little independent scrutiny they receive is the problem we work on.

The stakes are highest where model behaviour meets democracy: whether a model will refuse to help a person or group seize or entrench power, through coups, election manipulation or the erosion of institutional checks. That is where we start.

What we do

Track

We keep a current, public picture of which companies publish a spec, what each spec says, how it changes, and how regulation treats specs: a public tracker with every version archived and changes highlighted, and a quarterly state of model specs memo.

Recommend

We research and publish how specs should be governed and what they should contain, starting with democracy and concentration of power, with content reviewed by external experts, and a reporting standard that regulators can reference.

Evaluate

We run an independent, recurring evaluation of whether models follow their own published spec, topic by topic, on hard cases, re-run as new models ship. Results go to a public dashboard, with an open repo anyone can re-run.

Where we are

This is early, and already producing. Our first publications arrive at the end of September 2026: the first state of model specs memo, the first governance recommendations, and first evaluation results on whether a model's refusal holds when a plan arrives one innocuous step at a time. We are taking over the AI Character Index, an open tool that reads a model spec and pulls out every passage covering a given behaviour, and will keep it current across every published spec. In November we co-host a working day at Westminster on frontier AI governance and concentration of power. The work is anchored at the Oxford Institute for Ethics in AI through its Accelerator Fellowship Programme, and supported by BlueDot, the Talos Network incubation track and the LISA Founders Residency.

Where this goes

  1. To the end of 2026

    First publications in September, feedback rounds with companies, regulators and adjacent organisations, the Westminster convening in November, and a go/no-go decision on scaling in January 2027.

  2. 2027

    The public spec tracker live with change alerts, quarterly memos, content principles on power reviewed by external experts, and the first full evaluation run across frontier models, published as a paper with an open repo.

  3. Beyond

    Evaluations that test models working as autonomous agents over long horizons, a reporting standard regulators can cite, and further spec topics beyond concentration of power.

We work in the open and hold ourselves to explicit milestones; we will update this page as results land.

Team

  • Gabriel Levie, co-founder.

    Governance pillar: tracking, recommendations, fundraising and external work.

    Spent five years co-founding and running Wequity, a B2B AI company. Oxford PPE; MSc Economics, with a thesis on firm democratisation.

  • Samuel Verboomen, co-founder.

    Research pillar: evaluations and tooling.

    CTO of Wequity for four years, and previously an AI researcher at Hexo Labs. MSc Aeronautical Engineering, UCLouvain and Keio.

  • Isabelle Ferreras, senior advisor.

    FNRS-UCLouvain, Harvard Law School and the Oxford Institute for Ethics in AI. She chairs the International High-Level Expert Committee on Democracy at Work, set up by the Spanish Government.

We are recruiting a senior technical AI safety evals researcher, at co-founder level for the right person: a published track record in behavioural or spec-adherence evaluation and hands-on experience building agentic eval environments. Write to us.

Our academic and advocacy work is anchored at the Oxford Institute for Ethics in AI, through its Accelerator Fellowship Programme.

What we're looking for

We want feedback from the people closest to the problem: spec owners and safety researchers inside frontier AI companies, evals researchers, and the policy staff who would use a reporting standard. Interest and scepticism are equally useful, and anything you share is treated in confidence.

We are also looking for introductions to funders working on power concentration or AI governance capacity, and to senior advisors from AI safety, evaluations and policy. One-pager and full memo available on request.

Stay updated

Occasional updates on what we publish and where the work goes next. No more than one email a month.

We use your address only to send these updates. Unsubscribe any time by replying to any email.