• Towards AGI
  • Posts
  • Openness Just Became a Boardroom Metric

Openness Just Became a Boardroom Metric

Data Beats Compute Now.

Today, we’re diving into:

  • AI news: Your AI Model Is Only Data

  • Hot Tea: The OpenAI Debate Leaders Need

  • OpenAI: The Chat Box Era Has Ended

Dear Folks,

This briefing draws on three recent developments spanning AI training data economics, open source AI governance, and agent infrastructure strategy, reflecting the latest publicly reported industry developments. It is intended to support strategic planning, technology adoption, vendor evaluation, and informed decision-making across your organisation.

Your Next AI Model Is Only As Good As Its Data

You are betting millions on AI initiatives this year. But the real bottleneck isn't compute anymore. It's data, and the numbers behind that shift should change how you plan your next twelve months.

The $500 Million Wake-Up Call You Cannot Ignore

A four-year-old AI data-labeling startup just grew its gross annual run rate from $100 million to $500 million in eight months. That is not a typo. That is the speed at which demand for high-quality training data is accelerating across the industry.

You retain enterprise teams, doctors, lawyers, and scientists to validate and structure this data. Even at 60 to 70 percent net margins, the underlying business is scaling faster than most software categories you track today.

Signal to watch: Researchers now believe future AI spending on data could rival total spending on compute. If that plays out, your data strategy becomes as strategic as your cloud contract.

Why Compute Alone Won't Save Your AI Roadmap

For years, you were told that more GPUs meant better models. That era is ending. Multiple data-labeling companies are now scaling revenue faster than many infrastructure vendors, because raw compute without curated, expert-reviewed data produces mediocre outputs.

You have likely felt this already. Models trained on generic datasets plateau quickly. The leaders pulling ahead are the ones who treat data quality as a board-level priority, not an engineering afterthought.

What This Shift Actually Means For You

Contract sizes across the data industry are growing, and margins are expanding as providers generate synthetic data without human involvement. Some of that same data now sells to multiple buyers, pushing gross margins as high as 80 to 90 percent.

The Upside Nobody Is Talking About

Here is the part that should energize you, not worry you. As specialized data providers scale, your enterprise gets faster access to higher-quality, domain-specific training data at a lower relative cost than building this capability in-house.

  • Faster deployment cycles: Purpose-built datasets cut the time your teams spend on data cleanup and validation.

  • Lower total cost of ownership: Shared, off-the-shelf datasets reduce the need for you to fund duplicate data operations internally.

  • Better model reliability: Expert-reviewed data from real domain professionals reduces hallucinations in regulated, high-stakes workflows.

  • Competitive parity with faster movers: You no longer need a billion-dollar data operation to access enterprise-grade training pipelines.

This is the quiet advantage smart leaders are already using to close the gap with better-funded competitors, without building an internal data army.

The Trust Question You Cannot Skip

Not every data provider operates the same way. Some sell identical datasets to adversarial markets, a practice that has already drawn public criticism from industry founders themselves. If you are sourcing training data externally, provenance and buyer restrictions deserve a line item in your vendor due diligence, right next to security and compliance.

Where Enterprise Leaders Should Focus Next

Treat your data supply chain the way you treat any critical vendor relationship. Ask where the data originates, who else has access to it, and how quickly your provider can scale with you as your models mature.

Fix Your Data Blind Spot

Rivals are locking down data governance right now. Don't let a vendor gap cost you the next deal.

The organizations winning the next phase of AI adoption will not be the ones with the most GPUs. They will be the ones who understood, early, that data quality is the actual competitive moat.

The OpenAI Debate Every Leader Must Understand Now

You have probably picked a side in the open versus closed AI debate without realizing how much nuance you are missing. A new framework just changed the terms of that conversation, and it directly affects how you evaluate vendors.

The Framework Quietly Reshaping Vendor Trust

For two years, researchers, builders, and policy experts worked to answer a deceptively simple question. What does real openness in AI actually require, beyond marketing language?

That effort has now produced a peer-reviewed framework published through a major computing research body. It breaks the AI stack into its real components: data, code, model weights, and documentation, instead of treating openness as one yes-or-no label.

Signal to watch: A system can be open on its weights while staying closed on training data. That distinction alone should change how you read a vendor's "open source AI" claim.

Why Your Current Vendor Checklist Is Outdated

You have likely evaluated AI vendors using a simple open-or-closed filter. This new thinking proves that filter is too blunt for real procurement decisions.

The framework argues openness is a gradient, not a switch. A model might offer full code access but almost no evaluation documentation, leaving you unable to judge its real risk profile.

What This Means For Your Risk Assessment

Safety cannot be judged from model weights alone. Deployment environment, safeguards, and governance structures around a model shape its actual real-world risk, not the weights in isolation.

The Advantage This Creates For Your Organization

Here is where this shift works in your favor. A shared vocabulary for openness means you can finally compare AI vendors on consistent, defensible terms instead of vague claims.

  • Sharper procurement decisions: You can demand specifics on data, code, and documentation instead of accepting a blanket openness claim.

  • Reduced regulatory exposure: Clear openness criteria help you get ahead of emerging AI policy requirements in Washington and Brussels.

  • Stronger vendor accountability: Granular openness standards give your legal and procurement teams real leverage in contract negotiations.

  • Faster internal audits: A common framework simplifies how your compliance team documents AI system risk across departments.

Enterprises that adopt this layered view of openness early will negotiate from a position of knowledge, not guesswork, while competitors are still arguing over labels.

The Question You Should Ask Every Vendor Next

This framework deliberately avoids naming one correct level of openness for every system. That is intentional. Your risk tolerance in a regulated industry differs from a consumer app's.

Where To Start This Quarter

Ask every AI vendor which layers of their stack are genuinely open, and which are not. Data, code, weights, and documentation each deserve a separate answer, not one blanket claim.

The leaders who win this next phase of AI adoption will be the ones who stopped accepting vague openness claims and started asking layer-by-layer questions instead.

The Chat Box Era Of AI Just Quietly Ended

You have spent the past two years watching your teams paste context into chat windows and hope for the right answer. That entire model just became optional, and the shift favors leaders who move first.

Why Your Chatbot Strategy Is Already Behind

A major AI lab recently open-sourced the underlying engine that powers its most capable coding agents. That engine, often called a harness, is now free for any enterprise to build on.

You no longer need to force security analysts, support engineers, or product managers into a generic chat box. The intelligence layer can now live inside the dashboards and tools your teams already use daily.

Signal to watch: With harness-level optimization alone, one leading model's benchmark score nearly tripled while token usage dropped sixfold. The execution layer around a model matters as much as the model itself.

The Real Reason Your AI Tools Feel Generic

Most organizations assume a strong model plus a clever prompt equals a capable AI agent. That assumption is quietly costing you performance and differentiation against competitors.

A genuinely useful agent needs memory across long conversations, tool access, failure handling, and a way to pause for human approval. That underlying execution system is what separates a demo from a deployable product.

Codex can do far more than just power coding tools!

Greg Brockman, President, OpenAI

What Newly Open Components Give You

Three components now ship under a permissive open license: a command-line tool for automated pipelines, a developer SDK, and an app-server that embeds agent behavior directly into your own software.

Where This Creates Real Advantage For Your Business

This is the part worth pausing on. Early enterprise adopters are already proving measurable returns, not just technical novelty, and the pattern applies well beyond software teams.

  • Faster back-office throughput: One finance partnership processed roughly 7,000 tax filings while cutting preparation time by a third.

  • Lower integration cost: You can embed agent intelligence into existing dashboards instead of rebuilding workflows around a chat interface.

  • Stronger governance control: Human approval checkpoints let you decide exactly which actions an agent can execute unsupervised.

  • Vertical-specific deployment: A major networking vendor already built natural-language app creation directly into its cloud management platform.

Enterprises that embed this capability into existing systems now will operate faster and cheaper than competitors still relying on generic chat tools.

The Control You Have Been Missing Until Now

Open access to this execution layer, first released by OpenAI, gives you three forms of control you did not have before. Interface ownership, context ownership, and operational boundaries all stay with you.

What Leaders Should Do This Quarter

Identify one internal workflow currently trapped in a chat interface. Ask your engineering team whether embedding agent logic directly into that existing tool would remove friction.

Govern Your Agent Data

Agents are only as safe as the data behind them. See how leaders lock down access before agents go live.

The organizations that win this next phase will not be the ones with the flashiest chatbot. They will be the ones who quietly rebuilt their core workflows around it.

Journey Towards AGI

Research and advisory firm guiding on the journey to Artificial General Intelligence

Know Your Inference

Maximising GenAI impact on performance and Efficiency.

Model Context Protocol

Connect with us, and get end-to-end guidance on AI implementation.

Your opinion matters!

Hope you loved reading our piece of newsletter as much as we had fun writing it. 

Share your experience and feedback with us below ‘cause we take your critique very critically. 

How's your experience?

Login or Subscribe to participate in polls.

Thank you for reading

-Shen & Towards AGI team