When AI Exhausts the Internet: How DATA Network Is Rebuilding On-Chain Data Infrastructure

Market News
Updated: 07/02/2026 07:14

On June 25, 2026, a branding announcement in the crypto industry drew significant attention. Story Protocol, originally focused on on-chain intellectual property, officially rebranded as the DATA Foundation, while its underlying blockchain network was renamed DATA Network . The native token IP migrated to DATA on a 1:1 basis. Rather than a simple rebranding, this change represents a strategic shift from intellectual property infrastructure toward AI training data infrastructure.

According to DATA Foundation CEO Andrea Muttoni, "AI labs have effectively exhausted the internet of crawlable content. What remains is either expensive and custom-built, or legally complex and poorly documented." This statement highlights a broader industry challenge: high-quality, verifiable, and compliant training data is becoming a structural constraint for AI development. DATA Network positions itself as an attempt to address this issue by building a trust-oriented data infrastructure layer.

From IP Network to Data Trust Layer: A Strategic Repositioning

To understand DATA Network’s role, it is important to consider the context of its evolution.

Story Protocol was initially designed as a permissionless intellectual property network, aiming to provide on-chain provenance and licensing infrastructure for digital content. However, during its development cycle in 2025, the team identified increasing demand for AI training data infrastructure. This trend was further supported by external developments, including AI data infrastructure project Poseidon, which reportedly raised $15 million from a16z to support high-quality dataset creation for AI models.

From a market perspective, this strategic shift reflects structural pressures within the AI ecosystem. AI systems rely on two main data sources: publicly available internet data, which is increasingly limited in quality and availability, and proprietary datasets provided by commercial vendors, which are often costly and lack transparency. In parallel, regulatory scrutiny around data licensing and copyright compliance is increasing across multiple jurisdictions, raising operational complexity for AI developers.

Within this context, DATA Network is positioned as a "trust layer" for AI training data—aiming to provide verifiable records of data provenance, licensing status, and usage history through blockchain infrastructure, enabling more transparent and compliant data access processes.

The Role of Data in AI and Blockchain Systems

DATA Network operates at the intersection of two evolving technological paradigms.

In traditional AI pipelines, data is typically treated as a static input: collected, processed, and used in batch training cycles. However, as AI models scale toward increasingly large parameter sizes, this static model is becoming less effective. Modern AI systems require continuously updated, diverse, and validated datasets to improve performance and reduce bias over time.

At the same time, blockchain systems introduce new mechanisms for data verification and provenance tracking. On-chain records are immutable, time-stamped, and programmable, making them suitable for representing licensing conditions, ownership records, and usage permissions.

When these two trends converge, a new infrastructure layer emerges—often referred to as a Web3-based data layer. DATA Network is designed to operate within this emerging architecture, where data becomes not only a resource but also a verifiable and programmable asset.

From a broader infrastructure perspective, blockchain ecosystems in 2026 are increasingly moving toward modular architectures, separating execution, settlement, and data-related functions. In this environment, data can be treated as an independent infrastructure layer rather than a by-product of transactions.

Core Architecture: Storage, Indexing, and Transfer

The DATA Network protocol is structured around three primary components: storage, indexing, and data transfer.

Storage Layer

Rather than storing raw data directly on-chain, DATA Network introduces a registry system called Trace. Trace functions as an on-chain verification layer that records metadata for each dataset, including cryptographic hashes, licensing terms, contributor consent, payment records, and timestamps.

This design allows data itself to remain off-chain within the marketplace, while ensuring its provenance and authorization history can be independently verified. In this sense, Trace is often described as a verification interface for AI training datasets.

Indexing Layer

DATA Network also provides a structured indexing system for dataset discovery. This allows users to query datasets based on metadata attributes such as data type, licensing status, and contributor information, rather than relying on fragmented external sources.

As a result, the system effectively functions as a specialized search and discovery layer for AI training data, potentially reducing the complexity of dataset sourcing workflows.

Transfer Layer

Data access and licensing are managed through smart contracts. Contributors can define pricing models and usage conditions directly on-chain, while buyers receive automated access upon successful payment execution.

This mechanism removes the need for centralized intermediaries in data transactions and creates a transparent audit trail for all dataset usage and licensing activity.

According to project disclosures, DATA Network has indexed over 1 billion records and integrated with Kled, a large-scale voluntary data marketplace, which reportedly provides access to more than 1.5 billion datasets. The network is estimated to process around 10 million new records daily.

Key Differences from Traditional Data Infrastructure

DATA Network differs structurally from centralized cloud data systems such as AWS or Google Cloud across three key dimensions.

Ownership:
Traditional systems typically centralize data control within platform providers. In contrast, DATA Network is designed to allow contributors to retain ownership and define licensing conditions directly on-chain.

Verification:
Conventional databases rely on institutional trust. DATA Network instead uses cryptographic proofs to enable independent verification of data provenance and usage rights.

Economic model:
Traditional data platforms often rely on intermediary-based revenue models. DATA Network instead facilitates direct transactions between data contributors and buyers via smart contracts, enabling automated compensation flows.

Ecosystem Development and Market Overview

As of July 2, 2026 (Beijing time), DATA is trading on Gate at approximately $0.3035, reflecting a 24-hour increase of 4.73%. The token’s market capitalization is estimated at around $108 million, placing it approximately in the lower mid-tier of global crypto assets by ranking.

Intraday trading ranged between $0.2847 and $0.3673, with reported daily volume of approximately $475,600. Total supply stands at 1.029 billion tokens.

It is worth noting that the token previously known as IP reached an all-time high of $14.78 in September 2025. Since then, the price has declined significantly, while also rebounding from recent local lows.

From an ecosystem perspective, DATA Network has recently expanded integrations across multiple applications. The Poseidon project has entered a partnership with Toss, a widely used mobile financial application in South Korea. Its data contribution application, Numo, has been integrated as a mini-program within Toss, potentially reaching a large user base. In addition, DATA Foundation has raised approximately $140 million in total funding, led by a16z crypto.

Risks and Considerations

Like many early-stage infrastructure projects, DATA Network faces several uncertainties.

The first relates to data quality and scalability. While the network reports large-scale dataset coverage, the effectiveness of AI training depends heavily on data quality, consistency, and diversity. Low-quality or duplicated data may reduce model performance or introduce bias. Current data validation processes rely in part on partner systems, and their long-term scalability remains to be fully demonstrated.

The second relates to market adoption. The underlying assumption is that AI companies are willing to pay a premium for verifiable and compliant datasets. Whether this demand materializes at scale depends on regulatory pressure and industry adoption cycles, which remain evolving.

The third relates to token economic sustainability. Market valuation has experienced significant volatility, reflecting differing expectations regarding long-term utility and ecosystem growth. Future development will likely depend on continued protocol adoption and infrastructure expansion.

Conclusion

DATA Network represents an emerging attempt to build a verifiable data infrastructure layer at the intersection of AI and blockchain. By combining provenance tracking, licensing transparency, and programmable data access, it aims to support the growing demand for compliant and high-quality AI training datasets.

However, its long-term trajectory will depend on real-world adoption, regulatory alignment, and ecosystem scalability. For observers of AI infrastructure and blockchain-based data systems, DATA Network serves as an important case study in how data ownership and verification mechanisms may evolve in the era of large-scale AI systems.

FAQ

Q1: What is DATA Network?
DATA Network is a decentralized AI training data infrastructure protocol rebranded from Story Protocol in June 2026. It focuses on providing verifiable data provenance, licensing status, and usage history for AI training datasets.

Q2: What is the relationship between DATA and the previous IP token?
DATA is the successor token to IP, migrated on a 1:1 basis. The migration required no user action and accompanied a broader strategic shift toward AI data infrastructure.

Q3: What are Trace and Kled?
Trace is an on-chain registry system that records dataset metadata and provenance proofs. Kled is a large-scale data marketplace integrated into the ecosystem, providing access to a broad range of datasets.

Q4: How does DATA Network address compliance?
It records licensing, consent, and usage history on-chain, enabling verifiable proofs of data provenance that may assist AI companies in meeting regulatory requirements.

Q5: What is the current market status of DATA?
As of July 2, 2026, DATA trades at approximately $0.3035 on Gate, with a market capitalization of around $108 million. The token remains significantly below its historical peak.

Disclaimer: This is not investment advice. The information is provided for informational purposes only and should not be construed as a recommendation to buy, sell or hold any asset. Cryptocurrency trading involves a risk of loss. Gate EU services may be restricted in certain jurisdictions. For more information, please see our legal disclosures .
Like the Content

Share