AsiaTechDaily – Asia's Leading Tech and Startup Media Platform

  • Topics
    • AI & Big Data
    • AR & VR
    • Blockchain
    • Clean Technology
    • Content & Games
    • Cybersecurity
    • Enterprise & SaaS
    • Gadgets & Electronics
    • Health & Bio
    • FinTech
    • IoT
    • Transportation & Logistics
    • Marketplaces & E-commerce
    • Ecosystem
    • Robotics
    • Investments
    • Events
    • Innovasion Exchange Programme
    • Startup Program
    • EdTech
    • Featured
  • Deals
    • Private Equity
    • Venture Capital
    • IPO & Markets
  • Interviews
    • Investors’ interviews
    • Founders’ interviews
    • Unicorn interview
  • Governments
  • Events
  • Lists
Menu
  • Topics
    • AI & Big Data
    • AR & VR
    • Blockchain
    • Clean Technology
    • Content & Games
    • Cybersecurity
    • Enterprise & SaaS
    • Gadgets & Electronics
    • Health & Bio
    • FinTech
    • IoT
    • Transportation & Logistics
    • Marketplaces & E-commerce
    • Ecosystem
    • Robotics
    • Investments
    • Events
    • Innovasion Exchange Programme
    • Startup Program
    • EdTech
    • Featured
  • Deals
    • Private Equity
    • Venture Capital
    • IPO & Markets
  • Interviews
    • Investors’ interviews
    • Founders’ interviews
    • Unicorn interview
  • Governments
  • Events
  • Lists
Submit Article
Menu
  • Topics
    • AI & Big Data
    • AR & VR
    • Blockchain
    • Clean Technology
    • Content & Games
    • Cybersecurity
    • Enterprise & SaaS
    • Gadgets & Electronics
    • Health & Bio
    • FinTech
    • IoT
    • Transportation & Logistics
    • Marketplaces & E-commerce
    • Ecosystem
    • Robotics
    • Investments
    • Events
    • Innovasion Exchange Programme
    • Startup Program
    • EdTech
    • Featured
  • Deals
    • Private Equity
    • Venture Capital
    • IPO & Markets
  • Interviews
    • Investors’ interviews
    • Founders’ interviews
    • Unicorn interview
  • Governments
  • Events
  • Lists
Submit Article
Join Chat 💬
[the_ad id="20911"]
Analysis22 Jun 2026 8:28

Why Enterprise AI Can’t Reach Its Full Potential Until It Unlocks Unstructured Data

by Yong-Joon Bae
  • twitter
[the_ad id="20911"]
Bookmark (0)
Please login to bookmark Close

From emails and contracts to customer conversations and PDFs, the majority of enterprise knowledge remains trapped in unstructured data. As generative AI reshapes business operations, organizations are discovering that unlocking this information securely may determine the success of their AI strategies.


The enterprise AI conversation has long revolved around increasingly powerful large language models, faster chips, and expanding cloud infrastructure. Yet for many organizations, the biggest obstacle to realizing AI’s full potential is far less glamorous: their own data.

Industry estimates suggest that 80% to 90% of enterprise information exists in unstructured formats, including emails, documents, PDFs, contracts, customer support tickets, meeting transcripts, call recordings, images, and internal reports. Unlike structured data stored neatly in databases, this information has historically been difficult to analyze, making it an underutilized corporate asset.

Generative AI has fundamentally changed that equation. Large language models can now interpret natural language, summarize documents, retrieve institutional knowledge, analyze conversations, and power enterprise copilots capable of understanding vast collections of business information. What was once considered dormant data has suddenly become one of an organization’s most valuable AI assets.

But there is a catch. Much of this information also contains personally identifiable information (PII), financial records, healthcare information, intellectual property, confidential business documents, or commercially sensitive customer data. For enterprises operating under increasingly stringent privacy and AI governance regulations, using that information for AI development introduces significant legal, security, and compliance challenges.

As enterprise AI moves from experimentation to production, organizations are realizing that their biggest challenge is no longer accessing powerful AI models. It is determining how to safely use the data they already possess.

The enterprise AI opportunity hiding in plain sight

For years, enterprise analytics focused primarily on structured information stored inside databases, spreadsheets, and enterprise resource planning systems because these datasets were relatively easy to organize and analyze. However, much of an organization’s operational knowledge has always existed elsewhere. Customer complaints are captured in emails. Product feedback is buried inside support tickets. Legal obligations reside in contracts. Institutional knowledge lives in technical documentation. Sales insights emerge from call transcripts, while operational decisions are scattered across presentations, meeting notes, and internal communications.

Collectively, these documents represent the contextual knowledge that enables organizations to make informed decisions. The arrival of generative AI has dramatically increased their strategic value. Instead of simply analyzing rows and columns, enterprises can now deploy AI systems capable of understanding language, identifying relationships across documents, retrieving historical information, and generating context-aware responses. Technologies such as retrieval-augmented generation (RAG), enterprise search, AI assistants, and domain-specific copilots all depend on access to rich, high-quality unstructured information. In many ways, enterprise AI has shifted from being a model problem to becoming a data problem.

Why privacy has become AI’s biggest bottleneck

If unstructured information holds enormous business value, why has it remained largely inaccessible? The answer lies in privacy. Unlike structured datasets that can often be anonymized through relatively straightforward methods, unstructured documents frequently contain sensitive information embedded throughout natural language. Customer names, medical histories, financial details, addresses, legal discussions, and confidential commercial information can all appear within the same document. This makes preparing enterprise data for AI significantly more complicated.

While conversing with AsiaTechDaily, Grant de Leeuw, Co-Founder and CEO of DataMasque, explained why organizations have struggled to unlock this category of enterprise information despite its immense value.

“Enterprise information and data sits across structured, semi structured and unstructured data throughout the organization and valuable data, such as PDF applications and call transcripts, can hold significant value. The limitation to date is unstructured data can often hold sensitive information that enterprises have struggled to address.”

His observation reflects one of the defining challenges facing enterprise AI today. Organizations are no longer constrained by a lack of AI capability. They are constrained by uncertainty around how to safely expose sensitive business information to increasingly powerful AI systems.

Why preserving context matters as much as protecting privacy

Protecting sensitive information is not a new challenge. For decades, organizations have relied on techniques such as redaction, masking, and anonymization when sharing or testing sensitive datasets. However, generative AI introduces a different requirement. Large language models derive value from context. Removing names, locations, relationships, or other contextual elements too aggressively can reduce the usefulness of the data itself, limiting its effectiveness for model training, testing, retrieval, or fine-tuning. As enterprises increasingly seek to operationalize AI, preserving data utility has become almost as important as protecting privacy.

Speaking with AsiaTechDaily, de Leeuw said this balance between privacy and usability is precisely where many organizations continue to struggle.

“DataMasque de-identifies this data across all datastores, including unstructured data. Rather than redacting the data, we replace it with consistent, synthetically identical values ensuring the context and utility of the data are retained. This allows enterprises to experiment, train or fine-tune on this data without exposing any protected or personal data.”

The broader significance extends well beyond a single technology platform. Across the enterprise AI landscape, organizations are increasingly exploring approaches such as synthetic data, de-identification, and privacy-preserving data processing that enable AI development without unnecessarily exposing sensitive customer information.

AI governance is becoming a competitive advantage

As governments introduce new AI regulations and organizations strengthen internal governance frameworks, responsible data management is evolving from a compliance requirement into a strategic capability. Markets such as Singapore have actively promoted trusted AI governance frameworks, recognizing that responsible AI adoption depends not only on technological innovation but also on transparency, accountability, and effective data governance.

This shift is particularly significant for highly regulated industries. Financial institutions, insurers, healthcare providers, telecommunications companies, and government agencies possess enormous volumes of valuable unstructured information. Yet these sectors also operate under some of the world’s strictest privacy and regulatory requirements.

For them, enterprise AI adoption increasingly depends on answering a fundamental question: How can organizations extract intelligence from sensitive information without compromising customer trust or regulatory compliance? The answer will likely shape the pace of enterprise AI adoption across many industries.

Investors are beginning to recognize the enterprise data challenge

Growing interest in enterprise data infrastructure is also attracting investor attention. Earlier this year, DataMasque raised US$4 million in a funding round led by Wavemaker Ventures, with participation from existing investors OIF Ventures and Icehouse Ventures. Since its 2023 seed round, the company says it has achieved sixfold annual recurring revenue growth while expanding its customer base to include organizations such as New York Life, ADP, Best Western Hotels and Resorts, One NZ, TAL, and government agencies across New Zealand, Australia, and the United States.

The company is also expanding into Singapore, positioning itself within one of Asia’s leading AI governance ecosystems where enterprises face increasing pressure to balance AI innovation with regulatory compliance. The investment reflects a broader market reality. As organizations move beyond AI experimentation, technologies that enable secure access to enterprise data are becoming an increasingly important layer of the enterprise AI stack.

The next phase of enterprise AI will be defined by data readiness

The rapid evolution of foundation models has transformed what AI systems are capable of achieving. Increasingly, however, the limiting factor is no longer the intelligence of the models themselves. It is the quality, accessibility, and governance of the information organizations choose to place behind them.

For many enterprises, decades of accumulated documents, customer interactions, operational records, and institutional knowledge represent an extraordinary competitive advantage waiting to be unlocked. Yet realizing that value will require organizations to solve one of enterprise AI’s most complex challenges: making sensitive information both safe and useful.

The companies that succeed in the next phase of enterprise AI may not necessarily be those with access to the largest models or the greatest computing power.

Instead, they are likely to be the organizations that can responsibly unlock their vast reserves of unstructured knowledge while preserving privacy, maintaining regulatory compliance, and retaining the context that makes enterprise information valuable in the first place.

As generative AI becomes embedded across every business function, the race will increasingly be won not by who builds the smartest models, but by who can safely transform decades of untapped enterprise knowledge into actionable intelligence.


Quick Takeaways
  • Unstructured data is enterprise AI’s largest untapped resource. An estimated 80% to 90% of enterprise information exists in formats such as emails, PDFs, contracts, call transcripts, and internal documents, making it a critical yet underutilized asset for AI applications.
  • Privacy, not AI models, is becoming the biggest deployment challenge. As organizations move from AI experimentation to production, protecting sensitive information while maintaining data usability has emerged as a major barrier to enterprise AI adoption.
  • Context is essential for enterprise AI. Traditional redaction often removes valuable information needed by large language models. Enterprises are increasingly exploring de-identification and synthetic data techniques that preserve the context and utility of datasets while reducing privacy risks.
  • AI governance is evolving into a strategic business capability. Regulatory compliance, responsible AI practices, and strong data governance are becoming competitive differentiators, particularly for highly regulated industries such as finance, healthcare, insurance, and government.
  • Investor interest is shifting toward enterprise data infrastructure. Growing investment in technologies that enable secure, privacy-preserving AI reflects the increasing importance of data readiness as organizations operationalize AI across their businesses.
  • The next phase of enterprise AI will be driven by data readiness. While AI models continue to advance rapidly, organizations that can securely unlock their unstructured enterprise knowledge will be better positioned to build scalable, trustworthy, and high-value AI applications.

Tags: AnalysisDataStartup
[the_ad id="20911"]

Similar Articles

Analysis30 Aug 2026 7:56

When AI Knows More About Products, Brands Will Have to Prove More

More
Analysis29 Aug 2026 9:15

From Followers to Founders: Why Asia’s Creators Are Building Their Own Consumer Brands

More
Analysis27 Aug 2026 9:27

CTV Advertising Has Scaled Faster Than Its Underlying Infrastructure

More

[the_ad id=’22944′]

Topics

Menu
  • AI & Big Data
  • AR & VR
  • Blockchain
  • Clean Technology
  • Content & Games
  • Cybersecurity
  • Enterprise & SaaS
  • Gadgets & Electronics
  • Health & Bio

Program

Menu
  • Ecosystem
  • EdTech
  • Featured
  • FinTech
  • Investments
  • IoT
  • Marketplaces & E-commerce
  • Robotics
  • Transportation & Logistics

About

Menu
  • Home
  • About us
  • Privacy Policy
  • Collaborate with AsiaTechDaily
Facebook Instagram Linkedin
  • twitter

Subscribe and be informed first hand about the actual economic news.

All the day’s headlines and highlights, direct to you every morning.

[mc4wp_form id="5832"]

© 2023 asiatechdaily. All rights reserved.