What Is Driving the Struggle for Data Quality in Modern Enterprises?

Enterprises today are increasingly grappling with data quality challenges due to a combination of outdated data architectures and a general lack of innovative data culture within organizations. This issue is brought to light in the 2024 State of Analytics Engineering survey conducted by dbt Labs, which involved 456 data practitioners and leaders to offer a comprehensive overview of the current state of data management and quality in organizations.

Prevalence of Data Quality Concerns

A significant 57% of survey respondents identified poor data quality as their primary concern, a notable increase from 44% in 2022. This indicates a growing acknowledgment of data quality issues within enterprises, highlighting just how pervasive this problem has become. Additionally, the survey pointed out that low stakeholder data literacy and ambiguous data ownership were major concerns affecting nearly half and 44% of the respondents, respectively. These findings emphasize that data quality is intertwined with broader challenges in data governance and stakeholder engagement.

Time and Resource Allocation in Data Management

Data practitioners reportedly spend 55% of their time maintaining or organizing data sets. Coupled with this, 40% cited the integration of data from various sources as their biggest challenge, underlining the labor-intensive nature of current data management practices. These statistics reflect the substantial time and resources allocated to managing data rather than deriving valuable insights from it. It also highlights a critical need for streamlined processes and better tools to manage and integrate data effectively, allowing data teams to focus more on strategic data tasks.

Investment Trends

Despite these challenges, nearly 40% of companies plan to maintain their investment in data quality, platforms, and catalogs, with 10 to 37% intending to increase their investments, depending on the category. This investment trend underscores a recognition of the importance of data quality and the need for robust data management solutions. Companies are beginning to realize that without addressing the root issues, their data initiatives, including AI and machine learning applications, may not reach their full potential.

Three Root Causes of Data Quality Issues

Lack of Innovative Data Culture

Historically, there has been a neglect in focusing on data collection, management, and reuse, with enterprises often prioritizing application development over data quality. This has resulted in fragmented or trapped data within underutilized applications, hindering effective data utilization. An innovative data culture that emphasizes the importance of high-quality data and continuous improvement practices is essential for overcoming these issues.

Legacy Data Architectures

Many organizations continue to rely on outdated, costly legacy data architectures that do not scale efficiently. This creates what Cinchy terms an “integration tax,” consuming over 50% of IT budgets. The persistence of these old systems adds complexity and cost to data integration efforts, thereby impairing data quality.

Failure to Address Integration Complexity

Complicated architectures drive up integration costs, further exacerbating data quality issues. Effective solutions include adopting a standards-based semantic graph architecture to simplify integration. Simplifying these architectures can help mitigate costs associated with data integration and improve overall data quality.

Upstream Data Management Strategies

Pushing proactive data management efforts upstream—closer to the data source—can significantly improve data quality. This approach provides better control and understanding of data, reducing the challenges associated with repurposing data collected further downstream. By focusing on data quality at the point of origin, enterprises can ensure that data remains reliable and useful as it moves through the organization.

Industry Example: Direct Lithium Extraction

An analogy to effective upstream data management can be drawn from the process of Direct Lithium Extraction (DLE) in the oil industry. Just as DLE extracts valuable lithium from brine at oil field sites, proactive data management can extract valuable insights and ensure quality data by focusing on its collection and management at the source. This analogy underscores the importance of addressing data quality issues at the earliest stages of data collection.

Approach to Data Management

To overcome data quality challenges, it is vital for organizations to adopt a producer’s mindset towards data. This involves owning and managing data lifecycle processes effectively to create reliable, reusable data products. Treating data as a product ensures that it is consistently managed, monitored, and improved upon, leading to better data quality and more reliable insights.

Overarching Trends and Consensus Viewpoints

In today’s fast-paced business environment, enterprises are increasingly facing significant challenges related to data quality. This struggle stems primarily from outdated data architectures and a pervasive lack of innovative data culture within organizations. The magnitude of this problem is highlighted in the 2024 State of Analytics Engineering survey, conducted by dbt Labs. This extensive survey engaged 456 data practitioners and leaders to provide a detailed snapshot of the current landscape of data management and quality within businesses.

The findings underscore the criticality of transitioning from old data frameworks to more advanced and adaptable solutions to ensure accurate, reliable data. Organizations must recognize that fostering an innovative data culture is essential for not only addressing these data quality issues but also for driving business growth and efficiency. By embracing new technologies and cultivating a forward-thinking data culture, businesses can better harness the power of their data, leading to more informed decision-making and improved outcomes.

Explore more

How AI Agents Work: Types, Uses, Vendors, and Future

From Scripted Bots to Autonomous Coworkers: Why AI Agents Matter Now Everyday workflows are quietly shifting from predictable point-and-click forms into fluid conversations with software that listens, reasons, and takes action across tools without being micromanaged at every step. The momentum behind this change did not arise overnight; organizations spent years automating tasks inside rigid templates only to find that

AI Coding Agents – Review

A Surge Meets Old Lessons Executives promised dazzling efficiency and cost savings by letting AI write most of the code while humans merely supervise, but the past months told a sharper story about speed without discipline turning routine mistakes into outages, leaks, and public postmortems that no board wants to read. Enthusiasm did not vanish; it matured. The technology accelerated

Open Loop Transit Payments – Review

A Fare Without Friction Millions of riders today expect to tap a bank card or phone at a gate, glide through in under half a second, and trust that the system will sort out the best fare later without standing in line for a special card. That expectation sits at the heart of Mastercard’s enhanced open-loop transit solution, which replaces

OVHcloud Unveils 3-AZ Berlin Region for Sovereign EU Cloud

A Launch That Raised The Stakes Under the TV tower’s gaze, a new cloud region stitched across Berlin quietly went live with three availability zones spaced by dozens of kilometers, each with its own power, cooling, and networking, and it recalibrated how European institutions plan for resilience and control. The design read like a utility blueprint rather than a tech

Can the Energy Transition Keep Pace With the AI Boom?

Introduction Power bills are rising even as cleaner energy gains ground because AI’s electricity hunger is rewriting the grid’s playbook and compressing timelines once thought generous. The collision of surging digital demand, sharpened corporate strategy, and evolving policy has turned the energy transition from a marathon into a series of sprints. Data centers, crypto mines, and electrifying freight now press