"A lakehouse combines the scalability and flexibility of a data lake with the governance, structure, and performance of a data warehouse. It allows organizations to store both structured and unstructured data in one platform while supporting robust analytics, machine learning, and BI workloads. These lakehouses get by the data swamp problem by providing features that are database like. That includes support for ACID transactions." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Adopting a data mesh requires more than architecture; it involves changes to culture, roles, and accountability structures. [...] Many organizations will need to create new roles, such as data product managers, who combine stewardship, analytics, and product thinking. These individuals oversee the life cycle of a data product to ensure it delivers business value and meets usability standards. Success often depends on building a culture where data is treated as a shared asset across the enterprise and on putting clear lines of accountability in place. In practice, this can mean forming cross-domain councils to align on standards, tying performance objectives to data quality, or investing in training so domain teams are equipped to manage their products responsibly." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Agentic AI extends GenAI by embedding intelligence within autonomous or semi-autonomous systems that can plan, reason, and take actions within defined boundaries. Instead of simply generating a report, an agentic system might determine which data it needs, retrieve that data from multiple sources, perform analysis, summarize the results, and then trigger follow-up workflows, all while maintaining auditability and alignment with governance policies." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"AI maturity depends as much on people as it does on tools. Organizations need to build AI literacy across roles, from analysts to executives, so that teams can ask the right questions, interpret AI results, and act on them. Cross-functional teams with skills in data engineering, machine learning, business domain knowledge, and governance are essential. AI systems should be treated as products and not one-off projects." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"As part of the move to democratization and utilizing AI-infused tools and generative AI to gain insights into data, businesses need to improve their overall data literacy. Data literacy involves the awareness and recognition of the value of data, how well people understand and interact with data and analytics, and their ability to communicate data-driven insights to impact behavior and achieve business goals. It includes understanding the business and data elements, framing analytics, applying critical interpretation, and developing communication skills." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"As the need to store and analyze new data types (e.g., semi-structured, unstructured, real-time streams) grew, organizations began turning to data lakes. These systems are designed to store high volumes of raw data at scale and use a schema-on-read approach, offering greater flexibility. Initially built on platforms like Apache Hadoop, many early data lakes fell short due to lack of governance, poor performance, and inadequate metadata management. As a result, these early deployments often became so-called data swamps, where users struggled to find, trust, or use the data effectively. Such deployments delivered little to no business value." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Data observability platforms often consolidate various data health metrics, such as freshness, accuracy, completeness, and consistency, into a unified dashboard, offering a holistic view of data health. This may take the form of a data scorecard. These tools can also provide metrics such as system uptime/availability, data processing times, error rates in data processing, query performance metrics, resource utilization, and user engagement metrics. They provide customizable key performance indicators (KPIs), real-time alerts, and root cause analysis, and they can help assess the impact of data quality issues on downstream applications, reports, and business processes." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Data products are part of the enterprise data infrastructure. They are managed, versioned, and monitored much like APIs or microservices. This shift is supported by modern cloud-native infrastructure, which enables decentralized development, consistent governance, and scalable consumption. Data products don’t need to be stored in one location. They might be in a data lake, a data warehouse, a database, or a data lakehouse. They can become discoverable and accessible through a centralized data catalog or marketplace acting as a portal where users can find, understand, and request access to relevant data products." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Governance readiness is about ensuring trust. That involves trust in the data, trust in the models, and trust in how AI affects users and stakeholders. Without it, organizations risk reputational damage, regulatory penalties, and internal resistance." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Data readiness also involves the quality, integrity, and findability of the data. Is the data trustworthy and curated? Are data pipelines (to move data from source to target) and access controls (to ensure the right people get access to the right data) in place? Moreover, to be ready, the organization’s architectural components - data lakes, warehouses, lakehouses, etc. - must be coherent and aligned to support modern AI applications." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"GenAI, powered by large foundation models, is revolutionizing the way people interact with data and technology. Organizations are using GenAI to democratize AI through natural language interfaces, copilots, assistants, and creative tools that can draft text, summarize reports, or generate images on demand. Yet as transformative as this is, these systems remain largely reactive. They respond to prompts but do not decide what to do next. They do not plan, reason over time, or act toward goals. In short, they generate but they don’t take action." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"If AI is owned solely by IT or the data science team, it will fail. AI must be co-developed with business stakeholders. This means engaging stakeholders from the start, co-owning key performance indicators (KPIs), and embedding AI into workflows, not bolting it on after the fact. AI teams need domain experts, business sponsors, and clear lines of feedback. Use cases should be selected with the business, not handed down from a tech function." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"One of the most overlooked aspects of AI deployment is the operational side, what happens after a model is developed. Operational readiness is concerned with the systems, processes, and team structures required to bring AI into production. In many cases, organizations build prototypes that never make it past the lab. To succeed, organizations must have formalized processes for deploying models, integrating them into business workflows, and monitoring their performance over time (no model is good forever; they degrade as the external environment changes)." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Organizations should use a framework (such as business value vs. complexity) to identify high-impact, feasible AI projects. Organizations that succeed with AI don’t chase hype. Instead, they score and prioritize use cases based on criteria such as business value, data readiness, effort, and complexity. This helps avoid wasted time on technically interesting but low-value projects. Some create the frameworks themselves to determine the high-impact yet feasible projects. Others rely on consultants to help them with this." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"RAG applications must be built with semantics, metadata, and governance in mind. The retrieved information must be high-quality, secure, and appropriate for the user’s role. Equally important is monitoring and management: checking whether source data has changed, ensuring vector stores remain accurate, and watching for hallucinations or data leakage. Organizations are definitely starting to experiment with RAG models today; some are putting them into production applications. Some believe that using RAG helps mitigate hallucinations because it is grounded in trusted organizational data." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"The data architecture provides the blueprint for how the data is organized, accessed, and integrated. It defines the relationships between data sources, how data is modeled and transformed, and how the underlying technologies work together to support the business. While the terms data modeling and data architecture are sometimes used interchangeably, they are not the same. Data modeling is part of a data architecture and defines the specific structure and relationships within datasets (e.g., tables, columns, constraints), whereas data architecture is broader; it provides the framework and strategy for managing, storing, integrating, and utilizing data in an organization." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"The goal of a data product is to provide value, not just data. That value may come from enabling better decision-making, powering customer-facing applications, enriching AI models, or facilitating compliance. Importantly, data products are developed with users in mind, whether internal analysts, business stakeholders, external partners, or automated systems. This user orientation distinguishes data products from raw datasets or traditional reports." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"The modern data environment includes structured data as well as text, machine-generated logs and events, images, and audio. The rapid adoption of generative AI (GenAI) has accelerated both the creation and consumption of text data, increasing volume and variability. Additionally, hybrid and multicloud architectures are now standard. Data are stored and processed across warehouses, lakehouses, SaaS platforms, and edge services. Governance must operate consistently across these environments. In practice, this requires policies that can be enforced programmatically, with end-to-end lineage (where the data came from and how it has been changed) that is visible to both technical and business stakeholders." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"There are a few ways that organizations are trying put guardrails into place. First, they are providing enterprise-approved copilots and assistants that their teams can use. They are providing lightweight governance processes, such as AI councils or intake workflows to make it easy for employees to get approval for new use cases. They are implementing access controls if employees want to use certain tools against company data. Perhaps most importantly, some are implementing AI literacy programs so that employees understand both the risks and the responsible use of AI tools." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"[...] organizations are working to unify the data silos that often exist across their data ecosystem. Some are doing this through a centralized physical approach, such as implementing a data lakehouse pattern, which allows them to store and manage all types of data in one place. Many lakehouse implementations such as those from vendors such as Snowflake and Databricks have evolved into what is often referred to as the modern data platform. This is an architectural pattern that combines the lakehouse with tightly integrated tools for data ingestion, transformation, analytics, observability, and governance. This pattern reflects a shift toward cloud-native architectures that are designed to support end-to-end data workflows with scalability and flexibility." (Fern Halper, "Data Makes the World Go 'Round", 2026)
"Unreliable data is a problem because when data cannot be relied upon to deliver valid insights, organizations begin to question the value of their analytics and AI investments. It’s easy to underestimate the cost of bad data until it surfaces in painful ways. Inaccurate or incomplete data can derail projects, corrupt analytics, introduce legal risk, and erode trust, both internally and externally. Have you ever heard the phrase, 'garbage in, garbage out' related to building a model with poor quality data?" (Fern Halper, "Data Makes the World Go 'Round", 2026)





