News Feed Item

Dremio 2.0 Advances State of the Art Self-Service Data Platform With Enhancements to Data Reflections, New Learning Engine, Support for Looker

Dremio, the self-service data company, announced today a new major release of its open source platform, including its core acceleration technology, Data Reflections™, introduced the Dremio Learning Engine, and added support for Looker, a leading data analytics and business intelligence platform. With these new features, Dremio helps business analysts and data scientists to discover, curate, share, and analyze data from data lakes, NoSQL, and relational sources using their favorite tools at unprecedented speed.

Dremio’s open source, self-service data platform addresses the key data analysis issues that challenge modern organizations. This includes acceleration of queries on massive datasets for interactive analytics with BI and data science tools; the complexity of connecting data from disparate sources like data lakes, NoSQL databases, and relational databases; and long lead times for data engineering tasks. Dremio accelerates time-to-insight by empowering analysts and data scientists to be independent and self-directed in their use of data across the enterprise, while preserving governance and security, and providing insight into data lineage. Register for a technical deep dive on the release here.

“Dremio transparently enables unprecedented time-to-insight and, unlike traditional approaches which require building a data warehouse or rely on point-to-point single server designs, Dremio connects any analytical process to any data source and scales from one to 1,000 plus servers, running in the cloud, on dedicated hardware, or in a Hadoop cluster,” said Kelly Stirman, vice president strategy and CMO, Dremio. “With Dremio data consumers can perform critical data tasks themselves, without being dependent on IT. We bring self-service to the entire analytics stack. With this approach analysts and data scientists are more independent and self-directed.”

Enhancements to Data Reflections

Data Reflections™ provide a breakthrough in performance that transparently accelerates processing by up to a factor of 1000x. Unlike traditional cubes, extracts and data marts, Data Reflections leverage Apache Arrow for breakthrough performance gains and are entirely invisible to end users - Dremio’s query planner automatically selects the best reflections to accelerate queries for ad hoc requests, BI, and data science workloads using libraries such as TensorFlow and Scikit-learn. This release includes:

  • Starflake Data Reflections. Dremio can now automatically detect star and snowflake schemas, and other variations of joined datasets, and accelerate a wide range of queries that involve all tables or a subset thereof. This capability greatly improves user experience, simplifies administration, and lowers the cost of deploying workloads on Dremio. Learn more here.
  • 100x Improvement in Reflection Management Scalability. This release includes a new reflection management engine based on a state of the art relational algebra-based dependency graph. The engine automatically optimizes prioritizing, ordering and queueing reflection refreshes, as well as sophisticated error recovery. These improvements reduce management overhead, speed reflection updates, and significantly reduce resource utilization for maintenance tasks.
  • Native Support for Cloud Data Lakes. Now users can take advantage of cost effective object stores for Data Reflections, including Amazon S3 and Azure Data Lake Store. Unlike traditional technologies, Dremio separates compute and storage capabilities to allow for independent scaling of resources, and provides optimal in-memory performance with the cost advantages, elasticity, and unlimited scalability of cloud object stores.
  • Vectorized Processing with Apache Arrow. Enhancements to Apache Arrow in this release created by Dremio engineers provide up to a 60% reduction in query latency across a wide range of workloads. As a result, end user experience is improved and users can deploy greater workloads without increasing operational costs. Learn more.

Dremio Learning Engine

New in this release, the Dremio Learning Engine makes recommendations and adapts to evolving workloads, making it easier for users to be more productive in their use of the platform and for workloads to be deployed more efficiently with fewer infrastructure and administrative resources. This release includes:

  • Recommended Joins. Dremio will recommend complementary datasets to users as they work to curate data for analysis. Dremio learns how datasets can be joined based on observed behavior across all workloads with support for all join types.
  • Schema Learning. Dremio automatically observes data during query execution to detect schema changes in source systems, then adapts its Data Catalog automatically. This is essential for modern sources (e.g., Elasticsearch, MongoDB, JSON) where schema can vary from record to record, and for systems with evolving schemas. This enhancement makes data available to data consumers more quickly, with less intervention by administrators to manually synchronize schemas.
  • Predictive Metadata Caching. For large deployments with millions of tables, partitions, collections, and indexes, Dremio will intelligently cache and index metadata into Dremio’s Data Catalog, taking into account the access patterns to those datasets. These changes make large deployments more efficient, and allow data consumers to have the most up to date sense of source data for the most relevant datasets.

Simplified Operations

Dremio is designed to seamlessly integrate with existing data sources, as well as the favorite tools of data consumers, including BI and data science tools. With mission critical deployments at many organizations, Dremio has continued to simplify operation of production environments, including the following capabilities:

  • Automatic Failover. In the event of node and instance failures in a cluster, Dremio will automatically elect a new master coordinator node without requiring manual intervention. This allows companies to provide always-on availability to users while minimizing operational overhead.
  • Workload Management. New administrative controls in this release allow for more granular workload management, with new options for system and user concurrency, query queue size, and memory threshold controls that allow administrators to meet varying performance needs of different users without unnecessary spend on cluster resources.
  • Dynamic Granular Access Controls. Dremio integrates with centralized security controls like LDAP and Kerberos. New in this release, users can programmatically control access at both the row and column level based to dynamically secure, mask, and transform data for end user access, helping users to meet the stringent demands of new privacy controls such as GDPR.
  • REST APIs. Now users can interact with Dremio through a comprehensive set of REST APIs, allowing DevOps teams to orchestrate Dremio with other components of their technology stacks, and end users to more easily build web applications directly on top of Dremio.

Support for Looker

This release also adds support for Looker. Users of Looker can now take advantage of Dremio’s powerful data acceleration capabilities for relational databases, as well as NoSQL databases such as MongoDB, Elasticsearch, and data lakes built on Hadoop, Amazon S3, and Azure Data Lake Store. Learn more here.

“Dremio provides an easy-to-use solution for Looker customers wanting to analyze data directly from NoSQL data stores,” said Keenan Rice, vice president of alliances at Looker. “Dremio’s ability to query across multiple data stores without physically consolidating the data will save customers valuable time and resources.”


The latest release of Dremio’s self-service data platform is available immediately. Download Dremio here to connect data sources and start running analytics at unprecedented speed in minutes.

About Dremio

Dremio reimagines analytics for modern data. Created by veterans of open source and big data technologies, Dremio is a fundamentally new approach that dramatically simplifies and accelerates time to insight. Dremio empowers business users to curate precisely the data they need, from any data source, then accelerate analytical processing for BI tools, machine learning, data science, and SQL clients. Dremio starts to deliver value in minutes, and learns from your data and queries, making your data engineers, analysts, and data scientists more productive. For more information, visit www.dremio.com.

Founded in 2015, Dremio is headquartered in Mountain View, CA. Investors include Lightspeed Venture Partners, Norwest Venture Partners and Redpoint Ventures. Connect with Dremio on GitHub, LinkedIn, Twitter and Facebook.

More Stories By Business Wire

Copyright © 2009 Business Wire. All rights reserved. Republication or redistribution of Business Wire content is expressly prohibited without the prior written consent of Business Wire. Business Wire shall not be liable for any errors or delays in the content, or for any actions taken in reliance thereon.

Latest Stories
Moroccanoil®, the global leader in oil-infused beauty, is thrilled to announce the NEW Moroccanoil Color Depositing Masks, a collection of dual-benefit hair masks that deposit pure pigments while providing the treatment benefits of a deep conditioning mask. The collection consists of seven curated shades for commitment-free, beautifully-colored hair that looks and feels healthy.
The textured-hair category is inarguably the hottest in the haircare space today. This has been driven by the proliferation of founder brands started by curly and coily consumers and savvy consumers who increasingly want products specifically for their texture type. This trend is underscored by the latest insights from NaturallyCurly's 2018 TextureTrends report, released today. According to the 2018 TextureTrends Report, more than 80 percent of women with curly and coily hair say they purcha...
The textured-hair category is inarguably the hottest in the haircare space today. This has been driven by the proliferation of founder brands started by curly and coily consumers and savvy consumers who increasingly want products specifically for their texture type. This trend is underscored by the latest insights from NaturallyCurly's 2018 TextureTrends report, released today. According to the 2018 TextureTrends Report, more than 80 percent of women with curly and coily hair say they purcha...
We all love the many benefits of natural plant oils, used as a deap treatment before shampooing, at home or at the beach, but is there an all-in-one solution for everyday intensive nutrition and modern styling?I am passionate about the benefits of natural extracts with tried-and-tested results, which I have used to develop my own brand (lemon for its acid ph, wheat germ for its fortifying action…). I wanted a product which combined caring and styling effects, and which could be used after shampo...
The precious oil is extracted from the seeds of prickly pear cactus plant. After taking out the seeds from the fruits, they are adequately dried and then cold pressed to obtain the oil. Indeed, the prickly seed oil is quite expensive. Well, that is understandable when you consider the fact that the seeds are really tiny and each seed contain only about 5% of oil in it at most, plus the seeds are usually handpicked from the fruits. This means it will take tons of these seeds to produce just one b...
Steaz, the nation's top-selling organic and fair trade green-tea-based beverage company, announces its 2017 "Mind. Body. Soul." tour, which will bring authentic experiences inspired by the brand's signature Mind. Body. Soul. tagline to life across the country. The tour will inform, educate, inspire and entertain through events, digital activations and partner-curated experiences developed to support the three pillars of complete health and wellness.
The platform combines the strengths of Singtel's extensive, intelligent network capabilities with Microsoft's cloud expertise to create a unique solution that sets new standards for IoT applications," said Mr Diomedes Kastanis, Head of IoT at Singtel. "Our solution provides speed, transparency and flexibility, paving the way for a more pervasive use of IoT to accelerate enterprises' digitalisation efforts. AI-powered intelligent connectivity over Microsoft Azure will be the fastest connected pat...
There are many examples of disruption in consumer space – Uber disrupting the cab industry, Airbnb disrupting the hospitality industry and so on; but have you wondered who is disrupting support and operations? AISERA helps make businesses and customers successful by offering consumer-like user experience for support and operations. We have built the world’s first AI-driven IT / HR / Cloud / Customer Support and Operations solution.
ScaleMP is presenting at CloudEXPO 2019, held June 24-26 in Santa Clara, and we’d love to see you there. At the conference, we’ll demonstrate how ScaleMP is solving one of the most vexing challenges for cloud — memory cost and limit of scale — and how our innovative vSMP MemoryONE solution provides affordable larger server memory for the private and public cloud. Please visit us at Booth No. 519 to connect with our experts and learn more about vSMP MemoryONE and how it is already serving some of...
Darktrace is the world's leading AI company for cyber security. Created by mathematicians from the University of Cambridge, Darktrace's Enterprise Immune System is the first non-consumer application of machine learning to work at scale, across all network types, from physical, virtualized, and cloud, through to IoT and industrial control systems. Installed as a self-configuring cyber defense platform, Darktrace continuously learns what is ‘normal' for all devices and users, updating its understa...
Codete accelerates their clients growth through technological expertise and experience. Codite team works with organizations to meet the challenges that digitalization presents. Their clients include digital start-ups as well as established enterprises in the IT industry. To stay competitive in a highly innovative IT industry, strong R&D departments and bold spin-off initiatives is a must. Codete Data Science and Software Architects teams help corporate clients to stay up to date with the mod...
As you know, enterprise IT conversation over the past year have often centered upon the open-source Kubernetes container orchestration system. In fact, Kubernetes has emerged as the key technology -- and even primary platform -- of cloud migrations for a wide variety of organizations. Kubernetes is critical to forward-looking enterprises that continue to push their IT infrastructures toward maximum functionality, scalability, and flexibility. As they do so, IT professionals are also embr...
Platform9, the leader in SaaS-managed hybrid cloud, has announced it will present five sessions at four upcoming industry conferences in June: BCS in London, DevOpsCon in Berlin, HPE Discover and Cloud Computing Expo 2019.
At CloudEXPO Silicon Valley, June 24-26, 2019, Digital Transformation (DX) is a major focus with expanded DevOpsSUMMIT and FinTechEXPO programs within the DXWorldEXPO agenda. Successful transformation requires a laser focus on being data-driven and on using all the tools available that enable transformation if they plan to survive over the long term. A total of 88% of Fortune 500 companies from a generation ago are now out of business. Only 12% still survive. Similar percentages are found throug...
When you're operating multiple services in production, building out forensics tools such as monitoring and observability becomes essential. Unfortunately, it is a real challenge balancing priorities between building new features and tools to help pinpoint root causes. Linkerd provides many of the tools you need to tame the chaos of operating microservices in a cloud native world. Because Linkerd is a transparent proxy that runs alongside your application, there are no code changes required. I...