Skip to content

Data+ Practice Questions

Prepare for DA0-001 with more than an answer.

249 questions in the full set20 sample questionsUpdated Jul 16, 2026

Unlock the full exam and previous versions

  • v1CompTIA Data+ (V2) 159 questions Locked
  • DA0-001Legacy CompTIA Data+ 249 questions Current
Exam fee
$239 USD
Level
Entry-Level
Valid for
3 years
Domains covered on the exam 5
  1. Data Concepts and Environments15%
  2. Data Mining25%
  3. Data Analysis23%
  4. Visualization23%
  5. Data Governance, Quality, and Controls14%
  1. 1

    What is the median of the following dataset?

    [12, 45, 23, 18, 50, 31, 23]

    Show answer details

    Correct answer: A

    To find the median, you must first sort the dataset in ascending order: [12, 18, 23, 23, 31, 45, 50]. The median is the middle value. In this dataset of 7 numbers, the middle value is the 4th one, which is 23. If there were an even number of values, the median would be the average of the two middle numbers.

  2. 2

    An analyst is creating a dashboard for an executive audience who needs to see high-level Key Performance Indicators (KPIs) at a glance. Which visualization type is MOST suitable for displaying single, critical metrics like 'Total Revenue YTD' or 'Customer Growth %'?

    Show answer details

    Correct answer: B

    A scorecard, also known as a KPI card or metric visual, is specifically designed to display a single, important number in a large, prominent format. This allows executives to quickly absorb the most critical business metrics without needing to interpret a complex chart. They often include secondary information like a comparison to a target or a previous period.

  3. 3

    A data analyst presents a bar chart showing sales by region. The vertical axis (Y-axis) starts at $1,000,000 instead of $0. This choice makes small differences between regions appear very large. Which data visualization principle has been violated?

    Show answer details

    Correct answer: B

    Truncating the axis, specifically the Y-axis on a bar chart, is a common way to mislead viewers. Bar charts represent quantity through the length of the bars, and this length should be proportional to the value. By starting the axis at a value other than zero, this proportionality is broken, and the visual differences between the bars are exaggerated. It is a best practice for bar chart axes representing magnitude to always start at zero.

  4. 4

    Which of the following is a key characteristic of a dashboard, as opposed to a report?

    Show answer details

    Correct answer: C

    The primary differentiator of a dashboard is its ability to provide a high-level, visual overview of multiple key metrics on a single screen. Dashboards are designed for at-a-glance monitoring and are typically interactive, allowing users to filter and drill down into the data. Reports, in contrast, are often static and provide more detailed, specific information, frequently in a tabular format.

  5. 5

    The General Data Protection Regulation (GDPR) is a comprehensive data privacy law that primarily applies to:

    Show answer details

    Correct answer: C

    GDPR is a regulation in EU law on data protection and privacy for all individuals within the European Union and the European Economic Area. It also addresses the transfer of personal data outside the EU and EEA areas. Any organization, regardless of its location, that processes the personal data of EU residents must comply with GDPR.

  6. 6

    A company is establishing a data governance program. A key component is creating a centralized repository of definitions, business rules, and information about the company's data assets. What is this component called?

    Show answer details

    Correct answer: A

    A data dictionary is a centralized repository of information about data, such as its meaning, relationships to other data, origin, usage, and format. It is a critical component of data governance as it provides clarity and consistency in data definitions across the organization, ensuring everyone is using and interpreting data in the same way.

  7. 7

    A retail consortium is building a centralized data warehouse to analyze sales data from its various member stores. The primary analytical goal is to provide fast, simple queries on sales performance against multiple business dimensions like time, product, and store location. The current proposal involves a highly normalized schema which would require analysts to perform numerous complex joins. A data architect suggests denormalizing the data into a different structure to optimize for this type of analysis.

    Which schema design should the architect recommend to BEST meet the consortium's analytical requirements?

    graph TD subgraph "Proposed Star Schema" Fact_Sales --|> Dim_Time Fact_Sales --|> Dim_Product Fact_Sales --|> Dim_Store Fact_Sales --|> Dim_Customer end subgraph "Current Snowflake Schema (Simplified)" Fact_Sales_Snow --|> Dim_Time_Snow Fact_Sales_Snow --|> Dim_Product_Snow Dim_Product_Snow --|> Dim_Category_Snow Fact_Sales_Snow --|> Dim_Store_Snow Dim_Store_Snow --|> Dim_Region_Snow end style "Proposed Star Schema" fill:#d4edda,stroke:#155724 style "Current Snowflake Schema (Simplified)" fill:#f8d7da,stroke:#721c24

    Show answer details

    Correct answer: B

    A star schema is the optimal choice for this scenario. It is designed specifically for data warehousing and business intelligence applications where query performance and simplicity are paramount. By denormalizing dimensions (e.g., including region information directly in the store dimension), it reduces the number of joins required for queries, directly addressing the consortium's primary requirement for faster, simpler analysis compared to the more complex, normalized snowflake schema.

  8. 8

    A data analyst is working with a dataset containing customer feedback. The data is stored in a column named 'satisfaction_level' with values such as 'Very Satisfied', 'Satisfied', 'Neutral', 'Unsatisfied'. These values have an inherent order but the distance between them is not defined. Which data type BEST describes this column?

    Show answer details

    Correct answer: A

    Ordinal data is a categorical, statistical data type where the variables have natural, ordered categories and the distances between the categories are not known. 'Very Satisfied' is higher than 'Satisfied', establishing an order, but the exact difference in satisfaction is not measurable, making it ordinal. Nominal data has no order, interval data has order and a known difference but no true zero, and ratio data has all these properties plus a true zero.

  9. 9

    Which of the following database structures is characterized by a central fact table connected to multiple dimension tables, resembling a star shape, and is optimized for querying and analysis?

    Show answer details

    Correct answer: C

    A star schema is the fundamental structure of a dimensional model in a data warehouse. It consists of a central fact table containing quantitative data (measures) connected to a set of smaller dimension tables, which contain descriptive attributes. This denormalized structure is optimized for fast querying and is the most common schema used for data analysis and business intelligence.

  10. 10

    A retail company is designing a new data warehouse. The analytics team needs to frequently analyze sales data by store, product, and date. To optimize query performance for these common analyses, which approach would be MOST effective?

    Show answer details

    Correct answer: B

    A dimensional model using a star or snowflake schema is the industry standard for data warehousing and analytics. By creating a central fact table for sales metrics and separate dimension tables for store, product, and date, queries become simpler and much faster. This design directly supports the business requirement of slicing and dicing data by these key dimensions.

Create an account to continue.