_______ is a distributed database management system designed for large-scale data.
- Apache Hadoop
- MongoDB
- MySQL
- SQLite
Apache Hadoop is a distributed database management system specifically designed for handling large-scale data across multiple nodes. It is commonly used in big data processing. MongoDB, MySQL, and SQLite are database systems but are not specifically designed for distributed large-scale data.
If you are analyzing real-time social media data, which Big Data technology would you use to process and analyze data streams?
- Apache Flink
- Apache Hadoop
- Apache Kafka
- Apache Spark
Apache Kafka is a distributed streaming platform that is commonly used to handle real-time data streams. It allows for the processing and analysis of data in real-time, making it a suitable choice for analyzing social media data as it is generated.
Which machine learning technique is typically used for making predictions based on continuous data?
- Classification
- Clustering
- Dimensionality Reduction
- Regression
Regression is the machine learning technique used for making predictions based on continuous data. It models the relationship between the independent variables and the dependent variable, allowing for the prediction of numeric values.
_______ is a constraint in SQL that ensures unique values are inserted into a column.
- CHECK
- DEFAULT
- PRIMARY KEY
- UNIQUE
The "UNIQUE" constraint in SQL ensures that all values in a column are unique, meaning no two rows can have the same value in that column. It is often used to enforce data integrity and prevent duplicate entries.
For a business process improvement case study, the _______ framework is commonly applied to identify inefficiencies and areas for improvement.
- Agile
- PDCA
- SWOT
- Six Sigma
In business process improvement case studies, the Six Sigma framework is commonly applied to identify inefficiencies, reduce variability, and enhance overall process performance. Six Sigma focuses on data-driven decision-making and process optimization.
What is the output of print(list("123"[::-1])) in Python?
- ['1', '2', '3']
- ['3', '2', '1']
- [1, 2, 3]
- [3, 2, 1]
The output will be a list containing the characters of the string "123" in reverse order. The [::-1] slicing reverses the string, and list() converts it into a list of characters.
What type of database model is SQL based on?
- Hierarchical
- Network
- Object-Oriented
- Relational
SQL is based on the relational database model. It uses tables to organize data and relationships between them, making it a powerful and widely used language for managing relational databases.
If a company needs to process large volumes of unstructured data, which type of DBMS should they consider?
- Hierarchical
- NoSQL
- Object-Oriented
- Relational
In scenarios involving large volumes of unstructured data, a NoSQL database management system (DBMS) is well-suited. NoSQL databases offer flexibility and scalability, making them suitable for handling unstructured data types like documents, graphs, and key-value pairs.
When presenting a data-driven story about population growth in various regions, which visualization technique would best convey this information?
- Box-and-Whisker Plot
- Choropleth Map
- Gauge Chart
- Scatter Plot
A Choropleth Map would be the most effective visualization technique for conveying information about population growth in various regions. It uses color-coding to represent data values across geographical areas, making it ideal for displaying regional variations.
The process of arranging rows in a database table into a specific order is known as _______ in SQL.
- Indexing
- Ordering
- Sequencing
- Sorting
The process of arranging rows in a specific order in a database table is known as Ordering in SQL. It involves specifying the columns by which the result set should be sorted.