Quick Revision

GK One-Line Question & Answer

15541+ short questions with short answers, covering every category and sub-category on the site — no long articles to scroll through. Good for a fast recap before an exam, or a few minutes of daily practice.

DBMS → Functional Dependency 13

The FD preservation problem in BCNF decomposition means that:
Some FDs in the original relation cannot be enforced by checking just one decomposed relation - they would require joining multiple decomposed relations to verify making constraint enforcement expensive
click to copy
In relation R(A,B,C,D,E) with F={AB to C, D to E, C to B} is AD a candidate key?
AD+ = ADE only: start with A,D, apply D to E getting ADE. Cannot reach B or C without them already in the closure. AD does NOT determine all attributes, so AD is NOT a candidate key.
click to copy
What is a trivial functional dependency and why is it excluded from normal form violation checks?
X to Y where Y is a subset of X - it always holds in any relation (by reflexivity) because a set of attributes always determines its own subset providing no real constraint on data
click to copy
What is the union rule derived from Armstrongs axioms?
If X to Y and X to Z then X to YZ (you can combine right-hand sides when the left-hand side is the same)
click to copy
What is Fagans theorem about lossless join decomposition?
A theorem stating that the binary decomposition of R into R1 and R2 is lossless if and only if (R1 intersect R2) multidetermines (R1-R2) OR (R1 intersect R2) multidetermines (R2-R1) holds as an MVD in R
click to copy
In relation R(A,B,C,D,E) with F={A to BC, CD to E, B to D, E to A} what is the attribute closure of E?
E+ = EABCD by E to A to BC to D and CD to E (already have E): E determines all attributes in the relation
click to copy
What is the Armstrong completeness theorem?
Armstrongs axioms are both sound (only derive valid FDs) and complete (can derive ALL valid FDs) meaning F+ computed by Armstrong axioms exactly equals the set of all FDs logically implied by F
click to copy
What is the difference between a FD X to Y being satisfied by a relation instance vs being implied by a set of FDs?
Satisfied by an instance: this specific relation instance happens to have X to Y hold (extensional). Implied by a set F: X to Y holds in EVERY possible instance that satisfies F (intensional) - a much stronger condition capturing the semantics of the schema
click to copy
Given the FDs F={A to B, B to C, C to D, D to A} what is the canonical cover Fc?
Fc = A to B, B to C, C to D, D to A - each FD is individually necessary (no FD is derivable from the others) since removing any one breaks the cycle
click to copy
What is an independent set of FDs and why is it important?
A set of FDs where none of them can be derived from the others - important because it means each FD adds genuine new information about the schema and removing any one would change the closure F+
click to copy
What is the key lemma used to prove that Armstrongs axioms are sound?
For soundness: Reflexivity holds trivially by definition. Augmentation: if t1[X]=t2[X] then t1[XZ]=t2[XZ] which implies t1[Y]=t2[Y] gives t1[YZ]=t2[YZ]. Transitivity: if t1[X]=t2[X] implies t1[Y]=t2[Y] and t1[Y]=t2[Y] implies t1[Z]=t2[Z] then t1[X]=t2[X] implies t1[Z]=t2[Z]
click to copy
What is the chase algorithm used for in relational theory?
Testing whether a decomposition has the lossless join property and whether functional dependencies are preserved by applying FDs to a canonical table (tableau)
click to copy
What is the concept of dependency basis in the context of MVDs?
For a set of attributes X the dependency basis is the finest partition of (R-X) such that X multidetermines each block; this partitions the other attributes into independent groups that X independently multidetermines
click to copy

DBMS → SQL Basics 27

What is the correct logical execution order of SQL clauses in a SELECT statement?
FROM WHERE GROUP BY HAVING SELECT ORDER BY LIMIT
click to copy
What is the difference between WHERE and HAVING clauses in SQL?
WHERE filters individual rows BEFORE grouping; HAVING filters groups AFTER GROUP BY and aggregation - HAVING can reference aggregate functions, WHERE cannot
click to copy
What does the SQL clause NULLIF(expr1, expr2) return?
NULL if expr1 equals expr2 (returns expr1 otherwise); used to avoid division-by-zero errors: NULLIF(count, 0) returns NULL instead of causing error when count=0
click to copy
What is the behavior of aggregate functions (SUM, AVG, COUNT, MAX, MIN) with respect to NULL values?
Aggregate functions (except COUNT(*)) IGNORE NULL values - COUNT(*) counts all rows including NULLs; COUNT(column) counts only non-NULL values
click to copy
What is the purpose of the SQL WITH clause (Common Table Expression - CTE)?
To define named temporary result sets that can be referenced multiple times within a query improving readability and enabling recursive queries (WITH RECURSIVE)
click to copy
What is the difference between CHAR(n) and VARCHAR(n) data types?
CHAR(n) is fixed-length (always uses n bytes, padded with spaces if shorter); VARCHAR(n) is variable-length (uses only the space needed plus 1-2 bytes for length storage)
click to copy
What does the SQL CASE expression return when no WHEN condition matches and no ELSE clause is specified?
It returns NULL (the CASE expression evaluates to NULL when no WHEN matches and no ELSE is provided)
click to copy
What is a correlated subquery in SQL and how does it differ from a non-correlated subquery?
A subquery that references a column from the outer query causing it to be executed once for each row of the outer query (vs. non-correlated subquery which executes once independently)
click to copy
What is the SQL EXISTS operator and when should it be preferred over IN?
EXISTS returns TRUE if a subquery returns at least one row (stops at first match), preferred over IN when the subquery could return NULLs (IN with NULL has counterintuitive behavior) or when checking existence is more efficient
click to copy
What does SELECT * FROM employees WHERE department_id IN (SELECT department_id FROM departments WHERE location = NULL) return?
An empty result set - the condition WHERE location = NULL is always FALSE (must use IS NULL instead of = NULL)
click to copy
What is the difference between RANK(), DENSE_RANK(), and ROW_NUMBER() window functions?
ROW_NUMBER(): unique sequential number with no gaps or ties; RANK(): same rank for ties then skips numbers (1,1,3); DENSE_RANK(): same rank for ties no gaps (1,1,2)
click to copy
What are LEAD() and LAG() window functions used for?
Accessing values from subsequent rows (LEAD) or preceding rows (LAG) within the result partition useful for computing differences between consecutive rows without self-joins
click to copy
What is the OVER(PARTITION BY...ORDER BY...ROWS/RANGE BETWEEN...) clause used for?
Defining the window frame for window functions - specifying which rows to include in each computation relative to the current row
click to copy
What is the SQL PIVOT operation conceptually and how is it typically implemented?
Transforming row values into column headers converting a narrow table into a wide table - implemented via conditional aggregation (CASE + GROUP BY) in standard SQL
click to copy
What is the COALESCE(expr1, expr2, ..., exprN) function?
Returns the first non-NULL expression from left to right - short-circuits (stops evaluating) once a non-NULL value is found
click to copy
What does the SQL FETCH FIRST n ROWS ONLY clause do and which standard introduced it?
It limits the result set to the first n rows (equivalent to LIMIT n in MySQL/PostgreSQL), introduced by SQL:2008 standard
click to copy
What is the SQL MERGE statement (also called UPSERT) used for?
Performing INSERT, UPDATE, or DELETE operations in a single statement based on whether a match exists between source and target tables - useful for ETL and synchronization operations
click to copy
What is three-valued logic (3VL) in SQL and what are the three truth values?
TRUE, FALSE, and UNKNOWN - where UNKNOWN results from comparisons involving NULL values; logical operations follow specific rules: TRUE AND UNKNOWN = UNKNOWN, FALSE AND UNKNOWN = FALSE, TRUE OR UNKNOWN = TRUE
click to copy
What is the difference between TRUNCATE and DELETE without a WHERE clause in SQL?
TRUNCATE removes all rows without logging individual row deletions (faster, minimal logging, resets auto-increment), cannot be rolled back in some DBMS, and cannot have triggers. DELETE logs each row deletion (slower, fully transactional, triggers fire, can be rolled back)
click to copy
In SQL what does DISTINCT do when used inside an aggregate function like COUNT(DISTINCT column)?
It counts only unique non-NULL values of the column eliminating duplicates before counting - e.g. COUNT(DISTINCT dept_id) counts how many distinct departments have employees
click to copy
What is the purpose of SQL CHECK constraint and what are its limitations?
To enforce a condition that must be true for all rows in a table; limitations include: cannot reference other tables, cannot contain subqueries in standard SQL, and in some DBMS it was parsed but not enforced
click to copy
What does SELECT DISTINCT department_id FROM employees return differently from SELECT department_id FROM employees GROUP BY department_id?
DISTINCT returns unique values without aggregation capability; GROUP BY allows adding aggregate functions. However, for just listing unique values with no aggregation, they produce identical results - GROUP BY is more powerful but DISTINCT is cleaner for simple deduplication
click to copy
What is the BETWEEN operator in SQL and is it inclusive or exclusive of boundaries?
BETWEEN a AND b is INCLUSIVE of both endpoints - equivalent to >= a AND <= b (both a and b are included in the range)
click to copy
What is the SQL LIKE operator, and what do the wildcards % and _ represent?
% matches zero or more characters (any sequence); _ matches exactly one character (any single character) - used for pattern matching in strings
click to copy
What is the difference between INNER JOIN and CROSS JOIN in SQL?
INNER JOIN returns only rows with matching values in both tables based on a join condition; CROSS JOIN returns the Cartesian product (every combination of rows from both tables, no join condition)
click to copy
What is a recursive CTE (WITH RECURSIVE) in SQL and what problem does it solve?
A CTE that references itself enabling traversal of hierarchical/graph data (like org charts, bill of materials, file systems) without knowing the depth in advance - queries tree/graph structures iteratively until no more rows are added
click to copy
What is a window function in SQL and how does it differ from aggregate functions?
Window functions perform calculations across a set of rows related to the current row without collapsing them into a single result row - unlike aggregate functions which collapse groups into single rows
click to copy