Skip to article frontmatterSkip to article content
Site not loading correctly?

This may be due to an incorrect BASE_URL configuration. See the MyST Documentation for reference.

Open In Colab Binder

In the history of linear algebra, one fact looks anomalous: the early study of determinants came before the systematic development of matrix algebra in the mid-nineteenth century.

Determinants were at first closely tied to elimination and to simultaneous equations. Mathematicians studied how the coefficients combine and permute, and gradually turned an operation that originally served equation solving into an algebraic object in its own right.

At the end of the seventeenth century, Seki Takakazu and Leibniz each studied early forms of the determinant in problems of elimination and of equations. Their work was not yet the complete theory of determinants of arbitrary order that we have today, but it shows a similar motivation: how can the structure of a system of equations be read off from combinations of its coefficients?

In the eighteenth century, the work of Maclaurin, Cramer, and others further connected determinants of the coefficients with the solutions of linear equations. Today’s Cramer’s rule states this relationship with particular clarity: when the coefficient matrix is invertible, each unknown can be expressed as a ratio of determinants.

In the nineteenth century, the determinant gradually became a systematic algebraic theory. The work of Cauchy and others established and developed basic tools such as the multiplicative property det⁡(AB)=det⁡(A)det⁡(B)\det(\mathbf{AB})=\det(\mathbf{A})\det(\mathbf{B}); the connections among multilinearity, the sign change when two rows are exchanged, and the sign of a permutation also gave a unified structure to expansions that look intricate.

This winding history tells us that the most central modern geometric meaning of the determinant—the signed factor by which a linear transformation scales higher-dimensional volume—was not its historical starting point; it is a deeper essence that later generations recognized only through a long look back. Chapter 3 already introduced the effect of the determinant on area and volume intuitively in two and three dimensions (Definition 10).

The task of this chapter is to raise this geometric intuition to a rigorous axiomatic theory: starting from multilinearity and the alternating property, we extend the determinant to arbitrary nn-dimensional spaces, derive its Laplace expansion, the adjugate matrix, and its characteristic rules of computation, and forge the heavy algebraic machinery that is indispensable for the theory of eigenvalues and eigenvectors in Chapter 8.

Chapter Structure and Learning Objectives

The determinant can be defined from two entirely different angles, and the two eventually shake hands on the same formula—this tension is the core of §7.1. The permutation definition starts from the Leibniz formula and gives an explicit computational formula; the axiomatic definition goes the other way, starting from three geometric properties (multilinearity, the alternating property, normalization) and deriving the uniqueness of this formula. Each definition has its strengths: the former is convenient for computation, while the latter reveals the essence. Understanding the relationship between the two is the first step toward mastering the theory of determinants.

With the definitions in place, §7.2 builds the tools for computing determinants. Cofactor expansion (Laplace expansion) reduces a determinant of order nn recursively to smaller determinants and is a powerful tool for theoretical derivations; Gaussian elimination offers a far more efficient path in actual computation; and the Vandermonde determinant, with its elegant formula as “the product of all the differences,” serves as the theoretical high point of the section.

§7.3 turns to applications. Cramer’s rule, the protagonist of §7.3.1, shows how determinants solve systems of equations elegantly—although it is not the best choice for numerical computation, it is indispensable in theoretical analysis. §7.3.2 and §7.3.3 return to geometry and formally establish the determinant’s identity as the “volume scale factor”: the determinant is the signed volume scale factor, and its absolute value is the ordinary volume scale factor; in the nondegenerate case its sign records whether orientation is reversed. This viewpoint not only deepens geometric intuition but also prepares for understanding the geometric meaning of eigenvalues in Chapter 8. §7.4 closes the chapter with a Python implementation, using NumPy and a cofactor expansion function of our own to translate the theory of this chapter into code that can be executed and verified. Chapter 8 will use the characteristic polynomial det⁡(A−λI)\det(\mathbf{A} - \lambda\mathbf{I}) as the junction (see Definition 2), formally connecting the theory of determinants in this chapter to the theory of eigenvalues.

After reading this chapter, you will see how a number born from the leftover scraps of equation solving became, after two centuries of refinement, one of the most central tools of linear algebra.

7.1 Definition and Properties of the Determinant

This section discusses two different definitions of the determinant, together with some commonly used important properties.

7.1.1 The Permutation Definition of the Determinant

With a consistent, well-defined sign sgn(σ)\text{sgn}(\sigma), we can use the Leibniz formula with confidence, and be sure that when determinants are later computed by elementary column/row exchanges, the result does not depend on the reduction path.

7.1.2 The Axiomatic Definition of the Determinant

In §7.1.1 we started from the Leibniz formula of the permutation definition and verified that it satisfies three properties: normalization at the identity, the row-alternating property, and multilinearity in the rows. Mathematicians later discovered a deeper viewpoint: these three properties are not merely consequences of the Leibniz formula; conversely, they can even serve as “axioms”—among all functions, the determinant is the only one that satisfies all three requirements at once!

This is the axiomatic definition of the determinant (also called the Weierstrass axioms). The axiomatic method frees the determinant from tedious combinatorial computation with n!n! terms and lets us focus on its geometric and algebraic essence.

These three axioms completely determine the determinant function. Axiom 1 “anchors” the determinant at the identity matrix, while Axioms 2 and 3 describe how the determinant changes as the matrix changes. Starting from these three axioms, all other properties of the determinant can be derived, including the familiar formulas for computing determinants. This axiomatic approach is elegant and powerful: it lets us concentrate on the essential features of the determinant rather than on computational details. On this foundation, we can understand more deeply how the determinant is related to linear transformations, changes in volume, and the invertibility of matrices.

7.1.3 Basic Properties Derived from the Axioms

From a few basic principles, many important properties of the determinant can be derived.

The zero-row property is closely related to the rank of a square matrix: a matrix with a zero row is necessarily not of full rank, and its determinant is zero.

We may use the identical-rows property in place of the row-alternating property as one of the initial basic properties. The elementary row operation properties below embody the alternating nature of the determinant and are also directly linked to linear dependence: a matrix with linearly dependent rows has determinant zero.

This theorem is crucial for understanding and computing determinants; it is also the theoretical basis for computing determinants by Gaussian elimination.

By the same reasoning, together with Theorem 9, we see that a square matrix of full rank has a nonzero determinant‾\underline{\text{has a nonzero determinant}}.

The product formula for block upper triangular determinants proved in the basic problem also makes good on the unproved assertion in the property box of Chapter 5. It is only one block elimination away from proving the Schur complement determinant formula stated at the beginning of Chapter 5:

The determinant has an important symmetry property: a matrix and its transpose have equal determinants. Transpose invariance lets us choose freely whether to compute and analyze determinants by rows or by columns; choosing the direction with more zero entries simplifies the computation.

This property shows that every property of the determinant concerning rows applies equally to columns. For example:

  • Exchanging two columns changes the sign of the determinant.

  • Multiplying a column by a nonzero constant multiplies the determinant by that constant.

  • Adding a constant multiple of one column to another column leaves the determinant unchanged.

  • A matrix with a zero column or with linearly dependent columns has determinant zero.

7.2 Methods for Computing Determinants

This section introduces several commonly used methods for computing determinants.

Cofactor expansion is not only a practical method for computing determinants; it also connects the determinant with concepts such as the inverse matrix and the solutions of systems of linear equations, revealing the deep unity of the internal structure of linear algebra.

By Theorem Theorem 5, we can reduce a matrix to a triangular matrix by elementary row operations and then compute the determinant as the product of its diagonal entries.

Remark. Verification by Laplace expansion (along row 0):

det⁡(A)=2∣2114∣−1∣1134∣+3∣1231∣\det(\mathbf{A}) = 2\begin{vmatrix}2&1\\1&4\end{vmatrix} - 1\begin{vmatrix}1&1\\3&4\end{vmatrix} + 3\begin{vmatrix}1&2\\3&1\end{vmatrix}
=2(8−1)−1(4−3)+3(1−6)=14−1−15=−2✓= 2(8-1) - 1(4-3) + 3(1-6) = 14 - 1 - 15 = -2 \checkmark

The Vandermonde matrix is extremely important in modern mathematics and its applications: it appears in polynomial interpolation, the fast Fourier transform (the DFT matrix is a special Vandermonde matrix), the physical theory of the quantum Hall effect, and the theory of BCH codes and Reed–Solomon error-correcting codes. Here we prove this important mathematical property directly by fairly elementary mathematical induction; another common proof uses the factorization of polynomials from abstract algebra.

We close this section with three computational exercises covering cofactor expansion, reduction to triangular form by elementary operations, and a numerical check of the Vandermonde formula. When Chapter 8 expands the characteristic polynomial det⁡(A−λI)\det(\mathbf{A}-\lambda\mathbf{I}), these hand-computation skills will come into play directly.

7.3 Applications of the Determinant

7.3.1 Solving Systems of Equations

7.3.2 Geometric Applications of the Determinant

The determinant has rich applications in geometry, especially in computing areas and volumes.

We first prove that the determinant is invariant under a change of basis. In §4.2 (Definition 11) we discussed the following: suppose the matrix representations of a linear transformation T:V→VT: V \rightarrow V with respect to two different bases are A\mathbf{A} and B\mathbf{B}. By the theory of change of basis, there exists an invertible matrix P\mathbf{P} (the change-of-basis matrix) such that B=P−1AP\mathbf{B} = \mathbf{P}^{-1}\mathbf{A}\mathbf{P}, and by the multiplicative property of the determinant, clearly det⁡(B)=det⁡(P−1)det⁡(A)det⁡(P)=det⁡(A)\det(\mathbf{B}) = \det(\mathbf{P}^{-1})\det(\mathbf{A})\det(\mathbf{P})=\det(\mathbf{A}). This shows that the geometric quantities we discuss next are all independent of the choice of basis.

In particular, for n=2,3n = 2, 3 we return to the following formulas, already seen with the cross product in Chapter 2 (Theorem 2) and in the geometric introduction of Chapter 3; they are low-dimensional special cases of the theorem above.

7.3.3 Applications of the Determinant in Calculus

7.4 Computing Determinants in Python

In Python, determinants can be computed efficiently with the NumPy library.

7.4.1 Computing Determinants with NumPy

This example shows that the matrix A is singular (in fact because its row vectors are linearly dependent: row 2 = 2 × row 1 − row 0), while the matrix B is not singular.

7.4.2 Implementing the Cofactor Expansion Algorithm

Below we implement our own function that computes the determinant by cofactor expansion:

Our implementation agrees with NumPy’s result, but for large matrices the recursive cofactor expansion is inefficient.

7.4.3 A Practical Example of Using Determinants

Below is an example that uses a determinant to compute the area of a triangle:

7.4.4 Supplementary Checks: The PLU Decomposition and the Vandermonde Formula

We close with two short checks: the first obtains the PLU decomposition with scipy.linalg.lu and verifies det⁡(A)=det⁡(P)det⁡(U)\det(\mathbf{A}) = \det(\mathbf{P})\det(\mathbf{U}) (Theorem Theorem 8); the second builds a Vandermonde matrix with np.vander and compares it with the product formula of Theorem Theorem 16.

Determinants can be computed efficiently in Python and applied to all kinds of practical problems. NumPy provides optimized determinant computation that is well suited to scientific computing and engineering applications.

7.5 Chapter Summary

Review of the Theoretical Thread

Chapter 7 starts from an old question: how can a single number capture the essence of a square matrix? §7.1 answers this question in two strikingly different languages. The permutation definition gives a computational rule directly from the Leibniz formula, defining det⁡(A)\det(\mathbf{A}) as the sum of the sign-weighted products over all n!n! permutations; the axiomatic definition goes the other way, constraining a mapping by three geometric properties—normalization, the row-alternating property, and multilinearity—and then deriving its unique form. The equivalence of the two definitions (Theorem Theorem 10) is the deepest theoretical conclusion of this chapter: the determinant is both a rule of computation and a logical necessity of geometric properties, and the two paths shake hands on the same formula. On this basis, §7.1.3 systematically derives the zero-row and identical-rows properties, the rules for elementary row operations, the determinant of a permutation matrix, the product theorem det⁡(AB)=det⁡(A)det⁡(B)\det(\mathbf{AB}) = \det(\mathbf{A})\det(\mathbf{B}), and transpose invariance—each of these properties is a fruit that grows naturally from the three axioms, and together they form a complete theoretical arsenal for computing determinants.

§7.2 turns this arsenal into concrete computational methods. The Laplace expansion reveals the recursive structure of the determinant and leads to the adjugate matrix, a closed-form theoretical expression for the inverse matrix; the elementary row operation method lowers the computational complexity from O(n!)O(n!) to O(n3)O(n^3) and has become the practical standard for numerical computation; and the Vandermonde determinant theorem, with its elegant formula as “the product of all the differences,” hints at a deeper connection between determinants and the theory of polynomials. §7.3 widens the view further: Cramer’s rule expresses the solution of a system of linear equations in closed form; the geometric applications establish the determinant’s identity as the signed volume scale factor, and the fact that a change of basis does not change the determinant shows that this is an intrinsic invariant of the linear transformation; the Jacobian determinant extends the idea of scaling to the local linearization of nonlinear mappings and becomes the geometric foundation of the change-of-variables formula for multiple integrals.

Connections to Other Chapters

The connection between this chapter and Chapter 6 is both an inheritance of tools and a complementary viewpoint. Gaussian elimination in Chapter 6 centers on row operations and aims at solving systems of equations; in this chapter, the properties of elementary row operations (Theorem Theorem 5) and the PLU decomposition theorem (Theorem Theorem 8) use the same operations to compute determinants and give them an algebraic meaning. The two chapters share Gaussian elimination as their backbone but stand at different vantage points: Chapter 6 looks at the structure of the solution space, while Chapter 7 looks at the effect of a linear transformation on volume. The fact that a zero determinant is equivalent to the matrix not being of full rank (Theorem Theorem 7) is precisely where these two viewpoints meet.

The core tool of Chapter 8—the characteristic polynomial det⁡(A−λI)\det(\mathbf{A} - \lambda\mathbf{I})—is built directly on this chapter. The product theorem and the multilinear property are the basis for analyzing the coefficients of the characteristic polynomial; the invariance of the determinant under similarity transformations (the basis invariance of §7.3.2) guarantees that eigenvalues are attributes of the linear transformation rather than quantities that depend on coordinates. The permutations and the sign function introduced in §7.1 will also reappear when the expansion of the characteristic polynomial is discussed. It is fair to say that this chapter lays all the necessary computational groundwork for Chapter 8, while also previewing the geometric intuition of eigenvalue theory: eigenvalues describe the stretching ratios of a transformation along particular directions, and the determinant is the product of all the eigenvalues—the two are unified in the characteristic polynomial, with the fundamental theorem of algebra as the bridge.

The Role of This Chapter in the Book

Chapter 7 occupies a pivotal position in the book. The first six chapters successively established the language of sets and mappings, the geometry of vectors, linear transformations and matrices, abstract linear spaces, block matrices, and the theory of solutions of systems of linear equations—tools that are already powerful but still lack a quantity linking the algebraic structure of a matrix to the geometric properties of space. The determinant is exactly this link: with a single scalar, it encodes the overall effect on signed volume of the linear transformation that a square matrix represents.

This idea of “compressing a geometric quantity into a single number” is extremely common in nature and in engineering. In thermodynamics, the Jacobian determinant describes the local compression rate of phase space; in quantum mechanics, determinants appear in the antisymmetrization of many-particle wave functions (the Slater determinant); the controllability of a finite-dimensional linear time-invariant system is decided by whether the controllability matrix has full row rank (row rank equal to the dimension of the state); only in the single-input case is this matrix square, and only then is the criterion equivalent to a nonzero determinant. The determinant is so pervasive precisely because it captures one of the most essential invariants of a linear transformation. If you read on with the viewpoint of this chapter, you will see in Chapter 8 that eigendecomposition is essentially a search for a way to decompose the determinant “direction by direction”—and this decomposition is precisely the completion, at the highest level, of the geometric intuition introduced in Chapter 3.

Concept Map