Skip to article frontmatterSkip to article content
Site not loading correctly?

This may be due to an incorrect BASE_URL configuration. See the MyST Documentation for reference.

Open In Colab Binder

The word “vector” goes back to a name Hamilton gave in 1846 while organizing his theory of quaternions. He borrowed a Latin root meaning “carrier” to refer to the imaginary components that represent a geometric displacement in three-dimensional space. Although Hamilton gave the thing its name, he held to the end of his life that three-dimensional vectors could never exist apart from the quaternions that contained them. History, however, took exactly the opposite road: this geometric component, once attached to the quaternion system, went on to become the basic language of all of linear algebra.

The geometric intuition behind vectors, however, is two thousand years older than the name.

Around the fourth to third century BCE, the Aristotelian Mechanical Problems already observed that when an object takes part in two straight-line motions in different directions at once, its actual path is the diagonal of the parallelogram formed by the directions and distances of those two motions. The geometric intuition is correct, but it is in essence a kinematic composition based on the theory of proportions: it is not restricted to right angles, and it does not yet have the dynamical notion of “force.” In the first century CE, Hero of Alexandria used the construction of a parallelogram more explicitly to analyze compound motion and the decomposition of collisions—but this was still purely the language of geometric line segments, with no algebraic structure at all.

The real turning point came in 1586. The Dutch mathematician Simon Stevin used a thought experiment with a closed chain of beads hung over a double inclined plane to turn “force” from a vague experience of pushing into a line-segment length that could be quantified and operated on geometrically—the first time a physical quantity and a geometric operation were bound together rigorously. Half a century later, Galileo decomposed projectile motion into two independent components, uniform horizontal motion and accelerated vertical motion, which is in essence an orthogonal decomposition of vectors; he simply did not yet have the term. In 1687, in Corollary I of the Principia, Newton formally established the parallelogram law of forces as a foundational corollary of the entire system of classical mechanics—a universal, axiomatized construction, though its derivations still relied purely on the language of Euclidean geometry rather than on algebraic symbols.

From Aristotle to Newton, what accumulated over two thousand years was geometric and physical intuition, not a general algebraic tool.

The algebraic tool arrived from an entirely different direction. In 1843, Hamilton carved the quaternion multiplication formula into Brougham Bridge in Dublin (that is Chapter 3’s story). Seeking an algebraic system for three-dimensional space in which division is possible, he wrote a quaternion as a real scalar plus a combination of three imaginary directions. In 1846 he named the part with the three imaginary directions the vector—from the Latin vehere; the word means “carrier” or “bearer,” since a vector “carries” a point in space from the origin to its destination. Hamilton coined the term, but what he cared about most was always the quaternion as a whole, complete under multiplication and division; in his eyes the vector could never be separated out on its own.

Forty years later, this component, bound up within the quaternion system, staged a revolt.

In the 1880s the American physicist Josiah Willard Gibbs and the British engineer Oliver Heaviside independently took the quaternion apart: they kept vector addition, “brutally” split quaternion multiplication, once a seamless whole, into two independent operations—the scalar part, with its sign reversed, became the dot product (the inner product), and the vector part became the cross product—and then threw away entirely the quaternion framework that had held them together. Their reason was thoroughly practical: for a physicist, being able to solve problems simply matters more than preserving the closure of an algebraic structure. Both of the symbols “⋅\cdot” and “×\times” were introduced by Gibbs at this time.

This set off a fierce controversy, the “Great Vector Debate.” Hamilton’s loyalists, led by Peter Tait, denounced it as a desecration of the master’s legacy: they charged that the cross product obtained by taking things apart was a defective pseudovector (its sign behaves abnormally under a reflection of coordinates), and they condemned Gibbs’s system as a “monstrosity” that destroyed the completeness of the algebra. Gibbs and Heaviside retorted that physicists need handy geometric tools, not the obscure and cumbersome baggage of quaternions carried for the sake of mathematicians’ algebraic purity.

The argument ran for nearly twenty years. In 1901 Gibbs’s student E. B. Wilson compiled his lecture notes and published them as Vector Analysis, formally fixing the standard form of vectors found in every physics and engineering textbook today—addition, scalar multiplication, dot product, cross product, together with the symbol system we now know so well. Hamilton’s quaternions did not disappear; they are active today in computer graphics and robotics, where they describe three-dimensional rotations precisely and without any risk of gimbal lock. But the winner of that argument over a common language for physics was, in the end, Gibbs.

The geometric intuition behind vectors has a two-thousand-year history; the word “vector” has less than two hundred years; and the dot product and the cross product were settled as independent algebraic operations even later than matrix multiplication (1858). The tools we take for granted on the blackboard today are, in essence, the spoils left behind by a great collision of mathematical and physical thought.


Chapter Structure and Learning Objectives

The vector toolkit that Gibbs settled on is the core content of this chapter. §2.1 establishes the basic definition of a vector, moving from geometric arrow to algebraic coordinate tuple, and explains how the two formulations complement each other.

§2.2 starts from the linear operations and builds up the full picture of vector algebra step by step. Addition obeys the parallelogram law (the diagonal that Aristotle saw), and scalar multiplication corresponds to stretching—the axioms for these two operations are precisely the core of the definition of a vector space in Chapter 4. On that basis, linear combinations let us see how large a space a set of vectors can span, and the standard basis sets up a coordinate system for that space. The choice of basis is arbitrary, but once it is fixed, the correspondence between a vector and its coefficients in that basis is unique—an idea that recurs in the change of basis in Chapter 3 and in eigenvectors in Chapter 8.

§2.3 introduces the inner product, equipping the vector space with a metric structure: length, angle, and projection all follow from it. The dot product connects algebraic computation with geometric angle, and it is the finite-dimensional prototype of the inner product spaces of Chapter 9.

§2.4 takes up the standard three-dimensional cross product, whose component formula does not carry over directly to arbitrary dimensions; this chapter uses it to describe vector area and normal vectors.

By the end of this chapter, “vector” will mean to you far more than “a row of numbers”: it will mean what two thousand years of physical intuition and two hundred years of algebraic argument have jointly deposited.


2.1 Introduction to Vectors

In the geometric formulation, a vector has both a magnitude (a length) and a direction; in the algebraic formulation, a vector is usually a list of numbers arranged in order. Vectors can represent displacement, velocity, and force in physics, and they can equally represent a color (an RGB color, say). In the sections that follow we introduce the basic notions and operations of vectors a few at a time.

Vectors of higher dimension follow the same pattern. In general, we define the space in which nn-dimensional real vectors live as the nn-fold Cartesian product of the set of real numbers.

2.2 Linear Operations on Vectors

The most basic operations between vectors are addition and scalar multiplication. These operations constitute the foundation of linear algebra.

2.2.1 Linear Combinations

Linear combinations can be applied to define line segments in vector form.

2.2.2 The Standard Basis

In Rn\mathbb{R}^n, the standard basis is a particular set of unit vectors that are mutually orthogonal and can represent any vector in the space.

2.3 The Inner Product and Projection

The Cauchy-Schwarz inequality admits several different proofs. We give one of the simpler ones here.

2.4 The Cross Product

This chapter discusses the standard three-dimensional cross product. It has wide application in physics, engineering, and computer graphics, and it provides a natural tool for describing rotations, torques, and normal vectors.

2.5 Vector Operations in Python

The NumPy library in Python provides powerful facilities for vector operations. Here are some basic examples:

2.6 Chapter Summary

Review of the Theoretical Thread

This chapter set out from a question that looks like an everyday one: “what is a vector?” The answer given in §2.1 is twofold—a vector is both a geometric arrow and an ordered list of numbers, and once coordinates are introduced the two readings are completely equivalent. This dual identity is not decoration; it is the basic tension running through the whole book: geometric intuition tells us “why,” and algebraic structure tells us “how to compute.”

§2.2 equipped the set of vectors with two operations, addition and scalar multiplication, and immediately displayed their geometric content: addition is the diagonal of a parallelogram, and scalar multiplication is a stretch along the original direction. More importantly, these two operations gave rise to the notion of a linear combination: given several vectors, taking a weighted sum with arbitrary coefficients lets us “reach” a space spanned jointly by those vectors. The unique-decomposition theorem for the standard basis is the summit of this section: v=∑i=0n−1viei\mathbf{v} = \sum_{i=0}^{n-1} v_i \mathbf{e}_i is not merely a computational formula but the rigorous definition of the notion of a “coordinate.”

Addition and scalar multiplication, however, tell us only how to “move,” not how to “measure.” §2.3 introduced the inner product, hiding all of the information about length and angle inside the algebraic quantity ⟨u∣v⟩=∑iuivi\langle\mathbf{u}|\mathbf{v}\rangle = \sum_i u_i v_i. The Cauchy-Schwarz inequality ∣⟨u∣v⟩∣≤∥u∥∥v∥|\langle\mathbf{u}|\mathbf{v}\rangle| \le \|\mathbf{u}\|\|\mathbf{v}\| is the core of that section: it guarantees the legitimacy of the definition of angle in any dimension and makes orthogonal projection a precise geometric operation rather than merely an intuitive impression. §2.4 turned to the cross product, peculiar to three-dimensional space, adding through u×v\mathbf{u} \times \mathbf{v} a mechanism for generating normal vectors; computing the shortest distance between skew lines displayed the surprising economy of the cross product in geometric problems. §2.5 then put all of the above into practice with Python and NumPy, including a first look at a code fragment for Gram-Schmidt orthogonalization as a preview of the theory of orthogonality to come.

Connections to Other Chapters

There is a clear bridge between this chapter and the language of sets from Chapter 1: set theory supplied precise language for “a set of elements” and for “functions,” and this chapter added operations on top of that language—vector addition and scalar multiplication—promoting the set Rn\mathbb{R}^n to a “vector space” carrying algebraic structure. That promotion will be formalized completely in the abstract linear spaces of Chapter 4: the eight computational rules observed in §2.2 (commutativity, associativity, distributivity, and the rest) are precisely the prototype of the vector space axioms, and readers will discover there that function spaces and polynomial spaces satisfy the very same axioms, so that the notion of a vector is enormously generalized.

Looking ahead, this chapter supplies almost all of the conceptual material for the matrices and linear transformations of Chapter 3: the change of basis presupposes the unique representation in the standard basis, the geometric meaning of matrix multiplication requires vector addition and scalar multiplication working together, and the introduction of quaternions in the historical opening of Chapter 3 will echo the three-dimensional restriction on the cross product. Further ahead, the real inner product of §2.3 is the finite-dimensional prototype of the inner product spaces of Chapter 9, where the complex inner product and the polarization identity are generalized; and the Gram-Schmidt code fragment first met in §2.5 foreshadows the orthogonalization techniques required by the singular value decomposition of Chapter 11.

The Role of This Chapter in the Book

This chapter solved a basic and unavoidable problem: before we may speak of matrices, transformations, and decompositions, we must first be clear about what a vector is, what may be done to it, and how it is to be measured. Addition and scalar multiplication establish the linear structure, the inner product establishes the metric structure, and together they constitute the complete framework of Euclidean space. Carrying that framework, a reader who meets the matrices of Chapter 3 is not merely manipulating blocks of numbers but tracking how the coordinates of a vector transform under different bases; and a reader who meets the symmetric matrices of Chapter 9 will see at once that the orthogonality of eigenvectors originates in the geometric meaning of the inner product.

The ubiquity of the vector toolkit in nature, engineering, and information science also gives this chapter a significance beyond the classroom. Force and displacement in physics, spectral components in signal processing, embedding vectors in machine learning—these fields each developed a need for “vectors” independently, in their own contexts, and yet all converged on one and the same mathematical language. Historically, the birth of that language was not smooth: the twenty-year “quaternion war” between Hamilton’s quaternion school and Gibbs’s vector school produced as its spoils exactly the addition, inner product, and cross product that a linear algebra course now takes for granted. A reader who carries this history forward may come to hold each definition in a little more regard, and to feel a little more impulse to ask why each definition is as it is.

Concept Map