mm-Wave Measurements
Accurate characterization of millimeter-wave systems presents some of the most challenging measurement problems in electronics. At frequencies from 30 GHz to 300 GHz, wavelengths range from 10 mm down to 1 mm, making physical dimensions, connector repeatability, cable stability, and environmental factors critically important. Measurements that are routine at lower frequencies require sophisticated calibration procedures, specialized test fixtures, and careful attention to sources of uncertainty that can easily dominate the signal of interest.
The fundamental challenge of mm-wave measurements stems from the fact that nearly everything in the measurement system—connectors, cables, probes, fixtures, and even the air gap between surfaces—represents an electrically significant length. A 1 mm dimension that seems negligibly small represents a full wavelength at 300 GHz, creating phase shifts of 360 degrees and potentially introducing resonances, reflections, and coupling effects. A second theme runs alongside the first: as frequency rises, the accessible measurement plane retreats from the device. Coaxial connectors give way to waveguide flanges above 110 GHz, waveguide gives way to probe tips for on-wafer work, and highly integrated products with antennas in the package leave no conducted port at all. This article addresses the techniques that accuracy depends on at each of those planes, from VNA calibration, frequency extension, and probe station setup through de-embedding, fixturing, radiated testing, and uncertainty analysis.
Vector Network Analyzer Calibration at mm-Wave
Vector network analyzer (VNA) calibration establishes the reference planes for measurements and removes systematic errors introduced by the test equipment itself. At millimeter-wave frequencies, calibration becomes substantially more complex and critical than at lower frequencies. The calibration process defines the measurement reference plane—typically at the end of a coaxial connector, at wafer probe tips, or at the input to a test fixture—and characterizes the systematic errors of everything between the VNA's internal calibration plane and this external reference.
Calibration Standards and Methods
Standard calibration techniques include Short-Open-Load-Thru (SOLT), Thru-Reflect-Line (TRL), and the Line-Reflect-Match (LRM) family. Each method has specific advantages and limitations in the mm-wave regime:
- SOLT calibration relies on precisely characterized short, open, load, and thru standards. At mm-wave frequencies, opens can couple to nearby structures and radiate rather than store energy, shorts may not achieve perfect reflection because of skin effect and surface roughness, and loads must maintain 50-ohm impedance across extremely wide bandwidths. SOLT works well with high-quality coaxial connectors but becomes increasingly difficult to execute as frequency rises, because every standard must be modeled accurately over the full band.
- TRL calibration requires a thru connection, a reflective standard (usually a short or open), and one or more precision transmission lines. TRL is particularly effective at mm-wave because it needs no known load standard at all: the calibration reference impedance is set by the characteristic impedance of the line standards, and the reflect standard need only be highly reflective and identical on both ports, with its phase known to within a quarter wavelength. The penalty is bandwidth. The thru and line must differ in electrical length by roughly 90 degrees at band center, and the calibration degenerates as that difference approaches 0 or 180 degrees, so a single line covers only about an 8:1 frequency span—conventionally the region where the phase difference stays between about 20 and 160 degrees.
- LRM and LRRM calibration replace the line standard with a matched load, which removes the bandwidth limit that constrains TRL. LRRM (Line-Reflect-Reflect-Match) adds a second reflect standard and allows the load inductance to be determined from the measurement itself rather than assumed, which is why it has become the default method for broadband on-wafer work. These methods require a well-behaved match but tolerate imperfectly known reflects.
Multiline TRL (mTRL) extends TRL by using several line standards of different lengths, then combining them with a statistical weighting that favors whichever lines are best conditioned at each frequency. The result is high accuracy across a wide band plus redundancy that exposes a damaged or mis-fabricated standard. National metrology institutes use multiline TRL as the basis for traceable on-wafer S-parameter measurement, and public implementations—such as the NIST MultiCal and StatistiCAL software—made the technique widely available. mTRL is the natural choice at mm-wave, where no single line can hold a usable phase difference across the whole band of interest.
Connector and Interface Repeatability
Connector repeatability becomes a dominant source of measurement uncertainty at mm-wave frequencies. The reason is the short wavelength: at 100 GHz a full wavelength in an air-dielectric connector is only 3 mm, so 360 degrees of phase spans 3 mm and a 1-micron change in effective electrical length shifts the one-way phase by roughly 0.12 degree. A reflected (round-trip) path doubles that sensitivity, and small variations in mating depth, pin alignment, and contact resistance accumulate across the several connector pairs in a typical setup. High-quality precision connectors are air-dielectric interfaces named for their outer-conductor inner diameter—1.0 mm (rated to 110 GHz), 1.85 mm (V, to 67 GHz), 2.4 mm (to 50 GHz), and 2.92 mm (K, to 40 GHz)—and use tight mechanical tolerances to minimize this variation, but even the best connectors show measurable repeatability limits. IEEE Std 287 defines the mechanical and electrical requirements for this family of precision interfaces, and a smaller 0.8 mm interface extends coaxial operation to roughly 145 GHz. The general rule is that the usable frequency of a coaxial line is set by the onset of the first higher-order (TE11) mode, so every step up in frequency demands a smaller, more fragile, and less forgiving connector.
Best practices for connector care include:
- Using calibrated torque wrenches to ensure consistent mating force
- Regular inspection with microscopes or optical comparators to detect wear, damage, or contamination
- Careful cleaning procedures using appropriate solvents and lint-free materials
- Tracking connection cycles and retiring connectors before they reach end-of-life limits
- Never cross-connecting incompatible connector types that can cause permanent damage
On-wafer measurements eliminate many connector-related issues but introduce their own challenges with probe-to-pad contact repeatability. The overtravel distance, contact force, and probe tip condition all affect measurement repeatability.
Calibration Verification and Validation
After completing calibration, verification measurements confirm that the calibration was successful. Common verification checks include:
- Measuring the thru standard and confirming S21 magnitude near 0 dB and phase near 0 degrees
- Measuring a precision airline or verification standard with known characteristics
- Checking that the calibrated short shows S11 magnitude near 0 dB with phase near 180 degrees
- Verifying that the load shows S11 magnitude well below -20 dB across the band
These verification steps help identify calibration errors before proceeding to DUT measurements. At mm-wave frequencies, seemingly small calibration errors can produce large measurement uncertainties.
Frequency Extension and Banded Waveguide Measurements
Coaxial test ports run out of headroom near 110 GHz, yet much of the millimeter-wave and sub-terahertz spectrum lies above that limit. The standard solution is the frequency extender: a pair of modules that sit close to the device under test, multiply a microwave source up to the band of interest, and downconvert the reflected and transmitted signals with harmonic mixers back to an intermediate frequency the base VNA can process. Because the modules sit at the DUT rather than in the instrument rack, the lossy, phase-unstable path is confined to the low-frequency cabling, and the high-frequency path shrinks to a few centimeters of waveguide.
Waveguide Bands and Band Stitching
Extenders are banded: each module covers one rectangular waveguide band, defined by the waveguide's single-mode range. Common bands include WR-15 (50 to 75 GHz), WR-12 (60 to 90 GHz), WR-10 (75 to 110 GHz), WR-6.5 (110 to 170 GHz), WR-5.1 (140 to 220 GHz), WR-3.4 (220 to 330 GHz), and, in research settings, WR-2.2 and smaller waveguides that reach beyond 500 GHz. Covering a wide span therefore means physically swapping modules, calibrating each band separately, and stitching the resulting data sets together. Discontinuities at the seams between bands are a common artifact and a useful sanity check: a well-executed set of calibrations produces magnitude and phase that join smoothly across the boundary.
Banded operation also imposes a structural limitation. Rectangular waveguide has a low-frequency cutoff, so extender-based measurements yield no data at or near DC. Band-limited S-parameters complicate time-domain transformation, causal model extraction, and any analysis that assumes a continuous spectrum from DC upward, and they require careful extrapolation or model fitting when the results feed a broadband simulation.
Waveguide Calibration and Flange Alignment
Waveguide calibration kits are, in one respect, easier than coaxial kits: the standards are precisely machined metal geometry rather than assemblies of dissimilar materials. A typical waveguide calibration uses a flush short, an offset short realized with a quarter-wave shim, a matched load, and a thru, or a waveguide TRL set built from shim lines. The dominant error source shifts from standard definition to mechanical alignment. At WR-3.4 the broad wall of the guide is only 0.86 mm across, so a few micrometers of lateral flange offset represents a meaningful fraction of the aperture and produces reflection and loss that no calibration can remove.
Precision flanges with dowel-pin alignment substantially improve repeatability over plain flat flanges, and IEEE Std 1785 addresses waveguide dimensions and interfaces specifically for frequencies of 110 GHz and above. Practical technique matters as much as hardware: tighten flange screws in a balanced pattern to seat the faces evenly, inspect flange faces for burrs and debris, and avoid loading the joint with cable or module weight.
Power, Dynamic Range, and Noise
Available source power falls steeply with frequency because each multiplication stage is lossy. Extenders in the lower bands deliver perhaps a few milliwatts, while the highest bands may deliver only microwatts. Receiver conversion loss rises in parallel. The combined effect is that dynamic range, which can exceed 100 dB in the lower millimeter-wave bands, degrades markedly toward the sub-terahertz region. Recovering usable measurements at the top end means narrowing the IF bandwidth, averaging more, and accepting substantially longer sweep times—which in turn makes thermal drift during the sweep a more serious concern.
Probe Station Setup and On-Wafer Measurements
On-wafer measurements using probe stations enable direct characterization of devices, structures, and circuits on semiconductor wafers before packaging. This approach eliminates bond wire and package parasitics from the measurement, providing the most direct access to the device under test. However, probe-based measurements introduce their own set of challenges and require careful attention to setup and calibration.
Probe Types and Configurations
Ground-Signal-Ground (GSG) probes are standard for on-wafer mm-wave measurements, providing a well-controlled coplanar waveguide (CPW) interface to the device. The probe pitch (center-to-center spacing between adjacent signal and ground contacts) must match the DUT's pad layout, with common pitches of 50 µm, 75 µm, 100 µm, 125 µm, and 150 µm. Smaller pitches enable denser pad layouts but require more precise probe positioning and smaller probe tips that may wear more quickly.
Probe selection considerations include:
- Frequency range: Broadband probes carry a coaxial interface and cover DC to the limit of that connector—110 GHz for a 1.0 mm interface, roughly 145 GHz for a 0.8 mm interface. Above that, banded waveguide probes mate directly to an extender module and cover one waveguide band at a time, reaching several hundred gigahertz. The probe's coaxial-to-CPW or waveguide-to-CPW transition must maintain low reflection and low loss across the entire specified range.
- Impedance: Standard 50-ohm probes are most common, but 75-ohm and differential (ground-signal-signal-ground) probes are available for specific applications.
- Contact force and tip compliance: Contact force is not set directly; it follows from the overtravel applied after first touchdown, mediated by the compliance of the probe tip. Adequate force breaks through the native oxide on aluminum pads and yields contact resistance well below an ohm, while excessive force flattens tips and craters pads. The manufacturer's recommended overtravel range is the controlling specification.
- Tip material and construction: Tungsten, beryllium-copper, and plated nickel alloys offer different trade-offs among durability, contact resistance, and cost. Probes intended for soft aluminum pads use tip geometries that scrub through oxide with minimal pad damage, which matters when the same pads must be probed repeatedly.
Calibration Substrate and ISS Standards
On-wafer calibration typically uses Impedance Standard Substrates (ISS) that provide precision short, open, load, and thru structures fabricated on ceramic or semiconductor substrates. These standards are designed to match the DUT's substrate material and pad configuration as closely as possible, minimizing the effects of different dielectric constants, metal thicknesses, and pad geometries.
The substrate choice matters because electromagnetic fields extend into the substrate material, making the transmission line characteristics dependent on dielectric constant and loss tangent. An ISS calibrated on alumina (εr ≈ 9.9) will not accurately characterize devices on silicon (εr ≈ 11.9) or gallium arsenide (εr ≈ 12.9). Custom ISS substrates matching the DUT substrate provide the most accurate calibration.
Probe Positioning and Planarity
Precise probe positioning is essential for repeatable measurements. The probe station's positioning system must achieve micron-level repeatability in X, Y, and Z axes. The chuck (wafer holder) should be adjustable in tip/tilt to ensure the wafer surface is parallel to the probe travel plane. Poor planarity causes non-uniform contact across the probe tips, leading to variable contact resistance and possible open circuits on some probe tips.
Best practices for probe placement include:
- Using the probe station's microscope or camera system to align probes precisely to pad centers
- Approaching the wafer gradually to detect initial contact (visible scrubbing or resistance change)
- Applying the manufacturer's recommended overtravel consistently at every touchdown—commonly on the order of 25 to 75 µm for coplanar RF probes, which produces a horizontal scrub of roughly a third of that distance
- Checking that the scrub marks on the pads are of consistent size and position, which confirms that every tip made contact with comparable force
- Regularly cleaning probe tips to remove oxide, contaminants, or deposited material
Environmental Control
Temperature and humidity variations affect both the DUT and the calibration standards. Many probe stations include temperature-controlled chucks to maintain the wafer at a specific temperature, enabling characterization of temperature-dependent effects. The probe station enclosure should provide a stable thermal environment and may include active temperature control to maintain consistent conditions during long measurement sessions.
Humidity control prevents condensation, which can cause electrical leakage paths or corrosion. Atmospheric gases also absorb mm-wave energy along the open path between probes, fixtures, and antennas: water vapor has absorption lines near 22 GHz and 183 GHz, while molecular oxygen produces a broad band centered near 60 GHz and an additional line near 118 GHz. Free-space measurements that cross these bands can show humidity-dependent loss, so purging the probe station enclosure or the antenna range with dry nitrogen reduces measurement uncertainty.
Over-Temperature Testing
Many mm-wave devices and systems must operate across wide temperature ranges, from cryogenic conditions in space applications to high temperatures in automotive or industrial environments. Characterizing device performance across temperature reveals activation energies, thermal coefficients, and potential failure mechanisms while ensuring that specifications are met under all operating conditions.
Temperature-Controlled Test Fixtures
Over-temperature measurements require fixtures that can heat or cool the DUT while maintaining electrical performance. Probe stations with thermal chucks provide precise temperature control from -60°C to +300°C or more. The chuck temperature is controlled by resistive heaters working against liquid nitrogen or closed-cycle refrigeration, with feedback from thermocouples or resistance temperature detectors (RTDs). Dry-gas purge is essential below the dew point, since frost on a wafer ruins both the electrical measurement and the probe tips.
Key considerations for thermal measurements include:
- Thermal settling time: After changing chuck temperature, adequate time must elapse for the DUT to reach thermal equilibrium. Small devices may stabilize in seconds, while larger structures or packages may require minutes to reach steady-state temperature.
- Temperature uniformity: Temperature gradients across the DUT can cause measurement variations. High-quality thermal chucks use distributed heating elements and good thermal conductivity to minimize gradients.
- Probe thermal expansion: Probe positioners expand and contract with temperature, potentially moving probe tips relative to pads. Some systems include compensation mechanisms or require re-positioning probes at each temperature point.
- Calibration standards temperature: The ISS calibration substrate should ideally be at the same temperature as the DUT. Some advanced systems include temperature-controlled calibration substrates, but more commonly, calibration is performed at room temperature and measurements apply correction factors.
Temperature-Dependent Effects
Temperature affects virtually all device parameters. In semiconductors, carrier mobility decreases with increasing temperature, reducing transistor transconductance and maximum frequency. Bandgap energy decreases with temperature, shifting threshold voltages and leakage currents. In passive structures, conductor resistivity increases with temperature due to increased phonon scattering, increasing insertion loss. Dielectric constant and loss tangent also exhibit temperature dependence, shifting transmission line impedance and propagation velocity.
Proper characterization involves measuring S-parameters, noise figure, output power, and other key metrics at multiple temperature points spanning the intended operating range. The data reveals temperature coefficients that inform circuit design, compensation techniques, and specification limits.
De-embedding at mm-Wave Frequencies
De-embedding removes the effects of test fixtures, probe pads, transmission line feeds, and other parasitic structures from measurements, extracting the intrinsic performance of the device under test. At mm-wave frequencies, even microscopically small structures introduce significant phase shifts, reflections, and losses that must be accurately characterized and removed to obtain meaningful device parameters.
De-embedding Structures and Methods
The most common de-embedding approaches include:
- Open-Short de-embedding: This simple method uses an open structure (pads without the device) and a short structure (pads with a short circuit) to characterize shunt and series parasitics respectively. The method assumes parasitics can be modeled as simple lumped elements and works reasonably well at lower mm-wave frequencies but becomes less accurate as frequency increases.
- Open-Short-Load (Thru) de-embedding: Adding a thru or load structure improves accuracy by providing additional information about the parasitic network. The thru structure connects the probe pads directly without the DUT, characterizing the combined effect of all parasitics in the measurement path.
- Cascade de-embedding: This approach uses ABCD (chain) parameters to mathematically remove cascaded networks representing the access structures on each side of the DUT. It requires careful attention to reference plane definitions and can accumulate numerical errors if the de-embedding structures are not accurately characterized.
- 2x-Thru de-embedding: This method uses a companion structure consisting of the two fixture halves joined back to back, so that it contains exactly twice the access network that surrounds the DUT. Assuming symmetry and reciprocity, the technique splits the 2x-thru into two half-fixture models and removes them from the fixture-DUT-fixture measurement. It is the basis of automatic fixture removal in commercial software and is particularly useful when precise short and open standards are difficult to fabricate.
IEEE Std 370-2020 codified this practice for printed circuit board interconnects up to 50 GHz, defining both fixture design requirements and quantitative data-quality metrics. Its self-de-embedding consistency check is worth applying well above its nominal scope: de-embed each half-fixture model from the 2x-thru itself, and the result should approach an ideal zero-length thru. Residual insertion loss beyond roughly a tenth of a decibel or residual phase beyond about a degree indicates that the fixture model, and therefore the de-embedded DUT data, cannot be trusted. At mm-wave the same test remains diagnostic even though the tolerances that apply below 50 GHz are optimistic.
Pad Parasitics and Feed Structures
Probe pads introduce capacitance to ground and mutual capacitance between adjacent pads. Typical GSG probe pads might have 10-30 fF of capacitance per pad, which creates significant impedance at mm-wave frequencies. A 20 fF capacitance presents an impedance of only about 80 ohms at 100 GHz, shunting signal power to ground and creating reflections.
Transmission line feeds connecting probe pads to the DUT introduce propagation delay, loss, and potential impedance discontinuities. These structures must be long enough to accommodate probe placement but should be minimized to reduce their electrical impact. At 100 GHz a 100-micron line represents roughly 12 degrees of phase shift in air, and proportionally more on a dielectric—about 20 to 30 degrees for typical microstrip or coplanar waveguide on alumina or semiconductor substrates, where the effective permittivity slows the wave. Either way, the access lines add enough phase to significantly affect device characterization if they are not properly de-embedded.
Accuracy and Validation
De-embedding accuracy depends critically on how well the de-embedding structures match the actual parasitics around the DUT. Ideally, de-embedding structures are fabricated on the same wafer, in the same process, with identical geometry to the DUT's access structures. Even small variations in metal thickness, dielectric thickness, or lateral dimensions can introduce de-embedding errors.
Validation techniques include:
- Measuring devices with known characteristics (such as precision resistors or transmission lines) through the same access structures and verifying that de-embedded results match expected values
- Comparing results from different de-embedding methods to identify inconsistencies
- Checking for physical plausibility (passive devices should not show gain, impedances should be within reasonable ranges)
- Performing electromagnetic simulations of the complete structure and comparing to measured results
Fixturing Challenges
Test fixtures provide mechanical support and electrical connections to devices under test, but at mm-wave frequencies, even simple fixtures become complex electromagnetic structures. Connectors, transitions, transmission lines, and mounting hardware all interact with the signal in ways that can dominate the measurement if not carefully managed.
Connector Transitions and Launchers
Transitioning from coaxial connectors to planar transmission lines (microstrip, stripline, CPW) requires careful impedance matching across the entire frequency range. Commercial connector launchers are available for many substrate types and frequencies, but custom designs may be necessary for specific applications. The transition region must minimize reflections while maintaining consistent characteristic impedance.
Edge-launch connectors that mate directly to the edge of a PCB or substrate offer compact solutions but require precise control of substrate thickness, metal thickness, and edge preparation. Through-board connectors provide more mechanical stability but introduce via transitions that can cause resonances at mm-wave frequencies.
Shielding and Resonances
Test fixtures often include metal enclosures for mechanical protection and electromagnetic shielding. However, metal enclosures create cavity resonances that can dramatically affect measurements when the cavity dimensions correspond to half-wavelength or full-wavelength resonances. At 100 GHz, a 1.5 mm cavity dimension creates a half-wavelength resonance, causing sharp changes in S-parameters.
Mitigation strategies include:
- Careful dimensional control to ensure resonances fall outside the measurement band
- Using absorptive materials to dampen cavity modes
- Designing enclosures with non-parallel walls to break up standing wave patterns
- Including venting or cutouts that disrupt cavity modes while maintaining adequate shielding
Substrate Modes and Surface Waves
At mm-wave frequencies, electromagnetic energy can couple into substrate modes that propagate within the dielectric material rather than along the intended transmission line. These modes cause power loss, frequency-dependent behavior, and coupling to other structures on the substrate. Thinner substrates and lower dielectric constants reduce substrate mode coupling but may compromise mechanical strength or other design parameters.
Ground plane vias and fences can suppress substrate modes by creating electromagnetic barriers. Proper placement and spacing of these structures (typically much closer than λ/4) ensures effective mode suppression without introducing new resonances or impedance discontinuities.
Cable Effects and Stability
Coaxial cables connecting test equipment to fixtures or probes introduce loss, phase shift, and impedance variations that must be calibrated out or accounted for in measurements. At mm-wave frequencies, cable performance becomes increasingly challenging, and careful cable selection, handling, and stabilization are essential for accurate measurements.
Cable Loss and Dispersion
Coaxial cable loss rises with frequency, driven primarily by the skin effect, so attenuation increases roughly with the square root of frequency. Loss climbs steeply through the mm-wave bands: a high-quality 2.4 mm cable assembly might exhibit several dB/meter near its 50 GHz limit, while a 1.0 mm assembly rated to 110 GHz commonly exceeds 10 dB/meter at the top of its range. Because the cable sits in series with the device, this loss directly reduces measurement dynamic range and can make low-level or high-insertion-loss measurements difficult.
Phase-stable cables use special construction techniques—including solid or foam dielectrics with low temperature coefficients, mechanically stable center conductors, and reinforced outer conductors—to minimize phase variations with temperature and flexure. These cables are essential for applications requiring precise phase measurements or for systems where cables cannot remain absolutely stationary.
Cable Flexure and Repeatability
Flexing a coaxial cable changes its electrical length and characteristic impedance as the center conductor position shifts slightly within the dielectric. These changes cause phase and amplitude variations in measurements. The effect is more pronounced at higher frequencies where the wavelength is shorter.
Best practices for cable handling include:
- Minimizing cable movement during measurements
- Supporting cables to prevent stress on connectors
- Maintaining consistent cable bend radii (following manufacturer minimum bend radius specifications)
- Allowing cables to stabilize after movement before performing critical measurements
- Using cable conditioning (repeated flexing) on new cables to stabilize their electrical properties
Cable Alternatives
For applications where cable loss, phase stability, or repeatability are inadequate, alternatives include:
- Semi-rigid cables: These cables use a solid copper outer conductor and a solid PTFE dielectric, providing excellent electrical stability and low loss. However, they cannot be flexed repeatedly and must be carefully formed to the required shape.
- Waveguide: At lower frequencies the cross-section needed for single-mode operation makes rectangular waveguide bulky relative to coax, but in the upper mm-wave bands it offers markedly lower loss, and banded waveguide modules (for example WR-10 covering 75 to 110 GHz) are the standard interface for many systems. Waveguide requires careful flange alignment and is sensitive to mechanical tolerances.
- Direct probe mounting: Mounting VNA test port modules directly at the probe station eliminates cables entirely, providing the best possible phase stability and lowest loss. This approach requires careful thermal management and may limit accessibility to the DUT.
Beyond S-Parameters: Noise, Power, and Modulation
Linear network characterization answers only part of the question. A millimeter-wave amplifier, mixer, or transceiver must also be qualified for noise, large-signal behavior, and modulation fidelity, and each of those measurements has its own mm-wave complications.
Noise Figure
The classical Y-factor method compares the receiver output with a calibrated noise source switched between hot and cold states. Its weakness at millimeter-wave frequencies is the noise source itself: calibrated solid-state sources with well-characterized excess noise ratio become scarce and expensive as frequency rises, and in banded waveguide systems they may not exist at all for the band in question. Two alternatives dominate. The cold-source method measures the device's noise power with a matched termination at the input and uses a vector-corrected network analyzer to supply the gain and match data needed to solve for noise figure, which also removes the mismatch error that plagues Y-factor at high frequencies. Hot-and-cold-load radiometry substitutes physical loads at ambient temperature and at liquid-nitrogen temperature for the electronic noise source, trading convenience for traceability. Because mm-wave noise figures are often only a few decibels while the loss ahead of the DUT may be larger than that, accurately accounting for input-path loss is usually the dominant error term.
Large-Signal and Nonlinear Characterization
Load-pull measurement, which maps output power, efficiency, and linearity against the impedance presented to the device, is harder at mm-wave because passive tuners must place their reflection at the DUT reference plane through lossy intervening structure. Every decibel of loss between tuner and device shrinks the reflection coefficient the device actually sees, so passive tuners often cannot reach the impedances of interest. Active load-pull, which synthesizes the desired reflection by injecting a phase- and amplitude-controlled signal back into the output port, sidesteps the loss limit and is common in this frequency range. Harmonic tuning is frequently impractical at mm-wave simply because the second and third harmonics fall outside the usable band of the available hardware. Nonlinear vector network analyzers and behavioral formulations such as X-parameters extend the S-parameter concept to large-signal operation, capturing harmonic and compression behavior in a form that circuit simulators can use.
Spectrum, Modulation, and Phase Noise
Signal analyzers reach millimeter-wave frequencies through external harmonic mixers, which downconvert a waveguide band into the analyzer's native range. The price is conversion loss, degraded sensitivity, and image and spurious responses that must be identified and suppressed, typically by signal-identification algorithms or by preselection. Modulated-signal metrics add a further constraint: FR2 carriers occupy channel bandwidths of up to 400 MHz, and aggregated measurements are wider still, so the analyzer's amplitude and phase flatness across the full analysis bandwidth becomes a direct contributor to measured error vector magnitude. Careful frequency-response correction of the test path is mandatory, and the residual EVM of the instrumentation must be comfortably below the value being measured on the device. Oscillator and synthesizer phase noise is measured by downconverting against a reference source, since direct measurement at the carrier frequency is impractical; multiplication from a lower-frequency reference degrades phase noise by 20 log N decibels, which is why frequency-multiplied mm-wave sources are noisier than their microwave origins suggest.
Over-the-Air and Antenna-in-Package Testing
Many millimeter-wave products expose no usable RF connector. Phased arrays for 5G FR2 handsets, automotive radar sensors, and antenna-in-package transceivers integrate the radiating elements with the silicon, so the only accessible interface is the radiated field. Over-the-air (OTA) measurement becomes not a convenience but the sole option, and the antenna, package, and radio must be characterized as one inseparable unit.
Far-Field Criteria and Chamber Methods
Radiated measurements are meaningful only in the far field, conventionally taken as a range length of at least 2D²/λ, where D is the largest dimension of the radiating aperture. The quantity grows quickly: a 5 cm aperture at 28 GHz requires roughly 0.47 m, while the same aperture measured as part of a larger device at higher frequency can demand several meters. 3GPP's study of FR2 test methods recognizes three principal approaches. The direct far field method places the device at a true far-field distance in an anechoic chamber and is restricted to modest apertures. The indirect far field method, implemented as a compact antenna test range, uses an offset parabolic reflector to convert the spherical wave from a feed antenna into a locally planar wave, producing a far-field condition inside a chamber a fraction of the size a direct range would need. Near-field to far-field transformation samples amplitude and phase on a surface close to the device and computes the far-field pattern numerically, which is efficient for pattern work but demands accurate phase data and a stable positioner.
Radiated Metrics and Path Calibration
Because no conducted port exists, specifications are written in radiated terms: effective isotropic radiated power (EIRP) and effective isotropic sensitivity (EIS) for a single beam direction, total radiated power (TRP) integrated over the sphere, and spherical coverage expressed as a percentile of the cumulative distribution over all directions. Each requires the measurement path to be calibrated end to end. The usual approach is gain comparison, in which a reference antenna of known gain replaces the device and establishes the total path loss from the device position to the receiver. That path loss is substantial: free-space loss alone over one meter at 28 GHz is about 61 dB, before chamber cabling and any reflector or feed contribution.
Chamber quality sets the accuracy floor. Absorber performance, quiet-zone size and ripple, positioner accuracy and repeatability, and the alignment of the device's phase center to the axis of rotation all propagate directly into the reported numbers. Beam-steering adds another dimension, since a phased array must be measured across the beam states it will actually use, which multiplies test time and makes measurement speed a practical design constraint on the test system.
Measurement Repeatability
Repeatability—the ability to obtain the same measurement result when measuring the same DUT multiple times under the same conditions—is a critical quality metric for any measurement system. At mm-wave frequencies, many factors that are negligible at lower frequencies become significant sources of variation, making repeatability analysis essential for understanding measurement reliability.
Sources of Repeatability Variation
Common sources of repeatability variation include:
- Connector repeatability: Each connect/disconnect cycle introduces small variations in mating depth, alignment, and contact resistance. Manufacturers specify connection repeatability as a maximum magnitude and phase deviation; for premium 1 mm interfaces these figures are small fractions of a decibel and on the order of a degree, but they accumulate across every joint in the path and degrade as the connectors accumulate mating cycles.
- Probe contact variations: Differences in overtravel, contact force, and probe tip position from one touchdown to the next create measurement variations. Automated probe stations with precision positioners achieve better repeatability than manual systems.
- Thermal drift: Temperature changes affect cable phase, connector dimensions, and DUT characteristics. Measurements made hours apart may show variations due to room temperature cycling or equipment warm-up drift.
- Calibration stability: VNA calibrations drift over time due to temperature changes, connector wear, and other factors. Re-calibrating periodically improves repeatability for measurements made over extended periods.
Quantifying Repeatability
Standard practice for quantifying repeatability involves making multiple independent measurements of the same DUT, with full disconnect/reconnect cycles between measurements. Statistical analysis of the results provides mean values and standard deviations for each measured parameter. For S-parameters, this typically involves analyzing magnitude (in dB) and phase (in degrees) separately.
A typical repeatability test might include:
- 10 measurement cycles with complete disconnect/reconnect between each
- Calculation of mean and standard deviation for S11, S21, S12, and S22 at each frequency point
- Identification of frequency ranges where repeatability is best and worst
- Comparison against manufacturer specifications for test equipment
Improving Repeatability
Strategies to improve measurement repeatability include:
- Using high-quality connectors and replacing them before they reach end-of-life
- Implementing strict procedures for connection torque, probe placement, and handling
- Controlling environmental temperature and allowing equipment to reach thermal equilibrium
- Minimizing cable movement and using phase-stable cables where needed
- Calibrating more frequently when highest accuracy is required
- Using automated systems to eliminate operator variability
Measurement Uncertainty
Measurement uncertainty quantifies the doubt that exists about a measurement result. Unlike a simple repeatability study that examines variation under nominally identical conditions, uncertainty analysis considers all known sources of error—both random and systematic—and combines them to establish confidence intervals around reported values. At mm-wave frequencies, comprehensive uncertainty analysis is essential for comparing results between labs, qualifying measurement systems, and ensuring that devices meet specifications.
Uncertainty Sources and Components
Measurement uncertainty includes contributions from many sources:
- VNA instrumentation uncertainty: The VNA itself has specified accuracy limits for magnitude and phase measurements, which depend on factors like IF bandwidth, averaging, power level, and frequency. Manufacturer data sheets provide uncertainty specifications, often expressed as a function of measured magnitude.
- Calibration standard uncertainty: Imperfect knowledge of calibration standard characteristics—such as offset delays in short/open standards, characteristic impedance of line standards, or frequency-dependent loss—introduces systematic errors that propagate through the calibration and into subsequent measurements.
- Connector and interface repeatability: Random variations in connector mating contribute to measurement uncertainty. This component is typically characterized through repeatability studies as described in the previous section.
- Drift and stability: Long-term drift in cables, connectors, and VNA electronics creates time-dependent variations. Calibration intervals must be short enough that drift remains within acceptable limits.
- Mismatch uncertainty: Impedance mismatches between the VNA, cables, connectors, and DUT create multiple reflections that interfere constructively or destructively depending on electrical length. Ripple in measured S-parameters often indicates mismatch interactions.
- Noise and dynamic range: At mm-wave frequencies with high cable loss and DUT insertion loss, signal levels may approach the VNA's noise floor, increasing measurement noise and uncertainty.
Uncertainty Budgets and Analysis
A complete uncertainty analysis requires identifying all significant uncertainty sources, quantifying each contribution, and combining them according to established statistical methods (typically following the ISO Guide to the Expression of Uncertainty in Measurement, or GUM). The process involves:
- Identifying all input quantities that affect the measurement result
- Determining the standard uncertainty (one standard deviation) for each input
- Calculating sensitivity coefficients that describe how changes in each input affect the output
- Combining the individual uncertainty contributions using root-sum-square for uncorrelated sources
- Expressing the final result with an expanded uncertainty (typically k=2 for approximately 95% confidence)
For S-parameter measurements, uncertainty analysis often separates magnitude and phase, as they have different uncertainty sources and sensitivities. The result is typically expressed as an uncertainty bound that varies with frequency, such as: S21 magnitude ±0.15 dB (k=2) and S21 phase ±2.5° (k=2) over the frequency range 75-110 GHz.
Validation Through Inter-laboratory Comparisons
The ultimate validation of measurement uncertainty claims comes from inter-laboratory comparisons, where multiple independent labs measure the same artifacts using their own equipment and methods. If the reported values from different labs agree within their stated uncertainties, this provides confidence that uncertainty analyses are realistic. Significant disagreements indicate that one or more labs have underestimated uncertainty or have unrecognized systematic errors.
Formal round-robin measurement programs, often organized by standards bodies or industry consortia, provide structured frameworks for these comparisons. They establish measurement protocols, provide traveling standards, and analyze results statistically to identify outliers and evaluate overall measurement consistency across the community.
Best Practices for mm-Wave Measurements
Success in mm-wave characterization requires attention to many details that can be overlooked at lower frequencies. Key best practices include:
- Plan calibration carefully: Select calibration methods and standards appropriate for the measurement frequency, connector type, and required accuracy. Verify calibration quality before proceeding to DUT measurements.
- Minimize connections: Every connector pair adds uncertainty. Use the shortest practical cable lengths and minimize adapters and other connections in the signal path.
- Control the environment: Maintain stable temperature and humidity. Allow equipment to reach thermal equilibrium before calibration and measurements.
- Handle connectors with care: Use proper torque, inspect regularly for damage, keep threads clean, and retire worn connectors before they degrade measurement quality.
- Verify measurements: Use known standards or devices with expected characteristics to confirm that results are reasonable. Compare different measurement methods when possible.
- Document everything: Record calibration details, equipment serial numbers, cable identification, measurement conditions, and procedures. This documentation enables troubleshooting and supports uncertainty analysis.
- Check the seams: When a result is assembled from several waveguide bands or from separate calibrations, inspect the magnitude and phase where the segments meet. A visible step at a band boundary is evidence of a calibration or alignment problem, not of device behavior.
- Understand limitations: Recognize when measurements approach equipment limits for noise floor, dynamic range, or frequency coverage. Do not over-interpret results near these boundaries.
Conclusion
Millimeter-wave measurements demand rigorous attention to calibration, fixturing, environmental control, and uncertainty analysis. The techniques described here—from VNA calibration and banded frequency extension through probe station setup, de-embedding, cable management, radiated testing, and repeatability assessment—form the foundation for accurate mm-wave characterization. As integration tightens and frequencies climb into the sub-terahertz region, the trend is unmistakable: the accessible measurement plane keeps moving away from the device. Connectors give way to waveguide flanges, waveguide flanges give way to probe tips, and probe tips give way to radiated fields, with each step adding structure that must be calibrated out or characterized as part of the device.
Success in mm-wave measurements comes from understanding that nearly every physical detail matters at these frequencies. What appears to be a small imperfection—a worn connector, a slightly bent cable, a few degrees of temperature variation—can significantly impact measurement results. By applying the best practices and techniques outlined here, engineers can develop robust measurement capabilities that support development of next-generation mm-wave systems for communications, radar, sensing, and imaging applications.