Scene Modeling Photometry (SMP)
- Scene Modeling Photometry (SMP) is a pixel-level, generative approach that models a static host, variable point source, and background directly on original exposures for precise transient photometry.
- It employs a forward-modeling strategy using high-resolution spline or pixelized galaxy models and per-epoch PSF, astrometric, and photometric adjustments to fit the entire scene simultaneously.
- By avoiding image resampling and subtraction, SMP minimizes aliasing and calibration disparities, achieving millimagnitude-level precision in space-based Roman surveys and improved consistency in ZTF analyses.
Scene Modeling Photometry (SMP) is a pixel-level, generative approach to transient photometry in which the entire scene around a supernova is modeled directly on the original exposures rather than on resampled or subtracted images. In the Roman context, the method is explicitly identified with what Holtzman et al. (2008) called “scene modeling” photometry, and in the ZTF context it is described as a maximum-likelihood fit of a parametric model containing the SN PSF, a common host-galaxy light distribution, sky/background, per-epoch astrometric mapping, and photometric scaling derived from stars (Rubin et al., 2021, LaCroix et al., 4 Sep 2025). Across both implementations, the central idea is stable: a static host and a time-variable point source are fit jointly in pixel space over many epochs, so that host subtraction, PSF handling, calibration transfer, and uncertainty propagation are all embedded in one forward model.
1. Conceptual definition and mathematical structure
In its Roman formulation, the “scene” is the combination of a time-variable point source, a static host-galaxy light distribution, and a per-image spatially flat background. The forward model predicts the expected counts in every pixel of every exposure from a set of parameters that includes galaxy spline coefficients, SN fluxes, SN position, small astrometric shifts, and sky levels, with the observed image taken to be model plus noise (Rubin et al., 2021).
The galaxy component for image at detector pixel is written as
where is the underlying time-independent galaxy surface brightness, are coordinates on an 11× oversampled grid, are the WCS transforms, are small image-by-image astrometric corrections, and is the PSF including convolution with the pixel response. The full Roman pixel model is
with the SN flux in image 0, 1 the local pixel area on the sky, and 2 the constant sky level. Statistically, the model is 3, with fitting performed by minimizing a weighted 4 over all images and pixels (Rubin et al., 2021).
The ZTF implementation uses an equivalent scene-modeling logic, but with a different parameterization. Its core model is
5
where 6 indexes exposures and 7 pixels within the cutout. Here 8 is the SN flux, 9 the per-exposure PSF, 0 the static host galaxy in a reference frame, 1 the kernel mapping the reference-frame PSF to the exposure PSF, 2 an additive sky term, 3 a scalar photometric scale, and 4 the geometric transform from reference frame to exposure (LaCroix et al., 4 Sep 2025).
In both cases, SMP is therefore not merely a fitting convenience but a generative statistical model of the measurement process. A plausible implication is that its defining feature is the preservation of the native data model—PSF, WCS, pixel response, and background—rather than the creation of a derived difference image on which photometry is performed afterward.
2. Distinction from image subtraction and classical PSF photometry
Traditional supernova photometry commonly uses image resampling plus subtraction, or PSF/aperture photometry on host-subtracted stacked images. Those approaches require accurate resampling and PSF matching. The Roman study argues that they are problematic for the Wide Field Instrument because the data are undersampled, with PSF FWHM after pixel convolution ranging from 5 to 6, or 7 to 8 pixels depending on band, and because the survey will often have a single, undithered exposure per filter per epoch (Rubin et al., 2021).
The failure mode identified for resampling-based approaches is aliasing and correlated errors. In the Roman simulations, a drizzle-like resampling approach with a square kernel of 9 for references, followed by PSF photometry on subtracted images, performs much worse than scene modeling on the noise-free data: residuals correlate strongly with the galaxy second derivative, showing over-smoothing and aliasing (Rubin et al., 2021). The contrast is structural. Forward modeling constructs a high-resolution scene on the sky, convolves once with a high-resolution PSF and pixel response, then samples at the native pixel grid for each exposure, respecting exact pointing and rotation.
The ZTF comparison emphasizes a different set of limitations in difference imaging. In DR2 forced photometry, a single sky position is estimated from alerts, a ZOGY difference image is created for each exposure, and a fixed PSF at the fixed SN position is fit on the difference image with only the amplitude free. The paper identifies three limitations in that implementation: a fixed, non-spatial PSF on difference images; a fixed SN position based on the alert pipeline; and the inability to apply the same method to stars, since stars are not present in the difference image. As a result, stars are calibrated on science images with one pipeline, while SNe are measured on difference images with another, breaking the direct SN–star flux comparison (LaCroix et al., 4 Sep 2025).
This contrast addresses a common misconception: SMP is not simply a more elaborate variant of subtraction. The Roman and ZTF descriptions both treat it as a method that works directly on calibrated science frames and uses all available pixels, including host-galaxy morphology and background structure, rather than only the difference signal (Rubin et al., 2021, LaCroix et al., 4 Sep 2025).
3. Parameterizations, fitting strategy, and calibration logic
The Roman implementation fits, for each filter, a 2D spline galaxy model 0 on a uniform grid of spline nodes in sky coordinates, a single sky position for the SN, an independent flux parameter 1 for each image, a per-image constant sky level 2, and per-image small astrometric shifts 3. The node spacing is 2 nodes per PSF FWHM, described as roughly Nyquist sampling of the PSF-convolved galaxy. The modeled region is a circular patch around the SN, and the galaxy model is forced to go to zero at the edge of this patch to break degeneracy between galaxy pedestal and sky background (Rubin et al., 2021).
Roman PSFs are generated with WebbPSF and convolved with ideal square 4 pixels. The model uses 11× oversampling, with the galaxy defined on a 5 grid and the PSF sampled on the same oversampled grid. The supersampling strategy is stated to ensure the 2D spline is represented with 6 accuracy after convolution, sufficient for millimagnitude-level SN photometry. Optimization uses Levenberg–Marquardt least squares, with uncertainties and covariances of the flux parameters coming from the LM covariance matrix and validated empirically via pull distributions. The paper explicitly states that there is no MCMC and no explicit regularization term added to the cost function (Rubin et al., 2021).
The ZTF pipeline begins from detrended quadrant images, masks, and WCS. For each SN and band it selects on-frames in rest-frame phase 7 days and adds off-frames outside 8 days, with at least six times more off-frames and at least 100 in total, to constrain the host galaxy model. Pre-processing includes SExtractor source detection, sky background construction in object-masked regions, inverse-variance pixel weight maps based on sky variance and flat-field, and cross-matching to GAIA eDR3 and Pan-STARRS1 for PSF training, astrometric refinement, and photometric alignment (LaCroix et al., 4 Sep 2025).
Its PSF model follows Astier et al. (2013): an analytic core, usually a Moffat profile, plus a pixelized residual grid, both varying smoothly with position. For undersampled data, the analytic PSF part is integrated over each pixel using a 3-point Gaussian quadrature. Astrometric transforms are built from GAIA positions and proper motions, with relative transforms defined as 9. Photometric alignment is done through a global linear model
0
with the typical uncertainty on 1 reported as 2 mmag (LaCroix et al., 4 Sep 2025).
A distinctive feature of the ZTF formulation is the insistence that the same flux estimator be used for SNe and for calibration stars. In SMP, the same PSF model, astrometric transforms, kernels, and weighting scheme are used to fit the SN plus galaxy and to fit field stars with galaxy fixed to zero. The paper presents this symmetry as a key cosmology requirement because the average SMP flux of a star becomes the natural anchor for calibration, placing stars and SNe on the same instrumental system (LaCroix et al., 4 Sep 2025). This suggests that SMP is as much a calibration architecture as a host-subtraction method.
4. Validation on Roman undersampled imaging
The Roman validation uses simulated data designed to approximate about 3 of a full Roman SN survey: 2,061 simulated SNe Ia and 762,570 image cutouts around SNe or host galaxies in the 4, 5, 6, 7, and 8 filters, with a 5 observer-frame day cadence over a year and visits rotated by about 9 relative to the previous visit (Rubin et al., 2021). Host galaxies are drawn from the VELA cosmological simulations and provide oversampled, SED-informed mock images in Roman filters; SN light curves are generated with SALT2-Extended via SNCosmo, with redshifts uniform in 0 between 0.7 and 2.0 and SN positions drawn from a probability map proportional to the slightly smoothed 1 host image (Rubin et al., 2021).
The validation strategy is multi-layered. First, the authors examine galaxy modeling errors against local second derivatives of the host surface brightness, arguing that curvature rather than gradient is the relevant diagnostic. Figures 5 and 6 are summarized as showing no trend of noise-free residuals with local Laplacian, indicating that the spline grid with 2 nodes per FWHM is adequate and does not need to be finer (Rubin et al., 2021). Second, they study pull distributions,
2
whose means and medians are consistent with zero except in 3 at some magnitudes, while the dispersions are near unity at bright magnitudes and rise to about 4 at faint magnitudes in 5 and 6, implying that uncertainties are underestimated by 7 there (Rubin et al., 2021).
Third, they analyze flux bias versus magnitude. For noise-free images, all filters show biases at the 8 mmag level across most of the magnitude range, with somewhat larger biases at the very faint end, likely due to small galaxy subtraction residuals and the 9 limitation of the reprojection step. For images with noise, 0, 1, 2, and 3 remain extremely close to unity in the average scaling between true and observed flux, while 4 shows a consistent bias of order 2 mmag across much of the magnitude range (Rubin et al., 2021). The abstract summarizes the band-level result as biases at the millimagnitude level in the four red filters and larger 5–3 millimagnitude biases in 6, all meeting the 0.5% Roman inter-filter calibration requirement (Rubin et al., 2021).
At the light-curve level, SALT2-Extended fits to the recovered multiband photometry show distance-modulus residuals 7 whose mean and median in redshift bins are consistent with zero at the few mmag level, with dispersion increasing with redshift as expected. Sensitivity tests using 8 indicate that for 5-band Roman light curves the per-band sensitivities are typically 9, so a 5 mmag calibration bias in any single band would generally translate into 0 mmag distance-modulus bias per SN. Since the measured photometric biases are 1–3 mmag, their impact is described as safely below Roman’s systematic-error budget from calibration (Rubin et al., 2021).
These results also bear directly on the specific challenge of undersampling. Roman’s WFI will often have one undithered exposure per filter per epoch, so per-epoch images cannot be Nyquist-sampled by dithering. The paper’s conclusion is that undersampling can be handled by defining the model in sky coordinates, supersampling the galaxy and PSF, and fitting many epochs jointly, including numerous reference images (Rubin et al., 2021).
5. ZTF implementation and the requirement of a common instrumental system
The ZTF SN Ia DR2 paper places SMP in the setting of a large optical survey with 3628 spectroscopically confirmed Type Ia supernovae discovered during the first 2.5 years of operation, with DR2 light curves originally based on forced photometry on difference images. The motivation for reprocessing with SMP is explicit: precision cosmology requires control of total systematic error on SN fluxes at the few 2 level, or approximately 3 in flux per band, whereas the difference-image approach delivers 4–2% photometry, sufficient for population and rate studies but not for that calibration target (LaCroix et al., 4 Sep 2025).
Operationally, the SMP processing is large-scale. Across DR2, the selected data amount to approximately 5 million quadrants, about 6 TB, or 7 of all ZTF quadrants over that period. A complete SMP processing run requires 2–3 weeks at CC-IN2P3 using up to approximately 2000 cores. On a single 3 GHz core, the SN fit takes 8 s even for high-cadence SNe. Two full iterations produced about 9493 light curves for 3582 SNe, with failures at the 9 level occurring mainly in crowded Galactic-plane fields or poorly sampled 0-band light curves (LaCroix et al., 4 Sep 2025).
The paper’s technical emphasis falls on repeatability and calibration. Using SMP star light curves, the authors measure a weighted RMS as a function of magnitude. In 1 band, free-PSF repeatability for bright stars has a floor of approximately 10 mmag, while SMP repeatability is worse by approximately 2 mmag, which the paper identifies as the expected penalty of fixing positions under finite astrometric noise. The reported astrometric residuals relative to GAIA are 2 pixel and 3 pixel for bright stars, and for a Gaussian PSF the approximate flux bias is
4
yielding about 2 mmag flux bias on average and 3–4 mmag at best seeing (LaCroix et al., 4 Sep 2025).
The direct comparison between SMP and DR2 forced photometry is revealing. Differences in magnitudes per point, defined as SMP minus DR2, have mean offsets of approximately 30 mmag in 5, 50 mmag in 6, and 20 mmag in 7, with DR2 fluxes on average brighter than SMP. When both light-curve sets are fit with SALT2.4, stretch 8 and time of maximum 9 are consistent, and color 0 differs by only about 10 mmag. However, the derived distance modulus offset is about 90 mmag, with SMP distances larger than DR2. The paper interprets this as evidence that DR2 absolute calibration and/or linearity are not yet at the level required for precision cosmology (LaCroix et al., 4 Sep 2025).
The calibration chain itself is formulated through time-averaged SMP instrumental magnitudes of stars matched to PS1: 1 The fits reveal strong color terms consistent with ZTF bandpasses that are approximately 6 nm bluer in 2, 14 nm redder in 3, and 15 nm redder in 4 than PS1, together with 10–20 mmag non-linearities over 14–18 mag in calibration residuals versus magnitude (LaCroix et al., 4 Sep 2025).
6. Systematics, limitations, and broader significance
A central lesson from the ZTF study is that SMP does not eliminate instrumental systematics by itself. The paper identifies a new sensor effect, dubbed the “pocket effect,” which distorts the PSF in a flux-dependent manner and produces non-linearities of a few percent in photometry. Its signatures include serial-direction astrometric residuals that depend on star magnitude, strong variation of PSF skewness in 5 with brightness while 6-skewness remains essentially constant, strong CCD-to-CCD variation, background dependence that makes the effect strongest at low sky background and nearly absent above about 6000 ADU sky, and a worsening after a readout waveform upgrade in November 2019 (LaCroix et al., 4 Sep 2025).
The photometric consequences are substantial. On low-background frames, comparisons of PSF and aperture magnitudes show that some CCDs are nearly unaffected, while strongly affected CCDs exhibit 8% peak-to-peak variations across the magnitude range in 7, and even modestly affected CCDs show 1–2% peak-to-peak non-linearities. The paper concludes that the pocket effect alone can induce approximately 1–7% photometric non-linearities, far above the 0.1% target, and that both DR2 and current SMP are affected because they share the same raw images and detrending (LaCroix et al., 4 Sep 2025).
The stated remedy is a pixel-level correction applied during detrending, before PSF modeling and SMP. The correction must be sensor-dependent and time-dependent, reflecting pre- and post-2019 readout conditions. The authors do not provide the final model in the paper, but describe the intended strategy as measuring flux-dependent PSF shape and centroid shifts versus position and background, building a forward model of charge redistribution along the serial register, and inverting it to deconvolve deferred charges at the pixel level (LaCroix et al., 4 Sep 2025).
Roman’s limitations are different and more controlled. The validation assumes perfect knowledge of the PSF, perfect calibration of detector effects such as nonlinearity and IPC, no SED dependence in the PSF, and otherwise accurate WCS apart from fitted small shifts. The paper therefore positions its results as a validation of the forward-modeling algorithm under idealized instrumental knowledge, not as an end-to-end demonstration under all real-world systematics (Rubin et al., 2021). A plausible implication is that the Roman results establish the intrinsic adequacy of scene modeling for undersampled imaging, while the ZTF results demonstrate that cosmology-grade performance also requires sensor physics, bandpass modeling, and calibration control at the pixel level.
Within the broader SMP literature, the Roman paper explicitly connects its implementation to Holtzman et al. (2008) for SDSS-II, Astier et al. (2013) for SNLS, Brout et al. (2019) for DES, Rodet et al. (2008), and Bongard et al. (2011). The shared philosophy is joint fitting of a static galaxy plus variable SN directly in pixel space across multiple epochs. The differences lie in the galaxy basis and instrumental regime: Roman uses a 2D spline basis with nodes at about 8 FWHM spacing and is centered on undersampled, space-based PSFs with changing roll angles, whereas ZTF uses a pixelized host model in a ground-based survey and foregrounds the requirement that stars and SNe be measured with the same estimator (Rubin et al., 2021, LaCroix et al., 4 Sep 2025).
Taken together, these studies define SMP as a methodological framework for precision transient photometry rather than a single algorithmic template. In Roman simulations it delivers millimagnitude-level SN flux biases—9 mmag in 00, 01, 02, and 03, and about 2–3 mmag in 04—meeting the stated inter-filter calibration requirement (Rubin et al., 2021). In ZTF, it provides a statistically powerful and calibration-consistent alternative to difference-image forced photometry, but also exposes unresolved non-linearity, bandpass, and absolute-calibration problems that currently prevent both the legacy DR2 light curves and the present SMP measurements from being used for precise cosmological analyses (LaCroix et al., 4 Sep 2025).