1. Introduction
Spatial accessibility to healthcare services is a central concern in urban planning and public health research, as it directly influences healthcare utilization, equity, and health outcomes. Among different healthcare facility types, tertiary hospitals exhibit distinctive service characteristics: they provide specialized care, attract patients across large spatial ranges, and operate within competitive and hierarchical hospital systems. Accurately capturing accessibility to such facilities therefore requires models that reflect both spatial opportunity structures and residents’ healthcare-seeking behavior.
The Two-Step Floating Catchment Area (2SFCA) method and its variants have become widely used tools for measuring healthcare accessibility due to their intuitive structure and moderate data requirements [
1,
2,
3]. Extensions incorporating distance decay, such as Gaussian or kernel-based 2SFCA models [
4,
5], have improved realism by allowing service influence to decrease with travel cost. However, these approaches typically rely on exogenously defined catchment thresholds or globally fixed decay parameters [
6,
7,
8], implicitly assuming that residents evaluate healthcare opportunities within uniform spatial or temporal limits.
Such assumptions are problematic in the context of tertiary hospitals. Unlike primary care facilities, tertiary hospitals do not operate within locally bounded catchments; patients often travel well beyond conventional thresholds when perceived service quality or specialization justifies longer journeys. Empirical studies consistently show substantial heterogeneity in healthcare travel behavior across space and population groups, suggesting that fixed catchment definitions are conceptually misaligned with actual decision processes.
To address these limitations, studies have incorporated probabilistic choice models, such as the Huff model [
9], into accessibility analysis [
10,
11]. These approaches allow facilities to compete probabilistically based on attractiveness and distance, thereby capturing the substitution effects among alternative facilities. However, in most existing integrations, probabilistic choice is used only to reweight facility influence, while the effective search range remains externally imposed or implicitly constrained. As a result, the fundamental question of how far residents are behaviorally willing to search remains unresolved.
This study argues that accessibility assessment should be grounded in observed travel behavior rather than arbitrary thresholds. To achieve this, we leverage large-scale taxi trajectory data to empirically calibrate the Huff model parameters from observed hospital choices and to construct realistic travel impedance matrices from actual trips.
Taxi data capture actual road-network conditions, including congestion and route choices. We construct a complete origin–destination travel time matrix between all residential townships/subdistricts and tertiary hospitals based on observed taxi trips. Building on these data-driven foundations, we propose a Behavior-Calibrated Endogenous Choice 2SFCA (BCEC-2SFCA) framework. The model supplies origin-specific choice probabilities that reflect competition among hospitals, calibrated from observed taxi trips, while the 2SFCA structure accounts for supply–demand ratios and cumulative opportunities without imposing exogenous catchment thresholds. The resulting accessibility measure captures both the attractiveness of hospitals and the actual travel burden experienced by residents, without imposing exogenous catchment thresholds.
The proposed framework is empirically validated using multi-source data from Ningbo, China: hospital bed capacity, township-level population, over 1.58 million taxi trajectories, and annual outpatient volumes from 16 tertiary hospitals. By comparing modeled accessibility outcomes with observed travel patterns and outpatient distributions, we provide behavioral evidence for the validity of our approach.
This study makes three main contributions: (1) empirical calibration of Huff model parameters from observed taxi trips; (2) construction of high-resolution travel impedance matrices from real-world trajectories; and (3) integration of these elements into a behavior-calibrated 2SFCA framework (BCEC-2SFCA) that generates endogenous search ranges.
It is important to emphasize that the proposed framework is specifically designed for high-order medical services—namely tertiary hospitals—which are characterized by strong inter-facility competition, large service areas, and prevalent cross-regional healthcare-seeking behavior. The endogenous search range mechanism is predicated on the assumption that patients actively compare and choose among multiple alternative facilities, a condition that does not hold for proximity-based services such as community health centers, where travel is typically constrained to a local catchment. The applicability of the BCEC-2SFCA model to primary care or other locally bounded service contexts should therefore be considered limited, and the conventional Gaussian 2SFCA or similar threshold-based methods may remain more appropriate in those settings.
2. Literature Review
Accessibility is a crucial concept for estimating the ease with which people can overcome spatial barriers to access services or resources. W. G. Hansen first defined it in 1959 as “the potential of opportunities for interaction,” laying the foundation for accessibility research [
12]. Since then, the concept has expanded to encompass transportation, temporal, and psychological dimensions, becoming a core metric in geography and urban planning [
13].
Current methods for measuring healthcare resource accessibility primarily include the cumulative opportunity measure [
14], distance-based measure, network analysis [
15], isochrone method [
16], gravity model, and the two-step floating catchment area (2SFCA) method [
17]. These approaches have evolved by incorporating more realistic factors to enhance their explanatory power [
18].
The minimum distance method [
19] and buffer analysis evaluate accessibility by calculating the straight-line distance to the nearest facility or setting a fixed service radius. These methods are simple and intuitive, often used for preliminary delineation of hospital service areas. However, neither considers the constraints of the actual transportation network, which may lead to discrepancies with real-world conditions.
Network analyses can answer a range of questions related to linear networks such as roads, railways, rivers, facilities and utilities. This spatial analysis technique uses network data (usually linear features such as roads, footpaths) to calculate distances between points or nodes on the network. This approach underpins the satellite navigation systems found in many cars. Common applications are route finding, route planning, identifying the closest facility by travel time or distance, calculation of service areas (e.g., areas within 10 min’s walk of a bus stop), etc. There are various ways of parameterising the analysis based on typical road speeds, blockages, and minimising the use of smaller or remote parts of the network depending on the task [
20].
The gravity model is one of the most representative and classic methods in spatial interaction and accessibility analysis. Originating from Newton’s law of universal gravitation, it was first introduced into human geography and transportation planning by Zipf [
21,
22], and later endowed with a rigorous statistical mechanics foundation by Wilson within the framework of entropy maximization, thereby evolving from an empirical formula into a systematic theory of spatial interaction [
23]. In accessibility research, the fundamental principle of the gravity model is that the intensity of interaction between two locations (e.g., trip volume) is proportional to the “mass” of the origin (e.g., population size) and the “mass” of the destination (e.g., facility service capacity), and inversely proportional to a power of the spatial impedance (distance or time) between them. When applied to measure accessibility, the model is often used in its “potential model” form, which sums the service capacities of all medical facilities
j weighted by a distance decay function for a given residential location
i, thereby yielding a continuous accessibility score for that location. Compared with the cumulative opportunity method or the isochrone method, the gravity model offers the advantage of incorporating continuous distance decay, thus avoiding the issue of threshold discontinuity. In contrast to the two-step floating catchment area method, its traditional form typically does not explicitly address competition effects on the demand side. The two-step floating catchment area (2SFCA) method was first proposed by Radke and Mu (2000) [
17], and the version modified by Luo and Wang (2003) [
1] has been widely used. The 2SFCA method fully considers the influence of both supply-side and demand-side scales on accessibility and has become a primary research method for measuring spatial accessibility in the healthcare field. Extensions of the 2SFCA method by researchers can be summarized into four categories: (1) extensions of the distance decay function, (2) extensions of the catchment radius, (3) mode-of-transport-based extensions, and (4) extensions targeting demand or supply. The distance decay functions commonly used by researchers include a piecewise decay form that assigns different weights to different travel time intervals within the catchment area [
24], the gravity decay model, the kernel density form [
5], and the Gaussian form [
4].
Luo et al. proposed an enhanced two-step floating catchment area (E2SFCA) method to address the assumption of uniform accessibility within the catchment area in the traditional 2SFCA method. This method assigns differentiated weights to different travel time intervals in both steps to reflect the distance decay effect [
24]. The gravity-based 2SFCA method introduces the distance decay function from the well-established gravity model into the 2SFCA framework, thereby improving the discrete distance decay function of the E2SFCA method into a continuous function. The kernel density 2SFCA (KD2SFCA) method, proposed by Dai [
5], incorporates a kernel density-based distance decay function within the 2SFCA catchment radius. The kernel density decay function is concave: when the distance is small, accessibility decays slowly with increasing distance; as the distance increases, the decay accelerates. Based on the 2SFCA method, Dai introduced a Gaussian function to continuously weight the distance decay within the catchment area and formally proposed the Gaussian 2SFCA method. Using the Detroit metropolitan area as the study region, Dai investigated the impact of racial residential segregation and spatial accessibility to healthcare facilities on the rate of late-stage breast cancer diagnosis, aiming to more accurately assess the spatial accessibility of medical facilities [
4].
After its introduction, the 2SFCA method rapidly became one of the most widely recognized and commonly used approaches in healthcare accessibility research. Numerous subsequent studies have employed the 2SFCA method or its improved versions for accessibility assessment. Ahmed et al. used walking time and the 2SFCA method to evaluate accessibility to cyclone shelters in the coastal region of Bangladesh, and applied the Gini coefficient to analyze interregional inequalities [
25]. Cheng et al. took Nanjing, China as the study area, used the 2SFCA method to assess the spatial accessibility of multi-level hospital services for the elderly, and employed the Gini coefficient to examine spatial equity in accessibility [
26]. Mao et al. incorporated multiple different transportation modes into the mobile two-step floating catchment area method and compared it with the traditional single-mode two-step floating catchment area method. They found that the accessibility calculated using multi-modal transportation technology held more practical significance than results from single-mode calculations [
27]. Qu et al. systematically evaluated the influence of spatial resolution selection on measurement errors in public transit accessibility, using Melbourne, Australia as the study area. They selected representative points at different resolutions, calculated three accessibility metrics—proximity, cumulative opportunity, and Gaussian 2SFCA—and compared the measurement errors against building-level benchmark data [
28]. Juel et al., considering that travel time to medical facilities varies according to specific daily conditions, used an improved two-step floating catchment area method to evaluate the accessibility of public healthcare services in developing countries in the Caribbean region. They proposed that increasing the allocation of physicians could improve the accessibility of medical facilities [
29].
In addition, some scholars have incorporated competitive relationships into the 2SFCA framework to improve the method. Jin et al. proposed a 2SFCA approach integrating multiple transportation modes (driving, transit, and walking) with the Huff choice model to assess spatial accessibility to healthy and unhealthy food stores in Austin, Texas, USA, revealing significant disparities between the urban core and peripheral areas [
30]. Subal et al. improved the Huff 3SFCA method by replacing discrete zone-based weights with a continuous Gaussian distance-decay function to quantify the spatial accessibility of general practitioners in the Swabian region of Germany, identifying accessibility differences between urban and rural areas as well as within cities at the micro-scale [
31]. Mo et al. employed a Huff-model-based 2SFCA method to evaluate the accessibility of community care services in Hong Kong [
32]. Zhu et al. conducted a study in Luyang District, Hefei City, China, where they integrated the Gaussian distance-decay function and the Huff model into the traditional 2SFCA framework, while also considering two low-carbon travel modes (walking and public transit) to evaluate the accessibility of meal service facilities for the elderly [
33]. Chen et al. proposed a generalized flow-based 2SFCA method using Wuhan City as a case study. They extracted residents’ origin–destination flows for healthcare visits from taxi trajectory data, adaptively determining the catchment area, distance-decay coefficient, and attractiveness for each hospital. On this basis, they constructed a global popularity index based on the h-index concept and a local preference index reflecting community-level healthcare preferences, and incorporated these indices into the Huff model for accessibility assessment [
34].
While the integration of Huff models into accessibility analysis has gained increasing attention, the calibration of Huff model parameters—the attractiveness elasticity α and the distance-decay coefficient β—remains a critical challenge. In most existing Huff-FCA studies, these parameters are either assumed based on prior literature (e.g., α = 1, β = 2) or determined through ad hoc sensitivity tests. However, because α and β govern the sensitivity of residents to facility attractiveness and travel impedance, respectively, arbitrarily setting their values risks producing accessibility estimates that deviate systematically from actual travel behavior.
Recognizing this limitation, a growing body of research has sought to calibrate Huff model parameters using observed behavioral or location-based data, moving beyond assumed parameter values toward empirically grounded estimates. Yue et al. were the first to use taxi GPS trajectory data as a substitute for traditional questionnaire surveys to conduct an exploratory calibration of the Huff model, thereby demonstrating the feasibility of using large-scale trajectory data for retail trade area analysis [
35]. Gong et al. proposed using geographically and temporally weighted regression (GTWR) to simultaneously calibrate spatial and temporal parameter variations in the Huff model, revealing that attractiveness sensitivity is positively correlated with residents’ wealth and negatively correlated with leisure time [
36]. Gong et al. also developed an agent-based Destination-Aware Activity Simulation (DAS) model, incorporating geographically weighted regression (GWR) to calibrate agents’ destination choice parameters, enabling the dynamic simulation of residents’ multi-activity sequences [
37]. Liang et al. proposed a time-aware dynamic Huff model (T-Huff), employing large-scale mobile phone location data and particle swarm optimization to characterize the time-varying patterns of market shares across different business formats [
38]. Lu et al. investigated the influence of the number of sampling points on the calibration accuracy of the Huff model using mobile phone location data, finding that a small number of sampling points with high trip volumes can yield better calibration performance [
39].
In summary, three critical gaps in the existing Huff-based FCA literature are identified and addressed in this work. First, Huff parameters are typically assumed rather than empirically calibrated from local healthcare travel data. Second, impedance matrices are usually derived from idealized distances, failing to capture real-world travel conditions. Third, even with Huff probabilities, the effective service range remains constrained by exogenously imposed thresholds. This study addresses these gaps by integrating empirically calibrated Huff choice probabilities—derived from large-scale taxi trajectory data—into a 2SFCA framework to generate endogenous hospital search ranges, thereby eliminating reliance on both assumed behavioral parameters and exogenously imposed catchment thresholds. More precisely, the proposed BCEC-2SFCA framework makes three distinct contributions: (1) joint empirical calibration of α and β from observed taxi trips; (2) construction of an impedance matrix from actual travel times; and (3) a probabilistic 2SFCA structure in which catchment areas are endogenously generated by choice probabilities. The following section details the data and methods that operationalize this framework.
3. Methodology
3.1. Overview of the Framework
This study develops a Behavior-Calibrated Endogenous Choice 2SFCA (BCEC-2SFCA) framework for tertiary hospitals by integrating a Huff-type probabilistic choice model with the supply–demand logic of the two-step floating catchment area (2SFCA) method. The framework is designed to replace exogenously specified service catchments with an endogenously generated search range, inferred from observed travel behavior, hospital attractiveness, and inter-hospital competition.
The methodological workflow consists of three components. First, hospital-oriented taxi trajectories are used to construct an empirical origin–destination impedance matrix based on observed travel time and travel distance. Second, the parameters of a Huff choice model are calibrated from revealed destination choices so that the probability of selecting each hospital reflects the joint effects of hospital attractiveness and travel impedance. Third, the estimated choice probabilities are embedded into a probabilistic 2SFCA structure to compute hospital-specific supply–demand ratios and origin-specific accessibility scores.
The conceptual difference from conventional 2SFCA is important. In standard Gaussian 2SFCA models, the analyst predefines the effective service range using a fixed threshold or a globally imposed decay function. In the proposed framework, by contrast, the effective service range is not fixed a priori. Instead, it emerges from the estimated choice probabilities across all hospitals. A hospital with stronger attractiveness or weaker competitive pressure can exert influence over a wider area, whereas a less attractive hospital faces a narrower effective reach. Therefore, the search range is treated as endogenous to the healthcare system, rather than externally imposed.
This section presents the theoretical specification of the BCEC-2SFCA framework, including the Huff probabilistic choice model, the derivation of endogenous search ranges, and the two-step accessibility formulation. The empirical implementation—data sources, matrix construction, and parameter calibration—is detailed separately in
Section 4.
3.2. Endogenous Search Range Through Probabilistic Hospital Choice
3.2.1. Hospital Choice Model Based on Calibrated Huff-Type Probabilities
Residents’ choice of tertiary hospitals is modeled using a Huff-type probabilistic formulation:
where:
i = 1, 2, …, n denote demand locations, represented by townships/subdistricts;
j = 1, 2, …, m denote tertiary hospitals;
is the choice probability of the resident at demand point i for hospital j;
is the supply capacity of hospital j;
is the travel impedance from demand point i to hospital j, measured by observed travel time or travel distance;
is the service-capacity elasticity parameter;
is the impedance-decay parameter;
denotes the summation over the supply capacity of all alternative tertiary hospitals.
Equation (1) implies that the probability of choosing a hospital increases with hospital attractiveness and decreases with travel impedance. More importantly, it is normalized over all competing hospitals, so the attractiveness of any hospital is always evaluated relative to the alternatives available to the same origin. This allows inter-hospital competition to be represented directly within the accessibility model.
The term “endogenous search range” can now be defined rigorously. For any demand location
i, the effective set of hospitals
considered by residents is not determined by a fixed threshold such as 30 min or 20 km. Instead, it is generated by the set of hospitals with non-negligible choice probabilities:
where
is a small behavioral relevance threshold. Under this definition, the search range varies across origins and hospitals according to the joint effects of attractiveness, impedance, and competition. Thus, range is an outcome of the model rather than an external assumption.
3.2.2. Parameter Calibration from Observed Trips
To avoid imposing arbitrary parameter values,
α and
β are calibrated using observed hospital-oriented taxi trips. Let
denote the number of observed taxi trips from demand location
i to hospital
j, and let the observed hospital choice share be:
The Huff parameters are estimated by fitting the modeled choice probabilities
in Equation (1) to the observed shares
. This can be implemented using Log-centered transformation and Ordinary Least Squares (OLS) on the origin-hospital observations (See
Section 4.2 for details).
The estimated parameters quantify two distinct behavioral mechanisms:
α captures the sensitivity of residents to hospital capacity, while β measures how quickly hospital attractiveness declines with increasing travel impedance. Because these parameters are estimated from observed mobility rather than borrowed from previous studies, the accessibility model becomes more behaviorally grounded and more context-specific.
3.3. Construction of the Empirical Impedance Matrix
The spatial impedance matrix is built from taxi trajectory data linking demand locations to tertiary hospitals. For each observed origin-hospital pair, travel time and travel distance are extracted from actual taxi trips. These observations reflect real road-network conditions, including congestion, route choice, and temporal variation, thereby providing a more realistic measure of healthcare-related travel cost than Euclidean distance or idealized network paths. The impedance matrix is constructed solely from observed taxi trips, providing complete coverage for all OD pairs.
3.4. Probabilistic 2SFCA Accessibility Model
The calibrated choice probabilities are then incorporated into a probabilistic 2SFCA framework.
3.4.1. Step 1: Hospital Supply–Demand Ratio
For each hospital
j, the expected demand is defined as the population-weighted sum of origin-specific hospital choice probabilities:
where
denote the population at demand location i. This expression replaces the conventional fixed-catchment demand aggregation. In standard 2SFCA, only demand within a predefined threshold contributes to hospital congestion. Here, all demand locations may contribute, but their contributions are weighted by the estimated probability of choosing that hospital. Consequently, the effective catchment of each hospital becomes behavior-dependent and hospital-specific.
The effective supply–demand ratio of hospital
j is then given by:
A higher indicates that hospital j has greater effective service capacity relative to the demand it is expected to attract.
3.4.2. Step 2: Accessibility at Each Demand Location
The accessibility of origin
i is defined as the expected supply–demand ratio across all hospitals, weighted by the hospital choice probabilities:
Substituting Equation (5) into Equation (6) yields:
Equation (7) is the core accessibility measure used in this study. It retains the two-step logic of 2SFCA by accounting for both hospital congestion and demand-side access, but it eliminates the need for externally imposed catchment thresholds. Accessibility is instead determined by the endogenous hospital choice structure estimated from observed travel behavior.
3.5. Interpretation of the Endogenous Search Range
Under the proposed model, the effective influence range of a hospital is not fixed across space. For a large and attractive hospital, the term may offset higher travel impedance, allowing the hospital to retain meaningful choice probabilities over a wider area. Conversely, smaller or less attractive hospitals may have influence concentrated in nearby locations only. Likewise, for residents in peripheral areas with few competitive options, the dominant hospital may exhibit a broader effective search range than for residents in central areas with many alternatives.
Therefore, the endogenous search range is not represented by a single citywide threshold. Instead, it is embedded in the probability surface {Pij}, which varies across origins and facilities and is jointly shaped by attractiveness, impedance, and competition.
To provide a more intuitive geographical interpretation, the endogenous search range can be likened to a “probability cloud” surrounding each hospital, rather than a sharply bounded circle. In conventional threshold-based models, a hospital is either inside or outside a resident’s search range depending on whether the travel cost falls below a fixed cutoff—a binary, all-or-nothing designation. In the BCEC-2SFCA framework, by contrast, every hospital exerts some degree of influence on every origin, but the strength of that influence decays continuously with travel impedance and is further modulated by the hospital’s attractiveness relative to its competitors. A hospital with high attractiveness (e.g., a large tertiary center) may retain a non-negligible choice probability even for residents located 50 km away, meaning that it remains part of their effective search range. Conversely, a smaller hospital may fade into the background of the probability surface beyond a much shorter distance. The effective search range is therefore not a predefined geometric shape but an emergent property of the spatial pattern of choice probabilities—a gradient that captures both the friction of distance and the pull of service quality. This probabilistic, gradient-based conception of search range aligns more closely with the continuous nature of human spatial behavior than do conventional hard thresholds.
3.6. Standardization and Spatial Comparison
To facilitate spatial visualization and comparison across townships/subdistricts, the raw accessibility values
Ai are standardized using max-normalization:
where
∈(0,1]. Higher values indicate better relative access to tertiary hospital services.
3.7. Benchmark Model: Gaussian 2SFCA
To evaluate the effect of the proposed endogenous-range framework, its results are compared with those of a conventional Gaussian 2SFCA model. In the benchmark model, we use travel time as impedance, hospital supply–demand ratios are calculated using a predefined travel threshold and a Gaussian decay function:
Step 1:
where
is the fixed catchment threshold and
is the Gaussian decay function.
Unlike the proposed model, this benchmark relies on exogenously defined spatial influence. The comparison therefore highlights the extent to which behavior-calibrated, endogenous search ranges alter the resulting accessibility pattern.
3.8. Methodological Contribution
The proposed method contributes to accessibility assessment in three ways. First, it estimates hospital choice sensitivities from observed trips rather than assuming parameter values. Second, it constructs the impedance matrix from real trajectory data rather than idealized distances. Third, it embeds these empirically estimated choice probabilities into a supply–demand accessibility framework—the BCEC-2SFCA—in which the effective hospital search range is generated endogenously. The resulting model is particularly suitable for tertiary hospital systems characterized by long-distance patient movement, strong facility competition, and heterogeneous travel behavior.
5. Results and Analysis
5.1. Accessibility Spatial Pattern
Using the calibrated Huff model parameters () and the observed travel time matrix and travel distance matrix derived from taxi trajectories, we computed the healthcare accessibility index for each of the 81 townships/subdistricts in Ningbo, allowing a comparison of accessibility results under two impedance measures: travel time and travel distance. This section presents the spatial distribution of the normalized accessibility scores and compares the results with those obtained from the conventional Gaussian two-step floating catchment area (Gaussian 2SFCA) method. The comparison highlights how the incorporation of empirically calibrated choice probabilities and real-world travel times reshapes the accessibility landscape, revealing competitive effects in hospital-dense areas and moderating the accessibility gap between central and peripheral regions.
The BCEC-2SFCA model proposed in this study and the traditional Gaussian 2SFCA based on a fixed threshold exhibit systematic differences in characterizing the spatial pattern of accessibility to tertiary hospitals in Ningbo (
Figure 8).
The results calculated by the BCEC-2SFCA model reveal significant spatial heterogeneity in the accessibility of tertiary hospitals in Ningbo. High-accessibility areas are highly concentrated in central urban districts such as Haishu, Jiangbei, and Yinzhou, forming accessibility peak zones centered around key hospitals like First Affiliated Hospital of Ningbo University (Yuehu and Waitan campuses) and Ningbo Medical Center Lihuili Hospital. This region has a high density of hospital resources, where residents’ choice sets overlap significantly, leading to pronounced competitive effects and consequently higher effective accessibility calculated by the model. Areas with lower accessibility are primarily distributed in Fenghua District and Beilun District.
A comparison between the BCEC-2SFCA model (travel time) and the BCEC-2SFCA model (travel distance) reveals that the accessibility distributions of the two models are generally similar, with the main differences located in the western part of Haishu District and the northern part of Fenghua District. To further assess the robustness of the proposed framework, we examined the consistency between accessibility estimates derived from travel time and travel distance. The two sets of accessibility scores exhibit a strong positive correlation, with a Pearson correlation coefficient of 0.845 (p < 0.001) and a Spearman rank correlation coefficient of 0.817 (p < 0.001). This strong agreement indicates that the behavioral core of the model—the Huff probability structure—is not unduly sensitive to the choice of impedance metric. Whether travel time or travel distance is used as the measure of spatial separation, the resulting accessibility patterns remain remarkably consistent, underscoring the stability and reliability of the BCEC-2SFCA approach.
Crucially, even in peripheral areas, the minimum standardized accessibility calculated by the BCEC-2SFCA model is 0.31, with no residential points being absolutely inaccessible (value of 0). This reflects the reality of residents’ healthcare-seeking behavior: even those in outer suburbs have a possibility (albeit low probability) of choosing to travel to major tertiary hospitals in the urban center.
In contrast, the spatial pattern presented by the Gaussian 2SFCA model with a fixed catchment size is markedly different. In this study, the fixed threshold of 45 min was selected for the Gaussian 2SFCA method. This choice was guided by two complementary considerations. First, based on the complete origin–destination travel time matrix, 98.15% of all observed hospital-oriented taxi trips fell within 45 min, confirming that this threshold captures the vast majority of actual healthcare-seeking travel behavior. Second, following a maximum–minimum distance principle, we computed for each of the 81 townships/subdistricts the travel time to its nearest tertiary hospital. The maximum of these nearest-hospital travel times across all townships was 40.29 min. Any threshold below this value would leave at least one township with no hospital within range. The adopted threshold of 45 min exceeds this lower bound, thereby ensuring complete spatial coverage of the study area while remaining firmly grounded in the observed travel time distribution. However, even with this adjustment, the disparities in the model results clearly reveal the systematic spatial bias inherent in traditional fixed-threshold models. The Gaussian 2SFCA model exhibits the following issues:
- 1.
Overestimation of accessibility in areas proximate and at intermediate distances to hospitals: Within close and intermediate proximity to tertiary hospitals, the Gaussian 2SFCA model overestimates accessibility. This is primarily attributed to the relatively slow decay of the Gaussian decay function over short distances. Furthermore, the longer the preset exogenous fixed threshold, the larger the zone of slow decay becomes, thereby amplifying this effect. Additionally, the Gaussian 2SFCA model fails to capture the “choice overload” and competitive effects resulting from the high concentration of hospitals in central urban areas. In the BCEC-2SFCA model, although residents in these areas face a multitude of high-quality hospitals, the choice weights derived from the Huff model concentrate their selections on a few nearest or most attractive hospitals. The influence of other neighboring hospitals is diluted due to competition, consequently lowering the comprehensive accessibility score. In contrast, the Gaussian model simply performs a linear summation of the influence of all hospitals within a 45 min range, overlooking the concentration of resident decision-making in contexts of abundant choice.
- 2.
Underestimation of accessibility in areas distant from hospitals: The Gaussian 2SFCA model significantly underestimates accessibility in regions farther from hospitals. This represents the primary bias identified in this study. On one hand, due to its fixed hard cutoff, the Gaussian model completely disregards the long-distance travel behavior of residents in peripheral areas seeking quality medical resources. Conversely, the BCEC-2SFCA model does not set a fixed threshold, this incorporates large core hospitals, which, despite being distant, still have a probability of being selected, into the effective choice set, thereby assigning more realistic accessibility values to these areas. On the other hand, the Gaussian decay function decays very rapidly over intermediate distances, resulting in extremely low weights for hospitals at long distances. Consequently, the Gaussian model produces a substantial accessibility disparity between central and peripheral areas.
5.2. Analysis of Accessibility Ranking Changes
To quantitatively assess the differences in evaluation results between the BCEC-2SFCA Model (Travel Time) and Gaussian 2SFCA Model, this section further compares the accessibility rankings of each region under both models. Changes in ranking directly reflect how the relative assessment of regional healthcare resource accessibility is systematically reconfigured when residents’ behavioral decisions are incorporated.
Among the 81 townships/subdistricts, the accessibility rankings derived from the two models exhibit a substantial reconfiguration. Overall, 37 areas (45.7%) experienced an increase in rank under the BCEC-2SFCA model compared to the Gaussian 2SFCA model, indicating that their relative level of healthcare resource accessibility was reassessed as higher after incorporating residents’ behavioral decisions. Conversely, 43 areas (53.1%) experienced a decrease in rank, with these areas predominantly concentrated in the hospital-dense central city and inner suburbs. Rankings for only 1 area (1.2%) remained unchanged, highlighting the widespread corrective effect of the behaviorally endogenous model on the assessment outcomes.
Figure 9 provides a visual comparison of the accessibility rankings for all townships/subdistricts under the two models using a scatter plot. Regarding the magnitude of change, the ranking shifts are more pronounced than those observed in preliminary analyses. The average increase for areas that rose in rank was 16.14 places, while the average decrease for areas that fell in rank was 13.88 places. These figures indicate that the BCEC-2SFCA model not only reorders many areas but also does so with considerable magnitude, particularly upgrading peripheral locations and downgrading some central ones.
To further quantify the linear relationship between the accessibility rankings derived from the two models, we performed an ordinary least squares regression using the Gaussian 2SFCA rankings as the independent variable (
x) and the BCEC-2SFCA model rankings as the dependent variable (
y). The fitted linear equation is:
The slope of 0.6575, being substantially less than 1, indicates a notable compression effect: the BCEC-2SFCA model tends to assign relatively lower rankings (i.e., worse positions) to many of the top-ranked areas under the Gaussian model, while assigning relatively higher rankings (i.e., better positions) to many of the lower-ranked areas. This aligns with the earlier observation that the BCEC-2SFCA model moderates extreme rankings, narrowing the accessibility gap between central and peripheral regions. The positive intercept of 14.0435 implies that, on average, the BCEC-2SFCA rankings are shifted upward (i.e., numerically larger, meaning less favorable) relative to the Gaussian rankings, especially for areas with very low Gaussian rankings (near the top). Conversely, for areas with very poor Gaussian rankings (large x values), the BCEC-2SFCA rankings become comparatively better than the linear trend would predict. The R2 value of 0.4323 demonstrates a moderate linear association, confirming that the reconfiguration induced by the BCEC-2SFCA model is systematic, though with substantial scatter reflecting the model’s nuanced adjustments.
These regression results reinforce the conclusion that the BCEC-2SFCA model corrects the systematic biases of the fixed-threshold Gaussian model by moderating extreme rankings—downgrading overestimated central areas and upgrading underestimated peripheral regions—while still preserving a recognizable overall spatial order.
To further illustrate the spatial pattern of the adjustments introduced by the BCEC-2SFCA model,
Figure 10 presents a difference map of the standardized accessibility scores, computed as BCEC-2SFCA minus Gaussian 2SFCA for each township/subdistrict. Green shades indicate negative differences (BCEC < Gaussian), representing areas where the BCEC-2SFCA model assigns lower accessibility than the Gaussian benchmark. These areas are predominantly concentrated in the central urban districts of Haishu, Jiangbei, and Yinzhou, reflecting the competition effect that moderates accessibility in hospital-dense locations. In contrast, yellow, orange, and red shades indicate positive differences (BCEC > Gaussian), representing areas where the BCEC-2SFCA model yields higher accessibility. These positive adjustments are most pronounced in the peripheral townships of Fenghua, Beilun, and Zhenhai districts, confirming that the endogenous search range mechanism substantially elevates accessibility estimates in outlying areas. The spatial divergence revealed by the difference map underscores the extent to which the BCEC-2SFCA framework corrects the systematic overestimation of central accessibility and underestimation of peripheral accessibility inherent in fixed-threshold Gaussian models.
5.3. Implications for Planning and Policy
The model differences lead directly to divergent assessments of whether regional medical resources are adequate, carrying clear policy implications.
- 1.
Re-evaluating “Service Blind Spots”: Although the Gaussian 2SFCA model did not generate “service blind spots” (areas with zero accessibility) in this study, its fixed threshold and hard cutoff still assume that residents in peripheral areas can only access services from a very limited number of hospitals. Consequently, the accessibility level in these areas remains extremely low. In contrast, under the BCEC-2SFCA model, the accessibility disparity between peripheral and central areas is significantly narrowed. For example, in certain townships or subdistricts of Fenghua District and Beilun District, the accessibility scores under the Gaussian model are extremely low, even approaching zero. However, under the BCEC model, they still maintain accessibility scores of approximately 0.3–0.4 and exhibit non-negligible choice probabilities for major hospitals in the city center. This finding implies that, for regions identified by traditional methods as severely resource-deficient, policy priorities should shift away from simply constructing new hospitals to fill absolute gaps (which may be economically inefficient and ineffective), towards strategies aimed at reducing the actual barriers residents face in accessing existing high-quality resources located farther away. For instance, establishing a direct express bus route connecting the western townships of Fenghua District to the Ningbo First Hospital area could substantially reduce travel time for residents in these peripheral communities, thereby improving effective accessibility without requiring new hospital construction. Similarly, expanding telemedicine services and streamlining referral pathways between community health centers in Beilun’s eastern townships and major tertiary hospitals could mitigate the burden of long-distance travel.
- 2.
Transformation in Resource Allocation Assessment: In the central districts of Haishu, Jiangbei, and Yinzhou, the BCEC-2SFCA model assigns lower accessibility scores than the Gaussian model for many townships, despite their proximity to multiple hospitals. This downward adjustment reflects the competition effect captured by the Huff choice probabilities: when many attractive hospitals are available, residents’ selections concentrate on a few preferred facilities, diluting the effective demand at neighboring hospitals. This finding implies that the marginal benefit of constructing new large hospitals in the urban core may be overestimated. Instead, resource investment should focus on systemic integration and functional differentiation—for instance, strengthening specialized departments at existing hospitals to reduce redundant competition and improve overall service efficiency.
- 3.
Methodological Implications: The systematic differences between the BCEC-2SFCA and Gaussian 2SFCA rankings—particularly the fact that the average rank adjustment for certain townships exceeds 10 positions—indicate that when assessing the accessibility of high-level, wide-coverage service facilities, the primary source of measurement error is not data precision, but erroneous a priori assumptions regarding residents’ behavioral search ranges. Adopting endogenous, behavior-based choice models to define the search range can fundamentally avoid the systematic spatial bias (overestimating the periphery, underestimating the center) caused by exogenous, rigid thresholds, providing a scientific basis for more equitable and effective spatial planning of healthcare.
In summary, the comparative analysis in this chapter demonstrates that, compared to the traditional Gaussian 2SFCA model, the BCEC-2SFCA framework generates a continuously varying and behaviorally plausible spatial picture of accessibility. It not only eliminates unrealistic “service blind spots” but also reveals competitive effects in resource-concentrated areas, providing a more reliable analytical foundation for precisely identifying shortcomings in the spatial allocation of medical resources and formulating differentiated improvement strategies.
6. Model Validation Using Hospital Outpatient Volume
To objectively evaluate the predictive capability of the proposed BCEC-2SFCA model in reproducing residents’ hospital choice distribution, this chapter employs annual hospital outpatient volume—a dataset independent of the model construction process—for validation. By comparing predicted shares with actual outpatient shares, we quantify the predictive accuracy of different models (the travel time-based BCEC-2SFCA model and the Gaussian 2SFCA method) and examine their stability across hospital types.
6.1. Data Preparation and Processing
6.1.1. Actual Outpatient Share
Let
Vj denote the annual outpatient volume of hospital
j. The actual outpatient share
of hospital
j is defined as:
where
is the total number of tertiary hospitals for which annual outpatient volume data were obtained.
6.1.2. Model-Predicted Outpatient Share
For any model M∈{BCEC-2SFCA, Ga2SFCA}, the predicted share represents the proportion of total healthcare demand that the model allocates to hospital j.
- 1.
For the proposed model (BCEC-2SFCA), the predicted share is computed as:
is the number of demand points;
is the residential population of demand point i;
the calibrated Huff choice probability;
- 2.
For the comparative model Ga2SFCA, its predicted share is calculated in accordance with its internal logic. Specifically, it is defined as the proportion of the weighted demand population visiting hospital j, as represented by the denominator in the first step of the model, relative to the total similarly weighted demand population citywide.
The numerator represents the Gaussian-weighted demand population that can reach hospital j, and the denominator is the total weighted demand attracted by all hospitals citywide.
6.2. Validation Metrics
To comprehensively compare the consistency between predicted and actual shares, the following metrics are adopted:
6.2.1. Correlation Metrics
Calculate two types of correlation coefficients between the predicted share series for each model and the actual share series .
- 1.
Pearson Correlation Coefficient r: Measures the linear relationship between two series.
where
and
denote the means of the respective series.
- 2.
Spearman Rank Correlation Coefficient ρ: Assesses the monotonic relationship, robust to outliers.
where
and
are the ranks of
and
, respectively.
6.2.2. Error Metrics
Calculate the prediction errors of each model.
- 1.
Root Mean Square Error (RMSE): Reflects the average magnitude of prediction errors.
- 2.
Mean Absolute Percentage Error (MAPE): Expresses relative error in percentage terms.
6.3. Validation Results
This section validates the predictive accuracy of the BCEC-2SFCA model against the Gaussian 2SFCA model using annual outpatient volume data from 16 tertiary hospitals in Ningbo for which such data are available. The actual outpatient share is computed by Equation (27); the predicted shares and are obtained from Equation (28) and Equation (29), respectively. Four metrics are employed: Pearson correlation coefficient, Spearman rank correlation coefficient, root mean square error (RMSE), and mean absolute percentage error (MAPE).
Table 5 summarizes the validation results. The BCEC-2SFCA model achieves a Pearson correlation of 0.5682 (
p = 0.0216) and a Spearman rank correlation of 0.5265 (
p = 0.0362), indicating a moderately strong and statistically significant monotonic relationship with the actual outpatient shares. In contrast, the Gaussian 2SFCA model yields much weaker correlations: Pearson 0.1163 (
p = 0.6679) and Spearman 0.0529 (
p = 0.8456). This demonstrates that the BCEC-2SFCA model is considerably more capable of replicating the relative order of hospital attractiveness and the competitive effects captured by real-world choice behavior.
Regarding prediction errors, the BCEC-2SFCA model yields a lower MAPE (32.17% vs. 42.12%), suggesting that its empirically calibrated choice probabilities better capture the relative distribution of patient demand across hospitals. However, its RMSE (0.02700) is slightly higher than that of the Gaussian 2SFCA model (0.02211), indicating that while the BCEC-2SFCA model excels at replicating the rank order of hospital attractiveness, it does not uniformly achieve lower absolute errors across all facilities. This trade-off underscores that the model’s primary strength lies in capturing competitive effects and relative attractiveness rather than minimizing aggregate absolute prediction error.
Overall, the validation results confirm that the proposed BCEC-2SFCA framework, with its empirically calibrated parameters and behaviorally derived choice probabilities, provides a more plausible representation of residents’ hospital-seeking behavior in terms of ranking and spatial competition than the conventional fixed-threshold Gaussian 2SFCA model. However, this improved behavioral realism does not translate into uniformly superior absolute predictive accuracy. The framework is therefore particularly suited for evaluating the relative accessibility of high-order medical services, where capturing competition and patient choice is more critical than minimizing absolute prediction error.
To further examine whether model performance varies by hospital type, a supplementary subgroup analysis was conducted. The 16 hospitals with available outpatient data were divided into two groups based on the median number of beds: larger hospitals (≥1014 beds, N = 8) and smaller hospitals (<1014 beds, N = 8). For the larger hospitals, the BCEC-2SFCA model achieved a Pearson correlation of 0.192 and a Spearman rank correlation of 0.262, compared with 0.007 and −0.214 for the Gaussian model. The RMSE and MAPE for BCEC-2SFCA were 0.0219 and 23.9%, respectively, versus 0.0311 and 34.9% for the Gaussian model. For the smaller hospitals, BCEC-2SFCA yielded a Pearson correlation of 0.458 and a Spearman rank correlation of 0.548, whereas the Gaussian model produced values of −0.103 and 0.024. In this subgroup, BCEC-2SFCA had a lower MAPE (31.9% vs. 34.5%) but a slightly higher RMSE (0.0270 vs. 0.0242).
Although the reduced sample size in each subgroup precludes definitive statistical inference (none of the correlations reached conventional significance levels), the pattern of results is consistent with the full-sample findings: the BCEC-2SFCA model consistently outperforms the Gaussian benchmark in capturing the relative attractiveness of hospitals, regardless of hospital scale. The absolute prediction errors remain modest for larger hospitals but increase for smaller ones—a challenge common to both models and likely attributable to the greater volatility and lower absolute volumes of outpatient visits at smaller facilities. This exploratory analysis suggests that the BCEC-2SFCA framework maintains its behavioral advantage across heterogeneous hospital types, while also highlighting the inherent difficulty of predicting demand for smaller, lower-volume hospitals. Future work with larger validation samples could more rigorously assess the boundary conditions of the proposed approach.
6.4. Robustness Check: Model Calibration with All 22 Tertiary Hospitals
To assess the sensitivity of the parameter estimates to the composition of the calibration sample, we repeated the log-centered OLS calibration using all 22 tertiary hospitals located within the study area (1782 OD pairs; 81 townships × 22 hospitals). The estimated parameters were = 0.9586 (p < 0.001) and = 3.0685 (p < 0.001), with = 0.620. These values are comparable in magnitude and statistical significance to the estimates obtained from the 14-hospital subset used in the main analysis ( = 1.1758, = 2.9608, = 0.665).
Using the alternatively calibrated parameters, we recomputed the BCEC-2SFCA accessibility scores and the predicted outpatient shares for the 16 hospitals with available outpatient data. The resulting validation metrics are as follows: Pearson’s r = 0.505 (p = 0.046), Spearman’s ρ = 0.541 (p = 0.030), RMSE = 0.0254, and MAPE = 28.20%. For comparison, the 14-hospital calibration yielded Pearson’s r = 0.568, Spearman’s ρ = 0.527, RMSE = 0.027, and MAPE = 32.17%. The performance of the BCEC-2SFCA model under the 22-hospital calibration remains very close to that of the main calibration, and both substantially outperform the Gaussian 2SFCA benchmark (Pearson’s r = 0.116, Spearman’s ρ = 0.053, RMSE = 0.022, MAPE = 42.12%).
The slight variations in parameter estimates and validation metrics are expected, as the 14-hospital subset was selected to maximize the reliability of the taxi-based proxy. Nevertheless, the consistency of the results across the two calibration samples confirms that the core findings of this study are not sensitive to the specific composition of the calibration sample.
7. Discussion
The validation results presented in this study highlight a trade-off: the BCEC-2SFCA model excels at capturing the rank order of hospital attractiveness () and relative demand distribution (MAPE = 32.17%), but its RMSE (0.02700) is slightly higher than that of the Gaussian model (0.02211). This indicates that the model’s primary strength lies in replicating competitive dynamics rather than minimizing absolute prediction error.
The strong correlation between travel-time and travel-distance accessibility () further supports the robustness of the approach, suggesting that the model’s behavioral core is stable across different impedance specifications. This consistency enhances confidence in its applicability to diverse data contexts.
The findings of this study can be situated within the broader literature on Huff-based FCA methods. Compared to Jin and Lu [
30], who applied a multi-mode Huff-2SFCA to food accessibility, the present study extends the Huff-FCA framework to the healthcare domain while additionally calibrating the Huff parameters from observed travel behavior rather than relying on sensitivity tests. Relative to Subal et al. [
31], who incorporated a Gaussian distance-decay function into a Huff-3SFCA model, our approach eliminates the need for predefined decay parameters entirely by letting the probability surface emerge from empirical data. The generalized flow-based 2SFCA of Chen et al. [
34] shares with our work the use of taxi trajectory data; however, their method still relies on adaptive catchment delineation, whereas BCEC-2SFCA generates endogenous search ranges probabilistically without any spatial cutoff. These comparisons highlight that the key advancement of BCEC-2SFCA lies not in any single component but in the systematic integration of empirical calibration, real-world impedance, and endogenous catchment generation into a unified accessibility framework.
From a theoretical standpoint, this study contributes to the accessibility literature by demonstrating that the conventional dichotomy between “accessible” and “inaccessible”—embedded in all threshold-based 2SFCA variants—can be replaced with a continuous, competition-sensitive probability gradient. This shift from binary to probabilistic accessibility aligns the methodological framework more closely with the continuous nature of spatial behavior described in classical geography and with discrete choice theory in transportation research. Furthermore, the empirical calibration of Huff parameters from passive mobility data suggests a pathway toward behaviorally grounded accessibility models that reduce reliance on stated-preference surveys or ad hoc parameter assumptions, thereby enhancing the replicability and context-specificity of accessibility assessments.
Nevertheless, several limitations warrant acknowledgment. First, the use of taxi trajectory data does not capture all modes of healthcare travel; future work should integrate multi-source mobility data such as public transit smart card records or mobile phone signaling. Second, hospital attractiveness is represented solely by bed count, ignoring specialty mix, reputation, or insurance acceptance. Incorporating such attributes could refine the choice probabilities. Third, due to the unavailability of patient origin–destination flow data from hospital records—a common constraint in healthcare accessibility research—direct validation against actual patient flow patterns at the subdistrict level was not feasible. Nevertheless, the taxi trajectory data used for model calibration provide a large-scale proxy of revealed healthcare travel behavior, and the outpatient volume data offer an independent, facility-level benchmark for external validation. Fourth, the relatively small number of hospitals with available outpatient data (N = 16) precluded meaningful subsample analyses by region or hospital scale. Such analyses would require a substantially larger sample to retain statistical power within each subgroup. Future work could revisit these validation dimensions should more granular patient flow data or expanded outpatient datasets become available.
From a policy perspective, the findings suggest that resource allocation strategies should account for the competitive dynamics revealed by the BCEC-2SFCA model. In central urban areas, where accessibility is already high, further expansion of hospital capacity may yield diminishing returns due to competition; instead, efforts could focus on service integration and differentiation. In peripheral areas, improving transport connections to existing high-quality hospitals may be more effective than building new facilities, as residents already demonstrate a willingness to travel longer distances when necessary.
8. Conclusions
This study proposed a data-driven BCEC-2SFCA framework for assessing spatial accessibility to tertiary hospitals, using taxi trajectory data to empirically calibrate the Huff model parameters and to construct realistic travel impedance matrices. The framework was applied to Ningbo, China, with data from 22 tertiary hospitals, 81 townships/subdistricts, and over 1.58 million taxi trips. The main findings are as follows:
Empirically grounded parameters: The calibrated service capacity elasticity () and time decay coefficient () reflect residents’ strong preference for larger hospitals and high sensitivity to travel time, providing a behaviorally realistic foundation for accessibility modeling.
Competition effects and spatial equity: The BCEC-2SFCA model captures competition among hospitals, moderating accessibility estimates in hospital-dense central areas while avoiding the underestimation of accessibility in peripheral regions that plagues fixed-threshold models. This yields a continuously varying and equitable spatial distribution.
Robustness to impedance metrics: Accessibility results based on travel time and travel distance are highly correlated (), indicating that the model’s core mechanism is stable across different measures of spatial separation.
Predictive performance: Validation against annual outpatient volumes shows that the BCEC-2SFCA model excels at capturing the relative order of hospital attractiveness (Spearman ρ = 0.527) and relative demand distribution (MAPE = 32.17%), though its RMSE (0.02700) is slightly higher than that of the Gaussian model (0.02211). This trade-off underscores the model’s strength in replicating competitive dynamics rather than minimizing absolute prediction error.
Policy implications: The model highlights the need for differentiated strategies: in central areas, optimizing existing resources rather than expanding capacity; in peripheral areas, improving transport connectivity to overcome spatial barriers. The use of passive mobility data offers a scalable approach to dynamically monitor healthcare-seeking behavior.
In conclusion, the BCEC-2SFCA framework integrates empirically calibrated behavioral parameters and real-world travel data to produce more realistic and policy-relevant accessibility assessments. It should be reiterated that this framework is tailored specifically to high-order medical services such as tertiary hospitals, where facility competition and patient choice are central to travel behavior. For primary care or other services characterized by localized catchments, traditional threshold-based models may be more appropriate.
Future work could proceed in several directions. First, integrating additional mobility data sources—such as public transit smart card records, mobile phone signaling, and ambulance dispatch logs—would enable the calibration of mode-specific impedance matrices and better capture the healthcare travel behavior of underrepresented groups, including the elderly and low-income populations. Second, hospital attractiveness could be represented through a multi-dimensional index incorporating specialty mix, physician expertise, patient satisfaction, and insurance acceptance, rather than bed count alone. Third, the temporal dynamics of healthcare accessibility could be explored by extending the BCEC-2SFCA framework to account for diurnal and seasonal variations in traffic conditions and hospital service availability. Fourth, the generalizability of the framework should be tested in other urban contexts with different spatial configurations of healthcare resources, as well as for other types of high-order public services such as regional cultural centers and specialized elderly care facilities. Finally, coupling the BCEC-2SFCA model with spatial optimization algorithms could provide decision-support tools for equitable and efficient healthcare resource allocation.