SearcharxivSearch

arXiv subjects

Soumen Dey

Publications and source records attributed to Soumen Dey.

5 recordsLinked to original sources

Size biased Multinomial Modelling of detection data in Software testing

Estimation of software reliability often poses a considerable challenge, particularly for critical softwares. Several methods of estimation of reliability of software are already available in the literature. But, so far almost nobody used the concept of size of a bug for estimating software reliability. In this article we make used of the bug size or the eventual bug size which helps us to determine reliability of software more precisely. The size-biased model developed here can also be used for similar fields like hydrocarbon exploration. The model has been validated through simulation and subsequently used for a critical space application software testing data. The estimated results match the actual observations to a large extent.

cs.SE

Modelling spatially autocorrelated detection probabilities in spatial capture-recapture using random effects

Spatial capture-recapture (SCR) models are now widely used for estimating density from repeated individual spatial encounters. SCR accounts for the inherent spatial autocorrelation in individual detections by modelling detection probabilities as a function of distance between the detectors and individual activity centres. However, additional spatial heterogeneity in detection probability may still creep in due to environmental or sampling characteristics. if unaccounted for, such variation can lead to pronounced bias in population size estimates. Using simulations, we describe and test three Bayesian SCR models that use generalized linear mixed models (GLMM) to account for latent heterogeneity in baseline detection probability across detectors using: independent random effects (RE), spatially autocorrelated random effects (SARE), and a two-group finite mixture model (FM). Overall, SARE provided the least biased population size estimates (median RB: -9 -- 6%). When spatial autocorrelation was high, SARE also performed best at predicting the spatial pattern of heterogeneity in detection probability. At intermediate levels of autocorrelation, spatially-explicit estimates of detection probability obtained with FM where more accurate than those generated by SARE and RE. In cases where the number of detections per detector is realistically low (at most 1), all GLMMs considered here may require dimension reduction of the random effects by pooling baseline detection probability parameters across neighboring detectors ("aggregation") to avoid over-parameterization. The added complexity and computational overhead associated with SCR-GLMMs may only be justified in extreme cases of spatial heterogeneity. However, even in less extreme cases, detecting and estimating spatially heterogeneous detection probability may assist in planning or adjusting monitoring schemes.

stat.AP

Estimating Software Reliability Using Size-biased Modelling

Software reliability estimation is one of the most active areas of research in software testing. Since time between failures (TBF) has often been challenging to record, software testing data are commonly recorded as test-case-wise in a discrete set up. We have developed a Bayesian generalised linear mixed model (GLMM) based on software testing detection data and a size-biased strategy which not only estimates the software reliability, but also estimates the total number of bugs present in the software. Our approach provides a flexible, unified modelling framework and can be adopted to various real-life situations. We have assessed the performance of our model via simulation study and found that each of the key parameters could be estimated with a satisfactory level of accuracy. We have also applied our model to two empirical software testing data sets. While there can be other fields of study for application of our model (e.g., hydrocarbon exploration), we anticipate that our novel modelling approach to estimate software reliability could be very useful for the users and can potentially be a key tool in the field of software reliability estimation.

stat.AP

Bayesian Model Selection for a Class of Spatially-Explicit Capture Recapture Models

A vast amount of ecological knowledge generated recently has hinged upon the ability of model selection methods to discriminate among various ecological hypotheses. The last decade has seen the rise of Bayesian hierarchical models in ecology. Consequently, popular tools, such as the AIC, become largely inapplicable and other tools are not universally applicable. We focus on a class of competing Bayesian spatially explicit capture recapture (SECR) models and first apply some of the recommended Bayesian model selection tools: (1) Bayes Factor - using (a) Gelfand-Dey (b) harmonic mean methods, (2) DIC, (3) WAIC and (4) the posterior predictive loss function. In all, we evaluate 25 variants of model selection tools in our study. We evaluate these model selection tools from the standpoint of model selection and parameter estimation by contrasting the choice recommended by a tool with a `true' model. In all, we generate 120 simulated data sets using the true model and assess the frequency with which the true model is selected and how well the tool estimates N (population size). We find that when information content is low, no particular tool can be recommended to help realise, simultaneously, both the goals of model selection and parameter estimation. In such scenarios, we recommend that practitioners utilise our application of Bayes Factor for parameter estimation and recommend the posterior predictive loss approach for model selection when information content is low. When both the objectives are taken together, we recommend the use of our applications of Bayes Factor for Bayesian SECR models. Our study reveals that although new model selection tools are emerging (eg: WAIC) in the applied statistics literature, an uncritical absorption of these new tools (i.e. without assessing their efficacies for the problem at hand) into ecological practice may mislead inferences.

stat.AP

A spatially explicit capture recapture model for partially identified individuals when trap detection rate is less than one

Spatially explicit capture recapture (SECR) models have gained enormous popularity to solve abundance estimation problems in ecology. In this study, we develop a novel Bayesian SECR model that disentangles the process of animal movement through a detector from the process of recording data by a detector in the face of imperfect detection. We integrate this complexity into an advanced version of a recent SECR model involving partially identified individuals (Royle, 2015). We assess the performance of our model over a range of realistic simulation scenarios and demonstrate that estimates of population size $N$ improve when we utilize the proposed model relative to the model that does not explicitly estimate trap detection probability (Royle, 2015). We confront and investigate the proposed model with a spatial capture-recapture data set from a camera trapping survey on tigers (\textit{Panthera tigris}) in Nagarahole, southern India. Trap detection probability is estimated at 0.489 and therefore justifies the necessity to utilize our model in field situations. We discuss possible extensions, future work and relevance of our model to other statistical applications beyond ecology.

stat.AP