Skip to main content

Kishor S. Trivedi

Hudson Distinguished Professor Emeritus of Electrical and Computer Engineering
Pierre R. Lamond Department of Electrical and Computer Engineering
Box 90291, Durham, NC 27708-0291
534 Research Drive, 401 Wilkinson, Durham, NC 27708-0291

Scholarly Works - Conferences


Reliability and Availability Assessment

Conference IEEE Transactions on Reliability · March 1, 2024 Given heavy dependence on man-made systems in our daily lives, reliability and availability of these systems clearly gain great importance. Together with methods of enhancing reliability and availability of systems, methods of quantitative assessment of th ... Full text Cite

S-ADA: Software as an Autonomous, Dependable and Affordable System

Conference Proceedings 51st Annual IEEE IFIP International Conference on Dependable Systems and Networks Supplemental Volume Dsn S 2021 · June 1, 2021 ADA is a popular programming language that was named after Lady Ada Lovelace (1815-1852) and recommended by Department of Defense, USA, for development of large scale safety-critical software systems. In this Fast Abstract, ADA is reinterpreted as Autonomo ... Full text Cite

Transient Security and Dependability Analysis of MEC Micro Datacenter under Attack

Conference Proceedings Annual Reliability and Maintainability Symposium · January 1, 2021 A Multi-access Edge Computing (MEC) micro data center (MEDC) consists of multiple MEC hosts close to endpoint devices. MEC service is delivered by instantiating a virtualization system (e.g., Virtual Machines or Containers) on a MEC host. MEDC faces more n ... Full text Cite

A Multisite Characterization Study on Failure Causes in System and Applications Software

Conference Brazilian Symposium on Computing System Engineering Sbesc · January 1, 2021 A fundamental aspect of software reliability engineering is to understand how software failures manifest, identifying and comprehending their causes and effects. In this paper, we perform ex-post analyses of field software failure data, looking to characte ... Full text Cite

A Statistical Approach to Predict Operating System Failures Based on Multiple Failures Association

Conference Brazilian Symposium on Computing System Engineering Sbesc · November 24, 2020 Empirical studies have shown robust evidence of OS failure patterns characterized by multiple combinations of failure events composed of the same or different failure types. In this paper, we present a statistical approach to predict OS failures based on m ... Full text Cite

Chapter 1: Software Aging and Rejuvenation: A Genesis - Extended Abstract

Conference Proceedings 2020 IEEE 31st International Symposium on Software Reliability Engineering Workshops Issrew 2020 · October 1, 2020 This talk summarizes the genesis of software aging and rejuvenation as presented in the handbook of software aging and rejuvenation. It also lays out possible future directions to reflect the content of the concluding chapter of the handbook. ... Full text Cite

An Empirical Exploratory Analysis of Failure Sequences in a Commodity Operating System

Conference Brazilian Symposium on Computing System Engineering Sbesc · November 1, 2019 A fundamental need for software reliability engineering is to comprehend how software systems fail, which means understanding the dynamics that govern different types of failure manifestation. In this paper, we present an exploratory study on multiple-even ... Full text Cite

Supervised Representation Learning Approach for Cross-Project Aging-Related Bug Prediction

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · October 1, 2019 Software aging, which is caused by Aging-Related Bugs (ARBs), tends to occur in long-running systems and may lead to performance degradation and increasing failure rate during software execution. ARB prediction can help developers discover and remove ARBs, ... Full text Cite

Rejuvenation and the age of information

Conference Proceedings 2019 IEEE 30th International Symposium on Software Reliability Engineering Workshops Issrew 2019 · October 1, 2019 Two decades after the seminal paper on software aging and rejuvenation appeared in 1995, a new concept and metric referred to as the age of information (AoI) has been gaining attention from practitioners and the research community. In this vision paper, ou ... Full text Cite

2nd Workshop on Education and Practice ofPerformance Engineering: WEPPE'19 Chairs' Welcome

Conference Icpe 2019 Companion of the 2019 ACM Spec International Conference on Performance Engineering · April 4, 2019 Full text Cite

Software Aging and Software Rejuvenation

Conference Proceedings of the 2019 ACM/SPEC International Conference on Performance Engineering · April 4, 2019 Full text Cite

Performance Engineering Education

Conference Companion of the 2019 ACM/SPEC International Conference on Performance Engineering · March 27, 2019 Full text Cite

Performance modeling of hyperledger fabric (permissioned blockchain network)

Conference NCA 2018 2018 IEEE 17th International Symposium on Network Computing and Applications · November 26, 2018 Hyperledger Fabric (HLF) is an open-source implementation of a distributed ledger platform for running smart contracts in a modular architecture. In this paper, we present a performance model of Hyperledger Fabric v1.0+ using Stochastic Reward Nets (SRN). ... Full text Cite

Survivability model for security and dependability analysis of a vulnerable critical system

Conference Proceedings International Conference on Computer Communications and Networks ICCCN · October 9, 2018 This paper aims to analyze transient security and dependability of a vulnerable critical system, under vulnerability-related attack and two reactive defense strategies, from a severe vulnerability announcement until the vulnerability is fully removed from ... Full text Cite

Keynote Paper: Parametric Uncertainty Propagation through Dependability Models

Conference Proceedings 8th Latin American Symposium on Dependable Computing Ladc 2018 · July 2, 2018 The uncertainty propagation is to investigate the effect of errors in model input parameters on the system output measure in probability models. In this paper, we present a moment-based approach of the uncertainty propagation of model input parameters. The ... Full text Cite

Monitoring and mitigating software aging on IBM cloud controller system

Conference Proceedings 2017 IEEE 28th International Symposium on Software Reliability Engineering Workshops Issrew 2017 · November 14, 2017 As enterprises continue to move their workloads from traditional server-room environments to private cloud-based systems, there is an increasing desire and ability for companies like IBM to centrally monitor the systems on behalf of their customers to proa ... Full text Cite

Understanding the Impacts of Influencing Factors on Time to a DataRace Software Failure

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · November 14, 2017 Datarace is a common problem on shared-memory parallel computers, including multicores. Due to its dependence on the thread scheduling scheme of its execution environment, the time to a datarace failure is usually very long. How to accelerate the occurrenc ... Full text Cite

Experience Report: Fault Triggers in Linux Operating System: From Evolution Perspective

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · November 14, 2017 Linux operating system is a complex system that is prone to suffer failures during usage, and increases difficulties of fixing bugs. Different testing strategies and fault mitigation methods can be developed and applied based on different types of bugs, wh ... Full text Cite

Performance modeling of PBFT consensus process for permissioned blockchain network (hyperledger fabric)

Conference Proceedings of the IEEE Symposium on Reliable Distributed Systems · October 13, 2017 While Blockchain network brings tremendous benefits, there are concerns whether their performance would match up with the mainstream IT systems. This paper aims to investigate whether the consensus process using Practical Byzantine Fault Tolerance (PBFT) c ... Full text Cite

Epistemic uncertainty propagation in a Weibull environment for a two-core system-on-chip

Conference 2017 2nd International Conference on System Reliability and Safety Icsrs 2017 · July 2, 2017 Epistemic uncertainty analysis accounts for inaccurate input parameters and evaluates how such uncertainty propagates to output measures. In this work we will focus on Weibull distributions, in particular the one related to the reliability of multi-core sy ... Full text Cite

An empirical study of software reliability in SDN controllers

Conference 2017 13th International Conference on Network and Service Management Cnsm 2017 · July 1, 2017 Software Defined Networking (SDN) exposes critical networking decisions, such as traffic routing or enforcement of the critical security policies, to a software entity known as the SDN controller. Controller software, as written by humans, is intrinsically ... Full text Cite

An empirical investigation of fault triggers in android operating system

Conference Proceedings of IEEE Pacific Rim International Symposium on Dependable Computing Prdc · May 5, 2017 The growing popularity and complexity of Android operating system makes it prone to suffer failures during usage, which increases difficulties of fixing bugs. Different strategies and mitigation methods can be developed and applied based on different types ... Full text Cite

A novel approach for software vulnerability classification

Conference Proceedings Annual Reliability and Maintainability Symposium · March 29, 2017 Software vulnerability analysis plays a critical role in the prevention and mitigation of software security attacks, and vulnerability classification constitutes a key part of this analysis. This paper proposes a new approach for software vulnerability cla ... Full text Cite

Transient performance & availability modeling in high volume outpatient clinics

Conference Proceedings Annual Reliability and Maintainability Symposium · March 29, 2017 High volume outpatient clinics such as eye care centers cannot afford excessive delays, especially when due to limited resources, time, or overhead. Modeling tools from reliability & maintainability practice may provide the means to better assess where imp ... Full text Cite

Automated life cycle processing for complex medical imaging devices

Conference Proceedings Annual Reliability and Maintainability Symposium · March 29, 2017 Medical imaging systems from major modalities such as Magnetic Resonance Imaging or X-Ray Computed Tomography are complex devices subject to various types of maintenance. Medical device companies that develop these systems often monitor and maintain system ... Full text Cite

Availability modeling and analysis of a virtualized system using stochastic reward nets

Conference Proceedings 2016 16th IEEE International Conference on Computer and Information Technology CIT 2016 2016 6th International Symposium on Cloud and Service Computing IEEE Sc2 2016 and 2016 International Symposium on Security and Privacy in Social Networks and Big Data Socialsec 2016 · March 10, 2017 Availability is one of the key requirements for modern networked system. Availability of a virtualized system can be modelled and analyzed using stochastic models. In our previous work, availability of a virtualized system was modeled using a hierarchical ... Full text Cite

Application-level scheme to enhance VANET event-driven multi-hop safety-related services

Conference 2017 International Conference on Computing Networking and Communications Icnc 2017 · March 10, 2017 In this paper, we focus on the design and analysis of channel access in vehicular ad hoc networks (VANETs) for event-driven multi-hop safety services. First, a novel channel access scheme that incorporates an application-level distance (timer)-based rebroa ... Full text Cite

An approach for resiliency quantification of large scale systems

Conference Performance Evaluation Review · March 1, 2017 We quantify the resiliency of large scale systems upon changes encountered beyond the normal system behavior. Formal definitions for resiliency and change are provided together with general steps for resiliency quantification and a set of resiliency metric ... Full text Cite

Efficient computation of the mean time to security failure in cyber physical systems

Conference Valuetools 2016 10th Eai International Conference on Performance Evaluation Methodologies and Tools · January 1, 2017 In this paper, we present a computationally efficient technique for calculating the mean time to security failure (MTTSF) of a mobile cyber physical system (CPS). The CPS analyzed here has been comprehensively studied by other authors using stochastic rewa ... Full text Cite

Parametric sensitivity and uncertainty propagation in dependability models

Conference Valuetools 2016 10th Eai International Conference on Performance Evaluation Methodologies and Tools · January 1, 2017 Input parameters of dependability models are often not known accurately. Two principal methods of dealing with such parametric uncertainty are: sensitivity analysis and uncertainty propagation. This paper is an initial attempt to link the two approaches. T ... Full text Cite

Resiliency quantification for large scale systems: An IaaS cloud use case

Conference Valuetools 2016 10th Eai International Conference on Performance Evaluation Methodologies and Tools · January 1, 2017 We quantify the resiliency of large scale systems upon changes encountered beyond the normal system behavior. General steps for resiliency quantification are shown and resiliency metrics are defined to quantify the effects of changes. The proposed approach ... Full text Cite

Model-Based Survivability Analysis of a Virtualized System

Conference Proceedings Conference on Local Computer Networks LCN · December 22, 2016 Transient survivability analysis of a virtualized system (VS) is critical to the wide deployment of cloud services. The existing research of VS availability and/or reliability focused on the steady-state analysis. This paper presents a model and the closed ... Full text Cite

The Relationship between Software Bug Type and Number of Factors Involved in Failures

Conference Proceedings 2016 IEEE 27th International Symposium on Software Reliability Engineering Workshops Issrew 2016 · December 16, 2016 Previous studies have defined different types of software bugs based on their complexity and reproducibility. Simple bugs, which involve only direct factors and are often easy to reproduce, have been called 'Bohrbugs', while complex bugs, with at least one ... Full text Cite

Software Aging Detection Based on Differential Analysis: An Experimental Study

Conference Proceedings 2016 IEEE 27th International Symposium on Software Reliability Engineering Workshops Issrew 2016 · December 16, 2016 In this study we evaluate the applicability of the differential software analysis approach to detect memory leaks under a real workload. For this purpose, we used three different versions of a widely used software application, where one version was used as ... Full text Cite

DSN 2016 Tutorial: Reliability and Availability Modeling in Practice

Conference Proceedings 46th Annual IEEE IFIP International Conference on Dependable Systems and Networks Dsn W 2016 · September 22, 2016 Full text Cite

Analysis methods for performance & availability in critical care medicine

Conference Proceedings Annual Reliability and Maintainability Symposium · April 5, 2016 Operations of critical care departments in health systems are increasingly reliant on the availability of interoperable medical devices. Many large health care systems have fully transitioned in recent years to uniform electronic health record platforms, i ... Full text Cite

Reliability models of chronic kidney disease

Conference Proceedings Annual Reliability and Maintainability Symposium · April 5, 2016 With the rise in quantifiable approaches to health care, lessons from reliability modeling provide new avenues for improving patient outcomes. Describing the development of conditions leading to organ system failure provides visceral motivation for quantif ... Full text Cite

A Scalable Optimization Framework for Storage Backup Operations Using Markov Decision Processes

Conference Proceedings 2015 IEEE 21st Pacific Rim International Symposium on Dependable Computing Prdc 2015 · January 4, 2016 Explosive growth of data generation and increasing reliance of business analysis on massive data make data loss more damaging than ever before. Thus it has also become a critical issue for businesses to protect important data effectively. In a system with ... Full text Cite

Survivability analysis of a computer system under an advanced persistent threat attack

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2016 Computer systems are potentially targeted by cybercriminals by means of specially crafted malicious software called Advanced Persistent Threats (APTs). As a consequence, any security attribute of the computer system may be compromised: disruption of servic ... Full text Cite

Survivability quantification for networks

Conference Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) · January 1, 2016 Survivability is a critical attribute of modern computer and communication systems. The assessment of survivability is mostly performed in a qualitative manner and thus cannot meet the need for more precise and solid evaluation of service loss or degradati ... Cite

Software Reliability Analysis of NASA Space Flight Software: A Practical Experience.

Conference IEEE International Conference on Software Quality, Reliability and Security : proceedings. IEEE International Conference on Software Quality, Reliability and Security · January 2016 In this paper, we present the software reliability analysis of the flight software of a recently launched space mission. For our analysis, we use the defect reports collected during the flight software development. We find that this software was developed ... Full text Cite

Modeling of VANET for BSM safety messaging at intersections with non-homogeneous node distribution

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2016 This paper presents a new analytic model for the performance and reliability of safety-related message broadcast in vehicular ad hoc networks (VANETs) at intersections with non-homogeneous Poisson process (NHPP) for more general road traffic and node distr ... Full text Cite

Workshop on Model Based Design for Cyber-Physical Systems (MB4CP)

Conference Proceedings of the International Conference on Dependable Systems and Networks · September 14, 2015 This paper provides a summary of the First International Workshop on Model Based Design for Cyber- Physical Systems (MB4CP 2015) in conjunction with DSN 2015 conference in Rio de Janeiro, Brazil. ... Full text Cite

Emulating environment-dependent software faults

Conference Proceedings 1st International Workshop on Complex Faults and Failures in Large Software Systems Coufless 2015 · August 5, 2015 The interaction of software with its execution environment is an underestimated cause of complex faults activation and systems failure. This paper discusses a possible framework to emulate anomalous environment conditions in order to assess the impact of t ... Full text Cite

Survivability as a generalization of recovery

Conference 2015 11th International Conference on the Design of Reliable Communication Networks Drcn 2015 · July 2, 2015 Social infrastructure systems such as communication, transportation, power and water supply systems are now facing various types of threats including component failures, security attacks and natural disasters, etc. Whenever such undesirable events occur, i ... Full text Cite

An SRN-based resiliency quantification approach

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2015 Resiliency is often considered as a synonym for faulttolerance and reliability/availability. We start from a different definition of resiliency as the ability to deliver services when encountering unexpected changes. Semantics of change is of extreme impor ... Full text Cite

Software maintenance optimization based on stackelberg game methods

Conference Proceedings IEEE 25th International Symposium on Software Reliability Engineering Workshops Issrew 2014 · December 12, 2014 Application servers (AS) of virtualized platform may suffer from software aging problem. In this paper, we first formulate the system model including three virtual machines. Two of them act as the main servers, and the third machine acts as the backup node ... Full text Cite

Reproducibility of environment-dependent software failures: An experience report

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · December 11, 2014 We investigate the dependence of software failure reproducibility on the environment in which the software is executed. The existence of such dependence is ascertained in literature, but so far it is not fully characterized. In this paper we pinpoint some ... Full text Cite

Computing defects per million in cloud caused by virtual machine failures with replication

Conference Proceedings of IEEE Pacific Rim International Symposium on Dependable Computing Prdc · December 3, 2014 Virtual machines (VM) are used in cloud computing systems to handle user requests for service. A typical user request goes through several cloud service provider specific processing steps from the instant it is submitted until the service is completed. In ... Full text Cite

Message from the chairs

Conference 3rd International Workshop on Software Engineering Challenges for the Smart Grid Se4sg 2014 Proceedings · June 1, 2014 Cite

Defects per Million (DPM): A user-oriented perspective of telecommunication systems

Conference 2014 IEEE Globecom Workshops GC Wkshps 2014 · March 18, 2014 Defects Per Million (DPM), defined as the number of calls dropped out of a million calls due to failures, is used by the telecommunication systems community as a user-perceived dependability metric. As new standards evolve, with built-in mechanisms to hand ... Full text Cite

Message from the chairs

Conference 3rd International Workshop on Software Engineering Challenges for the Smart Grid, SE4SG 2014 - Proceedings · January 1, 2014 Cite

A Markov Decision Process Approach for Optimal Data Backup Scheduling

Conference 2014 44TH ANNUAL IEEE/IFIP INTERNATIONAL CONFERENCE ON DEPENDABLE SYSTEMS AND NETWORKS (DSN) · January 1, 2014 Full text Link to item Cite

The nature of the times to flight software failure during space missions

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · December 1, 2012 The growing complexity of mission-critical space mission software makes it prone to suffer failures during operations. The success of space missions depends on the ability of the systems to deal with software failures, or to avoid them in the first place. ... Full text Cite

SURVIVABILITY MODELING WITH STOCHASTIC REWARD NETS

Conference PROCEEDINGS OF THE 2009 WINTER SIMULATION CONFERENCE (WSC 2009 ), VOL 1-4 · January 1, 2009 Link to item Cite

Reliable system design: Models, metrics and design techniques

Conference 2008 IEEE/ACM International Conference on Computer-Aided Design · November 2008 Full text Cite

Achieving and assuring high availability

Conference 2008 IEEE INTERNATIONAL SYMPOSIUM ON PARALLEL & DISTRIBUTED PROCESSING, VOLS 1-8 · January 1, 2008 Link to item Cite

Availability Monitor for a Software Based System

Conference 10th IEEE High Assurance Systems Engineering Symposium (HASE'07) · November 2007 Full text Cite

Simulation versus analytic-numeric methods: Illustrative examples

Conference Valuetools 2007 2nd International Icst Conference on Performance Evaluation Methodologies and Tools · January 1, 2007 Performance along with dependability analysis is a tremendous challenge in the design or improvement of modern complex systems. Two different classes of solution methods are generally used: analytic-numeric methods and simulation methods. As most of the li ... Full text Cite

Survivability Quantification - Keynote.

Conference BROADNETS · 2007 Cite

Software rejuvenation - modeling and analysis

Conference IFIP Advances in Information and Communication Technology · January 1, 2004 Several recent studies have established that most system outages are due to software faults. Given the ever increasing complexity of software and the welldeveloped techniques and analysis for hardware reliability, this trend is not likely to change in the ... Full text Cite

SITAR: A scalable intrusion-tolerant architecture for distributed services

Conference Foundations of Intrusion Tolerant Systems Oasis 2003 · January 1, 2003 This paper presents a intrusion tolerant architecture for distributed services, especially COTS servers. An intrusion tolerant system assumes that attacks will happen, and some will be successful. However, a wide range of mission critical applications need ... Full text Cite

Maximizing interval reliability in operational software system with rejuvenation

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · January 1, 2003 Software aging often affects the performance of a software system and eventually causes it to fail. A novel approach to handle transient software failures is called software rejuvenation which can be regarded as a preventive and proactive solution that is ... Full text Cite

Specification-level integration of simulation and dependability analysis

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2003 Software architectural choices have a profound influence on the quality attributes supported by a system. Architecture analysis can be used to evaluate the influence of design decisions on important quality attributes such as maintainability, performance a ... Full text Cite

OPTIMAL WEBSERVER SESSION TIMEOUT SETTINGS FOR WEB USERS

Conference 28th International Computer Measurement Group Conference Cmg 2002 · January 1, 2002 From an end user’s point of view, too short a Webserver timeout implies too many forced logouts, and too long a timeout duration poses a higher security risk to users’ sensitive data. We propose cost functions to select the timeout value, which are based o ... Cite

Reliability prediction and sensitivity analysis based on software architecture

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · January 1, 2002 Prevalent approaches to characterize the behavior of monolithic applications are inappropriate to model modern software systems which are heterogeneous, and are built using a combination of components picked off the shelf, those developed in-house and thos ... Full text Cite

Software reliability and rejuvenation: Modeling and analysis

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2002 Several recent studies have established that most system outages are due to software faults. Given the ever increasing complexity of software and the well-developed techniques and analysis for hardware reliability, this trend is not likely to change in the ... Full text Cite

A framework for performability modeling of messaging services in distributed systems

Conference Proceedings of the IEEE International Conference on Engineering of Complex Computer Systems ICECCS · January 1, 2002 Messaging services are a useful component in distributed systems that require scalable dissemination of messages (events) from suppliers to consumers. These services decouple suppliers and consumers, and take care of client registration and message propaga ... Full text Cite

An approach for estimation of software aging in a Web server

Conference Isese 2002 Proceedings 2002 International Symposium on Empirical Software Engineering · January 1, 2002 A number of recent studies have reported the phenomenon of "software aging", characterized by progressive performance degradation or a sudden hang/crash of a software system due to exhaustion of operating system resources, fragmentation and accumulation of ... Full text Cite

Modeling and analysis of software rejuvenation in cable modem termination systems

Conference Proceedings International Symposium on Software Reliability Engineering ISSRE · January 1, 2002 In order to reduce system outages and the associated downtime cost caused by the "software aging" phenomenon, we propose to use software rejuvenation as a proactive system maintenance technique deployed in a CMTS (Cable Modem Termination System) cluster sy ... Full text Cite

All-terminal reliability analysis of the SRP-ring: The effect of enhanced intelligent protection switching

Conference Proceedings International Conference on Computer Communications and Networks ICCCN · January 1, 2002 Spatial reuse protocol (SRP) is a media access control (MAC)-layer protocol that operates over a double counter-rotating ring network topology. SRP is designed to enhance the SONET network so that it can handle data traffic more efficiently. We study the a ... Full text Cite

A new handoff scheme for decreasing both dropped calls and blocked calls in CDMA system

Conference Eurocon 2001 International Conference on Trends in Communications Proceedings · January 1, 2001 Soft handoff in the CDMA cellular system is analyzed. To improve performance degradation due to channel resource shortage during soft handoff, we propose a new scheme which converts channels occupied by some pseudo-handoff calls to new handoff calls. Stoch ... Full text Cite

Characterizing intrusion tolerant systems using a state transition model

Conference Proceedings Darpa Information Survivability Conference and Exposition II Discex 2001 · January 1, 2001 Intrusion detection and response research has so far mostly concentrated on known and well-defined attacks. We believe that this narrow focus of attacks accounts for both the successes and limitation of commercial intrusion detection systems (IDS). Intrusi ... Full text Cite

Analysis of periodic preventive maintenance with general system failure distribution

Conference Proceedings of IEEE Pacific Rim International Symposium on Dependable Computing Prdc · January 1, 2001 Preventive maintenance is applied to improve the system availability or decrease the operational cost. In this paper the preventive maintenance with generally distributed parameters are discussed, and the steady-state solution is obtained by solving the un ... Full text Cite

Reliable messaging using the CORBA Notification Service

Conference Proceedings 3rd International Symposium on Distributed Objects and Applications Doa 2001 · January 1, 2001 With the growing popularity of the CORBA architecture as a distributed computing infrastructure standard, the need for a reliable CORBA messaging solution is being increasingly felt. The Event Service, which is the first such solution, provides inadequate ... Full text Cite

Stochastic Petri nets and their applications

Conference PERFORMANCE AND QOS OF NEXT GENERATION NETWORKING · January 1, 2001 Link to item Cite

Performability analysis of TDMA cellular systems based on composite and hierarchical Markov chain models

Conference PERFORMANCE AND QOS OF NEXT GENERATION NETWORKING · January 1, 2001 Link to item Cite

Effects of failure correlation on software in operation

Conference Proceedings of IEEE Pacific Rim International Symposium on Dependable Computing Prdc · January 1, 2000 Since the early 1970's a number of models have been proposed for estimating software reliability. However, the realism of many of the underlying assumptions and the applicability of these models continue to be questioned. Our research work was motivated by ... Full text Cite

Statistical non-parametric algorithms to estimate the optimal software rejuvenation schedule

Conference Proceedings of IEEE Pacific Rim International Symposium on Dependable Computing Prdc · January 1, 2000 In this paper, we extend the classical result by Huang, Kintala, Kolettis and Fulton (1995), and in addition propose a modified stochastic model to determine the software rejuvenation schedule. More precisely, the software rejuvenation models are formulate ... Full text Cite

Building a reliable message delivery system using the CORBA Event Service

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 In this paper we study the suitability of the CORBA Event Service as a reliable message delivery mechanism. We first show that products built to the CORBA Event Service specification will not guarantee against loss of messages or guarantee order. This is n ... Full text Cite

SREPT: Software reliability estimation and prediction tool

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 Although several tools have been developed for the estima-tion of software reliability, they are highly specialized in the approaches they implement and the particular phase of the software lifecycle in which they are applicable. Also the conventional tech ... Full text Cite

Reliability and performability modeling using SHARPE 2000

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 The SHARPE package, Symbolic Hierarchical Automated Reliability and Performance Evaluator, is now 13 years old. A well known package in the field of reliability and performability, SHARPE is used in universities as well as in companies. Many important chan ... Full text Cite

SPNP: Stochastic petri nets. Version 6. 0

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 Full text Cite

Implementation of importance splitting techniques in stochastic petri net package

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 Stochastic Petri Net Package (SPNP) is a software package whose goal is to compute performance, availability or performability measures from Stochastic Petri Nets (SPN) and Fluid Stochastic Petri nets (FSPN). This software can use either analytic numeric m ... Full text Cite

Stochastic modeling formalisms for dependability, performance and performability

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 2000 Full text Cite

The optimal preventive maintenance policy for a software system with multi server station

Conference 6TH ISSAT INTERNATIONAL CONFERENCE ON RELIABILITY AND QUALITY IN DESIGN, PROCEEDINGS · 2000 Link to item Cite

Analysis of software cost models with rejuvenation

Conference Proceedings of IEEE International Symposium on High Assurance Systems Engineering · January 1, 2000 Software rejuvenation is a preventive maintenance technique that has been extensively studied in the recent literature. In this paper we extend the classical result by Huang et al. (1995), and in addition propose a modified stochastic model to generate the ... Full text Cite

Dependability modeling and evaluation of phased mission systems: A DSPN approach

Conference Dependable Computing for Critical Applications 7 · January 1, 1999 We focus on analytical modeling for the dependability evaluation of phased-mission systems. Because of their dynamic behavior, systems showing a phased behavior offer challenges in modeling. We propose the modeling and evaluation of phased-mission system d ... Full text Cite

A reliable CORBA-based network management system

Conference IEEE International Conference on Communications · January 1, 1999 Network management provides the central nervous system for the networks of telecommunications providers. A telco's network management system (NMS) needs to support uninterrupted management functionality of complex networks. The reliability of such systems ... Full text Cite

Availability and performance evaluation for automatic protection switching in TDMA wireless system

Conference Proceedings 1999 Pacific Rim International Symposium on Dependable Computing Prdc 1999 · January 1, 1999 In this paper, we compare the availability and performance of a wireless TDMA system with and without automatic protection switching. Stochastic reward net models are constructed and solved by SPNP (Stochastic Petri Net Package). Hierarchical decomposition ... Full text Cite

Dependability modelling and sensitivity analysis of scheduled maintenance systems

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1999 In this paper we present a new modelling approach for dependability evaluation and sensitivity analysis of Scheduled Maintenance Systems, based on a Deterministic and Stochastic Petri Net approach. The DSPN approach offers significant advantages in terms o ... Full text Cite

Performability analysis of fault tolerant RF link design in wireless communications networks

Conference ESM'99 - MODELLING AND SIMULATION: A TOOL FOR THE NEXT MILLENNIUM, VOL 1 · January 1, 1999 Link to item Cite

Locating program features using execution slices

Conference Proceedings 1999 IEEE Symposium on Application Specific Systems and Software Engineering and Technology Asset 1999 · January 1, 1999 An important step towards effective software maintenance is to locate the code relevant to a particular feature. We report a study applying an execution slice-based technique to a reliability and performance evaluator to identify the code which is unique t ... Full text Cite

Log-logistic software reliability growth model

Conference Proceedings 3rd IEEE International High Assurance Systems Engineering Symposium Hase 1998 · January 1, 1998 The finite-failure non-homogeneous Poisson process (NHPP) models proposed in the literature exhibit either constant, monotonic increasing or monotonic decreasing failure occurrence rates per fault, and are inadequate to describe the failure processes under ... Full text Cite

Srept: Software reliability estimation and prediction tool

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1998 Several tools have been developed for the estimation of soft- ware reliability. However, they are highly specialized in the approaches they implement and the particular phase of the software life-cycle in which they are applicable. There is an increasing n ... Full text Cite

An improved multiple variable inversion algorithm for reliability calculation

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1998 An improved algorithm based on the one proposed by Veeraraghavan and Trivedi(VT) to calculate system reliability using sum of disjoint products (SDP) and multiple variable inversion (MVI) techniques is presented. We compare the improved algorithm with seve ... Full text Cite

Performability analysis of channel allocation with channel recovery strategy in cellular networks

Conference Icupc 1998 IEEE 1998 International Conference on Universal Personal Communications Conference Proceedings · January 1, 1998 We propose and compare three channel recovery schemes for fixed channel assignment. In Scheme I, a failed channel is switched by an idle channel whenever it is available. In Scheme II, the switching strategy is employed only after an attempt to restore the ... Full text Cite

Model validation using simulated data

Conference Proceedings 1998 IEEE Workshop on Application Specific Software Engineering and Technology Asset 1998 · January 1, 1998 Effective and accurate reliability modeling requires the collection of comprehensive, homogeneous, and consistent data sets. Failure data required for software reliability modeling is difficult to collect, and even the available data tends to be noisy, dis ... Full text Cite

Toward accessibility enhancement of dependability modeling techniques and tools

Conference Digest of Papers 27th Annual International Symposium on Fault Tolerant Computing Ftcs 1997 · January 1, 1997 Although various dependability evaluation techniques and tools have been developed in the last two decades, no adequate attention has been paid to allow system designers not well versed in analytic modeling to easily employ these techniques and tools. In t ... Full text Cite

Minimizing completion time of a program by checkpointing and rejuvenation

Conference Sigmetrics 1996 Proceedings of the 1996 ACM Sigmetrics International Conference on Measurement and Modeling of Computer Systems · May 15, 1996 Checkpointing with rollback-recovery is a well known technique to reduce the completion time of a program in the presence of failures. While checkpointing is corrective in nature, rejuvenation refers to preventive maintenance of software aimed to reduce un ... Full text Cite

Important milestones in software reliability modeling

Conference SEKE '96: THE 8TH INTERNATIONAL CONFERENCE ON SOFTWARE ENGINEERING AND KNOWLEDGE ENGINEERING, PROCEEDINGS · January 1, 1996 Link to item Cite

Non-Markovian Petri Nets

Conference Proceedings of the 1995 ACM Sigmetrics Joint International Conference on Measurement and Modeling of Computer Systems Sigmetrics 1995 Performance 1995 · May 1, 1995 Non-Markovian models allow us to capture a very wide range of circumstances in which it is necessary to model phenomena whose times to occurrence is not exponentially distributed. Events such as timeouts in a protocol, service times at a machine performing ... Full text Cite

Time-dependent behavior of redundant systems with deterministic repair

Conference COMPUTATIONS WITH MARKOV CHAINS · January 1, 1995 Link to item Cite

From stochastic Petri nets to Markov regenerative stochastic Petri nets

Conference Proceedings IEEE Computer Society S Annual International Symposium on Modeling Analysis and Simulation of Computer and Telecommunications Systems Mascots · January 1, 1995 In this paper we survey the Petri net literature and focus on Petri nets with generally distributed transition firing times. In the framework of Markov regenerative stochastic Petri nets (MRSPN) we develop and solve two examples to illustrate the modeling ... Full text Cite

Steady state analysis of markov regenerative SPN with age memory policy

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1995 Non-Markovian Stochastic Petri Nets (SPN) have been developed as a tool to deal with systems characterized by non exponentially distributed timed events. Recently, some effort has been devoted to the study of SPN with generally distributed firing times, wh ... Full text Cite

TRANSIENT ANALYSIS OF REAL-TIME SYSTEMS USING DETERMINISTIC AND STOCHASTIC PETRI NETS

Conference QUALITY OF COMMUNICATION-BASED SYSTEMS · January 1, 1995 Link to item Cite

Approximate computation of sojourn time distribution in open queueing networks

Conference COMPUTATIONS WITH MARKOV CHAINS · January 1, 1995 Link to item Cite

Coverage Evaluation Through Fault Injection: Fault Sampling and Statistical Analysis

Conference 3rd IEEE International Workshop on Integrating Error Models with Fault Injection Wiem 1994 · January 1, 1994 Full text Cite

A stochastic reward net model for dependability analysis of real-time computing systems

Conference Proceedings of 2nd IEEE Workshop on Real Time Applications Rta 1994 · January 1, 1994 Dependability assessment plays an important role in the design and validation of fault-tolerant real-lime computer systems. Dependability models provide measures such as reliability, safety and mean time to failure as functions of the component failure rat ... Full text Cite

Techniques and tools for reliability and performance evaluation: Problems and perspectives

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1994 Modelling techniques and tools of the future must meet the challenges presented by today's highly demanding and schedule-oriented developing environment. With the emergence of high performance and reliability systems the problem of how to analyze such syst ... Full text Cite

Sensitivity analysis of Markov regenerative stochastic Petri nets

Conference Proceedings of 5th International Workshop on Petri Nets and Performance Models Pnpm 1993 · January 1, 1993 Sensitivity analysis, i.e., the analysis of the effect of small variations in system parameters on the output measures, can be studied by computing the derivatives of the output measures with respect to the parameter. An algorithm for parametric sensitivit ... Full text Cite

A methodology for formal expression of hierarchy in model solution

Conference Proceedings of 5th International Workshop on Petri Nets and Performance Models Pnpm 1993 · January 1, 1993 A methodology for formal specification of hierarchy both in model specification and model solution is presented. Hierarchy is allowed to exist among different model types used in performance and dependability modeling. This offers a lot of flexibility and ... Full text Cite

Transient analysis of deterministic and stochastic petri nets

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1993 Deterministic and stochastic Petri nets (DSPNs) are recognized as a useful modeling technique because of their capability to represent constant delays which appear in many practical systems. If at most one deterministic transition is allowed to be enabled ... Full text Cite

Integration of specification for modeling and specification for system design

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1993 This paper presents a procedure of transforming an Estelle specification into Stochastic Reward Net (SRN) formalism. Estelle is an ISO standard formal specification language which can help avoid ambiguity, incompleteness and inconsistency in system develop ... Full text Cite

FSPNs: Fluid stochastic petri nets

Conference Lecture Notes in Computer Science Including Subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics · January 1, 1993 In this paper we introduce a new class of stochastic Petri nets in which one or more places can hold fluid rather than discrete tokens. After defining the class of fluid stochastic Petri nets, we provide equations for their transient and steady-state behav ... Full text Cite

Dependability and performability analysis

Conference Lecture Notes in Computer Science · January 1, 1993 In this tutorial, we discuss several practical issues regarding specification and solution of dependability and performability models. We compare model types with and without rewards. Continuous-time Markov chains (CTMCs) are compared with (continuous-time ... Full text Cite

A TOOLCHEST FOR STOCHASTIC-MODELS

Conference INTERNATIONAL CONFERENCE ON SIMULATION IN ENGINEERING EDUCATION · 1992 Link to item Cite

DEPENDABILITY MODELING FOR COMPUTER-SYSTEMS

Conference PROCEEDINGS ANNUAL RELIABILITY AND MAINTAINABILITY SYMPOSIUM · January 29, 1991 Link to item Cite

Reliability modeling of the MARS system: A case study in the use of different tools and techniques

Conference Proceedings of the 4th International Workshop on Petri Nets and Performance Models Pnpm 1991 · January 1, 1991 Analytical reliability modeling is a promising method for predicting the reliability of different architectural variants and to perform trade-off studies at design time. However, generating a computationally tractable analytic model implies in general an a ... Full text Cite

A decomposition approach for stochastic Petri net models

Conference Proceedings of the 4th International Workshop on Petri Nets and Performance Models Pnpm 1991 · January 1, 1991 We present a decomposition approach for the solution of large stochastic Petri nets (SPNs). The overall model consists of a set of submodels whose interactions are described by an import graph. Each node of the graph corresponds to a parametrized SPN submo ... Full text Cite

Fixed Point Iteration in Availability Modeling.

Conference Fault-Tolerant Computing Systems · 1991 Cite

SPNP - THE STOCHASTIC PETRI NET PACKAGE

Conference NUMERICAL SOLUTION OF MARKOV CHAINS · 1991 Link to item Cite

SOLUTION OF LARGE GSPN MODELS

Conference NUMERICAL SOLUTION OF MARKOV CHAINS · 1991 Link to item Cite

Reliability analysis of the FDDI token ring

Conference Proceedings Conference on Local Computer Networks LCN · January 1, 1991 In this paper we develop reliability models and derive closed-form results for network reliability and network mean time to failure, including both node and link failures, for a very popular high speed LAN, the FDDI (Fiber Distributed Data Interface). We t ... Full text Cite

GSPN Models: Sensitivity analysis and applications

Conference Proceedings 28th Annual Southeast Regional Conference ACM Se 1990 · April 1, 1990 Sensitivity analysis of continuous time Markov chains has been considered recently by several researchers. This is very useful in performing bottleneck analysis and optimization on systems especially during the design stage. However the construction of the ... Cite

GSPM models: sensitivity analysis and applications.

Conference ACM Southeast Regional Conference · 1990 Cite

Transient Analysis of Markov and Markov Reward Models.

Conference Computer Performance and Reliability · 1987 Cite

A Measurement-Based Performability Model for a Multiprocessor System.

Conference Computer Performance and Reliability · 1987 Cite

PERFORMANCE ANALYSIS USING USER BEHAVIOR GRAPHS

Conference 12th International Computer Measurement Group Conference Cmg 1986 · January 1, 1986 The vser behavior graph is a graphical model for describing the behavior of the interactive users. The adequacy of user behavior graphs in several performance evaluation studies for workload characterization is investigated. In absence of memory constraint ... Cite

Issues in reliability modeling of fault-tolerant computers.

Conference Fehlertolerierende Rechensysteme · 1984 Cite

COMPUTER SCIENCE AND APPLIED PROBABILITY

Conference 9th International Computer Measurement Group Conference Cmg 1983 · January 1, 1983 This is an extended abstract for a recent textbook, Probability and S^at istics with Reliability* Qocoing. and Conpnter Science Applications -Kishor S. Trivedi, Prentice-Hall, Englewood Cliffs, N.J., 1982. ... Cite

Computer Science and Applied Probability (abstract).

Conference Int. CMG Conference · 1983 Cite

OPTIMAL FILE ALLOCATION, DEVICE CAPACITY AND CPU SPEED SELECTION DURING THE DESIGN OF INTERACTIVE COMPUTER SYSTEMS

Conference 8th International Computer Measurement Group Conference Cmg 1982 · January 1, 1982 This paper considers a computer configuration design problem. The computer is modeled as a closed queueinq network. The average resDonse time to an interactive user request is to be minimized. The decision variables are CPU speed, capacities of I/O devices ... Cite

Optimal Design of an Interactive System: File Allocation, Device Capacity Selection, and CPU Speed Selection

Conference 7th International Computer Measurement Group Conference Cmg 1981 · January 1, 1981 This paper considers a computer configuration design problem. The computer is modeled as a closed gueueing network. The average response time to an interactive user request is to be minimized. The decision variables are CPU speed, capacities of I/O devices ... Cite

Hardware configuration selection through discretizing a continuous variable solution

Conference Proceedings of the 1980 International Symposium on Computer Performance Modelling Measurement and Evaluation Performance 1980 · May 28, 1980 This paper extends a previous model for computer system configuration planning developed by the authors. The problem is to optimally select the CPU speed, the device capacities, and file assignments so as to maximize throughput subject to a fixed cost cons ... Full text Cite

Designing linear storage hierarchies so as to maximize reliability subject to cost and performance constraints

Conference Proceedings International Symposium on Computer Architecture · May 6, 1980 A geometric programming model is proposed to determine the optimal design of the CPU and its matching storage hierarchy. The objective function is the maximization of system reliability subject to performance and budgetary limitations. Examples illustratin ... Full text Cite

A performance comparison of optimally designed computer systems with and without virtual memory

Conference Proceedings International Symposium on Computer Architecture · April 23, 1979 In this paper, a comparison of the performance of optimally designed computer systems with and without virtual memory is made. The computer systems in question are modeled by closed queuing networks of the central server type. The design of the systems is ... Full text Cite

The status of investigations into the use of continued fractions for computer hardware

Conference Proceedings Symposium on Computer Arithmetic · January 1, 1972 The purpose of this paper is to demonstrate that representations of numbers other than positional notation may lead to practical hardware realizations for the digital calculation of classes of algorithms. It is the authors' opinion that practicality of the ... Full text Cite

Proactive fault-management in software systems

Conference Proceedings 33rd Annual Simulation Symposium (SS 2000) Full text Cite

An approach for combinatorial performance and availability analysis

Conference Proceedings of 1993 IEEE 12th Symposium on Reliable Distributed Systems Full text Cite

Dependency characterization in path-based approaches to architecture-based software reliability prediction

Conference Proceedings. 1998 IEEE Workshop on Application-Specific Software Engineering and Technology. ASSET-98 (Cat. No.98EX183) Full text Cite

Performance analysis of distributed real-time databases

Conference Proceedings. IEEE International Computer Performance and Dependability Symposium. IPDS'98 (Cat. No.98TB100248) Full text Cite

An analytical approach to architecture-based software reliability prediction

Conference Proceedings. IEEE International Computer Performance and Dependability Symposium. IPDS'98 (Cat. No.98TB100248) Full text Cite