Designing load balancing and admission control policies: lessons from NDS regime

Size: px
Start display at page:

Download "Designing load balancing and admission control policies: lessons from NDS regime"

Transcription

1 Designing load balancing and admission control policies: lessons from NDS regime VARUN GUPTA University of Chicago Based on works with : Neil Walton, Jiheng Zhang

2 ρ K θ is a useful regime to study the interplay of service distributions and control policies 1. Load balancing with PS servers 2. Concurrency control in a statedependent PS server 1 Poisson(λ) Load Balancer 2 FCFS buffer μ(n) K Concurrency Control

3 GOAL: Min. mean sojourn time E[T] Poisson(λ) Load Balancer K Processor Sharing servers Knows queue lengths Not job sizes Q: Optimal load balancing policy? A: Exponential service Shortest-Queue General service??

4 10 E[T] Det Exp Bim-1 Weib-1 Weib-2 Bim-2 SERVICE DISTRIBUTIONS (INC. VARIANCE ) RANDOM ROUND-ROBIN LEAST-WORK S. QUEUE OPT-0 Q: Is Shortest Queue insensitive and near-optimal? GOAL: a meaningful asymptotic regime

5 Option 1 Light traffic: K = fixed, λ 0 # Servers: K Arr. Rate: λ Mean service time: E S = 1 Load: ρ λ K Can not infer insensitivity! E T 1 Option 2 Conventional heavy traffic: K = fixed, λ K System collapses to a single PS server. Even Least-Work becomes insensitive.

6 Option 3 Halfin-Whitt: K, K λ~β K # Servers: K Arr. Rate: λ Mean service time: E S = 1 Load: ρ λ K Can not infer insensitivity! In fact, for K E T 1 K λ ~ ω(1) E T 1 K λ ~ o(1) Collapses to PS server

7 Option 4 K, K λ = α # Servers: K Arr. Rate: λ Mean service time: E S = 1 Load: ρ λ K Equivalently, ρ K θ e α Under the above setting, 1 < E[T] E[S] < Non-degenerate slowdown (NDS) regime

8 1. Shortest Queue insensitive under NDS? 2. Shortest Queue near-optimal under NDS?

9 Poisson(λ) K K K λ = α Analysis technique: Process for total jobs in system λ λ λ λ λ λ 0 1 K 3K/2 2K 2K+1 2K+2? = mean departure rate given 3K/2 jobs

10 K/2 N = 3K/2 Poisson(λ) Rate = K/2 K/2 Rate = K Departure rate = K 1 (not K) Key: Find the O(1) fluctuations O(1) idle queues

11 γk N = (1+γ)K (0<γ<1) Poisson(λ) (1- γ)k Departure rate = K (1-γ)/ γ Rate = (1-γ)K Rate = K O(1) idle queues

12 Poisson(λ) K K K λ = α Analysis technique: Process for total jobs in system λ λ λ λ λ λ 0 1 K (1+ )K 2K 2K+1 2K+2 Asymptotically negligible probability mass K- (1- )/ K K

13 More formally: N K Kt N(t) K where N(t) is a Brownian motion with drift: μ n = α + (2 n)+! n 1 variance: σ 2 = 2. COROLLARY Shortest-Queue: E N SQ = g(α) Central-Queue (M/M/K): E N CQ = α sup α E[N SQ ] E[N CQ ] 1.14

14 E[N SQ ] E[N CQ ] 1 K=64 K=16 K=4 Type equation here.

15 More formally: N K Kt N(t) K where N(t) is a Brownian motion with drift: μ n = α + (2 n)+ n 1 variance: σ 2 = 2. COROLLARY Shortest-Queue: E N SQ = g(α)! IQF up to factor 2 worse than Central Queue Central-Queue (M/M/K): E N CQ = α Idle-Queue-First: E N IQF = α

16 1. Shortest Queue insensitive under NDS? 2. Shortest Queue near-optimal under NDS?

17 Proposition: For Phase-type service distributions with SQ routing, lim K Intuitively: N (K) K is insensitive beyond E S. 1. Time-scale separation: N K (t) evolves slowly 2. Closed-system (fluid limit) under SQ is insensitive Proposition: For H 2 service distribution with Expected-Least-Work routing, lim N (K) K is sensitive to distribution. K

18 Conjecture: 1. For any non-anticipative (but possibly size-aware) load balancer, lim K E N Π (K) K E N CQ. 2. For OPT-0 (myopic size-aware) lim K (K) E N OPT 0 K = E N CQ. Conjecture + Propositions SQ is within 14% of the optimal size-aware policy in NDS regime!

19 ρ K θ is a useful regime to study the interplay of service distributions and control policies 1. Load balancing with PS servers 2. Concurrency control in a statedependent PS server 1 Poisson(λ) Load Balancer 2 FCFS buffer μ(n) K Concurrency Control

20 State-dependent Processor Sharing μ(n) μ(n) n jobs at server service rate μ(n)/n per job

21 FCFS buffer Sharing Limit Control (L) State-dependent Processor Sharing μ(n) μ(n) Motivation Congestion based friction e.g., server thread-pool management, traffic flow Service systems with human agents

22 FCFS buffer Sharing Limit Control (L) State-dependent Processor Sharing μ(n) μ(n) L* = 1 Straw-man: Set to maximum efficiency point (L*) High efficiency but FCFS dominates Trade-off between arrival rate and service variability

23 FCFS buffer Sharing Limit Control (L) State-dependent Processor Sharing μ(n) μ(n) L* = 1 GOAL: Diffusion approximation for static L, general service distribution Q: What is a meaningful sequence of system that faithfully approximates the original system?

24 1.3 μ(n) λ L Option 1 Conventional heavy traffic: λ (r) μ L Does not capture original system!

25 1.3 μ(n) λ L Option 2 Scale sharing limit: L (r) = Lr Stretch service rate curve: μ r (rx) μ x where μ x is a continuous extension of μ n In this example, system gets stuck at 0.

26 Proposal Feed the original system M/M/ input Let H n = Prob N n and H a continuous differentiable interpolation 1 H(n) H x Scaling: In the r th system L (r) = Lr Devise μ r rx so that under M/M/ input: H r (rx) H x L

27 Scaling: In the r th system L (r) = Lr Under M/M/ input: H r (rx) H x We guarantee a limit that depends on the entire μ n Reverse engineered service rates: H(n) L H x! lim r λ μ r rx = λ r ρ r θ Entire μ r curve collapses to λ at rate 1/r d log( h x ) dx α(x) μ(n) drift function λ

28 E N n=0 (n L) π(n) n=0 π(n) c 2 s +1 c 2 s +c2 a c 2 s +1 c 2 s +c2 a + c s n=0 (n L) + π(n) n=0 π(n) c 2 s +1 c 2 s +c2 a c 2 s +1 c 2 s +c2 a where π(n) is probability mass function for the original system under M/M/ input 12 E[N] Weibull Pareto (1.1) LogNormal Approx Sharing Limit (L)

29 1. ρ K θ regime useful for: qualitative insights into load balancing policies for finite systems designing service variability based control policies for finite systems 2. Shortest Queue Near-optimal ( 14%) without knowing job sizes Idle-Queue-First is a poor proxy for SQ but going until queue length 1 gains all benefits.

State-dependent Limited Processor Sharing - Approximations and Optimal Control

State-dependent Limited Processor Sharing - Approximations and Optimal Control State-dependent Limited Processor Sharing - Approximations and Optimal Control Varun Gupta University of Chicago Joint work with: Jiheng Zhang (HKUST) State-dependent Limited Processor Sharing Processor

More information

Dynamic Control of Parallel-Server Systems

Dynamic Control of Parallel-Server Systems Dynamic Control of Parallel-Server Systems Jim Dai Georgia Institute of Technology Tolga Tezcan University of Illinois at Urbana-Champaign May 13, 2009 Jim Dai (Georgia Tech) Many-Server Asymptotic Optimality

More information

VARUN GUPTA. Takayuki Osogami (IBM Research-Tokyo) Carnegie Mellon Google Research University of Chicago Booth School of Business.

VARUN GUPTA. Takayuki Osogami (IBM Research-Tokyo) Carnegie Mellon Google Research University of Chicago Booth School of Business. VARUN GUPTA Carnegie Mellon Google Research University of Chicago Booth School of Business With: Taayui Osogami (IBM Research-Toyo) 1 2 Homogeneous servers 3 First-Come-First-Serve Buffer Homogeneous servers

More information

Control of Fork-Join Networks in Heavy-Traffic

Control of Fork-Join Networks in Heavy-Traffic in Heavy-Traffic Asaf Zviran Based on MSc work under the guidance of Rami Atar (Technion) and Avishai Mandelbaum (Technion) Industrial Engineering and Management Technion June 2010 Introduction Network

More information

MODELING WEBCHAT SERVICE CENTER WITH MANY LPS SERVERS

MODELING WEBCHAT SERVICE CENTER WITH MANY LPS SERVERS MODELING WEBCHAT SERVICE CENTER WITH MANY LPS SERVERS Jiheng Zhang Oct 26, 211 Model and Motivation Server Pool with multiple LPS servers LPS Server K Arrival Buffer. Model and Motivation Server Pool with

More information

Routing and Staffing in Large-Scale Service Systems: The Case of Homogeneous Impatient Customers and Heterogeneous Servers 1

Routing and Staffing in Large-Scale Service Systems: The Case of Homogeneous Impatient Customers and Heterogeneous Servers 1 Routing and Staffing in Large-Scale Service Systems: The Case of Homogeneous Impatient Customers and Heterogeneous Servers 1 Mor Armony 2 Avishai Mandelbaum 3 June 25, 2008 Abstract Motivated by call centers,

More information

Load Balancing in Distributed Service System: A Survey

Load Balancing in Distributed Service System: A Survey Load Balancing in Distributed Service System: A Survey Xingyu Zhou The Ohio State University zhou.2055@osu.edu November 21, 2016 Xingyu Zhou (OSU) Load Balancing November 21, 2016 1 / 29 Introduction and

More information

Maximum pressure policies for stochastic processing networks

Maximum pressure policies for stochastic processing networks Maximum pressure policies for stochastic processing networks Jim Dai Joint work with Wuqin Lin at Northwestern Univ. The 2011 Lunteren Conference Jim Dai (Georgia Tech) MPPs January 18, 2011 1 / 55 Outline

More information

Positive Harris Recurrence and Diffusion Scale Analysis of a Push Pull Queueing Network. Haifa Statistics Seminar May 5, 2008

Positive Harris Recurrence and Diffusion Scale Analysis of a Push Pull Queueing Network. Haifa Statistics Seminar May 5, 2008 Positive Harris Recurrence and Diffusion Scale Analysis of a Push Pull Queueing Network Yoni Nazarathy Gideon Weiss Haifa Statistics Seminar May 5, 2008 1 Outline 1 Preview of Results 2 Introduction Queueing

More information

Servers. Rong Wu. Department of Computing and Software, McMaster University, Canada. Douglas G. Down

Servers. Rong Wu. Department of Computing and Software, McMaster University, Canada. Douglas G. Down MRRK-SRPT: a Scheduling Policy for Parallel Servers Rong Wu Department of Computing and Software, McMaster University, Canada Douglas G. Down Department of Computing and Software, McMaster University,

More information

State-dependent and Energy-aware Control of Server Farm

State-dependent and Energy-aware Control of Server Farm State-dependent and Energy-aware Control of Server Farm Esa Hyytiä, Rhonda Righter and Samuli Aalto Aalto University, Finland UC Berkeley, USA First European Conference on Queueing Theory ECQT 2014 First

More information

Advanced Computer Networks Lecture 3. Models of Queuing

Advanced Computer Networks Lecture 3. Models of Queuing Advanced Computer Networks Lecture 3. Models of Queuing Husheng Li Min Kao Department of Electrical Engineering and Computer Science University of Tennessee, Knoxville Spring, 2016 1/13 Terminology of

More information

Contact Centers with a Call-Back Option and Real-Time Delay Information

Contact Centers with a Call-Back Option and Real-Time Delay Information OPERATIOS RESEARCH Vol. 52, o. 4, July August 2004, pp. 527 545 issn 0030-364X eissn 526-5463 04 5204 0527 informs doi 0.287/opre.040.023 2004 IFORMS Contact Centers with a Call-Back Option and Real-Time

More information

This article was published in an Elsevier journal. The attached copy is furnished to the author for non-commercial research and education use, including for instruction at the author s institution, sharing

More information

Erlang-C = M/M/N. agents. queue ACD. arrivals. busy ACD. queue. abandonment BACK FRONT. lost calls. arrivals. lost calls

Erlang-C = M/M/N. agents. queue ACD. arrivals. busy ACD. queue. abandonment BACK FRONT. lost calls. arrivals. lost calls Erlang-C = M/M/N agents arrivals ACD queue Erlang-A lost calls FRONT BACK arrivals busy ACD queue abandonment lost calls Erlang-C = M/M/N agents arrivals ACD queue Rough Performance Analysis

More information

Design, staffing and control of large service systems: The case of a single customer class and multiple server types. April 2, 2004 DRAFT.

Design, staffing and control of large service systems: The case of a single customer class and multiple server types. April 2, 2004 DRAFT. Design, staffing and control of large service systems: The case of a single customer class and multiple server types. Mor Armony 1 Avishai Mandelbaum 2 April 2, 2004 DRAFT Abstract Motivated by modern

More information

On the Resource/Performance Tradeoff in Large Scale Queueing Systems

On the Resource/Performance Tradeoff in Large Scale Queueing Systems On the Resource/Performance Tradeoff in Large Scale Queueing Systems David Gamarnik MIT Joint work with Patrick Eschenfeldt, John Tsitsiklis and Martin Zubeldia (MIT) High level comments High level comments

More information

BIRTH DEATH PROCESSES AND QUEUEING SYSTEMS

BIRTH DEATH PROCESSES AND QUEUEING SYSTEMS BIRTH DEATH PROCESSES AND QUEUEING SYSTEMS Andrea Bobbio Anno Accademico 999-2000 Queueing Systems 2 Notation for Queueing Systems /λ mean time between arrivals S = /µ ρ = λ/µ N mean service time traffic

More information

Robustness and performance of threshold-based resource allocation policies

Robustness and performance of threshold-based resource allocation policies Robustness and performance of threshold-based resource allocation policies Takayuki Osogami Mor Harchol-Balter Alan Scheller-Wolf Computer Science Department, Carnegie Mellon University, 5000 Forbes Avenue,

More information

Electronic Companion Fluid Models for Overloaded Multi-Class Many-Server Queueing Systems with FCFS Routing

Electronic Companion Fluid Models for Overloaded Multi-Class Many-Server Queueing Systems with FCFS Routing Submitted to Management Science manuscript MS-251-27 Electronic Companion Fluid Models for Overloaded Multi-Class Many-Server Queueing Systems with FCFS Routing Rishi Talreja, Ward Whitt Department of

More information

Markovian N-Server Queues (Birth & Death Models)

Markovian N-Server Queues (Birth & Death Models) Markovian -Server Queues (Birth & Death Moels) - Busy Perio Arrivals Poisson (λ) ; Services exp(µ) (E(S) = /µ) Servers statistically ientical, serving FCFS. Offere loa R = λ E(S) = λ/µ Erlangs Q(t) = number

More information

Classical Queueing Models.

Classical Queueing Models. Sergey Zeltyn January 2005 STAT 99. Service Engineering. The Wharton School. University of Pennsylvania. Based on: Classical Queueing Models. Mandelbaum A. Service Engineering course, Technion. http://iew3.technion.ac.il/serveng2005w

More information

A STAFFING ALGORITHM FOR CALL CENTERS WITH SKILL-BASED ROUTING: SUPPLEMENTARY MATERIAL

A STAFFING ALGORITHM FOR CALL CENTERS WITH SKILL-BASED ROUTING: SUPPLEMENTARY MATERIAL A STAFFING ALGORITHM FOR CALL CENTERS WITH SKILL-BASED ROUTING: SUPPLEMENTARY MATERIAL by Rodney B. Wallace IBM and The George Washington University rodney.wallace@us.ibm.com Ward Whitt Columbia University

More information

Proactive Care with Degrading Class Types

Proactive Care with Degrading Class Types Proactive Care with Degrading Class Types Yue Hu (DRO, Columbia Business School) Joint work with Prof. Carri Chan (DRO, Columbia Business School) and Prof. Jing Dong (DRO, Columbia Business School) Motivation

More information

Introduction to Markov Chains, Queuing Theory, and Network Performance

Introduction to Markov Chains, Queuing Theory, and Network Performance Introduction to Markov Chains, Queuing Theory, and Network Performance Marceau Coupechoux Telecom ParisTech, departement Informatique et Réseaux marceau.coupechoux@telecom-paristech.fr IT.2403 Modélisation

More information

Queueing Theory I Summary! Little s Law! Queueing System Notation! Stationary Analysis of Elementary Queueing Systems " M/M/1 " M/M/m " M/M/1/K "

Queueing Theory I Summary! Little s Law! Queueing System Notation! Stationary Analysis of Elementary Queueing Systems  M/M/1  M/M/m  M/M/1/K Queueing Theory I Summary Little s Law Queueing System Notation Stationary Analysis of Elementary Queueing Systems " M/M/1 " M/M/m " M/M/1/K " Little s Law a(t): the process that counts the number of arrivals

More information

Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance

Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance In this technical appendix we provide proofs for the various results stated in the manuscript

More information

arxiv: v1 [cs.pf] 22 Dec 2017

arxiv: v1 [cs.pf] 22 Dec 2017 Scalable Load Balancing in etworked Systems: Universality Properties and Stochastic Coupling Methods Mark van der Boor 1, Sem C. Borst 1,2, Johan S.H. van Leeuwaarden 1, and Debankur Mukherjee 1 1 Eindhoven

More information

On the Partitioning of Servers in Queueing Systems during Rush Hour

On the Partitioning of Servers in Queueing Systems during Rush Hour On the Partitioning of Servers in Queueing Systems during Rush Hour This paper is motivated by two phenomena observed in many queueing systems in practice. The first is the partitioning of server capacity

More information

Blind Fair Routing in Large-Scale Service Systems with Heterogeneous Customers and Servers

Blind Fair Routing in Large-Scale Service Systems with Heterogeneous Customers and Servers Blind Fair Routing in Large-Scale Service Systems with Heterogeneous Customers and Servers Mor Armony Amy R. Ward 2 October 6, 2 Abstract In a call center, arriving customers must be routed to available

More information

Class 11 Non-Parametric Models of a Service System; GI/GI/1, GI/GI/n: Exact & Approximate Analysis.

Class 11 Non-Parametric Models of a Service System; GI/GI/1, GI/GI/n: Exact & Approximate Analysis. Service Engineering Class 11 Non-Parametric Models of a Service System; GI/GI/1, GI/GI/n: Exact & Approximate Analysis. G/G/1 Queue: Virtual Waiting Time (Unfinished Work). GI/GI/1: Lindley s Equations

More information

Stochastic Networks and Parameter Uncertainty

Stochastic Networks and Parameter Uncertainty Stochastic Networks and Parameter Uncertainty Assaf Zeevi Graduate School of Business Columbia University Stochastic Processing Networks Conference, August 2009 based on joint work with Mike Harrison Achal

More information

Queueing systems. Renato Lo Cigno. Simulation and Performance Evaluation Queueing systems - Renato Lo Cigno 1

Queueing systems. Renato Lo Cigno. Simulation and Performance Evaluation Queueing systems - Renato Lo Cigno 1 Queueing systems Renato Lo Cigno Simulation and Performance Evaluation 2014-15 Queueing systems - Renato Lo Cigno 1 Queues A Birth-Death process is well modeled by a queue Indeed queues can be used to

More information

SQF: A slowdown queueing fairness measure

SQF: A slowdown queueing fairness measure Performance Evaluation 64 (27) 1121 1136 www.elsevier.com/locate/peva SQF: A slowdown queueing fairness measure Benjamin Avi-Itzhak a, Eli Brosh b,, Hanoch Levy c a RUTCOR, Rutgers University, New Brunswick,

More information

Analysis of Round-Robin Implementations of Processor Sharing, Including Overhead

Analysis of Round-Robin Implementations of Processor Sharing, Including Overhead Analysis of Round-Robin Implementations of Processor Sharing, Including Overhead Steve Thompson and Lester Lipsky Computer Science & Engineering University of Connecticut Storrs, CT 669 Email: sat3@engr.uconn.edu;

More information

Operations Research Letters. Instability of FIFO in a simple queueing system with arbitrarily low loads

Operations Research Letters. Instability of FIFO in a simple queueing system with arbitrarily low loads Operations Research Letters 37 (2009) 312 316 Contents lists available at ScienceDirect Operations Research Letters journal homepage: www.elsevier.com/locate/orl Instability of FIFO in a simple queueing

More information

THE HEAVY-TRAFFIC BOTTLENECK PHENOMENON IN OPEN QUEUEING NETWORKS. S. Suresh and W. Whitt AT&T Bell Laboratories Murray Hill, New Jersey 07974

THE HEAVY-TRAFFIC BOTTLENECK PHENOMENON IN OPEN QUEUEING NETWORKS. S. Suresh and W. Whitt AT&T Bell Laboratories Murray Hill, New Jersey 07974 THE HEAVY-TRAFFIC BOTTLENECK PHENOMENON IN OPEN QUEUEING NETWORKS by S. Suresh and W. Whitt AT&T Bell Laboratories Murray Hill, New Jersey 07974 ABSTRACT This note describes a simulation experiment involving

More information

Asymptotic Coupling of an SPDE, with Applications to Many-Server Queues

Asymptotic Coupling of an SPDE, with Applications to Many-Server Queues Asymptotic Coupling of an SPDE, with Applications to Many-Server Queues Mohammadreza Aghajani joint work with Kavita Ramanan Brown University March 2014 Mohammadreza Aghajanijoint work Asymptotic with

More information

Routing and Staffing in Large-Scale Service Systems: The Case of Homogeneous Impatient Customers and Heterogeneous Servers

Routing and Staffing in Large-Scale Service Systems: The Case of Homogeneous Impatient Customers and Heterogeneous Servers OPERATIONS RESEARCH Vol. 59, No. 1, January February 2011, pp. 50 65 issn 0030-364X eissn 1526-5463 11 5901 0050 informs doi 10.1287/opre.1100.0878 2011 INFORMS Routing and Staffing in Large-Scale Service

More information

Periodic Load Balancing

Periodic Load Balancing Queueing Systems 0 (2000)?? 1 Periodic Load Balancing Gisli Hjálmtýsson a Ward Whitt a a AT&T Labs, 180 Park Avenue, Building 103, Florham Park, NJ 07932-0971 E-mail: {gisli,wow}@research.att.com Queueing

More information

SENSITIVITY OF PERFORMANCE IN THE ERLANG-A QUEUEING MODEL TO CHANGES IN THE MODEL PARAMETERS

SENSITIVITY OF PERFORMANCE IN THE ERLANG-A QUEUEING MODEL TO CHANGES IN THE MODEL PARAMETERS SENSITIVITY OF PERFORMANCE IN THE ERLANG-A QUEUEING MODEL TO CHANGES IN THE MODEL PARAMETERS by Ward Whitt Department of Industrial Engineering and Operations Research Columbia University, New York, NY

More information

Intro Refresher Reversibility Open networks Closed networks Multiclass networks Other networks. Queuing Networks. Florence Perronnin

Intro Refresher Reversibility Open networks Closed networks Multiclass networks Other networks. Queuing Networks. Florence Perronnin Queuing Networks Florence Perronnin Polytech Grenoble - UGA March 23, 27 F. Perronnin (UGA) Queuing Networks March 23, 27 / 46 Outline Introduction to Queuing Networks 2 Refresher: M/M/ queue 3 Reversibility

More information

MASSACHUSETTS INSTITUTE OF TECHNOLOGY 6.265/15.070J Fall 2013 Lecture 22 12/09/2013. Skorokhod Mapping Theorem. Reflected Brownian Motion

MASSACHUSETTS INSTITUTE OF TECHNOLOGY 6.265/15.070J Fall 2013 Lecture 22 12/09/2013. Skorokhod Mapping Theorem. Reflected Brownian Motion MASSACHUSETTS INSTITUTE OF TECHNOLOGY 6.265/15.7J Fall 213 Lecture 22 12/9/213 Skorokhod Mapping Theorem. Reflected Brownian Motion Content. 1. G/G/1 queueing system 2. One dimensional reflection mapping

More information

State Space Collapse in Many-Server Diffusion Limits of Parallel Server Systems. J. G. Dai. Tolga Tezcan

State Space Collapse in Many-Server Diffusion Limits of Parallel Server Systems. J. G. Dai. Tolga Tezcan State Space Collapse in Many-Server Diffusion imits of Parallel Server Systems J. G. Dai H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology, Atlanta, Georgia

More information

Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement

Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement Submitted to imanufacturing & Service Operations Management manuscript MSOM-11-370.R3 Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement (Authors names blinded

More information

ENGINEERING SOLUTION OF A BASIC CALL-CENTER MODEL

ENGINEERING SOLUTION OF A BASIC CALL-CENTER MODEL ENGINEERING SOLUTION OF A BASIC CALL-CENTER MODEL by Ward Whitt Department of Industrial Engineering and Operations Research Columbia University, New York, NY 10027 Abstract An algorithm is developed to

More information

Dynamic Matching Models

Dynamic Matching Models Dynamic Matching Models Ana Bušić Inria Paris - Rocquencourt CS Department of École normale supérieure joint work with Varun Gupta, Jean Mairesse and Sean Meyn 3rd Workshop on Cognition and Control January

More information

CPSC 531: System Modeling and Simulation. Carey Williamson Department of Computer Science University of Calgary Fall 2017

CPSC 531: System Modeling and Simulation. Carey Williamson Department of Computer Science University of Calgary Fall 2017 CPSC 531: System Modeling and Simulation Carey Williamson Department of Computer Science University of Calgary Fall 2017 Motivating Quote for Queueing Models Good things come to those who wait - poet/writer

More information

6 Solving Queueing Models

6 Solving Queueing Models 6 Solving Queueing Models 6.1 Introduction In this note we look at the solution of systems of queues, starting with simple isolated queues. The benefits of using predefined, easily classified queues will

More information

Beyond heavy-traffic regimes: Universal bounds and controls for the single-server (M/GI/1+GI) queue

Beyond heavy-traffic regimes: Universal bounds and controls for the single-server (M/GI/1+GI) queue 1 / 33 Beyond heavy-traffic regimes: Universal bounds and controls for the single-server (M/GI/1+GI) queue Itai Gurvich Northwestern University Junfei Huang Chinese University of Hong Kong Stochastic Networks

More information

DELAY, MEMORY, AND MESSAGING TRADEOFFS IN DISTRIBUTED SERVICE SYSTEMS

DELAY, MEMORY, AND MESSAGING TRADEOFFS IN DISTRIBUTED SERVICE SYSTEMS DELAY, MEMORY, AND MESSAGING TRADEOFFS IN DISTRIBUTED SERVICE SYSTEMS By David Gamarnik, John N. Tsitsiklis and Martin Zubeldia Massachusetts Institute of Technology 5 th September, 2017 We consider the

More information

BRAVO for QED Queues

BRAVO for QED Queues 1 BRAVO for QED Queues Yoni Nazarathy, The University of Queensland Joint work with Daryl J. Daley, The University of Melbourne, Johan van Leeuwaarden, EURANDOM, Eindhoven University of Technology. Applied

More information

Multi-layered Round Robin Routing for Parallel Servers

Multi-layered Round Robin Routing for Parallel Servers Multi-layered Round Robin Routing for Parallel Servers Douglas G. Down Department of Computing and Software McMaster University 1280 Main Street West, Hamilton, ON L8S 4L7, Canada downd@mcmaster.ca 905-525-9140

More information

Task Assignment in a Heterogeneous Server Farm with Switching Delays and General Energy-Aware Cost Structure

Task Assignment in a Heterogeneous Server Farm with Switching Delays and General Energy-Aware Cost Structure Task Assignment in a Heterogeneous Server Farm with Switching Delays and General Energy-Aware Cost Structure Esa Hyytiä a,,, Rhonda Righter b, Samuli Aalto a a Department of Communications and Networking,

More information

Scheduling for the tail: Robustness versus Optimality

Scheduling for the tail: Robustness versus Optimality Scheduling for the tail: Robustness versus Optimality Jayakrishnan Nair, Adam Wierman 2, and Bert Zwart 3 Department of Electrical Engineering, California Institute of Technology 2 Computing and Mathematical

More information

Other properties of M M 1

Other properties of M M 1 Other properties of M M 1 Přemysl Bejda premyslbejda@gmail.com 2012 Contents 1 Reflected Lévy Process 2 Time dependent properties of M M 1 3 Waiting times and queue disciplines in M M 1 Contents 1 Reflected

More information

Multi-layered Round Robin Routing for Parallel Servers

Multi-layered Round Robin Routing for Parallel Servers Multi-layered Round Robin Routing for Parallel Servers Douglas G. Down Department of Computing and Software McMaster University 1280 Main Street West, Hamilton, ON L8S 4L7, Canada downd@mcmaster.ca 905-525-9140

More information

Tales of Time Scales. Ward Whitt AT&T Labs Research Florham Park, NJ

Tales of Time Scales. Ward Whitt AT&T Labs Research Florham Park, NJ Tales of Time Scales Ward Whitt AT&T Labs Research Florham Park, NJ New Book Stochastic-Process Limits An Introduction to Stochastic-Process Limits and Their Application to Queues Springer 2001 I won t

More information

Analysis of Software Artifacts

Analysis of Software Artifacts Analysis of Software Artifacts System Performance I Shu-Ngai Yeung (with edits by Jeannette Wing) Department of Statistics Carnegie Mellon University Pittsburgh, PA 15213 2001 by Carnegie Mellon University

More information

A Fluid Approximation for Service Systems Responding to Unexpected Overloads

A Fluid Approximation for Service Systems Responding to Unexpected Overloads OPERATIONS RESEARCH Vol. 59, No. 5, September October 2011, pp. 1159 1170 issn 0030-364X eissn 1526-5463 11 5905 1159 http://dx.doi.org/10.1287/opre.1110.0985 2011 INFORMS A Fluid Approximation for Service

More information

Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement

Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement Dynamic Call Center Routing Policies Using Call Waiting and Agent Idle Times Online Supplement Wyean Chan DIRO, Université de Montréal, C.P. 6128, Succ. Centre-Ville, Montréal (Québec), H3C 3J7, CANADA,

More information

Fair Dynamic Routing in Large-Scale Heterogeneous-Server Systems

Fair Dynamic Routing in Large-Scale Heterogeneous-Server Systems OPERATIONS RESEARCH Vol. 58, No. 3, May June 2010, pp. 624 637 issn 0030-364X eissn 1526-5463 10 5803 0624 informs doi 10.1287/opre.1090.0777 2010 INFORMS Fair Dynamic Routing in Large-Scale Heterogeneous-Server

More information

Introduction to Queueing Theory

Introduction to Queueing Theory Introduction to Queueing Theory Raj Jain Washington University in Saint Louis Saint Louis, MO 63130 Jain@cse.wustl.edu Audio/Video recordings of this lecture are available at: http://www.cse.wustl.edu/~jain/cse567-11/

More information

Maximizing throughput in zero-buffer tandem lines with dedicated and flexible servers

Maximizing throughput in zero-buffer tandem lines with dedicated and flexible servers Maximizing throughput in zero-buffer tandem lines with dedicated and flexible servers Mohammad H. Yarmand and Douglas G. Down Department of Computing and Software, McMaster University, Hamilton, ON, L8S

More information

Stein s Method for Steady-State Approximations: Error Bounds and Engineering Solutions

Stein s Method for Steady-State Approximations: Error Bounds and Engineering Solutions Stein s Method for Steady-State Approximations: Error Bounds and Engineering Solutions Jim Dai Joint work with Anton Braverman and Jiekun Feng Cornell University Workshop on Congestion Games Institute

More information

Online Supplement to Delay-Based Service Differentiation with Many Servers and Time-Varying Arrival Rates

Online Supplement to Delay-Based Service Differentiation with Many Servers and Time-Varying Arrival Rates Online Supplement to Delay-Based Service Differentiation with Many Servers and Time-Varying Arrival Rates Xu Sun and Ward Whitt Department of Industrial Engineering and Operations Research, Columbia University

More information

Performance Evaluation of Queuing Systems

Performance Evaluation of Queuing Systems Performance Evaluation of Queuing Systems Introduction to Queuing Systems System Performance Measures & Little s Law Equilibrium Solution of Birth-Death Processes Analysis of Single-Station Queuing Systems

More information

A Heavy Traffic Approximation for Queues with Restricted Customer-Server Matchings

A Heavy Traffic Approximation for Queues with Restricted Customer-Server Matchings A Heavy Traffic Approximation for Queues with Restricted Customer-Server Matchings (Working Paper #OM-007-4, Stern School Business) René A. Caldentey Edward H. Kaplan Abstract We consider a queueing system

More information

Blind Fair Routing in Large-Scale Service Systems with Heterogeneous Customers and Servers

Blind Fair Routing in Large-Scale Service Systems with Heterogeneous Customers and Servers OPERATIONS RESEARCH Vol. 6, No., January February 23, pp. 228 243 ISSN 3-364X (print) ISSN 526-5463 (online) http://dx.doi.org/.287/opre.2.29 23 INFORMS Blind Fair Routing in Large-Scale Service Systems

More information

Elementary queueing system

Elementary queueing system Elementary queueing system Kendall notation Little s Law PASTA theorem Basics of M/M/1 queue M/M/1 with preemptive-resume priority M/M/1 with non-preemptive priority 1 History of queueing theory An old

More information

Dynamic resource sharing

Dynamic resource sharing J. Virtamo 38.34 Teletraffic Theory / Dynamic resource sharing and balanced fairness Dynamic resource sharing In previous lectures we have studied different notions of fair resource sharing. Our focus

More information

We consider how two networked large-scale service systems that normally operate separately, such as call

We consider how two networked large-scale service systems that normally operate separately, such as call MANAGEMENT SCIENCE Vol. 55, No. 8, August 2009, pp. 1353 1367 issn 0025-1909 eissn 1526-5501 09 5508 1353 informs doi 10.1287/mnsc.1090.1025 2009 INFORMS Responding to Unexpected Overloads in Large-Scale

More information

Congestion In Large Balanced Fair Links

Congestion In Large Balanced Fair Links Congestion In Large Balanced Fair Links Thomas Bonald (Telecom Paris-Tech), Jean-Paul Haddad (Ernst and Young) and Ravi R. Mazumdar (Waterloo) ITC 2011, San Francisco Introduction File transfers compose

More information

TOWARDS BETTER MULTI-CLASS PARAMETRIC-DECOMPOSITION APPROXIMATIONS FOR OPEN QUEUEING NETWORKS

TOWARDS BETTER MULTI-CLASS PARAMETRIC-DECOMPOSITION APPROXIMATIONS FOR OPEN QUEUEING NETWORKS TOWARDS BETTER MULTI-CLASS PARAMETRIC-DECOMPOSITION APPROXIMATIONS FOR OPEN QUEUEING NETWORKS by Ward Whitt AT&T Bell Laboratories Murray Hill, NJ 07974-0636 March 31, 199 Revision: November 9, 199 ABSTRACT

More information

On the Price of Anarchy in Unbounded Delay Networks

On the Price of Anarchy in Unbounded Delay Networks On the Price of Anarchy in Unbounded Delay Networks Tao Wu Nokia Research Center Cambridge, Massachusetts, USA tao.a.wu@nokia.com David Starobinski Boston University Boston, Massachusetts, USA staro@bu.edu

More information

Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance

Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance Technical Appendix for: When Promotions Meet Operations: Cross-Selling and Its Effect on Call-Center Performance In this technical appendix we provide proofs for the various results stated in the manuscript

More information

Overflow Networks: Approximations and Implications to Call-Center Outsourcing

Overflow Networks: Approximations and Implications to Call-Center Outsourcing Overflow Networks: Approximations and Implications to Call-Center Outsourcing Itai Gurvich (Northwestern University) Joint work with Ohad Perry (CWI) Call Centers with Overflow λ 1 λ 2 Source of complexity:

More information

OPEN MULTICLASS HL QUEUEING NETWORKS: PROGRESS AND SURPRISES OF THE PAST 15 YEARS. w 1. v 2. v 3. Ruth J. Williams University of California, San Diego

OPEN MULTICLASS HL QUEUEING NETWORKS: PROGRESS AND SURPRISES OF THE PAST 15 YEARS. w 1. v 2. v 3. Ruth J. Williams University of California, San Diego OPEN MULTICLASS HL QUEUEING NETWORKS: PROGRESS AND SURPRISES OF THE PAST 15 YEARS v 2 w3 v1 w 2 w 1 v 3 Ruth J. Williams University of California, San Diego 1 PERSPECTIVE MQN SPN Sufficient conditions

More information

Lecture 20: Reversible Processes and Queues

Lecture 20: Reversible Processes and Queues Lecture 20: Reversible Processes and Queues 1 Examples of reversible processes 11 Birth-death processes We define two non-negative sequences birth and death rates denoted by {λ n : n N 0 } and {µ n : n

More information

DIMENSIONING BANDWIDTH FOR ELASTIC TRAFFIC IN HIGH-SPEED DATA NETWORKS

DIMENSIONING BANDWIDTH FOR ELASTIC TRAFFIC IN HIGH-SPEED DATA NETWORKS Submitted to IEEE/ACM Transactions on etworking DIMESIOIG BADWIDTH FOR ELASTIC TRAFFIC I HIGH-SPEED DATA ETWORKS Arthur W. Berger * and Yaakov Kogan Abstract Simple and robust engineering rules for dimensioning

More information

Kendall notation. PASTA theorem Basics of M/M/1 queue

Kendall notation. PASTA theorem Basics of M/M/1 queue Elementary queueing system Kendall notation Little s Law PASTA theorem Basics of M/M/1 queue 1 History of queueing theory An old research area Started in 1909, by Agner Erlang (to model the Copenhagen

More information

Pooling in tandem queueing networks with non-collaborative servers

Pooling in tandem queueing networks with non-collaborative servers DOI 10.1007/s11134-017-9543-0 Pooling in tandem queueing networks with non-collaborative servers Nilay Tanık Argon 1 Sigrún Andradóttir 2 Received: 22 September 2016 / Revised: 26 June 2017 Springer Science+Business

More information

Fair Operation of Multi-Server and Multi-Queue Systems

Fair Operation of Multi-Server and Multi-Queue Systems Fair Operation of Multi-Server and Multi-Queue Systems David Raz School of Computer Science Tel-Aviv University, Tel-Aviv, Israel davidraz@post.tau.ac.il Benjamin Avi-Itzhak RUTCOR, Rutgers University,

More information

A Diffusion Approximation for the G/GI/n/m Queue

A Diffusion Approximation for the G/GI/n/m Queue OPERATIONS RESEARCH Vol. 52, No. 6, November December 2004, pp. 922 941 issn 0030-364X eissn 1526-5463 04 5206 0922 informs doi 10.1287/opre.1040.0136 2004 INFORMS A Diffusion Approximation for the G/GI/n/m

More information

A Study on Performance Analysis of Queuing System with Multiple Heterogeneous Servers

A Study on Performance Analysis of Queuing System with Multiple Heterogeneous Servers UNIVERSITY OF OKLAHOMA GENERAL EXAM REPORT A Study on Performance Analysis of Queuing System with Multiple Heterogeneous Servers Prepared by HUSNU SANER NARMAN husnu@ou.edu based on the papers 1) F. S.

More information

CDA5530: Performance Models of Computers and Networks. Chapter 4: Elementary Queuing Theory

CDA5530: Performance Models of Computers and Networks. Chapter 4: Elementary Queuing Theory CDA5530: Performance Models of Computers and Networks Chapter 4: Elementary Queuing Theory Definition Queuing system: a buffer (waiting room), service facility (one or more servers) a scheduling policy

More information

Designing a Telephone Call Center with Impatient Customers

Designing a Telephone Call Center with Impatient Customers Designing a Telephone Call Center with Impatient Customers with Ofer Garnett Marty Reiman Sergey Zeltyn Appendix: Strong Approximations of M/M/ + Source: ErlangA QEDandSTROG FIAL.tex M/M/ + M System Poisson

More information

On Dynamic Scheduling of a Parallel Server System with Partial Pooling

On Dynamic Scheduling of a Parallel Server System with Partial Pooling On Dynamic Scheduling of a Parallel Server System with Partial Pooling V. Pesic and R. J. Williams Department of Mathematics University of California, San Diego 9500 Gilman Drive La Jolla CA 92093-0112

More information

NICTA Short Course. Network Analysis. Vijay Sivaraman. Day 1 Queueing Systems and Markov Chains. Network Analysis, 2008s2 1-1

NICTA Short Course. Network Analysis. Vijay Sivaraman. Day 1 Queueing Systems and Markov Chains. Network Analysis, 2008s2 1-1 NICTA Short Course Network Analysis Vijay Sivaraman Day 1 Queueing Systems and Markov Chains Network Analysis, 2008s2 1-1 Outline Why a short course on mathematical analysis? Limited current course offering

More information

On the static assignment to parallel servers

On the static assignment to parallel servers On the static assignment to parallel servers Ger Koole Vrije Universiteit Faculty of Mathematics and Computer Science De Boelelaan 1081a, 1081 HV Amsterdam The Netherlands Email: koole@cs.vu.nl, Url: www.cs.vu.nl/

More information

arxiv: v1 [math.pr] 11 May 2018

arxiv: v1 [math.pr] 11 May 2018 FCFS Parallel Service Systems and Matching Models Ivo Adan a, Igor Kleiner b,, Rhonda Righter c, Gideon Weiss b,, a Eindhoven University of Technology b Department of Statistics, The University of Haifa,

More information

Fluid Models of Parallel Service Systems under FCFS

Fluid Models of Parallel Service Systems under FCFS Fluid Models of Parallel Service Systems under FCFS Hanqin Zhang Business School, National University of Singapore Joint work with Yuval Nov and Gideon Weiss from The University of Haifa, Israel Queueing

More information

Motivated by models of tenant assignment in public housing, we study approximating deterministic fluid

Motivated by models of tenant assignment in public housing, we study approximating deterministic fluid MANAGEMENT SCIENCE Vol. 54, No. 8, August 2008, pp. 1513 1527 issn 0025-1909 eissn 1526-5501 08 5408 1513 informs doi 10.1287/mnsc.1080.0868 2008 INFORMS Fluid Models for Overloaded Multiclass Many-Server

More information

DIMENSIONING BANDWIDTH FOR ELASTIC TRAFFIC IN HIGH-SPEED DATA NETWORKS

DIMENSIONING BANDWIDTH FOR ELASTIC TRAFFIC IN HIGH-SPEED DATA NETWORKS Submitted to IEEE/ACM Transactions on etworking DIMESIOIG BADWIDTH FOR ELASTIC TRAFFIC I HIGH-SPEED DATA ETWORKS Arthur W. Berger and Yaakov Kogan AT&T Labs 0 Crawfords Corner Rd. Holmdel J, 07733 U.S.A.

More information

Service-Level Differentiation in Many-Server Service Systems Via Queue-Ratio Routing

Service-Level Differentiation in Many-Server Service Systems Via Queue-Ratio Routing Submitted to Operations Research manuscript (Please, provide the mansucript number!) Service-Level Differentiation in Many-Server Service Systems Via Queue-Ratio Routing Itay Gurvich Kellogg School of

More information

Analysis of A Single Queue

Analysis of A Single Queue Analysis of A Single Queue Raj Jain Washington University in Saint Louis Jain@eecs.berkeley.edu or Jain@wustl.edu A Mini-Course offered at UC Berkeley, Sept-Oct 2012 These slides and audio/video recordings

More information

Chapter 10. Queuing Systems. D (Queuing Theory) Queuing theory is the branch of operations research concerned with waiting lines.

Chapter 10. Queuing Systems. D (Queuing Theory) Queuing theory is the branch of operations research concerned with waiting lines. Chapter 10 Queuing Systems D. 10. 1. (Queuing Theory) Queuing theory is the branch of operations research concerned with waiting lines. D. 10.. (Queuing System) A ueuing system consists of 1. a user source.

More information

A DIFFUSION APPROXIMATION FOR THE G/GI/n/m QUEUE

A DIFFUSION APPROXIMATION FOR THE G/GI/n/m QUEUE A DIFFUSION APPROXIMATION FOR THE G/GI/n/m QUEUE by Ward Whitt Department of Industrial Engineering and Operations Research Columbia University, New York, NY 10027 July 5, 2002 Revision: June 27, 2003

More information

arxiv: v1 [cs.pf] 5 Nov 2018

arxiv: v1 [cs.pf] 5 Nov 2018 Stabilizing the virtual response in single-server processor sharing queues with slowly -varying s arxiv:8.v [cs.pf] Nov 8 Abstract Yongkyu Cho, Young Myoung Ko Department of Industrial and Management Engineering

More information

A Queueing System with Queue Length Dependent Service Times, with Applications to Cell Discarding in ATM Networks

A Queueing System with Queue Length Dependent Service Times, with Applications to Cell Discarding in ATM Networks A Queueing System with Queue Length Dependent Service Times, with Applications to Cell Discarding in ATM Networks by Doo Il Choi, Charles Knessl and Charles Tier University of Illinois at Chicago 85 South

More information