On Continuous Counting and Learning

in a Distributed System episode artwork

EPISODE · Aug 2, 2012 · 1H 5M

On Continuous Counting and Learning in a Distributed System

from Hamilton Institute Seminars (HD / large) · host Hamilton Institute

Speaker: Dr. B. Radunović Abstract: Consider a distributed system that consists of a coordinator node connected to multiple sites. Items from a data stream arrive to the system one by one, and are arbitrarily distributed to different sites. The goal of the system is to continuously track a function of the items received so far within a prescribed relative accuracy and at the lowest possible communication cost. This class of problems is called a continual distributed stream monitoring. In this talk we will focus on two problems from this class. We will first discuss the count tracking problem (counter), which is an important building block for other more complex algorithms. The goal of the counter is to keep a track of the sum of all the items from the stream at all times. We show that for a class of input loads a randomized algorithm guarantees to track the count accurately with high probability and has the expected communication cost that is sublinear in both data size and the number of sites. We also establish matching lower bounds. We then illustrate how our non-monotonic counter can be applied to solve more complex problems, such as to track the second frequency moment and the Bayesian linear regression of the input stream. We will next discuss the online non-stochastic experts problem in the continual distributed setting. Here, at each time-step, one of the sites has to pick one expert from the set of experts, and then the same site receives information about payoffs of all experts for that round. The goal of the distributed system is to minimize regret with respect to the optimal choice in hindsight, while simultaneously keeping communication to the minimum. This problem is well understood in the centralized setting, but the communication trade-off in the distributed setting is unknown. The two extreme solutions to this problem are to communicate with everyone after each payoff, and not to communicate at all. We will discuss how to achieve the trade-off between these two approaches. We will present an algorithm that achieves a non-trivial trade-off and show the difficulties of further improving its performance.

Episode metadata supplied by the publisher feed · Published Aug 2, 2012

Embed this episode

Ready to play

On Continuous Counting and Learning in a Distributed System

0:00 1:05:53

No transcript for this episode yet

We transcribe on demand. Request one and we'll notify you when it's ready — usually under 10 minutes.

No similar podcasts found.

Frequently Asked Questions

How long is this episode of Hamilton Institute Seminars (HD / large)?

This episode is 1 hour and 5 minutes long.

When was this Hamilton Institute Seminars (HD / large) episode published?

This episode was published on August 2, 2012.

Can I download this Hamilton Institute Seminars (HD / large) episode?

Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!