Embodiments include a system and method for activity monitoring using video data from multiple dissimilar sources. The video data is processed to remove any dependency of the system on types of video input data. The video data is processed to yield useful human readable information regarding events in real time, such as how many people move through a line in a period of time.