Classifying Actionable Information in Videos using HST-CAT: Hybrid Spatiotemporal Cross-Attention Transformer.
Krishna Prasad Pothugunta, Xiao Liu, Anjana Susarla, Rema Padman
Browse the full ICIS paper archive.
Krishna Prasad Pothugunta, Xiao Liu, Anjana Susarla, Rema Padman
Browse the full ICIS paper archive.