| 2026 | ACL | Hemolix.TabGen: Optimized Table Generation from Documents. | Gyanendra Shrestha, Todor Ivanov, Karthik Vemireddy, Anna Pyayt, Michael N. Gubanov |
| 2026 | SIGMOD | Hemolix.Extract.V: LLM-based Information Extraction for Documents with AI-based Plan Selection. | Todor Ivanov, Gyanendra Shrestha, Karthik Vemireddy, Anna Pyayt, Michael N. Gubanov |
| 2025 | EDBT | Tabular Embeddings for Tables with Bi-Dimensional Hierarchical Metadata and Nesting. | Gyanendra Shrestha, Chutian Jiang, Sai Akula, Vivek Yannam, Anna Pyayt, Michael N. Gubanov |
| 2025 | ICDE | Scalable Tabular Hierarchical Metadata Classification in Heterogeneous Structured Large-Scale Datasets Using Contrastive Learning. | Bhimesh Kandibedala, Gyanendra Shrestha, Anna Pyayt, Todor Ivanov, Michael N. Gubanov |
| 2024 | CIKM | CancerKG.ORG - A Web-scale, Interactive, Verifiable Knowledge Graph-LLM Hybrid for Assisting with Optimal Cancer Treatment and Care. | Michael N. Gubanov, Anna Pyayt, Aleksandra Karolak |
| 2023 | DOLAP | Learning Circular Tabular Embeddings for Heterogeneous Large-scale Structured Datasets. | Michael N. Gubanov, Anna Pyayt, Sophie Pavia |
| 2023 | DOLAP | Scalable Hierarchical Metadata Classification in Heterogeneous Large-scale Datasets. | Bhimesh Kandibedala, Anna Pyayt, Chris Caballero, Michael N. Gubanov |
| 2023 | EDBT | COVIDKG.ORG - a Web-scale COVID-19 Interactive, Trustworthy Knowledge Graph, Constructed and Interrogated for Bias using Deep-Learning. | Bhimesh Kandibedala, Anna Pyayt, Nickolas Piraino, Chris Caballero, Michael N. Gubanov |
| 2023 | WWW | Learning Topical Structured Interfaces from Medical Research Literature. | Maitry Chauhan, Anna Pyayt, Michael N. Gubanov |
| 2022 | CIKM | Leveraging Scalable Profiling to Learn and Visualize the Latest Trustworthy COVID-19 Medical Research Findings. | Michael N. Gubanov, Sophie Pavia, Anna Pyayt, William Goble |
| 2022 | SIGMOD | Simplifying Access to Large-scale Structured Datasets by Meta-Profiling with Scalable Training Set Enrichment. | Sophie Pavia, Rituparna Khan, Anna Pyayt, Michael N. Gubanov |
| 2021 | DEXA | Scalable Tabular Metadata Location and Classification in Large-Scale Structured Datasets. | Kazi Islam, Michael N. Gubanov |
| 2020 | CIKM | WebLens: Towards Interactive Large-scale Structured Data Profiling. | Rituparna Khan, Michael N. Gubanov |
| 2019 | ICDE | Hybrid.Poly: A Consolidated Interactive Analytical Polystore System. | Maksim Podkorytov, Michael N. Gubanov |
| 2018 | WWW | Hybrid.AI: A Learning Search Engine for Large-scale Structured Data. | Sean Soderman, Anusha Kola, Maksim Podkorytov, Michael Geyer, Michael N. Gubanov |
| 2017 | CIDR | Hybrid: A Large-scale In-memory Image Analytics Engine. | Michael N. Gubanov |
| 2017 | ICDE | PolyFuse: A Large-Scale Hybrid Data Fusion System. | Michael N. Gubanov |
| 2017 | ICDE | Scalable Linear Algebra on a Relational Database System. | Shangyu Luo, Zekai J. Gao, Michael N. Gubanov, Luis Leopoldo Perez, Christopher M. Jermaine |
| 2017 | ICDM | Hybrid.poly: An Interactive Large-Scale In-memory Analytical Polystore. | Maksim Podkorytov, Dylan Soderman, Michael N. Gubanov |
| 2017 | WWW | CognitiveDB: An Intelligent Navigator for Large-scale Dark Structured Data. | Michael N. Gubanov, Manju Priya, Maksim Podkorytov |
| 2016 | EDBT | Type-aware Web-search. | Michael N. Gubanov, Anna Pyayt |
| 2015 | CIDR | Dataxformer: Leveraging the Web for Semantic Transformations. | Ziawasch Abedjan, John Morcos, Michael N. Gubanov, Ihab F. Ilyas, Michael Stonebraker, Paolo Papotti, Mourad Ouzzani |
| 2014 | EDBT | Large-scale Semantic Profile Extraction. | Michael N. Gubanov, Michael Stonebraker |
| 2014 | ICDE | Text and structured data fusion in data tamer at scale. | Michael N. Gubanov, Michael Stonebraker, Daniel Bruckner |
| 2013 | CIKM | READFAST: high-relevance search-engine for big text. | Michael N. Gubanov, Anna Pyayt |
| 2013 | IRI | A real-time classification algorithm for emotion detection using portable EEG. | Surya Cheemalapati, Michael N. Gubanov, Michael Del Vale, Anna Pyayt |
| 2013 | IRI | ReadFast: Optimizing structural search relevance for big biomedical text. | Michael N. Gubanov, Anna Pyayt |
| 2012 | IRI | MEDREADFAST: A structural information retrieval engine for big clinical text. | Michael N. Gubanov, Anna Pyayt |
| 2011 | IRI | READFAST: Browsing large documents through unified famous objects (UFO). | Michael N. Gubanov, Anna Pyayt, Linda G. Shapiro |
| 2011 | IRI | Learning Unified Famous Objects (UFO) to bootstrap information integration. | Michael N. Gubanov, Linda G. Shapiro, Anna Pyayt |
| 2008 | ICDE | Model Management Engine for Data Integration with Reverse-Engineering Support. | Michael N. Gubanov, Philip A. Bernstein, Alexander Moshchuk |