Zheda Mai
Computer Science · The Ohio State University
Publications
52
Citations
1,083
Est. group size
~3
Recurring co-author estimate
Active years
7
Publishing since 2020
Zheda Mai works in computer science on machine learning topics including how to adapt pretrained vision and language models to new tasks efficiently, continual and few-shot learning, and multimodal systems that combine vision, language, and time-series data. Recent work covers areas such as prompt-based model tuning, weakly-supervised image segmentation, embodied navigation agents, and biases in AI-driven recommendation systems. This research is largely methods-focused, aiming to make large pretrained models more adaptable, efficient, and reliable for downstream applications.
Publication output was minimal before 2020, rose to a steady 5-7 papers per year from 2020-2023, dipped in 2024, then increased sharply in 2025-2026, suggesting recent growth in activity.
Generated by claude-sonnet-5 from public bibliographic data · Jul 20, 2026
- BioCLIP 2
Zenodo (CERN European Organization for Nuclear Research) · 2026
- A Survey of Continual Learning for Robotics in the Foundation Model Era
2026
- A Survey of Continual Learning for Robotics in the Foundation Model Era
2026
- A Survey of Continual Learning for Robotics in the Foundation Model Era
2026
- Revisiting Model Stitching In the Foundation Model Era
arXiv (Cornell University) · 2026
- Revisiting Model Stitching In the Foundation Model Era
arXiv (Cornell University) · 2026
- A Study of Failure Modes in Two-Stage Human-Object Interaction Detection
arXiv (Cornell University) · 2026
- A Study of Failure Modes in Two-Stage Human-Object Interaction Detection
arXiv (Cornell University) · 2026
- TreeOfLife-200M
Hugging Face · 2026
- Prompt-CAM
Hugging Face · 2026
- Attention-Driven Causal Discovery: From Transformer Matrices to Granger Causal Graphs for Non-Stationary Time-series Data
2025
- Prompt-based Adaptation in Large-scale Vision Models: A Survey
arXiv (Cornell University) · 2025
- Efficiently Mitigating Video Content Misalignment on Large Vision Model with Time-Series Data Alignment
2025
- MLLM4TS: Leveraging Vision and Multimodal Language Models for General Time-Series Analysis
arXiv (Cornell University) · 2025
- IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation
arXiv (Cornell University) · 2025
- arXiv (Cornell University)×27
- Hugging Face×2
- Neurocomputing×1
- Journal of Visual Communication and Image Representation×1
- Information Processing & Management×1
- Wei‐Lun Chao
Computer Science · The Ohio State University
- Ruiqi Wang
Computer Science · Purdue University West Lafayette
- Gobinda Saha
Computer Science · Purdue University West Lafayette
- Ze Wang
Computer Science · Purdue University West Lafayette
- Cheng-Hao Tu
Computer Science · The Ohio State University
This profile was generated automatically from public scholarly data (OpenAlex). Group size and activity levels are estimates derived from co-authorship patterns.
Last updated Jul 19, 2026.
Claim or correct this profile