CV

Research Focus

Multimodal generative AI, controllable visual generation, diffusion models, multimodal reasoning, LLM/MLLM post-training, and agentic AI systems.

Education

Industry Research Experience

Additional Research Experience

Technical Skills

Publications

ViSTA: Visual Storytelling using Multi-modal Adapters for Text-to-Image Diffusion Models

2026

Sibo Dong, Ismail Shaheen, Maggie Shen, Rupayan Mallick, and Sarah Adel Bargal

IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) · Oral

Sidecar: Training-Free Semantic Reuse for Character-Consistent Free-form Visual Storytelling

2026

Sibo Dong and Sarah Adel Bargal

Under review

FreeStory: Training-Free Character Consistency for Free-Form Visual Storytelling

2026

Sibo Dong, Ismail Shaheen, and Sarah Adel Bargal

Under review

D-Feat Occlusions: Diffusion Features for Robustness to Partial Visual Occlusions in Object Recognition

2025

Rupayan Mallick, Sibo Dong, Nataniel Ruiz, and Sarah Adel Bargal

CVPR Workshop on Uncertainty Quantification for Computer Vision

Predicting Missing Response with BERT Model in Process Data

2024

Qiwei He and Sibo Dong

International Meeting of the Psychometric Society (IMPS)

SEINE: SEgment-based Indexing for NEural Information Retrieval

2022

Sibo Dong, Justin Goldstein, and Grace Hui Yang

SIGIR Workshop on Reaching Efficiency in Neural Information Retrieval (ReNeuIR)

Do We Really Need Everything Everywhere All at Once? Query-Specific Fine-Tuning for Transformer-Based Neural Retrievers

2022

Sibo Dong and Grace Hui Yang

Text REtrieval Conference (TREC)

GazBy: Gaze-Based BERT Model to Incorporate Human Attention in Neural Information Retrieval

2022

Sibo Dong, Justin Goldstein, and Grace Hui Yang

ACM SIGIR International Conference on Theory of Information Retrieval (ICTIR)

Professional Service & Honors