Scholar
Idan Schwartz
Google Scholar ID: 5V-yJT4AAAAJ
The Department of Computer Science at Bar-Ilan University
Multimodal
Attention
Computer Vision
Natural Language Processing
Deep Learning
Follow
Homepage
↗
Google Scholar
↗
Citations & Impact
All-time
Citations
1,067
H-index
14
i10-index
16
Publications
20
Co-authors
19
list available
Contact
CV
Open ↗
Twitter
Open ↗
GitHub
Open ↗
Publications
5 items
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation
2026
Cited
0
TempoControl: Temporal Attention Guidance for Text-to-Video Models
2025
Cited
0
Single Image Iterative Subject-driven Generation and Editing
2025
Cited
0
Iterative Object Count Optimization for Text-to-image Diffusion Models
arXiv.org · 2024
Cited
4
Discriminative Class Tokens for Text-to-Image Diffusion Models
IEEE International Conference on Computer Vision · 2023
Cited
5
Co-authors
13 total
Lior Wolf
The School of Computer Science at Tel Aviv University
Alexander Schwing
University of Illinois at Urbana-Champaign
Itai Gat
Meta AI, FAIR
Guy Yariv
The Hebrew University
Yossi Adi
The Hebrew University of Jerusalem
Sagie Benaim
Assistant Professor, Hebrew University of Jerusalem
Hila Chefer
PhD student at Tel Aviv University, Meta
Ariel Shamir
Professor of Computer Science, Reichman University (IDC Herzliya)