Preprint2023
BLIP 2
Unknown
BLIP-2 introduces Q-Former, a lightweight transformer module that bridges frozen image encoders and frozen LLMs for efficient vision-language pre-training.
0Jan 1, 2023Computer Vision
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.