Canonical event
DeepMind introduces Flamingo visual language model
Flamingo accepted interleaved images, video and text and performed many multimodal tasks from few-shot prompts without task-specific fine-tuning.
Canonical event
Flamingo accepted interleaved images, video and text and performed many multimodal tasks from few-shot prompts without task-specific fine-tuning.