N° 03 · AI Group

5 members · 1 resources

Vision Language Models

Multimodal encoders, contrastive learning, VQA, and grounding language in images and video.

§ 01

Members

5
  • Edward FengAdmin
  • Janam Ajay Patel
  • Kushal Patel
  • Mahek Kantharia
  • Manusha Malgareddy
§ 02

Resources

1
§ 03

Current projects

1
  • Lab-VQA demo

    Ask questions about images from the mechatronics lab in natural language.

§ 04

Discord

Daily chat, debugging, polls, and quick announcements happen in #ai-vlms. The website is the source of truth for membership, resources, and schedule. Discord is the source of truth for what's happening today.