SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation | allainews.com

April 23, 2024, 4:47 a.m. | Yuying Ge, Sijie Zhao, Jinguo Zhu, Yixiao Ge, Kun Yi, Lin Song, Chen Li, Xiaohan Ding, Ying Shan

cs.CV updates on arXiv.org arxiv.org

arXiv:2404.14396v1 Announce Type: new
Abstract: The rapid evolution of multimodal foundation model has demonstrated significant progresses in vision-language understanding and generation, e.g., our previous work SEED-LLaMA. However, there remains a gap between its capability and the real-world applicability, primarily due to the model's limited capacity to effectively respond to various user instructions and interact with diverse visual data. In this work, we focus on bridging this gap through integrating two enhanced features: (1) comprehending images of arbitrary sizes and ratios, …

arxiv cs.cv multimodal multimodal models seed type

More from arxiv.org / cs.CV updates on arXiv.org

PCLMix: Weakly Supervised Medical Image Segmentation via Pixel-Level Contrastive Learning and Dynamic Mix Augmentation an hour ago | arxiv.org

arxiv augmentation cs.cv dynamic +7

Retrieval-Augmented Egocentric Video Captioning an hour ago | arxiv.org

abstract arxiv benefit captioning +20

Geo-Localization Based on Dynamically Weighted Factor-Graph an hour ago | arxiv.org

abstract aerial arxiv cs.cv +12

Mesh Neural Cellular Automata an hour ago | arxiv.org

arxiv cellular cs.ai cs.cv +4

Mirror-Aware Neural Humans an hour ago | arxiv.org

abstract affordable alternative arxiv +14

MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing an hour ago | arxiv.org

arxiv cs.ai cs.cl cs.cv +5

A Foundation Model for Brain Lesion Segmentation with Mixture of Modality Experts an hour ago | arxiv.org

abstract arxiv brain complexity +13

MrRegNet: Multi-resolution Mask Guided Convolutional Neural Network for Medical Image Registration with Large Deformations an hour ago | arxiv.org

arxiv convolutional convolutional neural network cs.cv +8

Histopathology Foundation Models Enable Accurate Ovarian Cancer Subtype Classification an hour ago | arxiv.org

abstract artificial artificial intelligence arxiv +13

Software Engineer for AI Training Data (School Specific)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Python)

@ G2i Inc | Remote

View on ai-jobs.net

Software Engineer for AI Training Data (Tier 2)

@ G2i Inc | Remote

View on ai-jobs.net

Data Engineer

@ Lemon.io | Remote: Europe, LATAM, Canada, UK, Asia, Oceania

View on ai-jobs.net

Artificial Intelligence – Bioinformatic Expert

@ University of Texas Medical Branch | Galveston, TX

View on ai-jobs.net

Lead Developer (AI)

@ Cere Network | San Francisco, US

View on ai-jobs.net