論文まとめ:Generalized Decoding for Pixel, Image, and Language
847{icon} {views} タイトル:Generalized Decoding for Pixel, Image, and Language 著者:Xueyan Zou, Zi-Yi Dou, Jianwei Y […]...
論文まとめ:BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
9.7k{icon} {views} タイトル:BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large […]...
論文まとめ:InstructPix2Pix: Learning to Follow Image Editing Instructions
1.3k{icon} {views} タイトル:InstructPix2Pix: Learning to Follow Image Editing Instructions 著者:Tim Brooks, Aleksand […]...
論文まとめ:StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesis
1k{icon} {views} タイトル:StyleTTS: A Style-Based Generative Model for Natural and Diverse Text-to-Speech Synthesi […]...
論文まとめ:OCR-free Document Understanding Transformer
3.5k{icon} {views} タイトル:OCR-free Document Understanding Transformer 著者:Geewook Kim, Teakgyu Hong, Moonbin Yim, […]...
論文まとめ:Lightweight Attentional Feature Fusion: A New Baseline for Text-to-Video Retrieval
708{icon} {views} タイトル:Lightweight Attentional Feature Fusion: A New Baseline for Text-to-Video Retrieval 著者:F […]...
論文まとめ:Large Language Models are Zero-Shot Reasoners
7.2k{icon} {views} タイトル:Large Language Models are Zero-Shot Reasoners 著者:Takeshi Kojima, Shixiang Shane Gu, Ma […]...
論文まとめ:Extremely Simple Activation Shaping for Out-of-Distribution Detection
1.2k{icon} {views} タイトル:Extremely Simple Activation Shaping for Out-of-Distribution Detection 著者:Andrija Djuri […]...
論文まとめ:Domino: Discovering Systematic Errors with Cross-Modal Embeddings
366{icon} {views} タイトル:Domino: Discovering Systematic Errors with Cross-Modal Embeddings 著者:Sabri Eyuboglu, Ma […]...
論文まとめ:Exploring Visual Prompts for Adapting Large-Scale Models
2.1k{icon} {views} タイトル:Exploring Visual Prompts for Adapting Large-Scale Models 著者:Hyojin Bahng, Ali Jahanian […]...