GRIT: Faster and Better Image Captioning Transformer Using Dual Visual FeaturesOct 1, 2022·Quang Nguyen,Masanori Suganuma,Takayuki OkataniTypeConference paperPublicationEuropean Conference on Computer Vision (ECCV) 2022Last updated on Oct 1, 2022 AuthorsQuang Nguyen ← Exploring the Potential of Multi-Modal AI for Driving Hazard Prediction Apr 1, 2024Look Wide and Interpret Twice: Improving Performance on Interactive Instruction-following Tasks Aug 1, 2021 →