{
  "id": 223419,
  "title": "💥 💥   Image Captioning  - DL Research Papers + Code🔥🔥",
  "url": "/competitions/bms-molecular-translation/discussion/223419",
  "author_name": "",
  "post_date": "2021-03-03T17:48:55.028387200Z",
  "votes": 46,
  "comment_count": 6,
  "views": 0,
  "content": "<p>This github repository contains comprehensive collection of deep learning research papers from all premier conferences</p>\n<p><img src=\"https://drive.google.com/uc?id=1MFMVlR_0Oyt1DHWnpz9tFEH9RQPpOmLg\" alt=\"\"></p>\n<p><a href=\"https://github.com/zhjohnchan/awesome-image-captioning\" target=\"_blank\">https://github.com/zhjohnchan/awesome-image-captioning</a></p>\n<p>Some snippets from the repository below</p>\n<p>2020<br>\n<strong>AAAI 2020</strong><br>\n<strong>MemCap: Memorizing Style Knowledge for Image Captioning</strong> - Zhao et al, AAAI 2020.</p>\n<p><strong>Unified Vision-Language Pre-Training for Image Captioning and VQA</strong> - Zhou L et al, AAAI 2020.</p>\n<p><strong>Show, Recall, and Tell: Image Captioning with Recall Mechanism</strong> - Wang L et al, AAAI 2020.</p>\n<p><strong>Reinforcing an Image Caption Generator using Off-line Human Feedback</strong> - Hongsuck Seo P et al, AAAI 2020.</p>\n<p><strong>Interactive Dual Generative Adversarial Networks for Image Captioning</strong> - Liu et al, AAAI 2020.</p>\n<p><strong>Feature Deformation Meta-Networks in Image Captioning of Novel Objects</strong> - Cao et al, AAAI 2020.</p>\n<p><strong>Joint Commonsense and Relation Reasoning for Image and Video Captioning - Hou et al, AAAI 2020.</strong></p>\n<p><strong>Learning Long- and Short-Term User Literal-Preference with Multimodal Hierarchical Transformer Network for Personalized Image Caption</strong> - Zhang et al, AAAI 2020.</p>\n<p><strong>CVPR 2020</strong></p>\n<p>Normalized and Geometry-Aware Self-Attention Network for Image Captioning - Guo L et al, CVPR 2020.</p>\n<p>Object Relational Graph with Teacher-Recommended Learning for Video Captioning - Zhang Z et al, CVPR 2020.</p>\n<p>Say As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs - Chen S et al, CVPR 2020.</p>\n<p><strong>ACL 2020</strong></p>\n<p>Improving Image Captioning with Better Use of Caption - Shi Z et al, ACL 2020.</p>\n<p>Cross-modal Coherence Modeling for Caption Generation - Alikhani M et al, ACL 2020.</p>\n<p>Improving Image Captioning Evaluation by Considering Inter References Variance - Yi Y et al, ACL 2020.</p>\n<p>MART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning - Lei J et al, ACL 2020.</p>\n<p>Dense-Caption Matching and Frame-Selection Gating for Temporal Localization in VideoQA - Kim H et al, ACL 2020.</p>\n<p>Hope you found the post useful . If you come across useful resources like this , please share in the comment below and I will update the post .</p>\n<p>Good Luck with the competition</p>",
  "messages": [
    {
      "id": "1225562",
      "postDate": "03/03/2021 17:48:55",
      "content": "<p>This github repository contains comprehensive collection of deep learning research papers from all premier conferences</p>\n<p><img src=\"https://drive.google.com/uc?id=1MFMVlR_0Oyt1DHWnpz9tFEH9RQPpOmLg\" alt=\"\"></p>\n<p><a href=\"https://github.com/zhjohnchan/awesome-image-captioning\" target=\"_blank\">https://github.com/zhjohnchan/awesome-image-captioning</a></p>\n<p>Some snippets from the repository below</p>\n<p>2020<br>\n<strong>AAAI 2020</strong><br>\n<strong>MemCap: Memorizing Style Knowledge for Image Captioning</strong> - Zhao et al, AAAI 2020.</p>\n<p><strong>Unified Vision-Language Pre-Training for Image Captioning and VQA</strong> - Zhou L et al, AAAI 2020.</p>\n<p><strong>Show, Recall, and Tell: Image Captioning with Recall Mechanism</strong> - Wang L et al, AAAI 2020.</p>\n<p><strong>Reinforcing an Image Caption Generator using Off-line Human Feedback</strong> - Hongsuck Seo P et al, AAAI 2020.</p>\n<p><strong>Interactive Dual Generative Adversarial Networks for Image Captioning</strong> - Liu et al, AAAI 2020.</p>\n<p><strong>Feature Deformation Meta-Networks in Image Captioning of Novel Objects</strong> - Cao et al, AAAI 2020.</p>\n<p><strong>Joint Commonsense and Relation Reasoning for Image and Video Captioning - Hou et al, AAAI 2020.</strong></p>\n<p><strong>Learning Long- and Short-Term User Literal-Preference with Multimodal Hierarchical Transformer Network for Personalized Image Caption</strong> - Zhang et al, AAAI 2020.</p>\n<p><strong>CVPR 2020</strong></p>\n<p>Normalized and Geometry-Aware Self-Attention Network for Image Captioning - Guo L et al, CVPR 2020.</p>\n<p>Object Relational Graph with Teacher-Recommended Learning for Video Captioning - Zhang Z et al, CVPR 2020.</p>\n<p>Say As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs - Chen S et al, CVPR 2020.</p>\n<p><strong>ACL 2020</strong></p>\n<p>Improving Image Captioning with Better Use of Caption - Shi Z et al, ACL 2020.</p>\n<p>Cross-modal Coherence Modeling for Caption Generation - Alikhani M et al, ACL 2020.</p>\n<p>Improving Image Captioning Evaluation by Considering Inter References Variance - Yi Y et al, ACL 2020.</p>\n<p>MART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning - Lei J et al, ACL 2020.</p>\n<p>Dense-Caption Matching and Frame-Selection Gating for Temporal Localization in VideoQA - Kim H et al, ACL 2020.</p>\n<p>Hope you found the post useful . If you come across useful resources like this , please share in the comment below and I will update the post .</p>\n<p>Good Luck with the competition</p>",
      "rawMarkdown": "This github repository contains comprehensive collection of deep learning research papers from all premier conferences\n\n![](https://drive.google.com/uc?id=1MFMVlR_0Oyt1DHWnpz9tFEH9RQPpOmLg)\n\nhttps://github.com/zhjohnchan/awesome-image-captioning\n\nSome snippets from the repository below\n\n2020\n**AAAI 2020**\n**MemCap: Memorizing Style Knowledge for Image Captioning** - Zhao et al, AAAI 2020.\n\n**Unified Vision-Language Pre-Training for Image Captioning and VQA** - Zhou L et al, AAAI 2020.\n\n**Show, Recall, and Tell: Image Captioning with Recall Mechanism** - Wang L et al, AAAI 2020.\n\n**Reinforcing an Image Caption Generator using Off-line Human Feedback** - Hongsuck Seo P et al, AAAI 2020.\n\n**Interactive Dual Generative Adversarial Networks for Image Captioning** - Liu et al, AAAI 2020.\n\n**Feature Deformation Meta-Networks in Image Captioning of Novel Objects** - Cao et al, AAAI 2020.\n\n**Joint Commonsense and Relation Reasoning for Image and Video Captioning - Hou et al, AAAI 2020.**\n\n**Learning Long- and Short-Term User Literal-Preference with Multimodal Hierarchical Transformer Network for Personalized Image Caption** - Zhang et al, AAAI 2020.\n\n**CVPR 2020**\n\nNormalized and Geometry-Aware Self-Attention Network for Image Captioning - Guo L et al, CVPR 2020.\n\nObject Relational Graph with Teacher-Recommended Learning for Video Captioning - Zhang Z et al, CVPR 2020.\n\nSay As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs - Chen S et al, CVPR 2020.\n\n**ACL 2020**\n\nImproving Image Captioning with Better Use of Caption - Shi Z et al, ACL 2020.\n\nCross-modal Coherence Modeling for Caption Generation - Alikhani M et al, ACL 2020.\n\nImproving Image Captioning Evaluation by Considering Inter References Variance - Yi Y et al, ACL 2020.\n\nMART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning - Lei J et al, ACL 2020.\n\nDense-Caption Matching and Frame-Selection Gating for Temporal Localization in VideoQA - Kim H et al, ACL 2020.\n\nHope you found the post useful . If you come across useful resources like this , please share in the comment below and I will update the post .\n\nGood Luck with the competition",
      "votes": null
    },
    {
      "id": "1225645",
      "postDate": "03/03/2021 19:20:21",
      "content": "<p>Thanks for share more resources. I have a conceptual question, is this competition about \"image captioning\"?</p>",
      "rawMarkdown": "Thanks for share more resources. I have a conceptual question, is this competition about \"image captioning\"?",
      "votes": null
    },
    {
      "id": "1225834",
      "postDate": "03/04/2021 01:50:55",
      "content": "<p>Thanks for the awesome resources!</p>",
      "rawMarkdown": "Thanks for the awesome resources!",
      "votes": null
    },
    {
      "id": "1225835",
      "postDate": "03/04/2021 01:53:11",
      "content": "<p>I guess it can be thought of as \"image captioning\" -&gt; you feed in the formula-containing image and get the string (InChI) \"describing\" the image</p>",
      "rawMarkdown": "I guess it can be thought of as \"image captioning\" -> you feed in the formula-containing image and get the string (InChI) \"describing\" the image",
      "votes": null
    },
    {
      "id": "1226918",
      "postDate": "03/05/2021 02:46:27",
      "content": "<p>This post has all of the latest research! Thanks alot!</p>",
      "rawMarkdown": "This post has all of the latest research! Thanks alot!",
      "votes": null
    },
    {
      "id": "1230820",
      "postDate": "03/08/2021 13:36:57",
      "content": "<p>Nice Post……</p>",
      "rawMarkdown": "Nice Post......",
      "votes": null
    },
    {
      "id": "2300852",
      "postDate": "06/13/2023 13:43:15",
      "content": "<p>Thanks for sharing. Great learning ahead!</p>",
      "rawMarkdown": "Thanks for sharing. Great learning ahead!",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1225645,
      "author_name": "hiramcho",
      "author_url": "",
      "post_date": "03/03/2021 19:20:21",
      "content": "<p>Thanks for share more resources. I have a conceptual question, is this competition about \"image captioning\"?</p>",
      "votes": null,
      "replies": [
        {
          "id": 1225835,
          "author_name": "arka47",
          "author_url": "",
          "post_date": "03/04/2021 01:53:11",
          "content": "<p>I guess it can be thought of as \"image captioning\" -&gt; you feed in the formula-containing image and get the string (InChI) \"describing\" the image</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1225834,
      "author_name": "arka47",
      "author_url": "",
      "post_date": "03/04/2021 01:50:55",
      "content": "<p>Thanks for the awesome resources!</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 1226918,
      "author_name": "ligtfeather",
      "author_url": "",
      "post_date": "03/05/2021 02:46:27",
      "content": "<p>This post has all of the latest research! Thanks alot!</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 1230820,
      "author_name": "amolambkar",
      "author_url": "",
      "post_date": "03/08/2021 13:36:57",
      "content": "<p>Nice Post……</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 2300852,
      "author_name": "writabrata",
      "author_url": "",
      "post_date": "06/13/2023 13:43:15",
      "content": "<p>Thanks for sharing. Great learning ahead!</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "1225562": "This github repository contains comprehensive collection of deep learning research papers from all premier conferences\n\n![](https://drive.google.com/uc?id=1MFMVlR_0Oyt1DHWnpz9tFEH9RQPpOmLg)\n\nhttps://github.com/zhjohnchan/awesome-image-captioning\n\nSome snippets from the repository below\n\n2020\n**AAAI 2020**\n**MemCap: Memorizing Style Knowledge for Image Captioning** - Zhao et al, AAAI 2020.\n\n**Unified Vision-Language Pre-Training for Image Captioning and VQA** - Zhou L et al, AAAI 2020.\n\n**Show, Recall, and Tell: Image Captioning with Recall Mechanism** - Wang L et al, AAAI 2020.\n\n**Reinforcing an Image Caption Generator using Off-line Human Feedback** - Hongsuck Seo P et al, AAAI 2020.\n\n**Interactive Dual Generative Adversarial Networks for Image Captioning** - Liu et al, AAAI 2020.\n\n**Feature Deformation Meta-Networks in Image Captioning of Novel Objects** - Cao et al, AAAI 2020.\n\n**Joint Commonsense and Relation Reasoning for Image and Video Captioning - Hou et al, AAAI 2020.**\n\n**Learning Long- and Short-Term User Literal-Preference with Multimodal Hierarchical Transformer Network for Personalized Image Caption** - Zhang et al, AAAI 2020.\n\n**CVPR 2020**\n\nNormalized and Geometry-Aware Self-Attention Network for Image Captioning - Guo L et al, CVPR 2020.\n\nObject Relational Graph with Teacher-Recommended Learning for Video Captioning - Zhang Z et al, CVPR 2020.\n\nSay As You Wish: Fine-grained Control of Image Caption Generation with Abstract Scene Graphs - Chen S et al, CVPR 2020.\n\n**ACL 2020**\n\nImproving Image Captioning with Better Use of Caption - Shi Z et al, ACL 2020.\n\nCross-modal Coherence Modeling for Caption Generation - Alikhani M et al, ACL 2020.\n\nImproving Image Captioning Evaluation by Considering Inter References Variance - Yi Y et al, ACL 2020.\n\nMART: Memory-Augmented Recurrent Transformer for Coherent Video Paragraph Captioning - Lei J et al, ACL 2020.\n\nDense-Caption Matching and Frame-Selection Gating for Temporal Localization in VideoQA - Kim H et al, ACL 2020.\n\nHope you found the post useful . If you come across useful resources like this , please share in the comment below and I will update the post .\n\nGood Luck with the competition",
    "1225645": "Thanks for share more resources. I have a conceptual question, is this competition about \"image captioning\"?",
    "1225834": "Thanks for the awesome resources!",
    "1225835": "I guess it can be thought of as \"image captioning\" -> you feed in the formula-containing image and get the string (InChI) \"describing\" the image",
    "1226918": "This post has all of the latest research! Thanks alot!",
    "1230820": "Nice Post......",
    "2300852": "Thanks for sharing. Great learning ahead!"
  },
  "source": "meta"
}