{"id":27886205,"url":"https://github.com/mostafax/image-caption","last_synced_at":"2025-10-14T20:32:22.757Z","repository":{"id":46095307,"uuid":"151829612","full_name":"mostafax/Image-Caption","owner":"mostafax","description":"End to End Deep learning model that generate image captions ","archived":false,"fork":false,"pushed_at":"2018-12-25T16:50:03.000Z","size":13145,"stargazers_count":8,"open_issues_count":0,"forks_count":4,"subscribers_count":2,"default_branch":"master","last_synced_at":"2025-05-05T07:55:19.455Z","etag":null,"topics":["cnn","deep-learning","image-caption","image-captioning","image-classifier","keras","lstm","nuralnetwork","python3","rnn","text-from-image"],"latest_commit_sha":null,"homepage":"","language":"Python","has_issues":true,"has_wiki":null,"has_pages":null,"mirror_url":null,"source_name":null,"license":null,"status":null,"scm":"git","pull_requests_enabled":true,"icon_url":"https://github.com/mostafax.png","metadata":{"files":{"readme":"README.md","changelog":null,"contributing":null,"funding":null,"license":null,"code_of_conduct":null,"threat_model":null,"audit":null,"citation":null,"codeowners":null,"security":null,"support":null}},"created_at":"2018-10-06T10:37:54.000Z","updated_at":"2024-03-03T14:51:41.000Z","dependencies_parsed_at":"2022-08-29T16:40:12.638Z","dependency_job_id":null,"html_url":"https://github.com/mostafax/Image-Caption","commit_stats":null,"previous_names":[],"tags_count":0,"template":false,"template_full_name":null,"purl":"pkg:github/mostafax/Image-Caption","repository_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/mostafax%2FImage-Caption","tags_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/mostafax%2FImage-Caption/tags","releases_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/mostafax%2FImage-Caption/releases","manifests_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/mostafax%2FImage-Caption/manifests","owner_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners/mostafax","download_url":"https://codeload.github.com/mostafax/Image-Caption/tar.gz/refs/heads/master","sbom_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories/mostafax%2FImage-Caption/sbom","scorecard":null,"host":{"name":"GitHub","url":"https://github.com","kind":"github","repositories_count":279020905,"owners_count":26086948,"icon_url":"https://github.com/github.png","version":null,"created_at":"2022-05-30T11:31:42.601Z","updated_at":"2022-07-04T15:15:14.044Z","status":"online","status_checked_at":"2025-10-14T02:00:06.444Z","response_time":60,"last_error":null,"robots_txt_status":"success","robots_txt_updated_at":"2025-07-24T06:49:26.215Z","robots_txt_url":"https://github.com/robots.txt","online":true,"can_crawl_api":true,"host_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub","repositories_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repositories","repository_names_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/repository_names","owners_url":"https://repos.ecosyste.ms/api/v1/hosts/GitHub/owners"}},"keywords":["cnn","deep-learning","image-caption","image-captioning","image-classifier","keras","lstm","nuralnetwork","python3","rnn","text-from-image"],"created_at":"2025-05-05T07:52:54.007Z","updated_at":"2025-10-14T20:32:22.737Z","avatar_url":"https://github.com/mostafax.png","language":"Python","funding_links":[],"categories":[],"sub_categories":[],"readme":"# Image-caption Using End2end CNN , LSTM based Model!\n\u003ch3\u003eThe aim of the project is to generate a caption for images.\u003cbr\u003e\n Each image has a story, Image Captioning narrates it.\n\u003c/h3\u003e\n\u003cimg src = \"/PretrainedModel/Out.png\"\u003e\n\n \n this model is bases on [Show and Tell: A Neural Image Caption Generator\n](https://arxiv.org/pdf/1411.4555.pdf)\n\n📖 Documentation\n================\n## How to Run\n**Install the requirements:**\n```bash\npip3 install -r requirements.txt \n```\n**Running the Model**\n```bash\npython3 model.py\n```\n\n## Results\n\nThe results are not bad at all! a lot of test cases gonna be so realistic, but the model still needs more training\n\u003cimg src = \"/PretrainedModel/r1.png\"\u003e\n\u003cimg src = \"/PretrainedModel/r2.png\"\u003e\n\n## Paper\nThis project is an implementation of the [Show and Tell](https://arxiv.org/pdf/1411.4555.pdf), published 2015.\n\n## Dataset\n- Dataset used is Flicker8k each image have 5 captions.\n- you can request data from here [Flicker8k]\n(https://forms.illinois.edu/sec/1713398).\u003cbr\u003e\n**Sample of the data used**\n\u003cimg src = \"/PretrainedModel/dayaset.png\"\u003e\n\n## Model Used\n\u003cimg src = \"/PretrainedModel/model.png\"\u003e\n\n## Experiments\n\n\u003cimg src = \"/PretrainedModel/expermant.png\"\u003e\n\n## Future Work\n-Training, Training and more Training\u003cbr\u003e\n-Using Resnet instead of VGG16\u003cbr\u003e\n-Creating API for production level \u003cbr\u003e\n-Using Word2Vec embedding.\n\n","project_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fmostafax%2Fimage-caption","html_url":"https://awesome.ecosyste.ms/projects/github.com%2Fmostafax%2Fimage-caption","lists_url":"https://awesome.ecosyste.ms/api/v1/projects/github.com%2Fmostafax%2Fimage-caption/lists"}