Last updated
Updated on December 4byParameters
The tasks are to describe the image and to come up with a large set of keyword tags for it.
Write the Description using the active voice.
The Keywords must be one or two words each. Generate as many Keywords as possible using a controlled and consistent vocabulary.
For both Description and Keywords, make sure to include:
- Themes, concepts
- Items, animals, objects
- Structures, landmarks, setting
- Foreground and background elements
- Notable colors, textures, styles
- Actions, activities
If humans are present, include:
- Physical appearance
- Gender
- Clothing
- Age range
- Visibly apparent ancestry
- Occupation/role
- Relationships between individuals
- Emotions, expressions, body language
Use ENGLISH only. Generate ONLY a JSON object with the keys Description and Keywords as follows {"Description": str, "Keywords": []}
<EXAMPLE>
The example input would be a stock photo of two apples, one red and one green, against a white backdrop and is a hypothetical Description and Keyword for a non-existent image.
OUTPUT=```json{"Description": "Two apples next to each other, one green and one red, placed side by side against a white background. There is even and diffuse studio lighting. The fruit is glossy and covered with dropplets of water indicating they are fresh and recently washed. The image emphasizes the cleanliness and appetizing nature of the food", "Keywords": ["studio shot","green","fruit","red","apple","stock image","health food","appetizing","empty background","grocery","food","snack"]}```
</EXAMPLE>