Name: General Knowledge Text-Image Pair Corpus - DataoceanAI
SKU: King-IM-104
Availability: InStock

General Knowledge Text-Image Pair Corpus

Product Features: This corpus includes data from 23 categories such as cuisine, landscapes, architecture, cities, countryside, health, sports, medical, automobiles, backgrounds, finance, education, oil paintings, illustrations, watercolors, travel, fashion, romance, animals, plants, space, and technology.

Specifications:

ID:

King-IM-104

Size:

2,000,000 Sets

Image Specification

Text Specification

Includes labels, descriptions in both Chinese and English

People also searched for

Aesthetic Composition Training Corpus

Images are captured by professional photographers. Composition types include rule-of-thirds, horizontal, diagonal, triangular, and central composition. All images are evaluated and annotated by personnel with high aesthetic standards. Each image meets at least one composition type and at most three composition types.

Handheld Object Portrait Corpus

Data collection covers both indoor and outdoor environments, including offices, meeting rooms, parking lots, gardens, and other common work and daily-life scenarios. Lighting conditions include normal lighting, low-light, and backlit scenarios commonly encountered in real-world settings. Twenty video clips per participant, with each clip corresponding to the presentation of a single object and accompanied by three close-up images of the object from different angles. The subject appears in upper-body or full-body views, holding one object with one or both hands, recorded in standing or seated postures. The object is moved according to predefined actions. Each video clip includes a brief verbal description of the object provided by the model. The subject is clearly visible, and the face is not occluded for extended periods during recording.

English-Lao Parallel Corpus

Corpus Field: Most are inclined to fields such as news, transportation and tourism, daily life, sports and health, finance, and technology.

parallel corpus Multimodal Gujarati-English

Chinese-Lao Parallel Corpus