Load image sets using subreddit name(s) above or paste your image from clipboard
here (Ctrl+V).
MANUAL MEDIA INGESTION
Import individual images or batches directly into your active annotation session for dataset labeling and
multimodal AI prompt analysis.
DRAG & DROP IMAGES HERE
or click to browse from your device
PNGJPGWEBPAVIFMULTI-FILE
ⓘAccepts direct image links from Reddit, Imgur, Unsplash, or any public HTTPS server.
0STAGED FOR INGESTION
EXPORT DATASET ARCHIVE
Your complete dataset will be compiled into a standardized ZIP archive ready for training with images, YOLO
labels, COCO annotations, captions, and manifest. Ignored images are excluded.
[01] YOLOv8 / YOLOv11
labels/*.txt with normalized coords, data.yaml, and classes.txt.
[02] COCO JSON
Industry standard coco_annotations.json with bboxes & categories.
[03] STABLE DIFFUSION
Individual *.txt prompt caption files paired with each image.
[04] VLM / HUGGINGFACE
dataset.jsonl structured for multimodal vision-language fine-tuning.
SUBREDDIT(S): N/A
TOTAL IMAGES IN DATASET: 0
CLASSES: subject, foreground
⚙ AI VISION SETTINGS
Configure multimodal AI vision engines to generate descriptive captions
covering poses, lighting, subjects, background, and cultural/regional characteristics.
Hugging Face free serverless inference API. Tokens are stored securely in browser
localStorage only.
⚡ LIVE MODEL TEST BENCHINSTANT VERIFY
Send a diagnostic test request to verify API key authentication, model availability, and
response latency.
READY
MULTI-MODEL AI VISION COMPARISON & SELECTION
SELECT MODELS TO RUN CONCURRENTLY:
Select which models to compare above and click [RUN ALL SELECTED MODELS →] to generate side-by-side
prompt responses.
REDDIZ MANUAL & API KEYS // < nissshh / >
OPERATOR HANDBOOK // WORKFLOW INSTRUCTIONS
REDDIZ is a high-throughput dataset authoring workstation built for visual dataset generation, portrait
training, and fine-tuning. It harvests Reddit image boards, formats training captions, bounds objects, and
packages clean datasets.
01 / FETCH MEDIA
Type target subreddits in the top bar (e.g. streetphotography, cats, wallpapers) and press
Enter. All media streams into the vertical filmstrip ready for annotation.
02 / MULTIMODAL AI VISION
Use the inline model picker in the caption header or click ✦ PROMPT to analyze the
image with Google Gemini, OpenAI ChatGPT, Anthropic Claude, or Groq Vision. The engine automatically
outputs objective subject poses, facial features, Indian regional traits, lighting, shadows, and
environment.
03 / BOUNDING BOXES & TOKENS
Click and drag directly on the canvas to draw class boxes (subject, foreground,
background). Add specific trigger tags or tokens with Enter.
04 / SPEED KEYS & EXPORT
Press [ENTER] to Save & Advance • [X] to Ignore •
[↑ / ↓] or [A / D] to navigate. When done, click DONE
& DOWNLOAD ZIP for training-ready YOLO, COCO, or Stable Diffusion archives.