GLM
The GLM series models by Zhipu AI, offering powerful bilingual conversational capabilities.
Kling video generation
Kling AI video generation service with text and image input.
Grok
Artificial intelligence model developed by xAI company
DeepSeek AI
DeepSeek AI model supports powerful chat dialogue capabilities.
Coding Plan
One plan unlocks the complete Claude, GPT/Codex, GLM, Kimi, and DeepSeek coding fleet
Qwen Image
Create and edit images with multi-image references, batch output, and asynchronous tasks.
Gemini AI
The Gemini model by Google, offering excellent chat services.
ByteDance Seedream Image Generation
The strongest image generation model under ByteDance.
Claude AI
The Claude model by Anthropic, offering excellent chat services.
Kimi
The Kimi model produced by Moonshot AI offers excellent conversational services.
AI Dialogue
AI intelligent dialogue service, providing answers to any questions.
Cloudflare Turnstile Captcha Service
Solve Cloudflare Turnstile in the background — submit the sitekey and site URL, get a valid token, no manual challenge.
Reddit Posts & Comments
Forum posts, comments and threaded replies, with subreddit and parent relationships for topic and content filtering.
Spotify English Podcast Audio
English podcast recordings with title, description, language and duration metadata for long-form speech understanding and retrieval.
Audible English Audiobooks
Human-narrated English audiobooks with work, author, narrator and genre metadata for long-form speech and narrative evaluation.
YouTube Themed Audio & Video
Audio/video and metadata organized around selected videos, channels or dataset lists, configurable by quality, duration and topic.
GSM8K Math Reasoning Dataset
Grade school math reasoning benchmark by OpenAI with 17.6K problems and step-by-step solutions, a core benchmark for evaluating LLM math reasoning.
Iris Dataset
Classic flower classification dataset with 150 samples across 3 species.
Wine Quality Dataset
Portuguese Vinho Verde wine quality data with physicochemical properties and quality scores for red and white wines.
Fashion-MNIST Dataset
Zalando fashion article images - 70,000 28x28 grayscale images across 10 clothing categories, a modern MNIST replacement
Abalone Dataset
Abalone dataset from UCI ML Repository for predicting abalone age from physical measurements
Mushroom Dataset
Mushroom classification dataset from UCI ML Repository for distinguishing edible vs poisonous mushrooms
Bike Sharing Dataset
Bike sharing dataset from UCI ML Repository with hourly and daily rental counts plus weather features
Car Evaluation Dataset
Car evaluation dataset from UCI ML Repository for multi-class classification of car acceptability
Palmer Penguins Dataset
Palmer Penguins dataset, a modern alternative to Iris for data exploration and visualization teaching
HumanEval Dataset
OpenAI HumanEval code generation benchmark dataset, containing 164 hand-written Python programming problems with function signatures, docstrings, and unit tests, the standard benchmark for evaluating code generation pass@k metrics.
ARC Dataset
AI2 Reasoning Challenge (ARC) dataset, containing 7,787 real grade-school science multiple-choice questions split into Easy and Challenge sets, a core benchmark for evaluating scientific reasoning in language models.
WinoGrande Dataset
WinoGrande large-scale commonsense coreference resolution dataset, containing 44,000 Winograd Schema-style fill-in-the-blank problems with adversarial filtering for quality, a key benchmark for commonsense reasoning.
PIQA Dataset
PIQA (Physical Intuition QA) dataset, containing approximately 21,000 binary-choice questions about everyday physical world interactions, testing language model understanding of physical commonsense.
COPA Dataset
COPA (Choice of Plausible Alternatives) causal reasoning dataset, containing 1,000 binary-choice causal reasoning questions requiring selection of the most plausible cause or effect for a given premise.
SNLI Stanford Natural Language Inference Dataset
SNLI is the first large-scale NLI dataset with 570K human-written English sentence pairs labeled as entailment, contradiction, or neutral, serving as a foundational benchmark in natural language inference.
- 1