AI Models

  • Search
    Search the site's AI modules, directories, and articles by keyword to find the matching feature or content.
  • Google Gemini 3.8 Flash
    Google Gemini 3.8 Flash, Google's most intelligent Flash model with major gains in coding, agentic tasks, and multi-step reasoning.
  • Google Gemini 3.7 Flash
    Google Gemini 3.7 Flash, the newest fast general-purpose Gemini model with upgraded speed and quality.
  • Google Gemini 3.6 Flash
    Google Gemini 3.6 Flash, the newest fast general-purpose Gemini model at a lower cost.
  • Google Gemini 3.5 Flash
    Google Gemini 3.5 Flash, the latest fast and cost-efficient general-purpose Gemini model.
  • Google Gemini 3.5 Flash Lite
    Google Gemini 3.5 Flash Lite, the fastest and most affordable tier of the Gemini 3.5 family.
  • Google Gemini 3.1 Pro
    Google Gemini 3.1 Pro, the flagship Gemini model via Google AI Studio.
  • Google Gemini 3.1 Flash Lite
    Google Gemini 3.1 Flash Lite, the fastest and most affordable Gemini tier.
  • Anthropic Claude Fable 5.1
    Anthropic Claude Fable 5.1, the newest top-tier Claude model, with stronger long-horizon agentic work, multistep research, and document, spreadsheet, and slide work.
  • Anthropic Claude Fable 5
    Anthropic Claude Fable 5, the most powerful Claude model, a new tier above Opus.
  • Anthropic Claude Opus 5
    Anthropic Claude Opus 5, the new-generation flagship with a large jump in coding, reasoning, and computer use.
  • Anthropic Claude Opus 4.8
    Anthropic Claude Opus 4.8, the flagship of the Claude 4 family, with sharper judgment and improved honesty.
  • Anthropic Claude Sonnet 5
    Anthropic Claude Sonnet 5, a balanced mid-tier model with strong reasoning and writing quality.
  • Anthropic Claude Haiku 4.5
    Anthropic Claude Haiku 4.5, the fastest and cheapest Claude model.
  • OpenAI GPT-6 Astra
    OpenAI GPT-6 Astra — the most capable model, built for the hardest end-to-end work, 1.05M context and 128K max output.
  • OpenAI GPT-5.6 Sol
    OpenAI GPT-5.6 Sol — frontier flagship for complex professional work, 1.05M context and 128K max output.
  • OpenAI GPT-5.6 Terra
    OpenAI GPT-5.6 Terra — balances intelligence and cost for general-purpose tasks.
  • OpenAI GPT-5.6 Luna
    OpenAI GPT-5.6 Luna — built for cost-sensitive, high-volume workloads.
  • OpenAI ChatGPT 5.5
    OpenAI ChatGPT 5.5 — flagship with 1.05M context and 128K max output.
  • OpenAI ChatGPT 5.4 mini
    OpenAI ChatGPT 5.4 mini, a lighter and cheaper variant of GPT-5.4.
  • OpenAI ChatGPT 5.4 nano
    OpenAI ChatGPT 5.4 nano, the smallest and fastest 5.4-family model.
  • Perplexity Sonar Deep Research
    Perplexity Sonar Deep Research runs a multi-step research workflow: planning, iterative searching, cross-referencing, and synthesis.
  • Perplexity Sonar Reasoning Pro
    Perplexity Sonar Reasoning Pro combines live web search with deeper reasoning.
  • Perplexity Sonar Pro
    Perplexity Sonar Pro, a web-search-augmented LLM.
  • xAI Grok 4.6
    xAI Grok 4.6, the newest flagship model for coding, agentic tasks, and knowledge work, with a 500K context window.
  • xAI Grok 4.5
    xAI Grok 4.5, flagship model with strong reasoning and a 500K context window.
  • Meta Muse Spark 1.3
    Meta Muse Spark 1.3 — the latest Muse Spark model, tuned for agentic workflows and coding, with image understanding, reasoning effort, web search, and a 1M context window.
  • BytePlus Seed 2.0 Lite
    BytePlus Seed 2.0 Lite, the lightweight Seed tier.
  • BytePlus Seed 1.8
    BytePlus Seed 1.8, a strong Chinese-optimized LLM.
  • Kimi K3
    MoonshotAI Kimi K3 flagship model with vision and a 1M-token context.
  • GLM 5.2
    Z.ai GLM 5.2 large-scale reasoning model with a 1M-token context.
  • Qwen 3.5 397B A17B
    Alibaba Qwen 3.5 MoE (397B total, 17B active).
  • Site Default Model
    Site Default Model
  • AI Agent
    An autonomous AI agent with multi-round reasoning and tool use that breaks down and completes complex tasks.
  • Knowledge Chunk Splitter
    Upload documents or paste text, and the AI splits the content into knowledge blocks by chapter, paragraph, or concept, showing the result in JSONL format and producing a downloadable JSONL file.

Writing

  • Meeting Minutes
    Paste a meeting transcript (from speech_to_text or manual notes) and receive a clean summary with key decisions, action items, and follow-up dates.
  • Slide Outline
    One-click slide outline generator.
  • Slide Generation 2
    Plans and generates an A4-landscape editable PPTX deck with an AI Agent workflow.
  • Slide Generation
    Generate a 20-slide pptx deck via gamma.app.
  • AI Agent Notebook
    Card-based notebook manager for organising reading, research, and project notes.
  • AI Copywriter / Copywriting
    An AI copy-editor.
  • Official Document Writing
    Generates Taiwan-style official documents following the standard format and phrasing conventions.
  • Chat with PDF (Text Only)
    This module can only answer questions about the text in a PDF. To ask about images inside the PDF, please: 1. Use AI Agent instead 2. Use AI Agent Notebook instead 3. Use another module to extract the PDF content first
  • Homework Assistant
    An AI homework assistant that explains solutions step by step.
  • Social Post Optimization
    Takes your draft social-media post and rewrites it for higher engagement.
  • Contract Review
    Contract review assistant.
  • Report Generation
    Structured report writer.
  • Webpage Summary
    Paste a URL and receive a concise summary of the page's key points.
  • Google Search
    Enter keywords to find related web pages via Google Search — returns titles, snippets and clickable URLs.
  • Google News
    Enter keywords to find related news via Google News — returns titles, sources, publish times and clickable links.
  • Paper Writing
    Academic paper assistant.
  • Speech Recognition
    Cloud speech recognition powered by OpenAI's gpt-4o-transcribe model, with plain-text transcript and SRT subtitle output.
  • Speech Recognition gpt_transcribe
    Cloud speech recognition with OpenAI's next-generation gpt-transcribe model — upload audio, get a high-accuracy plain-text transcript.
  • Multi-Speaker Speech Recognition
    Multi-speaker speech recognition powered by OpenAI's cloud gpt-4o-transcribe-diarize model, labeling each speaker in the transcript automatically.
  • Live Voice Conversation gemini-3.1-flash-live
    Low-latency real-time voice conversation powered by Google Gemini 3.1 Flash Live, with barge-in support and transcripts for both sides.
  • Live Voice Translation gemini-3.5-live-translate
    Real-time speech translation across 70+ languages powered by Google Gemini 3.5 Live Translate, with low-latency translated audio and transcripts.
  • Live Transcription gpt_live_transcribe
    Low-latency streaming speech recognition with OpenAI's gpt-live-transcribe — the transcript appears as you speak.
  • Programming Assistant
    Programming assistant that answers coding questions, explains errors, reviews snippets, and suggests refactors.
  • Project Plan Writing
    Project plan generator.
  • Press Release
    Rewrites event material into formal written press-release style, suitable for media outreach, internal newsletters, and PR distribution.
  • Oral News Broadcast Script
    Adapts news material into a spoken-word broadcast script — short sentences, natural rhythm, easy to read aloud.
  • Image OCR
    Upload an image or screenshot; AI extracts editable plain text via OCR.
  • Document OCR to Markdown
    Upload documents and an AI Agent OCRs them page by page into an editable Markdown file, with figure crops, captions and formatted tables.
  • Document OCR to DOCX
    Upload documents and an AI Agent OCRs them page by page into an editable Word (.docx) file, with embedded figure crops, captions and real Word tables.
  • AI Content Detector
    Paste any article and receive an AI-vs-human authorship likelihood score with per-paragraph breakdown.
  • 500-Word Summary
    Long-form summarizer that distills your input down to approximately 500 Chinese characters (or equivalent English).
  • 1000-Word Summary
    Detailed summarizer targeting ~1,000 Chinese characters.
  • Writing Optimization
    General-purpose writing optimizer.
  • Essay Grading
    AI essay grader that assigns A / B / C / D / E with specific feedback on structure, content, and language.
  • Web Crawler
    Enter a URL to crawl the page and organize it into title, description, keywords, and content (with HTML and JavaScript stripped).
  • Document Text Extraction
    This module can only extract text from documents. For text recognition in images, please use other modules.

Media Creation

  • GPT Image 2.5 Flare Generation
    Image generation powered by OpenAI's gpt-image-2.5-flare (GPT Image 2.5 Flare), the fastest model in the GPT Image 2.5 family for high-quality everyday images, with selectable output quality (auto, low, medium, high, xhigh, max) and image editing with continuous refinement.
  • GPT Image 2.5 Sunburst Generation
    Image generation powered by OpenAI's gpt-image-2.5-sunburst (GPT Image 2.5 Sunburst), the next-generation flagship after GPT Image 2 with sharper detail, richer lighting and stronger prompt adherence, selectable output quality (auto, low, medium, high, xhigh, max), and image editing with continuous refinement.
  • Nano Banana Pro
    Nano Banana Pro image model with upload-based editing and continuous conversational refinement.
  • Nano Banana 3.1 Flash
    Nano Banana 3.1 Flash, the lightweight tier with continuous conversational refinement.
  • Image-to-Prompt
    Upload any image and the model reverse-engineers the prompt that would recreate it.
  • Gemini Omni Flash 1.1 Video Generation
    Google Gemini Omni Flash 1.1, the generally available natively multimodal video generation model: create short videos with native audio from text prompts at 720P, 1080P, or 4K in 16:9 or 9:16, upload one or two PNG/JPG images as references (image-to-video, first/last frame interpolation, or subject references) or one MP4 video to edit or extend, and refine results with multi-turn conversational editing.
  • Gemini Omni Flash 1.0 Video Generation
    Google Gemini Omni Flash 1.0 (preview), a natively multimodal video generation model: create short videos from text prompts, upload one PNG/JPG image as a reference (image-to-video), and refine results with multi-turn conversational editing.
  • Google Veo 3.1 Text-to-Video
    Google Veo 3.1 text-to-video, a premium cinematic video generator; you can also upload an image as the first-frame reference (image-to-video).
  • Music Generation
    AI music generator.
  • Extend Song
    Upload a song and AI continues it from the end, automatically merged into one full track.
  • Song Restyle
    Upload a song or instrumental track to change its style or instruments, or add vocals to instrumental music.
  • Sing with Your Voice
    Upload a voice sample and AI sings a song with your voice.
  • Vocal & Instrumental Separation
    Upload a song and split it into separate vocal and instrumental tracks.
  • AI Change Clothes
    Upload a person photo plus a clothing image (or text description) and the model dresses the subject in the target outfit while keeping pose, face, and background intact.
  • BytePlus Seedream 5.0 Pro Text-to-Image
    BytePlus Dola Seedream 5.0 Pro text-to-image, a flagship image generator with precise editing control, excellent Chinese prompt understanding, strong photo-realism, and continuous conversational refinement.
  • seedance 2.5 Generate Video
    BytePlus Seedance 2.5 multimodal video generation with up to 50 reference assets and video editing.
  • seedance 2.0 Generate Video
    BytePlus Seedance 2.0 text-to-video.
  • Cheer Video Generation
    A themed video generator that turns a name and a short cheer message into a personalized motivational clip.
  • Christmas Video Generation
    One-click Christmas greeting video generator.
  • New Year Video Generation
    Chinese New Year greeting video generator.
  • BytePlus Seed-TTS
    BytePlus seed-tts text-to-speech.
  • BytePlus SeedAudio 1.0 Audio Generation
    Generate sound effects and speech from a text description; optionally upload reference audio clips or one image to guide the timbre and mood.
  • BytePlus Voice Cloning
    Upload a 10-15 second sample of a person's voice and the model learns the timbre, then reads any text you provide in that cloned voice.

Lifestyle / Translation

  • Translate to Traditional Chinese
    Translates text in any supported language into Traditional Chinese suitable for Taiwan/HK readers.
  • Translate to English
    Translates text (especially Chinese) into natural American English.
  • English Grammar Check
    English grammar and style checker.
  • Multilingual Voice Tutor (Female) - Sophie
    Practice conversation with Sophie, a friendly female AI tutor. Replies are short text answers plus synchronized voice narration.
  • Multilingual Voice Tutor (Male) - Alex
    Practice conversation with Alex, a friendly male AI tutor. Replies are short text answers plus synchronized voice narration.
  • Knowledge Base
    Build a private knowledge base; AI conversations cite your internal docs automatically.
  • GPTs Custom Assistant
    Build your own chatbot; configure role prompts and knowledge bases for domain-specific AI.
  • Health Education
    Health-education assistant for generating patient-friendly explanations of medications, procedures, and care instructions.
  • Simplified Chinese Detector
    Scans your text for simplified-Chinese characters and mainland-specific phrasing, highlighting each finding.
  • Simplified to Traditional
    A quick Simplified-to-Traditional Chinese converter.
  • Traditional to Simplified
    A quick Traditional-to-Simplified Chinese converter.