一键重装系统工具 | U盘启动盘制作工具 | 误删文件恢复软件 | 硬盘数据抢救专家 | 电脑蓝屏修复助手 | C盘空间清理神器 | 电脑驱动离线安装工具 | 微信聊天记录恢复工具 | 照片误格式化恢复 | 电脑密码破解清除工具 | 系统崩溃紧急救援盘 | 电脑加速优化大师 | 电脑开不了机怎么重装系统 | 回收站清空了怎么恢复 | 硬盘分区丢失数据恢复 | 电脑卡顿重装系统有用吗 | U盘插入提示格式化数据恢复 | 电脑中毒文件被隐藏恢复 | 忘记电脑开机密码怎么办 | 新硬盘分区对齐工具 | 旧电脑装Win10流畅工具 | SD卡照片删除恢复免费版 | 移动硬盘打不开提示损坏修复 | 电脑无故重启系统修复工具 | 电脑小白一键重装神器 | 程序员电脑环境配置助手 | 设计师电脑字体/素材恢复工具 | 网吧网管系统维护工具箱 | 财务人员电脑发票备份恢复 | 学生党免费电脑系统安装包 | 电脑维修师傅必备工具盘 | 游戏玩家电脑性能优化助手 | 办公白领误删文档恢复软件 | 自媒体视频素材恢复工具 | 网课录制视频损坏修复工具 | 最好的U盘PE系统排名 | 数据恢复软件哪个最强 | 免费电脑助手与收费版区别 | 国产装机工具哪款无广告 | 离线版驱动助手推荐 | 轻量级电脑优化工具对比 | 支持NVMe驱动的PE工具 | 带网络功能的应急启动盘 | 2026最新版万能装机工具 | 支持Win11 24H2的PE工具 | 最新免激活系统重装工具 | 2026数据恢复软件破解版合集 | 纯净无捆绑装机助手V3.0 | 支持苹果M芯片的电脑助手 | 秋季更新版系统维护工具箱 | 电脑系统崩了怎么用U盘把重要资料拷贝出来 | 重装系统前哪些文件夹必须备份 | 固态硬盘误格式化还能恢复数据吗 | 如何制作一个既带PE又能存数据的双分区U盘 | 电脑总是弹窗广告用什么助手彻底拦截 后台管理
📢 欢迎访问系统之家!所有资源均经过安全检测。

Compare AI Models: Pricing, Context & Benchmarks

发布时间:2026-10-07 | 浏览:1
📥 下载地址(文章开头)
装机神器,在线重装利器,在线安装一切系统。
ElevenLabs: Eleven Multilingual v2 Eleven Multilingual v2 50% off Eleven Multilingual v2 is a text-to-speech model from ElevenLabs. It is suited for lifelike, consistent long-form narration such as voice-overs and audiobooks, supports 29 languages, and has a 10,000-character request limit. by elevenlabs Oct 7, 2026 $40/M characters Eleven Multilingual v2 is a text-to-speech model from ElevenLabs. It is suited for lifelike, consistent long-form narration such as voice-overs and audiobooks, supports 29 languages, and has a 10,000-character request limit. ElevenLabs: Eleven Flash v2.5 Eleven Flash v2.5 50% off Eleven Flash v2.5 is an ultra-low-latency text-to-speech model from ElevenLabs. It is suited for conversational and real-time use cases, supports 32 languages, and has a 40,000-character request limit. by elevenlabs Oct 7, 2026 $20/M characters Eleven Flash v2.5 is an ultra-low-latency text-to-speech model from ElevenLabs. It is suited for conversational and real-time use cases, supports 32 languages, and has a 40,000-character request limit. ElevenLabs: Eleven v3 Conversational Eleven v3 Conversational 50% off Eleven v3 Conversational is a text-to-speech model from ElevenLabs, a variant of Eleven v3 optimized for natural dialogue in conversational agents. It supports 70+ languages and has a 5,000-character request limit. by elevenlabs Oct 7, 2026 $20/M characters Eleven v3 Conversational is a text-to-speech model from ElevenLabs, a variant of Eleven v3 optimized for natural dialogue in conversational agents. It supports 70+ languages and has a 5,000-character request limit. ElevenLabs: Eleven v4 Turbo Eleven v4 Turbo 50% off Eleven v4 Turbo is a low-latency text-to-speech model from ElevenLabs. It keeps Eleven v4's expressive delivery and audio tags while being tuned for faster generation, with support for 90+ languages and a 10,000-character request limit. by elevenlabs Oct 7, 2026 $20/M characters Eleven v4 Turbo is a low-latency text-to-speech model from ElevenLabs. It keeps Eleven v4's expressive delivery and audio tags while being tuned for faster generation, with support for 90+ languages and a 10,000-character request limit. ElevenLabs: Eleven v3 Eleven v3 50% off Eleven v3 is a text-to-speech model from ElevenLabs. It produces emotionally rich, highly expressive speech with inline audio tags, supports 70+ languages, and has a 5,000-character request limit. by elevenlabs Oct 7, 2026 $40/M characters Eleven v3 is a text-to-speech model from ElevenLabs. It produces emotionally rich, highly expressive speech with inline audio tags, supports 70+ languages, and has a 5,000-character request limit. ElevenLabs: Eleven v4 Eleven v4 50% off Eleven v4 is a text-to-speech model from ElevenLabs. It is ElevenLabs' most expressive model, with inline audio tags for emotional and delivery control, support for 90+ languages, and a 10,000-character request limit. by elevenlabs Oct 7, 2026 $40/M characters Eleven v4 is a text-to-speech model from ElevenLabs. It is ElevenLabs' most expressive model, with inline audio tags for emotional and delivery control, support for 90+ languages, and a 10,000-character request limit. OpenAI: GPT-6 Luna Decisions GPT-6 Luna Decisions 6.33B tokens GPT-6 Luna Decisions is GPT-6 Luna served through OpenAI's Decisions API. Instead of generating text, it reads the content passed as state (text, JSON, or images) and returns typed, probabilistic answers to named questions in a single request: the probability of yes for a yes/no question ( noul ), a probability for every option plus the most likely one ( choice ), or a probability for every level of an ordered rubric plus the expected score ( score ). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 200 questions about the same content. by openai Oct 6, 2026 1.05M context $0.10 /M input tokens $0 /M output tokens GPT-6 Luna Decisions is GPT-6 Luna served through OpenAI's Decisions API. Instead of generating text, it reads the content passed as state (text, JSON, or images) and returns typed, probabilistic answers to named questions in a single request: the probability of yes for a yes/no question ( noul ), a probability for every option plus the most likely one ( choice ), or a probability for every level of an ordered rubric plus the expected score ( score ). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 200 questions about the same content. SpaceXAI: Grok Imagine Video 1.5 Lite Grok Imagine Video 1.5 Lite 4 hours Grok Imagine Video 1.5 Lite is a faster, lower-cost video generation model from SpaceXAI, distilled from Grok Imagine Video 1.5 . It supports text-to-video and image-to-video, trading some quality for speed and price. 1080p output is rendered at 720p and upscaled. by x-ai Oct 6, 2026 from $0.02/second Grok Imagine Video 1.5 Lite is a faster, lower-cost video generation model from SpaceXAI, distilled from Grok Imagine Video 1.5 . It supports text-to-video and image-to-video, trading some quality for speed and price. 1080p output is rendered at 720p and upscaled. Google: Nano Banana 2.1 Nano Banana 2.1 243M tokens Nano Banana 2.1 (Gemini Nano Banana 2.1) is Google's image generation and editing model on the Flash tier, succeeding Nano Banana 2 and Nano Banana Pro. It improves product recontextualization, mask- and ink-based editing, and factual accuracy, and renders photorealistic skin tones, detailed materials, lighting, and coherent backgrounds. It accepts text and image inputs, returns images with optional text, and supports 1K, 2K, and 4K output plus extended aspect ratios via the image_config API Parameter . by google Oct 6, 2026 66K context $1.50 /M input tokens $30 /M output tokens Nano Banana 2.1 (Gemini Nano Banana 2.1) is Google's image generation and editing model on the Flash tier, succeeding Nano Banana 2 and Nano Banana Pro. It improves product recontextualization, mask- and ink-based editing, and factual accuracy, and renders photorealistic skin tones, detailed materials, lighting, and coherent backgrounds. It accepts text and image inputs, returns images with optional text, and supports 1K, 2K, and 4K output plus extended aspect ratios via the image_config API Parameter . Mistral: Mistral Large 4 Mistral Large 4 50% off 30.1B tokens Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 512K-token context window with up to 256K output tokens, and supports tool calling and structured outputs. by mistralai Oct 6, 2026 524K context $0.68 /M input tokens $2.09 /M output tokens Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 512K-token context window with up to 256K output tokens, and supports tool calling and structured outputs.
📥 下载地址(文章中间)
装机神器,在线重装利器,在线安装一切系统。
Tencent: Hy Image 3.5 Preview Hy Image 3.5 Preview 121M tokens Hy Image 3.5 Preview is a unified image generation and editing model from Tencent. It handles text-to-image, image-to-image, and multi-turn editing through one endpoint, taking up to 20 reference images per request and producing output up to 4K. Built on the 80B mixture-of-experts Hy Image 3.0 base, it is particularly strong at subject consistency across edits, prompt-faithful composition, and rendering Chinese and English text inside images. by tencent Oct 5, 2026 100K context $1.60/M tokens Hy Image 3.5 Preview is a unified image generation and editing model from Tencent. It handles text-to-image, image-to-image, and multi-turn editing through one endpoint, taking up to 20 reference images per request and producing output up to 4K. Built on the 80B mixture-of-experts Hy Image 3.0 base, it is particularly strong at subject consistency across edits, prompt-faithful composition, and rendering Chinese and English text inside images. inclusionAI: Ling 3.1 Flash Ling 3.1 Flash 365B tokens Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total. by inclusionai Oct 2, 2026 262K context $0 /M input tokens $0 /M output tokens Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total. ByteDance Seed: Seedream 5.0 Flash Seedream 5.0 Flash 995M tokens Seedream 5.0 Flash is an image generation and editing model from ByteDance Seed. It is the fast, cost-efficient tier of the Seedream 5.0 family, suited for high-volume production and interactive editing workflows that need precise edits at low latency. by bytedance-seed Oct 1, 2026 from $0.018/image Seedream 5.0 Flash is an image generation and editing model from ByteDance Seed. It is the fast, cost-efficient tier of the Seedream 5.0 family, suited for high-volume production and interactive editing workflows that need precise edits at low latency. Perplexity: Decider V1 27B Decider V1 27B 8.63B tokens Legal (#40) SEO (#42) Trivia (#33) Decider V1 27B is a decision model from Perplexity. Instead of generating text, it reads content passed as state and returns typed, probabilistic answers to one or more named questions in a single request: the probability of yes for a yes/no question ( noul ), a probability for every option plus the most likely one ( choice ), or a probability for every level of an ordered rubric plus the expected score ( score ). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 128 questions about the same content. On OpenRouter it currently accepts text and JSON state ; image inputs are not yet supported. by perplexity Oct 1, 2026 262K context $0.04 /M input tokens $0 /M output tokens Decider V1 27B is a decision model from Perplexity. Instead of generating text, it reads content passed as state and returns typed, probabilistic answers to one or more named questions in a single request: the probability of yes for a yes/no question ( noul ), a probability for every option plus the most likely one ( choice ), or a probability for every level of an ordered rubric plus the expected score ( score ). It is built for classification, routing, moderation, and rubric grading where application code thresholds the returned numbers rather than parsing a chat reply. A request can carry up to 128 questions about the same content. On OpenRouter it currently accepts text and JSON state ; image inputs are not yet supported. Black Forest Labs: FLUX.3 Image FLUX.3 Image 50% off 238M tokens FLUX.3 Image is Black Forest Labs' flagship image generation and editing model. It handles text-to-image and multi-reference editing with up to 10 input images, and renders at fixed resolution tiers from 768 up to 4K with a selectable aspect ratio. Pricing is a flat per-image rate that scales with the chosen resolution tier. by black-forest-labs Oct 1, 2026 47K context from $0.0205/image FLUX.3 Image is Black Forest Labs' flagship image generation and editing model. It handles text-to-image and multi-reference editing with up to 10 input images, and renders at fixed resolution tiers from 768 up to 4K with a selectable aspect ratio. Pricing is a flat per-image rate that scales with the chosen resolution tier. LiquidAI: d1 d1 8.15B tokens d1 is Liquid AI's structured decision model, served as a System One endpoint. Send a state along with typed questions, and it returns a choice, a score, or a yes/no answer, each with a probability taken directly from the model rather than written out as text. It uses the same /v1/systemone schema as other OpenRouter Decisions models, so it suits routing, classification, and policy checks that need a fast, scored answer instead of prose. by liquid Oct 1, 2026 66K context $0.04 /M input tokens $0 /M output tokens d1 is Liquid AI's structured decision model, served as a System One endpoint. Send a state along with typed questions, and it returns a choice, a score, or a yes/no answer, each with a probability taken directly from the model rather than written out as text. It uses the same /v1/systemone schema as other OpenRouter Decisions models, so it suits routing, classification, and policy checks that need a fast, scored answer instead of prose. Apodex: Apodex 1.1 Mini (free) Apodex 1.1 Mini (free) 192B tokens Apodex 1.1 Mini is a reasoning-first model from Apodex, built for complex, long-horizon research and forecasting tasks. It works directly with files, data, code, and tools to produce verifiable results, and is designed for agentic research workflows where answers need to be grounded in evidence. by apodex Oct 1, 2026 262K context $0 /M input tokens $0 /M output tokens Apodex 1.1 Mini is a reasoning-first model from Apodex, built for complex, long-horizon research and forecasting tasks. It works directly with files, data, code, and tools to produce verifiable results, and is designed for agentic research workflows where answers need to be grounded in evidence. Cloudflare: Clef Flash Clef Flash 4.93B tokens Trivia (#6) Clef-flash is the fast 9B member of Cloudflare's open-source Clef decision model family, a fine-tune of Qwen3.5-9B served on Workers AI. It turns a state (text or structured JSON) plus a schema of typed questions into decisions, returning a calibrated probability for every allowed option of every question in a single forward pass instead of generating tokens. Use it for low-latency classification, routing, scoring, and guardrails through the Decisions API. Note: Workers AI currently truncates long text state to roughly the first 2K tokens, so content beyond that is not read; images are counted separately. by cloudflare Oct 1, 2026 66K context $0.09 /M input tokens $0 /M output tokens Clef-flash is the fast 9B member of Cloudflare's open-source Clef decision model family, a fine-tune of Qwen3.5-9B served on Workers AI. It turns a state (text or structured JSON) plus a schema of typed questions into decisions, returning a calibrated probability for every allowed option of every question in a single forward pass instead of generating tokens. Use it for low-latency classification, routing, scoring, and guardrails through the Decisions API. Note: Workers AI currently truncates long text state to roughly the first 2K tokens, so content beyond that is not read; images are counted separately. Cloudflare: Clef Clef 3.64B tokens Trivia (#9) Clef is Cloudflare's open-source 27B multimodal decision model, a fine-tune of Qwen3.8-27B served on Workers AI. It turns a state (text or structured JSON) plus a schema of typed questions into decisions, returning a calibrated probability for every allowed option of every question in a single forward pass instead of generating tokens. Use it for classification, routing, scoring, guardrails, and agentic control flow through the Decisions API. Note: Workers AI currently truncates long text state to roughly the first 2K tokens, so content beyond that is not read; images are counted separately. by cloudflare Oct 1, 2026 66K context $0.24 /M input tokens $0 /M output tokens Clef is Cloudflare's open-source 27B multimodal decision model, a fine-tune of Qwen3.8-27B served on Workers AI. It turns a state (text or structured JSON) plus a schema of typed questions into decisions, returning a calibrated probability for every allowed option of every question in a single forward pass instead of generating tokens. Use it for classification, routing, scoring, guardrails, and agentic control flow through the Decisions API. Note: Workers AI currently truncates long text state to roughly the first 2K tokens, so content beyond that is not read; images are counted separately. Microsoft AI: MAI-Voice-2.1-Flash MAI-Voice-2.1-Flash 2.4M tokens MAI-Voice-2.1-Flash is a low-latency text-to-speech model from Microsoft AI, optimized for real-time responsiveness. It produces natural, expressive speech across 23 languages, with human-like intonation, rhythm, and emotional nuance. It is suited for voice agents, assistants, call centers, and other interactive applications where latency and cost matter most. On OpenRouter, set voice to a full voice ID with the model suffix, such as "en-US-Harper:MAI-Voice-2.1-Flash" . A voice's locale sets the synthesis language. Set response_format to "mp3" or "pcm" (24 kHz mono). Harper and Grant support the agent , customer-call-center , educational , and narrator speaking styles, and many locale voices add emotion styles such as excited , happy , sad , and whispering . The full list of voices is in the supported_voices field of the models API . See the text-to-speech guide . by microsoft Oct 1, 2026 $15/M characters MAI-Voice-2.1-Flash is a low-latency text-to-speech model from Microsoft AI, optimized for real-time responsiveness. It produces natural, expressive speech across 23 languages, with human-like intonation, rhythm, and emotional nuance. It is suited for voice agents, assistants, call centers, and other interactive applications where latency and cost matter most. On OpenRouter, set voice to a full voice ID with the model suffix, such as "en-US-Harper:MAI-Voice-2.1-Flash" . A voice's locale sets the synthesis language. Set response_format to "mp3" or "pcm" (24 kHz mono). Harper and Grant support the agent , customer-call-center , educational , and narrator speaking styles, and many locale voices add emotion styles such as excited , happy , sad , and whispering . The full list of voices is in the supported_voices field of the models API . See the text-to-speech guide .
📥 下载地址(文章结尾)
装机神器,在线重装利器,在线安装一切系统。