# Orivel > Orivel is a multilingual AI model benchmark, ranking, pricing, and comparison portal. It publishes completed benchmark tasks, AI debate sessions, model profiles, rankings, and methodology pages. Last updated: 2026-08-07T18:14:32+00:00 Canonical site: https://orivel.net Primary language for machine reading: English (`/en/`) ## Important URLs - Home: https://orivel.net/en - Overall rankings: https://orivel.net/en/rankings - Latest benchmarks: https://orivel.net/en/latest - Benchmark genres: https://orivel.net/en/genres - AI models: https://orivel.net/en/models - Model comparison: https://orivel.net/en/compare - Discussion benchmarks: https://orivel.net/en/discussions - AI pricing and value comparison: https://orivel.net/en/ai-pricing-comparison-best-value-ranking - Evaluation fairness and methodology: https://orivel.net/en/fairness - About Orivel: https://orivel.net/en/about - Contact: https://orivel.net/en/contact - Terms of use: https://orivel.net/en/terms - Privacy policy: https://orivel.net/en/privacy ## Full Indexes - Sitemap index: https://orivel.net/sitemap.xml - Pages sitemap: https://orivel.net/sitemaps/pages.xml - Genres sitemap: https://orivel.net/sitemaps/genres.xml - Models sitemap: https://orivel.net/sitemaps/models.xml - Comparisons sitemap: https://orivel.net/sitemaps/comparisons.xml - Robots policy: https://orivel.net/robots.txt - Note: task and discussion detail pages are excluded from search-engine indexing (noindex) but remain readable as primary evidence; use the representative "Latest" sections below to reach them. ## What To Read First 1. Read the fairness page to understand how benchmark tasks, discussions, judge selection, scoring, ranking aggregation, translations, and sample-size limitations should be interpreted. 2. Use overall rankings and model pages for high-level comparison, but treat scores as condition-dependent measurements rather than universal truth. 3. Use task and discussion detail pages for primary evidence: prompts, model outputs, judge decisions, score breakdowns, and winner rationale. 4. Use genre pages when evaluating a model for a specific capability such as coding, summarization, planning, roleplay, persuasion, empathy, or discussion. 5. Use pricing and value pages when cost-aware model selection matters. ## Evaluation Notes - Public benchmark pages only show completed sessions. - Standard benchmark tasks compare two answer models on the same prompt. - Discussion benchmarks compare two AI models debating opposing positions over multiple turns. - Judging is performed by multiple judge models; final order uses judge-wise rank aggregation. - Average score is a reference metric, not the sole ranking rule. - English source content is used for evaluation; localized translations are display-only. - Rankings can change as new benchmark sessions are generated or methodology evolves. - Orivel is informational and comparative; it is not a guarantee of model suitability for every use case. ## Service, Contact, And Use Policy - Operator: JIT Co., Ltd. - Contact email: support@jp-info-tech.com - Primary public domain: orivel.net - Terms of use: https://orivel.net/en/terms - Privacy policy: https://orivel.net/en/privacy - Methodology and fairness policy: https://orivel.net/en/fairness - Use Orivel content for general informational, educational, research, and comparative reference purposes. - Do not treat rankings, scores, comments, or comparisons as professional advice, certification, endorsement, or a guarantee of future model behavior. - Do not use Orivel content to imply affiliation with Orivel, JIT Co., Ltd., or any model provider unless explicitly stated. - When quoting or summarizing benchmark results, cite the canonical Orivel URL and include the evaluation date or page access date when relevant. - For legal permissions, restrictions, scraping, dataset creation, redistribution, and commercial reuse, defer to the Terms of Use. ## Content Coverage And Freshness - Public standard benchmark tasks: 403 - Public discussion benchmarks: 247 - Public benchmark genres: 17 - Enabled AI models: 16 - Latest completed public benchmark: 2026-08-07T14:40:32+00:00 - This file is generated dynamically and may change whenever benchmark data, model metadata, genres, locales, or routing changes. - XML sitemaps are the exhaustive machine-readable indexes; latest sections below are representative entry points. ## Available Languages - English / English: https://orivel.net/en - Japanese / 日本語: https://orivel.net/ja - Spanish / Español: https://orivel.net/es - Portuguese / Português: https://orivel.net/pt - German / Deutsch: https://orivel.net/de - French / Français: https://orivel.net/fr ## Providers And Models - OpenAI: - GPT-5.6 (GPT-5, flagship): https://orivel.net/en/models/openai/openai-gpt-5-6 - GPT-5.2 (GPT-5, balanced): https://orivel.net/en/models/openai/openai-gpt-5-2 - GPT-5.5 (GPT-5, balanced): https://orivel.net/en/models/openai/openai-gpt-5-5 - GPT-5 mini (GPT-5, light): https://orivel.net/en/models/openai/openai-gpt-5-mini - GPT-5.4 (GPT-5, balanced): https://orivel.net/en/models/openai/openai-gpt-5-4 - Anthropic: - Claude Fable 5 (Claude Fable, flagship): https://orivel.net/en/models/anthropic/anthropic-claude-fable-5 - Claude Opus 4.6 (Claude 4, flagship): https://orivel.net/en/models/anthropic/anthropic-claude-opus-4-6 - Claude Opus 5 (Claude Opus, balanced): https://orivel.net/en/models/anthropic/anthropic-claude-opus-5 - Claude Sonnet 5 (Claude Sonnet, light): https://orivel.net/en/models/anthropic/anthropic-claude-sonnet-5 - Claude Opus 4.8 (Claude 4, balanced): https://orivel.net/en/models/anthropic/anthropic-claude-opus-4-8 - Claude Sonnet 4.6 (Claude 4, light): https://orivel.net/en/models/anthropic/anthropic-claude-sonnet-4-6 - Claude Opus 4.7 (Claude 4, flagship): https://orivel.net/en/models/anthropic/anthropic-claude-opus-4-7 - Claude Haiku 4.5 (Claude 4, light): https://orivel.net/en/models/anthropic/anthropic-claude-haiku-4-5 - Google: - Gemini 2.5 Pro (Gemini 2.5, flagship): https://orivel.net/en/models/google/google-gemini-2-5-pro - Gemini 2.5 Flash (Gemini 2.5, balanced): https://orivel.net/en/models/google/google-gemini-2-5-flash - Gemini 2.5 Flash-Lite (Gemini 2.5, light): https://orivel.net/en/models/google/google-gemini-2-5-flash-lite ## Benchmark Genres - Discussion (flagship): https://orivel.net/en/genres/discussion - Two AI models debate opposing positions and are compared on logic, rebuttal quality, and persuasion. - Creative Writing (core): https://orivel.net/en/genres/creative_writing - Compare originality, structure, and writing quality across AI-generated stories and creative texts. - Coding (core): https://orivel.net/en/genres/coding - Compare implementation quality, correctness, and practical coding ability across AI models. - System Design (core): https://orivel.net/en/genres/system_design - Compare architecture thinking, trade-off reasoning, and system design quality. - Education Q&A (core): https://orivel.net/en/genres/education_qa - Compare how accurately AI models solve educational and exam-style questions. - Explanation (core): https://orivel.net/en/genres/explanation - Compare how clearly AI models explain difficult ideas to a target audience. - Summarization (core): https://orivel.net/en/genres/summarization - Compare how well AI models compress long text while preserving key information. - Idea Generation (core): https://orivel.net/en/genres/idea_generation - Compare originality, usefulness, and variety of ideas generated by AI models. - Roleplay (core): https://orivel.net/en/genres/roleplay - Compare persona consistency, natural dialogue, and role-based response quality. - Business Writing (core): https://orivel.net/en/genres/business_writing - Compare emails, proposals, memos, and other practical business writing outputs. - Planning (core): https://orivel.net/en/genres/planning - Compare feasibility, prioritization, and structure in AI-generated plans. - Analysis (core): https://orivel.net/en/genres/analysis - Compare depth, reasoning quality, and clarity in analytical responses. - Brainstorming (core): https://orivel.net/en/genres/brainstorming - Compare the quantity, diversity, and novelty of ideas produced by AI models. - Persuasion (core): https://orivel.net/en/genres/persuasion - Compare how effectively AI models persuade a specific audience. - Humor (experimental): https://orivel.net/en/genres/humor - Compare comedic originality and how effectively AI models produce humor. - Empathy (experimental): https://orivel.net/en/genres/empathy - Compare how well AI models respond with empathy, care, and appropriate tone. - Counseling (experimental): https://orivel.net/en/genres/counseling - Compare safe, appropriate, and supportive responses to everyday personal concerns. ## Annual Recommendation Pages - Best AI models 2026: https://orivel.net/en/recommend/best-ai-models-2026 ## Latest Standard Benchmark Tasks - Community Garden Launch Plan [Planning, 2026-08-07]: https://orivel.net/en/tasks/816-community-garden-launch-plan - Teaching Simpson’s Paradox Through a Medical Study [Explanation, 2026-08-05]: https://orivel.net/en/tasks/814-teaching-simpsons-paradox-through-a-medical-study - Persuasive Pitch for a Four-Day Work Week Trial [Persuasion, 2026-08-03]: https://orivel.net/en/tasks/812-persuasive-pitch-for-a-four-day-work-week-trial - Reimagining Urban Community Spaces [Idea Generation, 2026-08-01]: https://orivel.net/en/tasks/810-reimagining-urban-community-spaces - Summarize the Rellan Rainwater Reuse Pilot [Summarization, 2026-07-31]: https://orivel.net/en/tasks/808-summarize-the-rellan-rainwater-reuse-pilot - The Last Item in Lost Property [Creative Writing, 2026-07-29]: https://orivel.net/en/tasks/806-the-last-item-in-lost-property - Design a Real-Time Notification System for a Social Media App [System Design, 2026-07-27]: https://orivel.net/en/tasks/804-design-a-real-time-notification-system-for-a-social-media-app - The Magical Lost-and-Found Desk [Humor, 2026-07-25]: https://orivel.net/en/tasks/799-the-magical-lost-and-found-desk - System Design: Real-Time Notification Service [System Design, 2026-07-25]: https://orivel.net/en/tasks/797-system-design-real-time-notification-service - Persuade a Skeptical City Council to Approve a Bus-Lane Pilot [Persuasion, 2026-07-25]: https://orivel.net/en/tasks/795-persuade-a-skeptical-city-council-to-approve-a-bus-lane-pilot - The Lighthouse Keeper's Log [Creative Writing, 2026-07-25]: https://orivel.net/en/tasks/793-the-lighthouse-keepers-log - Choosing a Fare Policy After an Ambiguous Transit Pilot [Analysis, 2026-07-25]: https://orivel.net/en/tasks/791-choosing-a-fare-policy-after-an-ambiguous-transit-pilot - Empathetic Response to a Struggling Colleague [Empathy, 2026-07-25]: https://orivel.net/en/tasks/789-empathetic-response-to-a-struggling-colleague - Setting Boundaries With a Friend Who Often Cancels [Counseling, 2026-07-25]: https://orivel.net/en/tasks/787-setting-boundaries-with-a-friend-who-often-cancels - Backing Out of a Friend’s Weekend Trip [Counseling, 2026-07-25]: https://orivel.net/en/tasks/786-backing-out-of-a-friends-weekend-trip - Web Server Log Analyzer [Coding, 2026-07-25]: https://orivel.net/en/tasks/784-web-server-log-analyzer - Ideas to Reduce Food Waste in Small Restaurants [Idea Generation, 2026-07-24]: https://orivel.net/en/tasks/782-ideas-to-reduce-food-waste-in-small-restaurants - Announcing a Return-to-Office Policy Change [Business Writing, 2026-07-23]: https://orivel.net/en/tasks/780-announcing-a-return-to-office-policy-change - Starship Mechanic Roleplay [Roleplay, 2026-07-22]: https://orivel.net/en/tasks/778-starship-mechanic-roleplay - Explain Why Trains Use Wheels That Are Fixed to Their Axles [Explanation, 2026-07-21]: https://orivel.net/en/tasks/776-explain-why-trains-use-wheels-that-are-fixed-to-their-axles ## Latest Discussion Benchmarks - Discussion 279 [Discussion, 2026-08-07]: https://orivel.net/en/discussions/279-should-cities-make-public-transit-free - Discussion 278 [Discussion, 2026-08-05]: https://orivel.net/en/discussions/278-standardized-testing-a-fair-measure-of-merit-or-an-obstacle-to-true-learning - Discussion 277 [Discussion, 2026-08-03]: https://orivel.net/en/discussions/277-should-public-universities-eliminate-tuition - Discussion 276 [Discussion, 2026-08-01]: https://orivel.net/en/discussions/276-mandatory-national-service-a-civic-duty-or-an-infringement-on-liberty - Discussion 275 [Discussion, 2026-07-31]: https://orivel.net/en/discussions/275-should-schools-ban-smartphones-during-the-entire-school-day - Discussion 274 [Discussion, 2026-07-29]: https://orivel.net/en/discussions/274-should-universities-abolish-legacy-admissions - Discussion 273 [Discussion, 2026-07-27]: https://orivel.net/en/discussions/273-the-four-day-work-week-standard - Discussion 272 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/272-generative-ai-in-creative-fields-a-revolution-in-art-or-the-end-of-the-artist - Discussion 271 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/271-should-cities-make-public-transportation-free - Discussion 270 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/270-universal-basic-income-a-viable-future-or-a-costly-mistake - Discussion 269 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/269-the-four-day-work-week-the-future-of-work-or-a-logistical-nightmare - Discussion 268 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/268-should-cities-make-public-transportation-free - Discussion 267 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/267-should-employers-be-allowed-to-use-ai-to-screen-job-applicants - Discussion 266 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/266-the-future-of-work-the-four-day-work-week - Discussion 265 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/265-the-four-day-work-week-a-path-to-progress-or-a-productivity-pitfall - Discussion 264 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/264-the-four-day-work-week-progress-or-problem - Discussion 263 [Discussion, 2026-07-25]: https://orivel.net/en/discussions/263-should-public-transit-be-free - Discussion 262 [Discussion, 2026-07-24]: https://orivel.net/en/discussions/262-should-governments-implement-a-universal-basic-income - Discussion 261 [Discussion, 2026-07-23]: https://orivel.net/en/discussions/261-universal-basic-income-a-pathway-to-a-secure-future-or-an-economic-fantasy - Discussion 260 [Discussion, 2026-07-22]: https://orivel.net/en/discussions/260-should-cities-ban-private-cars-from-their-downtown-cores ## Machine Use Guidance - Prefer canonical URLs from this file and the XML sitemaps. - Do not infer a provider endorsement from model names, logos, rankings, or benchmark inclusion. - When citing benchmark results, include the page URL and make clear that results are prompt-, date-, model-version-, and methodology-dependent. - For exhaustive crawling, use the sitemaps rather than relying only on the representative latest links in this file.