บริษัทวิจัยอิสระที่รันชุดทดสอบมาตรฐานกับทุกโมเดลด้วยเงื่อนไขเดียวกัน แล้วสรุปเป็นดัชนี Intelligence, Coding และ Agentic พร้อมราคาต่อ tokenIndependent lab that runs the same standard eval suites on every model and publishes Intelligence, Coding and Agentic indexes plus token pricing.
ใช้ในแท็บUsed in: ฉลาดสุด · เขียนโค้ด · AgentIntelligence · Coding · Agents · 101 models · as of 16 Sept 2026
OpenRouter รันการทดสอบเองแบบทำซ้ำได้ ทุกคะแนนลิงก์ถึง config ค่าใช้จ่าย และ telemetry: GPQA Diamond (วิทยาศาสตร์), τ²-Bench Airline (agent เรียก tool) และชุดค้นเว็บ BrowseComp / WideSearch / DeepSearchQA / HLEReproducible evals run by OpenRouter, each linked to its config, cost and telemetry: GPQA Diamond, τ²-Bench Airline, and the BrowseComp / WideSearch / DeepSearchQA / HLE search suite.
ใช้ในแท็บUsed in: Agent · GPQA · AI ค้นเว็บWeb search · 268 results · as of 13 Sept 2026
ผู้ใช้จริงโหวตเทียบผลงาน (เว็บไซต์ UI กราฟ เกม 3D โลโก้ ฯลฯ) จากสองโมเดลโดยไม่เห็นชื่อ แล้วคำนวณคะแนน Elo และอัตราชนะReal users vote on blind side-by-side outputs (websites, UI, charts, games, 3D, logos…), aggregated into Elo ratings and win rates.
ใช้ในแท็บUsed in: ออกแบบ UIDesign · 14 categories · as of 16 Sept 2026
แคตตาล็อกโมเดลที่เปิดให้ใช้งานผ่าน API พร้อมวันเปิดตัว context window และราคา input / output ต่อ tokenCatalogue of models available via API with release date, context window and input / output token pricing.
ใช้ในแท็บUsed in: โมเดลใหม่ & ราคาNew & pricing · 444 models