Data month 2026-07

Best AI Observability & Eval ToolsPage 2

#ProductIndustryVisitsMoM
51
Cyara
Cyara
Cyara is represented by cyara.com. Power your CX with Agentic AI. Cyara helps you test, monitor, and optimize customer journeys at scale with speed and precision.
General AI44.7K+12.8%
52
World of AI Bench
World of AI Bench
World of AI Bench 是 World of AI Bench 旗下与 woaibench.ai 对应的产品官网,主要提供:Compare AI models on real coding tasks with private benchmarks, live HTML previews, cost tracking, ELO battles, and public model rankings.
Programming42.5K+79.9%
53
Sherlocks.ai
Sherlocks.ai
Sherlocks.ai 对应 sherlocks.ai。Cut MTTR by 10x with AI SREs that investigate incidents 24/7, automate root cause analysis, and prevent outages before they happen. Try Sherlocks.ai free.
General AI39.2K+1022.6%
54
Balto
Balto
Balto is represented by balto.ai. Balto is the #1 Contact Center AI software. Explore agent guidance, automated QA, business insights, coaching, call summarization, and more.
Business38.9K+22.1%
55
Headroom
Headroom
Headroom 是 Garm Tech 旗下与 extraheadroom.com 对应的产品官网,主要提供:Headroom is a menu bar app that trims prompt bloat before it reaches Claude Code and Codex, cutting token spend ~50% so the plan you already pay for lasts 2x longer.
Programming38.1K+143.4%
56
Hamming
Hamming
Hamming AI is represented by hamming.ai. Complete voice & chat agent QA platform. Auto-generate scenarios, replay production calls, 50+ metrics. SOC 2 certified. Start in under 10 minutes.
Business37.7K+7.2%
57
Openlayer
Openlayer
评估机器学习工作区数据预处理和特征选择工具与各种机器学习算法集成模型评估和可视化功能
Programming37.1K+1%
58
HoneyHive
HoneyHive
评估提示和模型监控生产环境性能管理和版本控制提示调试代理和RAG管道以评估和微调标记数据集
General AI36.8K-7.8%
59
Chainlit
Chainlit
评估LLM应用程序的可观察性和分析平台
General AI36.1K-24.1%
60
Lunary
Lunary
Lunary 对应 lunary.ai。Observability and prompt management platform for LLM-based apps. Take your LLMs to the next level.
Programming33.8K-8.3%
61
Agentops
Agentops
AgentOps is represented by agentops.ai. Trace, Debug, & Deploy Reliable AI Agents.
General AI30.5K+7.6%
62
Agenta
Agenta
Open-source prompt management & evals for AI teams
Productivity29.1K-15.6%
63
Triple Session
Triple Session
Triple Session 对应 triplesession.com。Triple Session is an AI sales coaching platform that automatically reviews every sales call against your playbook, scores rep performance, and recommends targeted training. Built for sales managers and enablement teams. From $45/user/month.
Education28.9K+7.3%
64
Guild Control Plane
Guild Control Plane
Guild Control Plane 是 Guild.ai, Inc. 旗下与 guild.ai 对应的产品官网,主要提供:The control plane for AI agents. Guild lets engineering teams manage, govern, and share every agent in production — with security and observability built in.
Business28.2K+4.3%
65
SourceBae
SourceBae
将资助的初创公司或企业与远程开发人员联系起来允许初创公司或企业创建工作职位使开发人员能够浏览和申请远程职位简化双方的招聘流程
Programming27.9K-19.8%
66
Picsellia Atlas
Picsellia Atlas
Your vision AI agent to build VisionAI applications
General AI27.1K+48.5%
67
Klu.ai
Klu.ai
在几分钟内构建,评估和优化GPT-4应用程序。使用Klu Studio prototype连接CRM、数据库、知识库和工单系统,动态生成提示。生成式AI与上下文,内存和会话聊天,以连接多个操作。基于检索生成交互式设计环境。连接SQL、Snowflake、Elasticache、Redis CRM、知识库和工单,集成一键发布。部署高级数据引擎以扩展生成式AI以观察使用情况,成本和性能见解通过API或UI过滤真实用户行为和反馈通过上下文元数据高级过滤器添加上下文文档,导入/导出收集真实世界的学习使用您的数据微调GPT-4等。使用Klu Python、TypeScript和React SDK构建和上下文联合编辑使用文档引用生成动态内容使用品牌语音创建聊天体验创建按需辅导汇总大型文档快速识别潜在客户使用动态提示创建内容分析用户反馈和情感数据的提取、清理和转换
Programming26.5K-26.4%
68
Treblle
Treblle
Runtime Intelligence Platform is represented by treblle.com. Discover, Govern, and Secure APIs, Agents, and AI Across Any Cloud, Gateway or Technology.
Programming25.8K-9.6%
69
Solidroad
Solidroad
AI对话模拟器定制反馈实践销售报表改进销售技巧节省时间和金钱见解和报告
AI Marketing25.6K-6.9%
70
Cleanlab Studio
Cleanlab Studio
多标签数据纠错工具
Programming24.7K+13.6%
71
LangWatch
LangWatch
LLM幻觉分析-使用任何LLM模型构建您自己的图表
Productivity24.1K+37.7%
72
Metaplane
Metaplane
自动故障检测、影响分析和根本原因诊断
Productivity23.9K-17.1%
73
Replay QA
Replay QA
Replay QA — AI wrote the app. Replay QA finds what broke. 是 Replay QA — AI wrote the app. Replay QA finds what broke. 旗下与 replay.io 对应的产品官网,主要提供:Add your GitHub repo for continuous testing, or drop in a URL for a one-time check. Replay QA explores your app, finds real bugs, and gives your coding agent the root cause and fix.
Programming23.6K+63.5%
74
Rival
Rival
Rival is represented by rival.tips. Rival official website.
General AI20.0K-35.7%
75
XSCT Bench
XSCT Bench
发现最适合你的 AI 模型。
General AI19.8K+19.3%
76
Permiso
Permiso
Detect and protect against human and non-human identity threats in real-time 对应 permiso.io。Permiso creates a single pane of glass for your identities across IdPs, IaaS, SaaS and CI/CD pipelines to find the riskiest actors in your environments.
Business19.3K+40.8%
77
AI Ping
AI Ping
一站式大模型服务评测与 API 调用平台。
General AI18.5K+61.2%
78
Keep
Keep
Keep 对应 keephq.dev。The open-source AIOps (AI for IT operations) and alert management platform.
Business17.2K+33.8%
79
Getdynamiq
Getdynamiq
Dynamiq is represented by getdynamiq.ai. Build, deploy and monitor on-premise GenAI applications to solve specific business needs, reducing costs, driving revenue and boosting competitiveness - all in one platform
Business16.3K-8.8%
80
Anomalo
Anomalo
自动化数据质量监控无监督机器学习根本原因分析数据血缘数据验证规则
Productivity14.9K+4.3%
81General AI13.9K-5.2%
82
systemprompt MCP
systemprompt MCP
Voice controlled iOS & Android MCP client
Audio13.5K-10%
83
Loamly
Loamly
AI Recommendation Intelligence 对应 loamly.ai。We diagnose why AI platforms recommend your competitors instead of you. Forensic analysis across ChatGPT, Claude, Gemini & Perplexity. 2,000+ citations traced.
Productivity13.5K+29.3%
84
Stax
Stax
Move your LLM evals from vibes to data
General AI12.3K-5.9%
85
GenRank
GenRank
GenRank 对应 genrank.io。Start tracking your brand mentions in ChatGPT with our Free plan. Track, measure, and optimize with Genrank, the original ChatGPT rank tracker.
AI Marketing12.2K+1.9%
86
Captum
Captum
基于PyTorch的多模式可扩展
Programming11.6K-20.1%
87
Edgee
Edgee
The AI Gateway that TL;DR tokens
General AI11.5K-6.9%
88
Openpipe
Openpipe
OpenPipe is represented by openpipe.ai. OpenPipe | RL For Agents
General AI11.4K+5.9%
89
AthinaAI
AthinaAI
完全可见的RAG pipeline 40 + 预设评估指标,用于检测幻觉并测量性能
Productivity11.3K+19.7%
90
Upsolve AI
Upsolve AI
Upsolve AI 对应 upsolve.ai。Upsolve AI: the platform to build, deploy, and evaluate grounded, governed, trustworthy data agents. Agent Context Studio is the context layer agentic analytics needs.
AI Marketing11.1K-16.1%
91
Digma.ai
Digma.ai
用于识别风险代码的运行时代码检查器实时代码问题突出显示功能可伸缩性分析性能基线分析
Programming10.1K-11.9%
92
Cloudidr
Cloudidr
Cloudidr LLM Ops — Save 75–90% on AI API Costs 对应 cloudidr.com。Cut your LLM API costs by 75–90% with intelligent model routing across OpenAI, Anthropic, Google Gemini, and AWS Bedrock — without changing a line of code. Real-time spend visibility, hard budget controls, and a free starter plan to get started.
Business9.1K-21%
93
OCR Arena
OCR Arena
The world's first OCR leaderboard
AI Marketing9.0K-9.3%
94
Signum.AI
Signum.AI
AI-增强的客户参与度和保留率
Writing&Design8.7K-16.9%
95
MobileGym
MobileGym
MobileGym 是 MobileGym 旗下与 mobilegym.dev 对应的产品官网,主要提供:MobileGym is a verifiable and highly parallel simulation platform for mobile GUI agent research — programmatic state verification, structured state snapshots, identical-state replication, and scalable online RL across 28 mobile apps.
General AI8.4K-43.8%
96
Siteline
Siteline
Growth analytics for the agentic web
AI Marketing8.1K+25.6%
97
Raga
Raga
RagaAI is represented by raga.ai. RagaAI provides purpose-built, production-grade AI Agent Suites for Healthcare, Lifesciences, and Aerospace. Powered by Prism and Catalyst for maximum reliability.
Life Assistant7.7K-42.6%
98
Openlit
Openlit
OpenLIT is represented by openlit.io. OpenLIT is an open-source LLM observability platform built on OpenTelemetry. Monitor token usage, cost, and latency. Self-host for free under Apache 2.0.
General AI7.6K-17.8%
99
Plurai
Plurai
Plurai 是 AI Agent Trust Platform 旗下与 plurai.ai 对应的产品官网,主要提供:Production-ready AI agents with simulation, evaluation, and protection. Trusted by Microsoft, Google, NVIDIA. 15x edge-case coverage, 7x faster deployment.
General AI7.5K+56%
100
Tokenomy
Tokenomy
Predict. Optimize. Ship AI with confidence. is represented by tokenomy.ai. Tokenomy: Advanced AI token calculator and cost estimator for LLMs. Optimize your AI prompts, analyze token usage, and save money on OpenAI, Anthropic, and other LLM APIs.
Programming7.3K+7.2%
Browse all 253 AI Observability & Eval products

Related tags

Data month: 2026-07 · Visit figures are third-party traffic estimates aggregated monthly — best for scale and trends. Multiple domains of the same product are merged. See methodology.