Langfuse
AI Evaluation & Testing
Berlin-based open-source LLM observability platform providing continuous execution logging and trace audits in Frankfurt.
Post-market monitoring capabilities can support applicable high-risk provider obligations. The exact monitoring plan depends on the system and role.
Navigator mappings are a discovery aid. They do not establish that an organisation or system is legally in scope, compliant or certified.
These software matches come from the requirement tags recorded in the Navigator catalogue. Review each vendor profile and its source evidence before making a procurement or compliance decision.
Run assessment11 source-linked profiles currently mapped to this requirement.
AI Evaluation & Testing
Berlin-based open-source LLM observability platform providing continuous execution logging and trace audits in Frankfurt.
AI Evaluation & Testing
Developer-centric LLM observability platform providing automated tracing, regression evaluation, and runtime prompt logging.
AI Evaluation & Testing
AI evaluation and observability platform tracking embedding drift, real-time evals, and prompt telemetry.
AI Evaluation & Testing
AI quality management suite evaluating explainability, relevance, and hallucination scores across LLMs and predictive ML.
AI Governance & Inventory
Austrian AI governance system mapping enterprise AI risk tiers directly to EU statutory obligations.
AI Governance & Inventory
AI assurance engine providing automated risk verification and performance warranties for enterprise AI models.
AI Evaluation & Testing
Continuous AI validation platform providing automated test suites across data integrity, model behavior, and LLM evaluation.
AI Evaluation & Testing
Evaluation and guardrails platform providing automated chain-of-thought metrics, prompt debugging, and production safety scoring.
AI Evaluation & Testing
Continuous ML and data quality monitoring platform providing automated anomaly detection, performance tracking, and confidence scoring.
AI Evaluation & Testing
Open-source standard and platform for tracing LLM execution spans, token metrics, and runtime errors.
AI Evaluation & Testing
Open-source Python library focusing on outlier, adversarial, and concept drift detection for tabular, image, and text models.