Sharing Performance Metrics for De-Generative AI Models

As part of our ongoing commitment to transparency and responsible AI development, we are sharing the performance metrics of our de-generative AI models.

Performance indicators

These metrics, benchmarked against a diverse dataset comprising both academic and proprietary (privately generated) samples, allow us—and others—to assess how well our models perform across various tasks and use cases, and provide a clear baseline for future enhancements. This transparency is especially important in the context of degenerative AI, where outputs are often complex, nuanced, and subject to interpretation.

To offer a consistent and interpretable view of model performance, we will report the following standard evaluation metrics. Each of these metrics offers unique insights into model behavior and collectively provide a robust framework for evaluating and comparing degenerative model performance over time.

Accuracy

This metric measures the overall correctness of the model's outputs by calculating the proportion of total predictions that are correct. It is a broad indicator of performance but can be misleading in the presence of class imbalance.

Precision

Precision evaluates the proportion of correct positive predictions relative to the total number of positive predictions made. It answers the question: Of all the outputs labeled as positive by the model, how many were actually correct? This is particularly important in scenarios where false positives carry a high cost.

Recall

Also known as sensitivity, recall measures the proportion of actual positive cases that the model correctly identified. It reflects the model’s ability to capture all relevant cases and is critical in contexts where missing a true positive is particularly undesirable.

F1 Score

The F1 score is the harmonic mean of precision and recall, providing a single metric that balances both concerns. It is especially useful when there is an uneven class distribution or when a balance between precision and recall is essential.

By making these results available, we invite constructive feedback from the research community, foster shared learning, and ultimately strive for more responsible and effective deployment of degenerative AI technologies.

latest model performance
REVELIO FACE II
The latest deepfake detection model for faces with 99.5% of detection accuracy
Frontal view of a wireframe human head and neck model on a black background.
93.8%
artificial precision
98.7%
human precision
99.3%
artificial recall
89.1%
human recall
96.5%
artificial F1 score
93.6%
human F1 score
REVELIO V PLUS
The latest deepfake detection for videosand images model with 99.8% of accuracy
Frontal view of a wireframe human head and neck model on a black background.
100%
artificial precision
95.9%
human precision
99.9%
artificial recall
99.2%
human recall
99.9%
artificial F1 score
97.5%
human F1 score
QUIET
The latest deepfake detection for audios with 91,83% of accuracy
Frontal view of a wireframe human head and neck model on a black background.
93,18%
artificial precision
90,57%
human precision
90,27%
artificial recall
93,39%
human recall
91,70%
artificial F1 score
91,96%
human F1 score
See our latest models performance
Performance →
Discover our technology and let one of our experts guide you to learn how easy it can be to know the truth regarding online contents. Book a demo now!
Accuracy
Teal bullseye target
Precision
Teal target symbol

talk to a human expert

Tell us about your business. We'll come back to you within one business day.

Thank you!
Your submission has been successfully sent to our team
Oops! Something went wrong while submitting the form.

No sales pitch. Just a conversation.

We stand for truth