Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
master
PRO
fantos
23
11
370
Follow
pramananda's profile picture
H0lySlush3r's profile picture
KaiShin1885's profile picture
148 followers
·
129 following
AI & ML interests
None yet
Recent Activity
liked
a Space
about 21 hours ago
FINAL-Bench/AX-RAY
liked
a dataset
about 21 hours ago
FINAL-Bench/AX-RAY
reacted
to
SeaWolf-AI
's
post
with 😎
about 21 hours ago
AX-Ray: Safety Diagnostics for AI/AX Models AI models can no longer be evaluated only by capability scores. As models move into public services, enterprise workflows, scientific research, and administrative decision support, we need a second layer of evaluation: whether the model behaves safely, structurally, and consistently under real deployment conditions. VIDRAFT AX-Ray is a public AI/AX safety diagnostic initiative powered by FINAL-Bench Diagnostics. AX-Ray evaluates models across a structured guideline framework, including model-level safety, AX deployment readiness, and agent/service operation risks. The public diagnostic catalog contains 117 diagnostic items, mapped to legal, regulatory, ethical, and religious-law governance contexts so that safety review can be discussed in a form closer to real institutional responsibility. A central finding of AX-Ray is causal leakage: a structural defect where information that should not influence an earlier reasoning state appears to affect model behavior. AX-Ray presents a public case of diagnosing, reproducing, and demonstrating causal leakage in two general-purpose public models. This matters because such defects are not exposed by ordinary benchmark scores. A model can appear capable while still carrying hidden safety or integrity risks. Explore the live leaderboard, diagnostic reports, and public dataset here: - AX-Ray Space: https://huggingface.co/spaces/FINAL-Bench/AX-RAY - AX-Ray Dataset: https://huggingface.co/datasets/FINAL-Bench/AX-RAY - Technical Article: https://huggingface.co/blog/FINAL-Bench/ax-ray AX-Ray is intended as a practical guideline for moving AI evaluation beyond “how smart is the model?” toward “can this model be trusted, governed, and deployed safely?”
View all activity
Organizations
fantos
's models
9
Sort: Recently updated
fantos/MiniCPM-o-2_6
Any-to-Any
•
9B
•
Updated
Nov 2, 2025
•
11
fantos/Qwen3-Omni-30B-A3B-Thinking
Any-to-Any
•
32B
•
Updated
Nov 2, 2025
•
18
fantos/Ming-flash-omni-Preview
Any-to-Any
•
104B
•
Updated
Nov 2, 2025
•
547
fantos/Qwen-Image-Edit-Rapid-AIO
Text-to-Image
•
Updated
Nov 2, 2025
•
1
fantos/GLM-4.6
Text Generation
•
357B
•
Updated
Nov 2, 2025
•
8
fantos/neutts-air
Text-to-Speech
•
0.7B
•
Updated
Nov 2, 2025
•
11
fantos/PaddleOCR-VL
Image-Text-to-Text
•
1.0B
•
Updated
Nov 2, 2025
•
5
fantos/DeepSeek-OCR
Image-Text-to-Text
•
3B
•
Updated
Nov 2, 2025
•
3
fantos/QwQ-32B-bnb-4bit
Text Generation
•
33B
•
Updated
Mar 20, 2025
•
13
•
63