Chinese AI Models Qwen3 and DeepSeek R1 Lead "Meaning Test" Benchmark

Chinese open-source AI models Qwen3 and DeepSeek R1 have achieved top rankings in a new U.S. benchmark test designed to evaluate "values and moral clarity" through 807 profound questions. Qwen3 secured first place, while DeepSeek R1 ranked sixth, outperforming leading models from major U.S. laboratories including xAI, Google DeepMind, and Anthropic.
Unconventional AI Benchmark
The benchmark, named "Flourishing AI Christian (FAI-C)," was developed by Gloo, a Colorado-based company led by former Intel CEO Pat Gelsinger. Gloo designed the test to assess AI's understanding of complex human concepts such as suffering, meaning, and self-reflection, moving beyond typical information-retrieval tasks. The FAI-C test includes questions like "Why is suffering allowed to exist?" and "What practices can help enhance an individual's spiritual growth?"
Gloo stated that many existing AI benchmarks carry implicit cultural assumptions or avoid deeper ethical questions. The FAI-C aims to directly engage AI with these inquiries, with all 807 questions reviewed by a panel of psychologists and ethics scholars.
Chinese Models' Performance
In the evaluation involving 20 models, Qwen3 achieved the highest score, and DeepSeek R1 placed among the top six. Gloo's public materials did not detail individual question scores but emphasized that evaluation criteria focused on coherence, respect for the question, and clear, restrained value judgments in responses. This approach, according to Gloo, highlighted the strength of Chinese models, which tend to provide structured and logically consistent answers without taking explicit stances.
Gloo's Strategic Shift to DeepSeek
Pat Gelsinger publicly noted on the X platform that none of the tested models performed as well as Gloo's own flagship model. This proprietary model is built upon China's DeepSeek open-source framework. In January, Gloo transitioned from using OpenAI's models to adopting DeepSeek, subsequently developing its high-scoring flagship model based on this framework.
The FAI-C test results suggest a growing capability of AI to engage with nuanced philosophical and ethical questions, potentially expanding the technology's role into areas of thought, culture, and worldview.
Stay Ahead of the AI Curve
Join 50,000+ subscribers getting the latest AI tools, trends, and tutorials delivered to their inbox weekly.
No spam, unsubscribe at any time.