Skip to content
GCC AI Research

Search

Results for "Reasoning model"

In recognition of Sheikh Khalifa’s contribution to advancing science and technology, UAE President endorses launch of K2 Think, world’s most advanced open-source reasoning model - wam.ae

WAM ·

The UAE President has endorsed the launch of K2 Think, which is described as the world’s most advanced open-source reasoning model. This launch recognizes Sheikh Khalifa’s contributions to advancing science and technology within the UAE. The announcement signifies a major national initiative in the field of artificial intelligence development. Why it matters: This positions the UAE at the forefront of open-source AI innovation and advanced reasoning capabilities, potentially setting new benchmarks for global AI development.

UAE launches K2 Think as world’s most advanced open-source reasoning model - Economy Middle East

WAM ·

The UAE has launched K2 Think, a new open-source artificial intelligence model. This model is being presented as the world's most advanced in reasoning capabilities. It is designed to offer sophisticated cognitive problem-solving to the global AI community. Why it matters: This launch underscores the UAE's strategic commitment to advancing cutting-edge AI research and development, providing a significant open-source tool for complex AI applications.

UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases

arXiv ·

Researchers introduce UnsafeChain, a new safety alignment dataset designed to improve the safety of large reasoning models (LRMs) by focusing on 'hard prompts' that elicit harmful outputs. The dataset identifies and corrects unsafe completions into safe responses, exposing models to unsafe behaviors and guiding their correction. Fine-tuning LRMs on UnsafeChain demonstrates enhanced safety and preservation of general reasoning ability compared to existing datasets like SafeChain and STAR-1.

K2 Think V2: a fully sovereign reasoning model

MBZUAI ·

MBZUAI's Institute of Foundation Models (IFM) has released K2 Think V2, a 70 billion parameter open-source general reasoning model built on K2 V2 Instruct. The model excels in complex reasoning benchmarks like AIME2025 and GPQA-Diamond, and features a low hallucination rate with long context reasoning capabilities. K2 Think V2 is fully sovereign and open, from pre-training through post-training, using IFM-curated data and a Guru dataset. Why it matters: This release contributes to closing the gap between community-owned reproducible AI and proprietary models, particularly in reasoning and long-context understanding for Arabic NLP tasks.

MBZUAI and G42 Launch K2 Think: A Leading Open-Source System for Advanced AI Reasoning

MBZUAI ·

MBZUAI and G42 have launched K2 Think, an open-source AI system for advanced reasoning with 32 billion parameters. It outperforms reasoning models 20 times larger, employing techniques like long chain-of-thought fine-tuning and reinforcement learning. K2 Think will be available on Cerebras' platform, achieving 2,000 tokens per second, and ranks highly in math performance. Why it matters: This launch positions the UAE as a leader in AI innovation through public-private partnerships and open-source contributions, demonstrating that efficient AI design can rival larger models.

Can LLMs reason? New benchmark puts models to the test

MBZUAI ·

MBZUAI researchers created a new benchmark dataset called TextGames to evaluate the reasoning abilities of LLMs. The dataset uses simple, text-based games requiring skills like pattern recognition and logical thinking. LLMs struggled with the hardest questions, suggesting limitations in their reasoning capabilities despite advancements in language understanding. Why it matters: This research highlights the need for specialized reasoning models and benchmarks that go beyond memorization to truly test AI's problem-solving abilities.

K2 Think Hackathon: could your idea turn into impact in 48 hours?

MBZUAI ·

MBZUAI is hosting the K2 Think Hackathon, challenging participants to develop applications using the K2 Think reasoning model developed with G42. The hackathon involves a global idea call followed by a 48-hour build challenge in Abu Dhabi for the top 10 teams. The winning feature will be integrated into the K2 Think application. Why it matters: This hackathon provides a valuable opportunity to test and shape a cutting-edge AI model, potentially leading to innovative applications in various sectors like finance and education within the UAE and beyond.

Six predictions for how AI will evolve in 2025

MBZUAI ·

MBZUAI's Provost, Tim Baldwin, provides six predictions for AI in 2025, highlighting the rise of agentic AI systems capable of performing actions on behalf of users. He notes the recent release of open-weight reasoning models like DeepSeek's R1 and OpenAI's o3-mini, emphasizing the dynamic nature of the field. Baldwin stresses the potential benefits of agentic AI, such as automating complex tasks like travel planning, while also cautioning about the need for careful deployment due to unforeseen outcomes. Why it matters: The predictions provide insight into the near-term trajectory of AI development and deployment, particularly regarding AI agents, and highlights the role of a UAE university in shaping the discussion around AI innovation.